Leveling up AI building and compute
Start for free, upgrade for higher quotas and premium storage, rent real GPUs by the second, and move to Enterprise for governance and compliance controls.
Explore the platform
Free
- 10GB private storage
- Unlimited public repos
- Community support
- Public Spaces
For individual builders
Pro
- 100GB private storage
- 100 GPU hours / month
- 8× Spaces quota + priority queue
- Private dataset viewer
For growing organizations
Team
- 500GB private storage / user
- Unlimited private repos
- Centralized billing
- Organization-wide access controls
- Priority support response queue
For mission-critical workloads
Enterprise
- SSO (OIDC) and enterprise RBAC
- Audit logs and compliance exports
- GPU data-residency enforcement
- Dedicated support and onboarding
Need team governance, SSO, and audit controls? Talk to our Enterprise team →
Transparent, volume-based pricing
Store models, datasets, and artifacts at a flat, honest per-TB rate — with automatic volume discounts as you grow. No egress surprises, no per-request line items.
Comparative per-TB/mo list prices are indicative. Inferix figure is its real base storage rate.
Base
Up to 50 TB
Scale
−20%50 TB+
Growth
−25%200 TB+
Enterprise
−33%500 TB+
Per-TB rates shown monthly. Base private rate ($0.03/GB) is live; volume-tier discounts and public rates are indicative.
Storage pricing calculator
Estimate what you'll pay in storage overage beyond your plan's included quota. Overage is billed at $0.03/GB per month.
Included with Pro
100 GB
Overage this month
$1.50
Need team governance, SSO, and audit controls?
Move to Team or Enterprise when you need organization-wide policy controls and compliance workflows.
Rent GPUs by the second. No commitments.
Spin up a real GPU in seconds, pay only for the time you use, and shut it down whenever you like.
| GPU | VRAM | Price / hr | |
|---|---|---|---|
RTX 5090Live rate | 32 GB | $0.64 / GPU-hr | Rent now |
RTX 4090 | 24 GB | $0.44 / GPU-hr | Rent now |
L40S | 48 GB | $1.10 / GPU-hr | Rent now |
RTX PRO 6000 Blackwell | 96 GB | $1.79 / GPU-hr | Rent now |
A100 | 80 GB | $1.89 / GPU-hr | Rent now |
H100 | 80 GB | $2.49 / GPU-hr | Rent now |
H200 | 141 GB | $2.99 / GPU-hr | Rent now |
On-demand hardware for your Spaces. Starting at $0.
Spaces are the easiest way to host and share ML apps and demos — build with Gradio, Streamlit, Docker, or static, attach on-demand hardware, and share with a single URL.
| Name | CPU | Memory | Accelerator | VRAM | Hourly |
|---|---|---|---|---|---|
| CPU Basic | 2 vCPU | 16 GB | — | — | FREE |
| CPU Upgrade | 8 vCPU | 32 GB | — | — | $0.03 |
| ZeroGPU | dynamic | dynamic | Nvidia RTX Pro 6000 Blackwell | up to 96 GB | FREE |
| Nvidia T4 - small | 4 vCPU | 15 GB | Nvidia T4 | 16 GB | $0.40 |
| Nvidia T4 - medium | 8 vCPU | 30 GB | Nvidia T4 | 16 GB | $0.60 |
| 1x Nvidia L4 | 8 vCPU | 30 GB | Nvidia L4 | 24 GB | $0.80 |
| 4x Nvidia L4 | 48 vCPU | 186 GB | Nvidia L4 | 96 GB | $3.80 |
| 1x Nvidia L40S | 8 vCPU | 62 GB | Nvidia L40S | 48 GB | $1.80 |
| 4x Nvidia L40S | 48 vCPU | 382 GB | Nvidia L40S | 192 GB | $8.30 |
| 8x Nvidia L40S | 192 vCPU | 1534 GB | Nvidia L40S | 384 GB | $23.50 |
| Nvidia A10G - small | 4 vCPU | 15 GB | Nvidia A10G | 24 GB | $1.00 |
| Nvidia A10G - large | 12 vCPU | 46 GB | Nvidia A10G | 24 GB | $1.50 |
| 2x Nvidia A10G - large | 24 vCPU | 92 GB | Nvidia A10G | 48 GB | $3.00 |
| 4x Nvidia A10G - large | 48 vCPU | 184 GB | Nvidia A10G | 96 GB | $5.00 |
| Nvidia A100 - large | 12 vCPU | 142 GB | Nvidia A100 | 80 GB | $2.50 |
| 4x Nvidia A100 | 48 vCPU | 568 GB | Nvidia A100 | 320 GB | $10.00 |
| 8x Nvidia A100 | 96 vCPU | 1136 GB | Nvidia A100 | 640 GB | $20.00 |
| Custom | on demand | on demand | on demand | on demand | on demand |
Production endpoints, billed by the GPU.
Deploy any model to a dedicated, auto-scaling HTTPS endpoint. You pay the same hourly GPU rate as above — and endpoints can scale to zero, so an idle endpoint costs nothing.
Starter
RTX 4090 · 24 GB
Small models, embeddings, and prototypes. Scales to zero when idle so you only pay while serving.
Standard
L40S / A100 · 48–80 GB
7B–70B models in production with auto-scaling replicas and a stable HTTPS endpoint.
Frontier
H100 / H200 · 80–141 GB
The largest open models and highest throughput. Multi-GPU capacity on demand.
Usage-based: an endpoint's cost is the hourly rate of the GPU it runs on, prorated per second. Non-5090 GPU rates are indicative drafts.
Deploy an endpointDedicated inference, starting at $0.033/hour.
Deploy any model on dedicated, autoscaling infrastructure — production-ready, with no cold starts.
Learn more| Instance | vCPU | Memory | Hourly |
|---|---|---|---|
| Intel Sapphire Rapids (aws) | 1 | 2GB | $0.03 |
| Intel Sapphire Rapids (aws) | 2 | 4GB | $0.07 |
| Intel Sapphire Rapids (aws) | 4 | 8GB | $0.13 |
| Intel Sapphire Rapids (aws) | 8 | 16GB | $0.27 |
| Intel Sapphire Rapids (aws) | 16 | 32GB | $0.54 |
| Intel Xeon (azure) | 1 | 2GB | $0.06 |
| Intel Xeon (azure) | 2 | 4GB | $0.12 |
| Intel Xeon (azure) | 4 | 8GB | $0.24 |
| Intel Xeon (azure) | 8 | 16GB | $0.48 |
| Intel Sapphire Rapids (gcp) | 1 | 2GB | $0.05 |
| Intel Sapphire Rapids (gcp) | 2 | 4GB | $0.10 |
| Intel Sapphire Rapids (gcp) | 4 | 8GB | $0.20 |
| Intel Sapphire Rapids (gcp) | 8 | 16GB | $0.40 |
| Instance | Size | Memory | Hourly |
|---|---|---|---|
| Inf2 Neuron (aws) | x1 | 14.5GB | $0.75 |
| Inf2 Neuron (aws) | x1-large | 124GB | $1.95 |
| Inf2 Neuron (aws) | x12 | 760GB | $12.00 |
| TPU v5e (gcp) | 1x1 | 16GB | $1.20 |
| TPU v5e (gcp) | 2x2 | 64GB | $4.75 |
| TPU v5e (gcp) | 2x4 | 128GB | $9.50 |
| Instance | Count | Memory | Hourly |
|---|---|---|---|
| NVIDIA T4 (aws) | 1 | 14GB | $0.50 |
| NVIDIA T4 (aws) | 4 | 56GB | $3.00 |
| NVIDIA L4 (aws) | 1 | 24GB | $0.80 |
| NVIDIA L4 (aws) | 4 | 96GB | $3.80 |
| NVIDIA L40S (aws) | 1 | 48GB | $1.80 |
| NVIDIA L40S (aws) | 4 | 192GB | $8.30 |
| NVIDIA L40S (aws) | 8 | 384GB | $23.50 |
| NVIDIA A100 (aws) | 1 | 80GB | $2.50 |
| NVIDIA A100 (aws) | 2 | 160GB | $5.00 |
| NVIDIA A100 (aws) | 4 | 320GB | $10.00 |
| NVIDIA A100 (aws) | 8 | 640GB | $20.00 |
| NVIDIA H100 (aws) | 1 | 80GB | $4.50 |
| NVIDIA H100 (aws) | 2 | 160GB | $9.00 |
| NVIDIA H100 (aws) | 4 | 320GB | $18.00 |
| NVIDIA H100 (aws) | 8 | 640GB | $36.00 |
| NVIDIA H200 (aws) | 1 | 141GB | $5.00 |
| NVIDIA H200 (aws) | 2 | 282GB | $10.00 |
| NVIDIA H200 (aws) | 4 | 564GB | $20.00 |
| NVIDIA H200 (aws) | 8 | 1128GB | $40.00 |
| NVIDIA B200 (aws) | 1 | 179GB | $9.25 |
| NVIDIA B200 (aws) | 2 | 358GB | $18.50 |
| NVIDIA B200 (aws) | 4 | 716GB | $37.00 |
| NVIDIA B200 (aws) | 8 | 1432GB | $74.00 |
| NVIDIA RTX PRO 6000 (aws) | 1 | 96GB | $2.75 |
| NVIDIA RTX PRO 6000 (aws) | 2 | 192GB | $5.50 |
| NVIDIA RTX PRO 6000 (aws) | 4 | 384GB | $11.00 |
| NVIDIA RTX PRO 6000 (aws) | 8 | 768GB | $22.00 |
Indicative rates — final pricing may vary.
Volume & committed use
Reserve capacity or commit to volume for up to ~40% off on-demand rates. Discount is indicative — let's find the right fit for your workload.
Compare plans in detail
Every feature, side by side.
| Feature | Free | Pro | Team | Enterprise |
|---|---|---|---|---|
| Private storage | 10GB | 100GB | 500GB / user | Custom |
| GPU inference hours | Pay as you go | 100 / month | 500 / month | Custom |
| Spaces quota | Standard | 8× priority | 8× priority | Custom |
| Private dataset viewer | ||||
| Organization-wide access controls | ||||
| SSO (OIDC) | ||||
| Audit logs | ||||
| Dedicated support |
Do more with Inferix Pro
Upgrade your Inferix experience for $9.99/month.
Private storage
10× private storage capacity
Public storage
2× public storage capacity
Inference Providers
20× included inference credits
Spaces
Host your own ZeroGPU, Gradio & Docker Spaces on compute
ZeroGPU
8× ZeroGPU quota and highest queue priority
Spaces Dev Mode
Fast iterations via SSH/VS Code
Blog
Personal blog publishing
Dataset Viewer
For private datasets
Features Preview
Get early access to upcoming features
PRO Badge
Show your support with a PRO badge
Explore, experiment, collaborate and build with Machine Learning
The Inferix Hub is the central place to explore, experiment, collaborate and build with real GPUs — join the open ML movement.
Sign upsentence-transformers/paraphrase-multilingual-MiniLM-L12-v2
Sentence Similarity
amazon/chronos-2
Time-Series
facebook/opt-125m
Text Generation
Rixy Playground
Gradio · ZeroGPU
Image Studio
Docker · L40S
Inferix brand assets
Logos, colors, and the Rixy render-cube mascot — everything you need to build with the Inferix brand.