Modal vs Together AI
- GPUs
- 11 4
- Cheapest GPU
- $0.59/hr $1.99/hr
- Price wins
- 0/4 4/4
Together AI is cheaper on all 4 GPUs both rent. Modal carries 7 more GPUs.
Serverless GPU platform where code runs as functions for inference, batch jobs and training, billed per second with no commitments.
Best for: Serverless inference and batch jobs
Inference cloud with serverless and batch APIs priced per token, plus dedicated endpoints and rentable GPU clusters.
Best for: Token-priced inference and fine-tuning
Price by GPU
The 4 GPUs both rent. Cheapest live rate per GPU-hour, most powerful first. The cheaper side is in purple.
| GPU | Modal | Together AI | Gap |
|---|---|---|---|
| NVIDIA B200 SXM 180GB | $6.25/hr | $4.09/hr | 53% |
| NVIDIA B300 SXM 262GB | $7.10/hr | $4.99/hr | 42% |
| NVIDIA H100 SXM5 80GB | $3.95/hr | $1.99/hr | 98% |
| NVIDIA H200 SXM 141GB | $4.54/hr | $2.99/hr | 52% |
Prices last checked: Modal 1 Oct 2026, Together AI 4 Oct 2026.
Only on one side
Only on Modal 7
Only on Together AI 0
Every GPU Together AI rents is also on the other side.
Company facts
| Fact | Modal | Together AI |
|---|---|---|
| Headquarters | ||
| Type | GPU cloud | GPU cloud |
| Founded | n/a | 2022 |
| GPUs priced | 11 | 4 |
| Fastest GPU | NVIDIA B200 SXM 180GB | NVIDIA B200 SXM 180GB |
| Services | Serverless GPU compute, inference, batch jobs, training, sandboxes | Serverless and batch inference, provisioned throughput, dedicated inference, GPU clusters (Accelerated Compute), fine-tuning, managed storage, sandboxes |
| Pricing models | Pay-per-second, no commitments | Pay-as-you-go, token-based, committed capacity, hourly GPU |
Compliance
Frameworks each provider states it holds, linked to its own statement. Not published means we found no statement, not that the provider fails it. Check the scope with the provider.
Pick your provider
Questions
- What is the difference between Modal and Together AI?
- Modal: Serverless GPU platform where code runs as functions for inference, batch jobs and training, billed per second with no commitments. Together AI: Inference cloud with serverless and batch APIs priced per token, plus dedicated endpoints and rentable GPU clusters.
- Is Modal cheaper than Together AI?
- Together AI is cheaper on all 4 GPUs both rent. Modal carries 7 more GPUs.
- How much does a NVIDIA B200 SXM 180GB cost on Modal and Together AI?
- Modal charges $6.25 per GPU-hour and Together AI charges $4.09 per GPU-hour, the cheapest live rate on each.
- Which GPUs can I rent on both Modal and Together AI?
- 4 GPUs, including NVIDIA B200 SXM 180GB, NVIDIA B300 SXM 262GB, NVIDIA H100 SXM5 80GB, NVIDIA H200 SXM 141GB.
- How are these prices compared?
- Each figure is the cheapest live rate per GPU-hour for that GPU. We use on-demand pricing where a provider sells it and fall back to community or reserved terms, labelled, where that is all they sell. Spot capacity is left out because it can be reclaimed. Rates are refreshed by our daily ingest.
How we define a live price: methodology.





