GPU cloud head to head

Modal vs Together AI

GPUs
11 4
Cheapest GPU
$0.59/hr $1.99/hr
Price wins
0/4 4/4

Together AI is cheaper on all 4 GPUs both rent. Modal carries 7 more GPUs.

Serverless GPU platform where code runs as functions for inference, batch jobs and training, billed per second with no commitments.

Best for: Serverless inference and batch jobs

Inference cloud with serverless and batch APIs priced per token, plus dedicated endpoints and rentable GPU clusters.

Best for: Token-priced inference and fine-tuning

Price by GPU

The 4 GPUs both rent. Cheapest live rate per GPU-hour, most powerful first. The cheaper side is in purple.

GPU Modal Together AI Gap
NVIDIA B200 SXM 180GB $6.25/hr $4.09/hr 53%
NVIDIA B300 SXM 262GB $7.10/hr $4.99/hr 42%
NVIDIA H100 SXM5 80GB $3.95/hr $1.99/hr 98%
NVIDIA H200 SXM 141GB $4.54/hr $2.99/hr 52%

Prices last checked: Modal 1 Oct 2026, Together AI 4 Oct 2026.

Only on one side

Company facts

Fact Modal Together AI
Headquarters United States New York City, United States United States San Francisco, CA
Type GPU cloud GPU cloud
Founded n/a 2022
GPUs priced 11 4
Fastest GPU NVIDIA B200 SXM 180GB NVIDIA B200 SXM 180GB
Services Serverless GPU compute, inference, batch jobs, training, sandboxes Serverless and batch inference, provisioned throughput, dedicated inference, GPU clusters (Accelerated Compute), fine-tuning, managed storage, sandboxes
Pricing models Pay-per-second, no commitments Pay-as-you-go, token-based, committed capacity, hourly GPU

Compliance

Frameworks each provider states it holds, linked to its own statement. Not published means we found no statement, not that the provider fails it. Check the scope with the provider.

Framework Modal Together AI
SOC 2 Type II Yes Yes
HIPAA Yes Not published
ISO 27001 Not published Yes

Pick your provider

Modal logo
Modal

11 GPUs, from $0.59/hr

Visit website
Together AI logo
Together AI

4 GPUs, from $1.99/hr

Visit website

Questions

What is the difference between Modal and Together AI?
Modal: Serverless GPU platform where code runs as functions for inference, batch jobs and training, billed per second with no commitments. Together AI: Inference cloud with serverless and batch APIs priced per token, plus dedicated endpoints and rentable GPU clusters.
Is Modal cheaper than Together AI?
Together AI is cheaper on all 4 GPUs both rent. Modal carries 7 more GPUs.
How much does a NVIDIA B200 SXM 180GB cost on Modal and Together AI?
Modal charges $6.25 per GPU-hour and Together AI charges $4.09 per GPU-hour, the cheapest live rate on each.
Which GPUs can I rent on both Modal and Together AI?
4 GPUs, including NVIDIA B200 SXM 180GB, NVIDIA B300 SXM 262GB, NVIDIA H100 SXM5 80GB, NVIDIA H200 SXM 141GB.
How are these prices compared?
Each figure is the cheapest live rate per GPU-hour for that GPU. We use on-demand pricing where a provider sells it and fall back to community or reserved terms, labelled, where that is all they sell. Spot capacity is left out because it can be reclaimed. Rates are refreshed by our daily ingest.

How we define a live price: methodology.

More matchups

Choose your own

© 2026 Flopper.io - Compare the GPUs Powering AI