Blank white background with no objects or features visible.

TrueFoundry Named Frost & Sullivan's 2026 Global Transformational Innovation Leader. Read report

TrueForgeのご紹介:オープンソースでベンダーフリーなエージェントハーネス。コストを50%削減します。今すぐ試す→

Baseten vs Modal: Pricing, Cold Starts and What the Rate Card Hides

By アシシュ・ドゥベイ

Published: September 28, 2026

Baseten vs Modal — TrueFoundry

‍

Try now.

One gateway for all your models, MCP servers, and agents.
No credit card needed.

Start free
Table of Contents

One Gateway for Every LLM, Agent and MCP Server

Book a 30-min with our AI expert

Book a Demo

The fastest way to build, govern and scale your AI

Book Demo
Summarize with
ChatGPT logo by OpenAI
Perplexity AI logo
Blurry red snowflake on white background, symmetrical frosty design with soft edges and abstract shape.

Discover More

No items found.
September 28, 2026
|
5 min read

Langfuse Alternatives: 7 Options Compared on Licence, Price and Limits

No items found.
September 28, 2026
|
5 min read

Data Loss Prevention for LLM Traffic: Where It Has to Sit

No items found.
September 28, 2026
|
5 min read

Data Masking in the AI Gateway: What Actually Works

No items found.
September 28, 2026
|
5 min read

API Rate Limiting for LLMs: Count Tokens, Not Requests

No items found.
No items found.

Recent Blogs

Black left pointing arrow symbol on white background, directional indicator.
Black left pointing arrow symbol on white background, directional indicator.

Frequently asked questions

Baseten vs Modal: which is cheaper?

It depends on the GPU, and the published rates mislead. On headline numbers Modal looks 37-40% cheaper on H100 and A100 80GB. Add the CPU and memory Baseten bundles and the H100 gap narrows to about 19%, A100 80GB is parity, and on T4 and A10 Baseten is 9-22% cheaper. Modal’s pinning and non-preemptible multipliers push further toward Baseten, and its $250/mo Team fee is fixed cost Baseten does not charge.

What is Modal pricing, exactly?

Starter $0/mo with $30/mo free compute; Team $250/mo with $100/mo free compute; Enterprise custom. Compute is per second and itemised: GPU at its own rate, CPU at $0.04716 per physical core-hour, memory at $0.007992 per GiB-hour. Sandbox and Notebook compute is 3x the Function rate, and pinning and non-preemptible capacity carry multipliers.

What is Baseten pricing, exactly?

Basic is $0/month pay-as-you-go with no platform fee, including SOC 2 Type II and HIPAA; Pro and Enterprise are on quote. GPU compute bills per minute against all-in instance SKUs bundling vCPU and RAM — $6.4998/hr for an H100 80GB. Per-token Model API rates are published; the free-credit amount is not.

How fast are Modal’s cold starts?

Containers boot in roughly a second, and Memory Snapshots deliver 3-10x faster restores per Modal’s docs. The “100x” figure in Modal’s Series C post is not supported by its own documentation, so plan against 3-10x. GPU snapshots are Alpha and do not speed up loading weights from storage.

Is Baseten valued at $26 billion?

Not confirmed. The last valuation Baseten has published is $13B, from a $1.5B Series F on 22 June 2026. Press reports in late September 2026 describe talks around a larger round, but that is talks-stage reporting and the company has published nothing.

Take a quick product tour
Start Product Tour
Product Tour