Modal

Modal pricing charges by the second for three resources at once: CPU cores, memory and GPU each carry their own rate, and a single function's bill is the sum.

Pricing Model:

Pricing Model:

Usage, with a platform fee on paid plans

Usage, with a platform fee on paid plans

Usage, with a platform fee on paid plans

Packaging Model:

Packaging Model:

Freemium, Good / Better / Best (GBB)

Freemium, Good / Better / Best (GBB)

Freemium, Good / Better / Best (GBB)

Credit Model:

Credit Model:

Monthly compute allowance, $30 on Starter, $100 on Team

Monthly compute allowance, $30 on Starter, $100 on Team

Monthly compute allowance, $30 on Starter, $100 on Team

Updated on:

Modal pricing: three meters running on one container

Modal pricing charges by the second for three resources at once: CPU cores, memory and GPU each carry their own rate, and a single function's bill is the sum. There's no instance type to pick and no minimum billing increment. Modal charges max(request, actual) for every second the container is alive. That decomposition is the whole design, and it's rare.

Key takeaways

  • Modal charges $0.0000131 per physical core-second ($0.0472 per core-hour) and $0.00000222 per GiB-second ($0.0080 per GiB-hour), independent of any GPU attached.

  • GPU rates run from $0.000164 per second for a T4 ($0.59/hour) to $0.001972 for a B300 ($7.10/hour), added on top of the CPU and memory charge.

  • Every container carries a minimum request of 0.125 cores and 128 MiB, and Modal bills the higher of your request or your real usage.

  • Startup time bills, and so does the default 60 seconds of idle time before a container scales down.

Modal pricing in 2026

Tier

Price

What's included

Metered limit

 

Starter

$0 + compute

$30/month credit, 3 seats, 200 deployed apps, 1 day log retention

100 containers, 10 GPU concurrency

Team

$250/month + compute

$100/month credit, unlimited seats, 30 day logs, environment budgets, static IP proxy

5,000 containers, 50 GPU concurrency

Enterprise

Custom

Volume discounts, SAML SSO, audit logs, HIPAA, private Slack support

Custom

What Modal actually meters

Modal meters three resources on separate clocks and adds them up. CPU is billed per physical core per second, and Modal notes that one physical core equals two AWS or GCP vCPUs. Memory is billed per GiB-second. GPUs are billed per GPU-second by chip. Because the axes are independent, a job that's memory-hungry and GPU-light costs what it actually consumes rather than what the nearest instance type would have charged.

The rate you pay is max(request, actual) on CPU and memory, measured continuously, with a floor of 0.125 cores and 128 MiB per container. Disk isn't a separate line: requesting disk raises your memory request at a 20:1 ratio, so 500 GiB of disk bills as 25 GiB of memory.

Three multipliers sit on top. Pinning a container to a broad region such as us costs 1.15x base price, and a narrow region such as us-west costs 1.75x. Running CPU functions on non-preemptible capacity costs 3x on CPU and memory. Sandboxes and Notebooks get their own block on the pricing page at $0.00003942 per core-second and $0.00000667 per GiB-second, which is the 3x non-preemptible rate applied to base figures the page itself rounds, because Modal documents Sandboxes as non-preemptible unless they request a GPU. The pricing page prints those numbers without connecting them to the multiplier.

From 1 October 2026 network egress joins the meter at $0.04 per GiB, after 1 TiB included on Starter, 10 TiB on Team and 100 TiB on Enterprise.

What happens when you hit the limit

Modal auto-charges. There's no included compute beyond the monthly credit, so once credit runs out you pay list rate and get billed at the end of the cycle, plus incremental charges the first time you cross certain thresholds. Modal doesn't publish what those thresholds are.

Spend control is a first-class feature rather than an alert. A Workspace budget caps usage before credits apply. A separate spend limit caps net charges after credits, and Modal stops workloads that would push past it. If you don't set one, the default spend limit is the cycle's usage limit minus credits. Team and Enterprise workspaces can add Environment budgets, though those exclude storage and reservations, so an environment budget isn't an invoice total.

How Modal pricing has changed across all these years

Date

Milestone

Source

 

1 Oct 2026

Network egress starts billing at $0.04 per GiB beyond the plan allowance

Vendor

1 Sep 2026

Network egress appears on the billing page, metered but not charged

Vendor

12 Aug 2026

Workspace.billing.rates() and modal billing rates let you query your own rate card

Vendor

23 Jul 2026

Billing summary APIs expose spend by category plus credit usage

Vendor

23 Jun 2026

Workspace and Environment billing report APIs added

Vendor

12 Feb 2026

modal billing report CLI reaches general availability for Team and Enterprise

Vendor

20 Nov 2025

Non-preemptible CPU capacity ships with a 3x multiplier on CPU and memory pricing

Vendor

Flexprice’s Take

Modal decomposes the bill the way the machine actually works, and almost nobody else does.

Replicate is the closest structural neighbour. Its gpu-h100 costs $0.001525 per second and bundles 13 CPUs and 144 GiB of RAM whether you touch them or not. Modal sells the resources instead, so an H100 job that needs one core pays for one core. That difference compounds on spiky inference, and the roughly one-second container boot means you're not paying for provisioning either.

The cost of a function is a calculation, not a lookup. Three meters plus three multipliers do that: pinning to us-west costs 1.75x, and Sandbox compute quietly runs at 3x without the pricing page saying why.

Startup time bills, and so do the 60 seconds of idle scaledown before a container stops. Both are easy to miss when you're modelling cost per request. Egress also arrives as a new charge on 1 October at $0.04 per GiB, landing on teams already committed.

The rate card itself is complete, published to six decimal places, and queryable from the CLI.

Best For

Bursty GPU workloads with uneven CPU and memory needs.

Watch Out For

Long scaledown_window settings and narrow region pinning, which stack quietly.

Manish Choudhary

CEO & Co-founder, Flexprice

Reselling GPU compute and need to meter CPU, memory and GPU per customer?

Flexprice does multi-dimensional metering out of the box.

Flexprice’s Take

Modal decomposes the bill the way the machine actually works, and almost nobody else does.

Replicate is the closest structural neighbour. Its gpu-h100 costs $0.001525 per second and bundles 13 CPUs and 144 GiB of RAM whether you touch them or not. Modal sells the resources instead, so an H100 job that needs one core pays for one core. That difference compounds on spiky inference, and the roughly one-second container boot means you're not paying for provisioning either.

The cost of a function is a calculation, not a lookup. Three meters plus three multipliers do that: pinning to us-west costs 1.75x, and Sandbox compute quietly runs at 3x without the pricing page saying why.

Startup time bills, and so do the 60 seconds of idle scaledown before a container stops. Both are easy to miss when you're modelling cost per request. Egress also arrives as a new charge on 1 October at $0.04 per GiB, landing on teams already committed.

The rate card itself is complete, published to six decimal places, and queryable from the CLI.

Best For

Bursty GPU workloads with uneven CPU and memory needs.

Watch Out For

Long scaledown_window settings and narrow region pinning, which stack quietly.

Manish Choudhary

CEO & Co-founder, Flexprice

Reselling GPU compute and need to meter CPU, memory and GPU per customer?

Flexprice does multi-dimensional metering out of the box.

Customer
Sentiment Highlights

"Now we’re running the same workloads for under $100, with fast autoscaling and no lag."

botirk, Hacker News, October 2025

"Cold boot times are around 5m but if your usage periods are predictable it can work out ok. Works out at $2 an hour."

siquick, Hacker News, February 2026

Frequently Asked Questions

Frequently Asked Questions

How much does Modal cost per hour?

Does Modal have a free tier?

Does Modal bill for cold starts and idle time?

Do Modal credits roll over?

Launch usage-based billing this week, not next quarter

Launch usage-based billing this week, not next quarter

Get Instant Feedback on Your Pricing | Join the Flexprice Community with 400+ Builders on Slack

Join the Flexprice Community on Slack