Runtime, the conference for engineers running AI in production. Oct. 1 in SF
PRICING PLANS
Starter
$0
+
compute / month
Built for small teams and independent developers looking to level up.
$30 / month free credits
3 workspace seats included
100 containers + 10 GPU concurrency
Scheduled and Web Functions (limited)
Real-time metrics and logs
Region selection
Team
$250
+
compute / month
Built for startups and larger organizations looking to scale quickly.
$100 / month free credits
Unlimited seats
5000 containers + 50 GPU concurrency
Unlimited Scheduled Functions
Custom domains
Static IP proxy
Deployment rollbacks
Environment-level budgets
Enterprise
Custom
For organizations prioritizing security, support, and everlasting confidence.
Volume-based discounts
Unlimited seats
Higher GPU concurrency
Embedded ML engineering services
Environment-level budgets
Support via private Slack
Audit logs, Okta SSO, and HIPAA
Credit grants for startups
Early-stage startups can get free compute credits on Modal.
Credit grants for academics
Graduate students, labs, and researchers can get up to $10k free compute credits on Modal
Use committed spend on Modal
Transact through the AWS and GCP marketplace to use committed spend on Modal.
Modal Sandbox + Notebooks Pricing
Only pay for what you use. Burst up to what you need without over-allocating CPU or memory in advance.
CPU
Physical core
(2 vCPU equivalent)
$0.00003942 / core / sec
*minimum of 0.125 cores per container
Memory
$0.00000667 / GiB / sec

Why serverless?
Serverless pricing vs. traditional cloud pricing
Modal is serverless, which means that we instantly autoscale up and down for you based on request volume. For spiky or unpredictable workloads, we are more cost-effective than fixed on-demand/reserved compute.
Traditional cloud: $5,400
75 GPUs * 24 hrs * $3 / GPU-hr
Modal serverless cloud: $4,740
Avg 50 GPUs * 24 hrs * $3.95 / GPU-hr
