Flexible pricing for every workload. Choose serverless or on-demand GPUs, pay-as-you-go sandboxes, or BYOC.
An endpoint on an RTX 4090 with 2 physical cores (4 vCPU) and 16 GiB RAM costs $1.77/hr. Billed by the millisecond, only while your code is running.
Pay by the millisecond, only when the sandbox is running.
A 1-core (2 vCPU) / 8 GiB sandbox costs $0.319/hr, or $0.0000887 per second.
8 hours on an H100 machine is $13.92. That price already includes the 26 vCPU, 200 GB RAM, and 1 TB NVMe drive.
Clusters
Multi-node GPU clusters with InfiniBand, reserved monthly or yearly. Every cluster connects to Beam, so one SDK runs your cluster, serverless workloads, and your own cloud.
Storage
Persistent volumes for models, weights, and any data stored.
A g5.2xlarge instance on AWS (8 vCPU · 32 GB RAM) costs $0.44/hr in Beam fees. The instance itself is billed directly from your AWS account using your credits.
Cost calculator
Pick a GPU (or not), add CPU cores and RAM, and set your monthly usage.
Usage Tiers
$0Per month, plus usage | $89Per month, plus usage | Contact Us | |
|---|---|---|---|
| Usage | |||
| Monthly Credits | $30 | $30 | $30 |
| Apps | Unlimited apps | Unlimited apps | Unlimited apps |
| Cloud Storage Volumes | Unlimited storage volumes | Unlimited storage volumes | Unlimited storage volumes |
| Custom Images | Unlimited custom images | Unlimited custom images | Unlimited custom images |
| GPU Concurrency | 5 Containers | 50 Containers | 1,000+ Containers |
| CPU Concurrency | 30 Containers | 1000 Containers | Unlimited Containers |
| API Requests | Unlimited requests | Unlimited requests | Unlimited requests |
| Log Retention | 30 Days | 30 Days | 1 Year |
| Seats | 1 seat | 3 seats included, $25 per additional seat | Unlimited seats |
| Log Export | Included | Included | Included |
| Bandwidth | Included | Included | Included |
| Features | |||
| Cloud Storage Volumes | |||
| Secrets Manager | |||
| Deployment Logs | |||
| Export Your Code | |||
| Keep Warm | |||
| Autoscaling | |||
| Compliance | |||
| SOC2 Compliance | |||
| HIPAA Compliance | |||
| Support | |||
| Community Support | |||
| Live Chat Support | |||
| Private Slack Channel | |||
Running more than $10K/mo? Volume discounts available — book a call
Beam is powering hands-down the best developer experience to run models on GPUs easily at scale. Best decision on the infra side for us this year so far.

Frase is running language models exclusively on Beam and it was surprisingly easy to migrate, less maintenance, and is saving us money.

TRUSTED BY THE BEST AI COMPANIES
Frequently asked questions
How much does storage cost?
Storage is included up to 1TB, free of charge. Beyond that, it costs $0.021 per GB per month. Snapshots are included.
I don't want serverless. Can I use Beam while keeping the server running 24/7?
Yes. By default, apps will spin down automatically after each request, but you can control how long your apps are active. Many customers with latency-sensitive workloads choose to keep their servers running 24/7 on Beam.
Am I billed for cold start?
No. We only charge for the time to load your application code. We don’t charge for the time to spin up a server or load your container image.
I have an existing Docker image. Can I use it on Beam?
Yes. You can import any base image from a third-party image registry onto Beam.
How secure is Beam?
Workloads are isolated from one another and run in non-root containers. We also offer a self-hosted product which runs entirely in your own environment, ensuring that no data leaves your VPC.
Can I run Beam in my own cloud account?
Yes. With Bring Your Own Cloud, Beam manages workloads directly on instances in your AWS or GCP account, so you can use your existing cloud credits and negotiated rates. You pay Beam a flat management fee of $0.019/hr per vCPU and $0.009/hr per GB of RAM — the compute itself is billed to your cloud account.
Do you charge for egress or bandwidth?
No. There are no egress or bandwidth fees on any plan.
Start shipping on infra
you won’t outgrow.
Run sandboxes and GPU workloads on your cloud, and scale out to ours when you need to. No infra to manage.