Beam
the six Ws · specification
Developers deploying AI inference endpoints, task queues, and code sandboxes without managing servers.
Beam is a serverless GPU compute platform offering inference endpoints, task queues, and sandboxes with sub second cold starts via memory snapshots.
Deployed through Beam's Python SDK and CLI, running across AWS, GCP, Azure, and other backing clouds Beam orchestrates.
Founded as a Y Combinator company and actively operating as a commercial platform as of 2026.
Lets teams scale AI workloads from zero to many GPU workers with pay as you go pricing and minimal cold start latency.
Requires a Beam account and API key, and deployments are defined using Beam's Python decorators around existing model code.
Closed source commercial hosted service; requires account credentials and billing, workloads execute on cloud GPUs orchestrated by Beam.