Baseten is an AI / GPU Cloud company founded in 2019 and based in San Francisco, United States. It has raised $2.1B in total funding, most recently a Series F in 2026 at a $13B valuation.
| Date | Stage | Amount | Valuation | Lead investors |
|---|---|---|---|---|
| Jan 20, 2026 | Series E | $300M | $5B | IVP, CapitalG |
A serverless inference platform that converts open-source and custom machine-learning models into production-ready, autoscaling API endpoints with GPU-backed serving. It handles model deployment, autoscaling, observability, and optimization so teams can run low-latency inference at scale without operating their own GPU fleet. Baseten serves more than 100 enterprise customers across high-throughput production workloads and offers both dedicated and self-serve deployment options, billing on usage-based GPU compute.
Truss is Baseten's open-source framework for packaging machine-learning models into standardized, reproducible deployment artifacts. It lets developers define a model's dependencies, pre- and post-processing, and serving configuration in code, then deploy the packaged model to Baseten or other environments. Truss reduces the friction of moving a custom model from a notebook to a production endpoint and underpins Baseten's broader inference platform, supporting rapid integration of bespoke and open-source AI models.
We don't have a live feed for this company's ATS. Their careers page has every open role.
View all careers ↗