
Stability AI is a Foundation Models - Image & Audio company founded in 2020 and based in London, United Kingdom. It has raised $256M in total funding, most recently a Growth in 2025 at a $1B valuation.
| Date | Stage | Amount | Valuation | Lead investors |
|---|---|---|---|---|
| Mar 5, 2025 | Growth | — | — | WPP |
| Jun 25, 2024 | Seed | $80M | $1B |
| Coatue Management, Lightspeed Venture Partners, Greycroft, Sound Ventures |
| Nov 9, 2023 | Convertible | $50M | — | Intel |
| May 1, 2023 | Convertible | $25M | — | — |
| Oct 17, 2022 | Seed | $101M | $1B | Coatue, Lightspeed Venture Partners, O'Shaughnessy Ventures LLC |

An open-source TRELLIS.2 Studio tool generates 3D assets in under 7 minutes on consumer NVIDIA GPUs.

A researcher releases SesquiLSR, a 3M-parameter learned latent upscaler for Flux, Anima, SDXL and other image models, as an open-source ComfyUI node.

Clark Air releases a compressed 1.58-bit version of Stability AI's Sana 1.6B model achieving 8x compression with minimal quality loss.

The Atlantic publishes a searchable database exposing music datasets used to train AI models, including datasets used by Google and Stability.

Community developer releases texture-to-albedo LoRA for Flux Klein 9B on Hugging Face, enabling lighting-independent texture extraction.

Forge Neo adds SDXL GGUF support, expanding inference options for Stable Diffusion models.
Stable Diffusion is a text-to-image generative AI model that creates photorealistic images from natural language descriptions. Released in August 2022, it democratized access to image generation technology by being open-source and free, with downloads exceeding 350 million. The latest version includes improved multi-subject prompt handling, image quality, and spelling abilities. It operates via API, web interface, and downloadable models for researchers, artists, and creative professionals globally.
Stable Audio 2.5 is an enterprise-grade AI music and audio generation model that creates custom-length music, sound effects, and audio tracks from text descriptions. It generates full compositions up to three minutes in stereo at 44.1 kHz quality, with advanced audio-to-audio and audio inpainting capabilities. Trained on licensed datasets with opt-out protections for creators, it serves musicians, producers, gaming studios, and advertisers seeking high-quality, customizable audio generation without copyright concerns.
Stable Video 3D is a generative model for creating detailed three-dimensional objects and volumetric content from images or videos in seconds. Building on Stable Video Diffusion, it delivers improved quality and multi-view generation capabilities, enabling content creators to rapidly prototype 3D assets for games, films, virtual environments, and design applications. Available for both commercial and non-commercial use via API or self-hosted deployment.

We don't have a live feed for this company's ATS. Their careers page has every open role.
View all careers ↗