Stable Diffusion 1.0
Public release of the first open-weight text-to-image model. Democratized image generation overnight and launched the broader open-source diffusion ecosystem.
Source · Stability AI archiveA definitive 2022–2026 timeline of Stability AI — Stable Audio 3.0, Stable Diffusion 3.5, the Mostaque-to-Akkaraju leadership reshuffle, the $80M Series, and the cleanest way to actually use every Stable Diffusion model without a dedicated Stability subscription.
TL;DR — Stability AI in 2026
Stability AI news in 2026 centers on Stable Audio 3.0 (four models, the largest generating 6 minutes 20 seconds of music) and Stable Diffusion 3.5 — released under a new open-weight Community License — backed by a quiet turnaround under CEO Prem Akkaraju, Executive Chairman Sean Parker, and an $80M Series round at a roughly $1B valuation. ZeroTwo bundles Stable Diffusion 3.5 variants alongside Flux Pro, DALL·E, and Imagen for $0 free / $29.99 Pro, so you don't need a dedicated Stability subscription to actually use the models.
Ten dated checkpoints — every major model release, leadership change, and funding event — sourced and linked. Scroll the carousel below.
Public release of the first open-weight text-to-image model. Democratized image generation overnight and launched the broader open-source diffusion ecosystem.
Source · Stability AI archiveFirst commercially viable AI music generation tool, producing 44.1 kHz stereo audio. Set the foundation for Stable Audio 2.0 and 3.0 in the next 30 months.
Source · Stability AIFounder and CEO Emad Mostaque steps down. Interim co-CEOs Shan Shan Wong and Christian Laforte take over as the board searches for permanent leadership.
Source · SiliconANGLEThree-minute audio generation with audio-to-audio transformation. Trained on a fully licensed AudioSparx dataset of 806,284 files — a deliberate move away from prior copyright fights.
Source · Stability AISD3 and SD3 Turbo go live via API. Introduces the MMDiT (Multi-Modal Diffusion Transformer) architecture that anchors every subsequent SD release.
Source · Stability AISD3 (2B parameters, full text-to-image) released to public beta with open weights for non-commercial research and a separate commercial license tier.
Source · Stability AIPrem Akkaraju (former Weta Digital CEO) appointed CEO. Sean Parker becomes Executive Chairman. $80M Series round led by Coatue, Lightspeed, Greycroft, and Sound Ventures values the company at roughly $1B post-round.
Source · BloombergAcademy Award–winning filmmaker James Cameron joins Stability AI's Board of Directors, alongside Sean Parker, Greycroft's Dana Settle, and Coatue's Colin Bryant. Signals deeper push into film and CGI workflows.
Source · Stability AIThree variants — Large (8B), Large Turbo, and Medium (2.5B) — released with open weights under the Stability AI Community License. Medium variant adds a path for VRAM-constrained consumer hardware.
Source · Stability AIFour model variants (459M–2.7B parameters) ship with the large model generating up to 6 minutes 20 seconds of music. Open weights for the small and medium tiers. Partnerships announced with WPP, Warner Music Group, and Universal Music Group; Ethan Kaplan joins to lead the pro music offering.
Source · TechCrunch← Swipe / scroll →
The most consequential Stability AI news of the last two years was not a model release — it was a complete leadership overhaul. On March 23, 2024, founder and CEO Emad Mostaque resigned amid investor pressure over the company's burn rate. Interim co-CEOs Shan Shan Wong and Christian Laforte ran the company through the spring while the board searched for permanent leadership and an investor consortium negotiated a debt restructuring.
On June 25, 2024, Prem Akkaraju — former CEO of Weta Digital, the visual-effects studio behind Avatar and The Lord of the Rings — was appointed permanent CEO. The same day, Sean Parker (founding president of Facebook, founder of Sound Ventures) joined as Executive Chairman, and Stability announced an $80M Series round at roughly a $1B post-money valuation. Lead investors included Coatue, Lightspeed Venture Partners, Greycroft, and Sound Ventures. Per SiliconANGLE, the investor group also persuaded suppliers to forgive over $100M in debt and roughly $300M in future obligations, effectively giving the company a clean balance sheet.
The turnaround worked. Fortune reported in December 2024 that the company had moved from a $30M+ Q1 2024 loss under prior leadership to triple-digit growth and a debt-free balance sheet. Then, on September 24, 2024, Stability AI announced that filmmaker James Cameron had joined the Board of Directors — a strong signal of where Akkaraju's team intends to push the company next: deeper into film, gaming, and CGI workflows.
Akkaraju's biography is part of the story. Before Stability AI, he was CEO of Weta Digital — the New Zealand visual effects company behind Avatar, The Lord of the Rings trilogy, King Kong, and dozens of other landmark releases — and he led the 2021 sale of Weta Digital's technology arm to Unity for $1.625 billion. That background matters because Stability AI's open-weight model strategy intersects directly with the film and gaming production pipeline, and the company needs leadership that can speak fluently to studio buyers and creative directors, not just to ML researchers and developers.
The fully reconstituted board reflects the same thesis. Sean Parker brings consumer and investor reach via Sound Ventures; Greycroft's Dana Settle brings B2B SaaS and developer-tools expertise; Coatue's Colin Bryant brings late-stage capital discipline; and James Cameron brings unmatched creative authority on visual storytelling. Per the PR Newswire release announcing the funding round, the company's stated focus under the new leadership is "the development of cutting-edge AI models" with a clear path to enterprise adoption.
"I've spent my career seeking out emerging technologies that push the very boundaries of what's possible, all in the service of telling incredible stories… the intersection of generative AI and CGI image creation is the next wave. The convergence of these two totally different engines of creation will unlock new ways for artists to tell stories in ways we could have never imagined."
Stable Diffusion 3.5 dropped on October 22, 2024, with three model variants released simultaneously under the Stability AI Community License: SD 3.5 Large (the flagship), SD 3.5 Large Turbo (a few-step distilled variant for speed), and SD 3.5 Medium. The Medium variant in particular was a deliberate accessibility choice — it runs on consumer GPUs with as little as 8 GB of VRAM, which is the hardware most independent artists actually own.
The underlying architecture is MMDiT (Multi-Modal Diffusion Transformer), first introduced in the SD 3 API release in April 2024. MMDiT is the same family of architecture used in Stable Audio 3.0, and the shared backbone is part of why Stability AI's release cadence accelerated through 2024 and 2026. Compared to SDXL's U-Net, MMDiT delivers stronger prompt adherence — especially for typography, multi-subject composition, and long prompts with multiple attribute bindings.
The licensing nuance is important to understand. The Stability AI Community License is not a fully unrestricted open-source license in the OSI sense. It allows research use, personal use, and commercial use up to a defined annual revenue threshold; over that threshold, customers need an enterprise license from Stability AI. That structure — free for research and small creators, paid for enterprise — is the financial model that supports the open-weight releases.
Stable Diffusion 3.5 also extends the broader open-weight ecosystem in ways that matter for downstream creators. Because weights are publicly available, community fine-tunes ship quickly: Animagine XL and Illustrious for anime-style art, Pony Diffusion for vivid digital illustration, and a long tail of LoRA adapters for specific characters, artists, and aesthetics. The ControlNet conditioning ecosystem — including Canny edge, depth, openpose, and tile conditioning — works across the SD 3.x family and is what enables practical workflows like sketch-to-photo, pose transfer, and consistent multi-shot character generation.
For most artists and small teams, comparing Stable Diffusion 3.5 head-to-head with Flux Pro and Imagen 3 is more useful than running it standalone — different prompts land better on different models, and a multi-model platform lets you A/B test without paying multiple subscriptions. You can also try anime-style Stable Diffusion generation with Animagine XL and Illustrious checkpoints on the same workspace.
On May 20, 2026, TechCrunch broke the news of Stable Audio 3.0, the most significant audio release in Stability AI's history. The family ships in four sizes — a 459M Small SFX model tuned for short sound effects, a 459M Small music model, a 1.4B Medium model, and a 2.7B Large model capable of generating up to 6 minutes 20 seconds of music in one continuous run. Open weights ship for the Small and Medium tiers under the Community License; the Large model is available through Stability's commercial API and partner integrations.
The training set continues Stable Audio 2.0's licensing-first approach. Stable Audio 3.0 was trained on 806,284 fully licensed audio files from the AudioSparx production library — a deliberate departure from the copyright fights that shaped early Stability AI history and a clear strategic signal toward enterprise music partnerships.
Those partnerships matter. Alongside the model release, Stability AI confirmed collaboration agreements with WPP (one of the world's largest advertising holding companies), Warner Music Group, and Universal Music Group. The hire of Ethan Kaplan to lead Stability's professional music offering — paired with the WMG/UMG/WPP relationships — is the company's most direct enterprise push to date in audio.
Worth a note on context: a third pillar — video generation — continues to advance through Stable Video Diffusion (SVD), which is a 1.52B-parameter latent video diffusion model with 656M parameters devoted to temporal processing, according to its HuggingFace model card. SVD remains a key building block for short clip generation while the company's image and audio releases dominate headlines.
The 6-minute song generation length is the headline number, but the more interesting technical detail is what changed in the training pipeline. Stable Audio 3.0 keeps the latent diffusion approach pioneered in 2.0 but extends the temporal context window dramatically and adds finer-grained conditioning signals — tempo, key, instrument-set, mood — so prompts can target a specific musical direction with much less hand-tuning. Combined with the open-weight Small and Medium tiers, that opens the door for indie producers and game studios to fine-tune Stable Audio for genre-specific soundtracks, sound-effect libraries, and in-game adaptive scoring without licensing the Large variant from Stability.
The numbers behind the 2024 recapitalization and the 2026 audio release — sourced and linked. Every figure here links back to its primary or tier-1 secondary source.
Skip the subscription juggling
Start free. Upgrade to Pro at $29.99/mo when you want unlimited use across every bundled image and chat model.
Free · Pro $29.99 · Pro 2x $59.98 · Ultra $120
There are three practical ways to use Stable Diffusion in 2026: a multi-model platform (the fastest path), a local install (the geekiest path), or Stability AI's own Stable Assistant product (the official-but-paid path). For most people, the multi-model platform path is the right answer because it lets you run the same prompt across Stable Diffusion 3.5, Flux Pro, Imagen 3, and DALL·E 3 inside one workspace — then pick the output that landed best — without paying four subscriptions.
No credit card required. The free tier covers daily Stable Diffusion generation alongside Flux, Imagen, and DALL·E. Upgrade to Pro at $29.99/mo when you outgrow the daily quota.
SD 3.5 Large for the highest-quality output; SD 3.5 Large Turbo for fast iteration; SD 3.5 Medium for lighter compute. Community fine-tunes — Animagine XL, Illustrious, Pony — are also bundled for genre-specific aesthetics.
A 4-up generation across Stable Diffusion 3.5, Flux Pro, Imagen 3, and DALL·E 3 typically takes under two minutes and gives you four distinct aesthetic directions to choose between. No tool-switching, no copy-paste between platforms.
Important: ZeroTwo is not Stability AI
ZeroTwo is an independent multi-model AI platform. We use Stable Diffusion under its Community License alongside Flux Pro, Imagen 3, DALL·E 3, and other models — we don't build or own Stable Diffusion. For first-party product news, enterprise licensing, and official model weights, go directly to stability.ai. For multi-model creative work using Stable Diffusion plus every other major model in one workspace, ZeroTwo is the fastest path.
The questions Google, ChatGPT, Perplexity, and Claude get asked most about Stability AI — answered with sourced facts and the licensing nuance most articles skip.
Stable Diffusion 3.5 head-to-head with Flux Pro, Imagen 3, and DALL·E 3. Scorecards, pricing, sample outputs.
Type a prompt, pick a model, get pictures. Multi-model image generation including Stable Diffusion.
Animagine XL, Illustrious, and Pony checkpoints built on Stable Diffusion for anime-style characters.
Test Stable Diffusion + Flux + Imagen + DALL·E with no credit card on ZeroTwo's free tier.
Painterly fantasy portraits, creature concepts, and high-fantasy landscapes across Stable Diffusion variants.
Tabletop characters, NPCs, and campaign maps via Stable Diffusion checkpoints tuned for the genre.
ControlNet conditioning on Stable Diffusion to render photoreal images from a drawing or layout.
Spotify-ready 3000×3000 cover art across Stable Diffusion, Flux, Imagen 4, and GPT-image-1.
Methodology-driven scorecard of every major AI platform — pricing, model breadth, workflow tools.
No credit card. No setup. Stable Diffusion plus every other major image and chat model in one workspace.
Free · Pro $29.99 · Pro 2x $59.98 · Ultra $120
Published 2026-05-21 · By ZeroTwo Editorial · Updated 2026-05-21