Stable Diffusion 3 early preview announced β strengthens text rendering and prompt comprehension
Stability AI has announced an early preview of its next-generation flagship model, βStable Diffusion 3.β With image and video generation AI from other companies like Sora and Gemini gaining prominence, Stability AI emphasized improvements to image quality and prompt adherence, as well as better accuracy in spelling out text within images β long a weak point for AI image generation.
Details
- Model sizes: Offered as multiple derivative models ranging from 800 million to 8 billion parameters, letting developers choose and tune based on their use case
- Improved text rendering: Focused heavily on improving the ability to accurately generate text within images
- Prompt comprehension: Improved handling of complex, long prompts, making it easier to reproduce an intended composition
- Availability: Initially released as a limited early preview to those on a waitlist
What happened next
About four months after this early preview announcement, in June 2024, βSD3 Medium,β a 2-billion-parameter version, was released publicly with open weights. The SD3 series became an important step in demonstrating Stability AIβs technical presence amid its ongoing corporate restructuring.