Stability AI Launches Advanced Stable Diffusion XL 1.0 Image Generator
Stability AI unveils Stable Diffusion XL 1.0, a highly capable open-source text-to-image model that produces vivid 1-megapixel images in seconds. The new system excels at rendering legible text and understanding complex, short prompts.
Stability AI introduces Stable Diffusion XL 1.0, calling it the most advanced text-to-image model in its lineup. The open-source system features 3.5 billion parameters and generates vibrant, high-contrast 1-megapixel resolution images in seconds. Unlike its predecessor, this updated model achieves these high-quality results with less computational power.
The new model handles complex designs using simple, short natural language prompts instead of requiring long, detailed text. It also shows significant improvement in generating legible text within images, solving a common struggle for similar AI systems. Furthermore, users can edit existing photos through inpainting, outpainting, and image-to-image prompt capabilities.
Stability AI makes Stable Diffusion XL 1.0 available across multiple platforms, including GitHub, its API, and consumer apps like ClipDrop and DreamStudio. The company designs the model to be easily customizable and fine-tuned for specific styles and concepts. However, releasing this powerful open-source tool continues to bring up familiar ethical concerns regarding generative AI.