Stable Diffusion XL
Open-weight latent diffusion model by Stability AI for high-resolution image generation.
About
Stable Diffusion XL, usually shortened to SDXL, is Stability AI's open-weight latent diffusion model for generating 1024x1024 images from text. It uses a two-stage pipeline: a base model produces the initial image latents and an optional refiner model sharpens fine detail, with prompts interpreted through a pair of text encoders, OpenCLIP-ViT/G and CLIP-ViT/L. The weights live in the Stability-AI generative-models repository alongside SDXL-Turbo, a distilled variant that generates in very few steps using adversarial diffusion distillation, and the Stable Video Diffusion models. Released under the CreativeML Open RAIL++-M license, SDXL runs locally on GPUs with around 8 GB of VRAM and remains one of the most broadly supported checkpoints in the open image ecosystem, with fine-tunes, LoRAs, and ControlNets available across tools like ComfyUI and the AUTOMATIC1111 web UI. Typical users range from hobbyists generating art locally to studios building image features on open weights.
Reviews (0)
Leave a Review
No reviews yet. Be the first to review!
Details
- Category
- Image Generation
- Price
- Free
- Platform
- Local/Desktop
- Difficulty
- Intermediate (3/5)
- License
- Open RAIL-M
- Minimum VRAM
- 8 GB
- Added
- Apr 3, 2026
Related Tools
State-of-the-art open image generation model by Black Forest Labs with rectified flow transformers.
Next-generation image generation model by Black Forest Labs.
Bilingual text-to-image diffusion transformer by Tencent with Chinese and English support.
Zero-shot identity-preserving image generation from a single face photo.
Image prompt adapter for pre-trained text-to-image diffusion models.
Neural network architecture for adding spatial control to diffusion models.