Kolors

Bilingual text-to-image model by Kuaishou with strong Chinese prompt understanding.

Open SourceSelf HostedOffline CapableGPU Required (8GB+ VRAM)
0.0 (0)

About

Developed by Kuaishou's Kolors team, Kolors is a large scale latent diffusion text-to-image model trained on billions of text and image pairs, with genuinely bilingual understanding of Chinese and English prompts including the ability to render Chinese characters inside generated images. It accepts prompt context up to 256 tokens and scored strongly in human preference evaluations against contemporary open models. The surrounding ecosystem includes IP-Adapter-Plus for image prompts, an IP-Adapter-FaceID variant for identity preserving portraits, ControlNet conditioning on Canny edges, depth, and pose, an inpainting model, image-to-image generation, and LoRA and DreamBooth fine-tuning, with integration into Hugging Face Diffusers. The code is Apache 2.0 licensed, while the weights are free for academic research and require registration with Kuaishou for commercial use above certain scale thresholds. Teams needing accurate Chinese prompt handling are frequent adopters.

Reviews (0)

Leave a Review

No reviews yet. Be the first to review!

Details

Price
Free
Platform
Local/Desktop
Difficulty
Intermediate (3/5)
License
Apache-2.0
Minimum VRAM
8 GB
Added
Apr 3, 2026

Related Tools

Featured

State-of-the-art open image generation model by Black Forest Labs with rectified flow transformers.

Open SourceSelf HostedOfflineGPU 12GB+
Intermediate
0.0 (0)
Featured

Next-generation image generation model by Black Forest Labs.

Open SourceSelf HostedOfflineGPU 12GB+
Intermediate
0.0 (0)

Bilingual text-to-image diffusion transformer by Tencent with Chinese and English support.

Open SourceSelf HostedOfflineGPU 12GB+
Intermediate
0.0 (0)

Zero-shot identity-preserving image generation from a single face photo.

Open SourceSelf HostedOfflineGPU 8GB+
Intermediate
0.0 (0)

Image prompt adapter for pre-trained text-to-image diffusion models.

Open SourceSelf HostedOfflineGPU 8GB+
Intermediate
0.0 (0)
Featured

Neural network architecture for adding spatial control to diffusion models.

Open SourceSelf HostedOfflineGPU 8GB+
Intermediate
0.0 (0)
Browse all Image Generation tools