Researchers have introduced LLaDA-Image, a novel framework for generating high-quality images using a 6B Diffusion Transformer trained from scratch. This model leverages image-only pre-training and a specialized optimizer to achieve photorealistic results and precise editing capabilities. A distilled version, LLaDA-Image-Turbo, enables rapid inference. LLaDA-Image has set a new state-of-the-art among open-source models on the Qwen-Image-Bench, with its weights, code, and training recipes being publicly released to foster further research. AI
IMPACT Sets new open-source SOTA for image generation and editing, encouraging further research and development in generative models.
RANK_REASON The cluster describes a research paper detailing a new image generation model and its performance on a benchmark, with associated code and weights released.
Read on Hugging Face Daily Papers →
- Diffusion Transformer
- inclusionAI/LLaDA-Image-Turbo
- LLaDA2.0-mini
- LLaDA-Image
- LLaDA-Image-Turbo
- Muon optimizer
- Qwen Image Bench
- Hugging Face
- RMSNorm
- Diffusion Transformer (DiT)
- muon
AI-generated summary · Google Gemini · from 4 sources. How we write summaries →