Image
A high‑resolution text‑to‑image generation model that achieves state‑of‑the‑art visual quality and prompt alignment. Outperforms leading open‑source models like FLUX.1[dev] and Qwen‑Image in aesthetic realism and compositional accuracy
- RL‑finetuned model — delivers the highest visual fidelity and realism
- SFT‑soup model — excels in prompt following and overall visual quality
- Pretrain checkpoint — for researchers to conduct further fine‑tuning and experimentation
- Native Cyrillic text rendering — trained on 10M+ Russian‑text images in different styles, from print to handwriting





























































