Images in Sentences: Scaling Interleaved Instructions for Unified Visual Generation Paper • 2605.12305 • Published May 12 • 2
RefineAnything: Multimodal Region-Specific Refinement for Perfect Local Details Paper • 2604.06870 • Published Apr 8 • 44
Running on Zero MCP 3.75k Z Image Turbo 🖼 3.75k Generate high-quality images from text prompts in seconds
docling-project/SmolDocling-256M-preview Image-Text-to-Text • 0.3B • Updated Sep 17, 2025 • 30.2k • 1.62k