OCR deepseek-ai/DeepSeek-OCR-2 Image-Text-to-Text • 3B • Updated Feb 3 • 1.41M • 1.08k zai-org/GLM-OCR Image-Text-to-Text • 1B • Updated May 19 • 3.1M • • 1.99k uv-scripts/ocr Updated 20 days ago • 2.77k • 155 numind/NuMarkdown-8B-Thinking Image-to-Text • 8B • Updated Jun 5 • 217k • 493
Language tencent/Hunyuan-MT-7B Translation • 8B • Updated Dec 30, 2025 • 3.14k • 737 tencent/HunyuanWorld-Voyager Image-to-Video • Updated Oct 17, 2025 • 116 • 609 moonshotai/Kimi-K2-Instruct-0905 Text Generation • 1T • Updated Jan 30 • 42.8k • • 786 Qwen/Qwen3.8-27B Image-Text-to-Text • 28B • Updated 5 days ago • 666k • • 11.2k
Voice microsoft/VibeVoice-1.5B Text-to-Speech • 3B • Updated Jan 22 • 116k • 2.46k Running Featured 447 FastVLM WebGPU 🍎 447 Real-time video captioning powered by FastVLM openbmb/VoxCPM-0.5B Text-to-Speech • Updated 1 day ago • 15k • 814 Paused 85 MiMo-Audio-Chat 💬 85 Chat with Xiaomi MiMo-Audio using voice
Model training merve/smol-vision Image-Text-to-Text • Updated 9 days ago • 195 HiDream-ai/HiDream-E1-1 Any-to-Any • 17B • Updated Jul 17, 2025 • 97 • 216 Jackrong/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled Image-Text-to-Text • 28B • Updated Jul 7 • 44.2k • • 2.94k netflix/void-model Video-to-Video • Updated Apr 6 • 963
Jackrong/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled Image-Text-to-Text • 28B • Updated Jul 7 • 44.2k • • 2.94k
music ACE-Step/acestep-v15-base Text-to-Audio • 2B • Updated Feb 6 • 2.33k • 67 Running on Zero MCP 34 BS-Roformer Leap Audio Separator 🎵 34 Separate audio into vocals and instruments with BS-Roformer Running on Zero MCP 17 StuPASE Speech Enhancement 🎙 17 Studio-quality generative speech enhancement
Running on Zero MCP 34 BS-Roformer Leap Audio Separator 🎵 34 Separate audio into vocals and instruments with BS-Roformer
Papers Group Sequence Policy Optimization Paper • 2507.18071 • Published Jul 24, 2025 • 322 MatrAIx: Simulating the World with 8.3 Billion Persona Agents Paper • 2608.04205 • Published 15 days ago • 47
MatrAIx: Simulating the World with 8.3 Billion Persona Agents Paper • 2608.04205 • Published 15 days ago • 47
music ACE-Step/acestep-v15-base Text-to-Audio • 2B • Updated Feb 6 • 2.33k • 67 Running on Zero MCP 34 BS-Roformer Leap Audio Separator 🎵 34 Separate audio into vocals and instruments with BS-Roformer Running on Zero MCP 17 StuPASE Speech Enhancement 🎙 17 Studio-quality generative speech enhancement
Running on Zero MCP 34 BS-Roformer Leap Audio Separator 🎵 34 Separate audio into vocals and instruments with BS-Roformer
OCR deepseek-ai/DeepSeek-OCR-2 Image-Text-to-Text • 3B • Updated Feb 3 • 1.41M • 1.08k zai-org/GLM-OCR Image-Text-to-Text • 1B • Updated May 19 • 3.1M • • 1.99k uv-scripts/ocr Updated 20 days ago • 2.77k • 155 numind/NuMarkdown-8B-Thinking Image-to-Text • 8B • Updated Jun 5 • 217k • 493
Language tencent/Hunyuan-MT-7B Translation • 8B • Updated Dec 30, 2025 • 3.14k • 737 tencent/HunyuanWorld-Voyager Image-to-Video • Updated Oct 17, 2025 • 116 • 609 moonshotai/Kimi-K2-Instruct-0905 Text Generation • 1T • Updated Jan 30 • 42.8k • • 786 Qwen/Qwen3.8-27B Image-Text-to-Text • 28B • Updated 5 days ago • 666k • • 11.2k
Voice microsoft/VibeVoice-1.5B Text-to-Speech • 3B • Updated Jan 22 • 116k • 2.46k Running Featured 447 FastVLM WebGPU 🍎 447 Real-time video captioning powered by FastVLM openbmb/VoxCPM-0.5B Text-to-Speech • Updated 1 day ago • 15k • 814 Paused 85 MiMo-Audio-Chat 💬 85 Chat with Xiaomi MiMo-Audio using voice
Papers Group Sequence Policy Optimization Paper • 2507.18071 • Published Jul 24, 2025 • 322 MatrAIx: Simulating the World with 8.3 Billion Persona Agents Paper • 2608.04205 • Published 15 days ago • 47
MatrAIx: Simulating the World with 8.3 Billion Persona Agents Paper • 2608.04205 • Published 15 days ago • 47
Model training merve/smol-vision Image-Text-to-Text • Updated 9 days ago • 195 HiDream-ai/HiDream-E1-1 Any-to-Any • 17B • Updated Jul 17, 2025 • 97 • 216 Jackrong/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled Image-Text-to-Text • 28B • Updated Jul 7 • 44.2k • • 2.94k netflix/void-model Video-to-Video • Updated Apr 6 • 963
Jackrong/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled Image-Text-to-Text • 28B • Updated Jul 7 • 44.2k • • 2.94k