OCR deepseek-ai/DeepSeek-OCR-2 Image-Text-to-Text • 3B • Updated Feb 3 • 1.8M • 1.07k zai-org/GLM-OCR Image-Text-to-Text • 1B • Updated May 19 • 3.59M • • 1.97k uv-scripts/ocr Updated 8 days ago • 3.18k • 155 numind/NuMarkdown-8B-Thinking Image-to-Text • 8B • Updated Jun 5 • 58.5k • 493
Language tencent/Hunyuan-MT-7B Translation • 8B • Updated Dec 30, 2025 • 1.8k • 736 tencent/HunyuanWorld-Voyager Image-to-Video • Updated Oct 17, 2025 • 157 • 609 moonshotai/Kimi-K2-Instruct-0905 Text Generation • 1T • Updated Jan 30 • 57k • • 787
Voice microsoft/VibeVoice-1.5B Text-to-Speech • 3B • Updated Jan 22 • 78.6k • 2.45k Configuration error Featured 446 FastVLM WebGPU 🍎 446 Real-time video captioning powered by FastVLM openbmb/VoxCPM-0.5B Text-to-Speech • Updated Sep 19, 2025 • 2.8k • 808 Paused 85 MiMo-Audio-Chat 💬 85 Chat with Xiaomi MiMo-Audio using voice
Model training merve/smol-vision Image-Text-to-Text • Updated Nov 5, 2025 • 194 HiDream-ai/HiDream-E1-1 Any-to-Any • 17B • Updated Jul 17, 2025 • 127 • 216 Jackrong/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled Image-Text-to-Text • 28B • Updated Jul 7 • 59.9k • • 2.93k netflix/void-model Video-to-Video • Updated Apr 6 • 961
Jackrong/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled Image-Text-to-Text • 28B • Updated Jul 7 • 59.9k • • 2.93k
music ACE-Step/acestep-v15-base Text-to-Audio • 2B • Updated Feb 6 • 2.22k • 67 Running on Zero MCP 32 BS-Roformer Leap Audio Separator 🎵 32 Separate audio into vocals and instruments with BS-Roformer
Running on Zero MCP 32 BS-Roformer Leap Audio Separator 🎵 32 Separate audio into vocals and instruments with BS-Roformer
music ACE-Step/acestep-v15-base Text-to-Audio • 2B • Updated Feb 6 • 2.22k • 67 Running on Zero MCP 32 BS-Roformer Leap Audio Separator 🎵 32 Separate audio into vocals and instruments with BS-Roformer
Running on Zero MCP 32 BS-Roformer Leap Audio Separator 🎵 32 Separate audio into vocals and instruments with BS-Roformer
OCR deepseek-ai/DeepSeek-OCR-2 Image-Text-to-Text • 3B • Updated Feb 3 • 1.8M • 1.07k zai-org/GLM-OCR Image-Text-to-Text • 1B • Updated May 19 • 3.59M • • 1.97k uv-scripts/ocr Updated 8 days ago • 3.18k • 155 numind/NuMarkdown-8B-Thinking Image-to-Text • 8B • Updated Jun 5 • 58.5k • 493
Language tencent/Hunyuan-MT-7B Translation • 8B • Updated Dec 30, 2025 • 1.8k • 736 tencent/HunyuanWorld-Voyager Image-to-Video • Updated Oct 17, 2025 • 157 • 609 moonshotai/Kimi-K2-Instruct-0905 Text Generation • 1T • Updated Jan 30 • 57k • • 787
Voice microsoft/VibeVoice-1.5B Text-to-Speech • 3B • Updated Jan 22 • 78.6k • 2.45k Configuration error Featured 446 FastVLM WebGPU 🍎 446 Real-time video captioning powered by FastVLM openbmb/VoxCPM-0.5B Text-to-Speech • Updated Sep 19, 2025 • 2.8k • 808 Paused 85 MiMo-Audio-Chat 💬 85 Chat with Xiaomi MiMo-Audio using voice
Model training merve/smol-vision Image-Text-to-Text • Updated Nov 5, 2025 • 194 HiDream-ai/HiDream-E1-1 Any-to-Any • 17B • Updated Jul 17, 2025 • 127 • 216 Jackrong/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled Image-Text-to-Text • 28B • Updated Jul 7 • 59.9k • • 2.93k netflix/void-model Video-to-Video • Updated Apr 6 • 961
Jackrong/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled Image-Text-to-Text • 28B • Updated Jul 7 • 59.9k • • 2.93k