music ACE-Step/acestep-v15-base Text-to-Audio • 2B • Updated Feb 6 • 2.71k • 67 Running on Zero MCP 35 BS-Roformer Leap Audio Separator 🎵 35 Separate audio into vocals and instruments with BS-Roformer Running on Zero MCP 18 StuPASE Speech Enhancement 🎙 18 Studio-quality generative speech enhancement
Running on Zero MCP 35 BS-Roformer Leap Audio Separator 🎵 35 Separate audio into vocals and instruments with BS-Roformer
Image Qwen/Qwen-Image-Edit Image-to-Image • 20B • Updated Aug 25, 2025 • 130k • • 2.5k sensenova/SenseNova-U1.5-8B-MoT Any-to-Any • 18B • Updated 13 days ago • 7.93k • 216
Papers Group Sequence Policy Optimization Paper • 2507.18071 • Published Jul 24, 2025 • 323 MatrAIx: Simulating the World with 8.3 Billion Persona Agents Paper • 2608.04205 • Published Aug 4 • 53
MatrAIx: Simulating the World with 8.3 Billion Persona Agents Paper • 2608.04205 • Published Aug 4 • 53
OCR deepseek-ai/DeepSeek-OCR-2 Image-Text-to-Text • 3B • Updated Feb 3 • 987k • 1.09k zai-org/GLM-OCR Image-Text-to-Text • 1B • Updated May 19 • 2.01M • • 2.02k uv-scripts/ocr Updated 9 days ago • 2.59k • 157 numind/NuMarkdown-8B-Thinking Image-to-Text • 8B • Updated Jun 5 • 361k • 494
Language tencent/Hunyuan-MT-7B Translation • 8B • Updated Dec 30, 2025 • 2.59k • 739 tencent/HunyuanWorld-Voyager Image-to-Video • Updated Oct 17, 2025 • 135 • 609 moonshotai/Kimi-K2-Instruct-0905 Text Generation • 1T • Updated Jan 30 • 36.8k • • 786 Qwen/Qwen3.8-27B Image-Text-to-Text • 28B • Updated 23 days ago • 6.19M • • 14.1k
Voice microsoft/VibeVoice-1.5B Text-to-Speech • 3B • Updated Jan 22 • 253k • 2.48k Running Featured 447 FastVLM WebGPU 🍎 447 Real-time video captioning powered by FastVLM openbmb/VoxCPM-0.5B Text-to-Speech • Updated 19 days ago • 20k • 816 Paused 85 MiMo-Audio-Chat 💬 85 Chat with Xiaomi MiMo-Audio using voice
Model training merve/smol-vision Image-Text-to-Text • Updated 27 days ago • 195 HiDream-ai/HiDream-E1-1 Any-to-Any • 17B • Updated Jul 17, 2025 • 82 • 217 Jackrong/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled Image-Text-to-Text • 28B • Updated Jul 7 • 4.74k • • 2.94k netflix/void-model Video-to-Video • Updated Apr 6 • 965
Jackrong/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled Image-Text-to-Text • 28B • Updated Jul 7 • 4.74k • • 2.94k
music ACE-Step/acestep-v15-base Text-to-Audio • 2B • Updated Feb 6 • 2.71k • 67 Running on Zero MCP 35 BS-Roformer Leap Audio Separator 🎵 35 Separate audio into vocals and instruments with BS-Roformer Running on Zero MCP 18 StuPASE Speech Enhancement 🎙 18 Studio-quality generative speech enhancement
Running on Zero MCP 35 BS-Roformer Leap Audio Separator 🎵 35 Separate audio into vocals and instruments with BS-Roformer
OCR deepseek-ai/DeepSeek-OCR-2 Image-Text-to-Text • 3B • Updated Feb 3 • 987k • 1.09k zai-org/GLM-OCR Image-Text-to-Text • 1B • Updated May 19 • 2.01M • • 2.02k uv-scripts/ocr Updated 9 days ago • 2.59k • 157 numind/NuMarkdown-8B-Thinking Image-to-Text • 8B • Updated Jun 5 • 361k • 494
Language tencent/Hunyuan-MT-7B Translation • 8B • Updated Dec 30, 2025 • 2.59k • 739 tencent/HunyuanWorld-Voyager Image-to-Video • Updated Oct 17, 2025 • 135 • 609 moonshotai/Kimi-K2-Instruct-0905 Text Generation • 1T • Updated Jan 30 • 36.8k • • 786 Qwen/Qwen3.8-27B Image-Text-to-Text • 28B • Updated 23 days ago • 6.19M • • 14.1k
Image Qwen/Qwen-Image-Edit Image-to-Image • 20B • Updated Aug 25, 2025 • 130k • • 2.5k sensenova/SenseNova-U1.5-8B-MoT Any-to-Any • 18B • Updated 13 days ago • 7.93k • 216
Voice microsoft/VibeVoice-1.5B Text-to-Speech • 3B • Updated Jan 22 • 253k • 2.48k Running Featured 447 FastVLM WebGPU 🍎 447 Real-time video captioning powered by FastVLM openbmb/VoxCPM-0.5B Text-to-Speech • Updated 19 days ago • 20k • 816 Paused 85 MiMo-Audio-Chat 💬 85 Chat with Xiaomi MiMo-Audio using voice
Papers Group Sequence Policy Optimization Paper • 2507.18071 • Published Jul 24, 2025 • 323 MatrAIx: Simulating the World with 8.3 Billion Persona Agents Paper • 2608.04205 • Published Aug 4 • 53
MatrAIx: Simulating the World with 8.3 Billion Persona Agents Paper • 2608.04205 • Published Aug 4 • 53
Model training merve/smol-vision Image-Text-to-Text • Updated 27 days ago • 195 HiDream-ai/HiDream-E1-1 Any-to-Any • 17B • Updated Jul 17, 2025 • 82 • 217 Jackrong/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled Image-Text-to-Text • 28B • Updated Jul 7 • 4.74k • • 2.94k netflix/void-model Video-to-Video • Updated Apr 6 • 965
Jackrong/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled Image-Text-to-Text • 28B • Updated Jul 7 • 4.74k • • 2.94k