CohereLabs/cohere-transcribe-03-2026 Automatic Speech Recognition • 2B • Updated Jun 10 • 1.07M • • 1.06k
Running on Zero Agents 29 Fibo-Edit-RMBG Background Removal 🎨 29 Background removal with Fibo-Edit-RMBG
Jackrong/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled Image-Text-to-Text • 28B • Updated 16 days ago • 109k • • 2.92k
Running on Zero Agents Featured 179 ReconViaGen 🖥 179 High-fidelity 3D Geometry Generation from multi-view images
VibeVoice Collection Frontier Text-to-Speech Models https://microsoft.github.io/VibeVoice/ • 8 items • Updated Mar 2 • 251
Spatial-MLLM: Boosting MLLM Capabilities in Visual-based Spatial Intelligence Paper • 2505.23747 • Published May 29, 2025 • 69
Distilling LLM Agent into Small Models with Retrieval and Code Tools Paper • 2505.17612 • Published May 23, 2025 • 82
Runtime error Agents 61 TRELLIS - Multiple Imagen a 3D 🚀 61 Scalable and Versatile 3D Generation from images
docling-project/SmolDocling-256M-preview Image-Text-to-Text • 0.3B • Updated Sep 17, 2025 • 30.3k • 1.62k
view article Article Llama can now see and run on your device - welcome Llama 3.2 +5 merve, philschmid, osanseviero, reach-vb, lewtun, ariG23498, pcuenq • Sep 25, 2024 • 192