view article Article Introducing @huggingface/kernels: 200+ WebGPU Kernels for Local AI nico-martin, Xenova • 9 days ago • 59
view article Article Extremely Fast and Accurate Transcription with Granite Speech 5.0 Turbo CTC ibm-granite • 16 days ago • 33
UI-Mate Collection Open-weight CUA models and office-centric CUA benchmark • 4 items • Updated 2 days ago • 21
view article Article Build Low-Latency Multilingual Voice Agents: Open Weights & Full Deployment Control with NVIDIA Magpie TTS nvidia • about 1 month ago • 36
view article Article State of Open Models: Summer 2026 Observations +1 AdinaY, multimodalart, irenesolaiman • 27 days ago • 190
LTX-2.5 Collection LTX-2.5 base models, quantized models and accompanying LoRAs and IC-LoRAs • 5 items • Updated about 13 hours ago • 57
NVIDIA Nemotron v3 Collection Open, Production-ready Enterprise Models • 33 items • Updated 26 days ago • 371
Muse Glimmer Collection Muse Glimmer 30B: multimodal agentic model for local deployment. BF16 weights, GGUF k-quants, ExecuTorch builds, DFlash drafter. • 4 items • Updated about 1 month ago • 107
MerchantBench: Benchmarking LLM Agents for Long-Term Coherence in E-Commerce Operations Paper • 2607.28956 • Published Jul 31 • 112