view article Article FineBooks: are open OCR models good enough to unlock historical knowledge? finebooks • 8 days ago • 24
SenseNova-U1 Collection SenseNova-U1: Unifying Multimodal Understanding and Generation with NEO-Unify Architecture • 12 items • Updated 4 days ago • 75
Multimodal Implementations Collection Comprehensive Demo of Multimodal VLMs on the Hub • 27 items • Updated about 12 hours ago • 14
view article Article We’re open-sourcing our text-to-image model and the process behind it Photoroom • Nov 12, 2025 • 101
view article Article Train AI models with Unsloth and Hugging Face Jobs for FREE +4 burtenshaw, danielhanchen, shimmyshimmer, mlabonne, davanstrien, evalstate • Feb 20 • 107
BitDance Collection BitDance: Open-source autoregressive model with binary visual tokens. A research project for building powerful multimodal autoregressive model. • 10 items • Updated Mar 2 • 11
Alterbute: Editing Intrinsic Attributes of Objects in Images Paper • 2601.10714 • Published Jan 15 • 31
YOLO26 Models Collection YOLO26 models: detection, segmentation, classification, pose, and OBB variants with demos and ONNX variants. • 42 items • Updated Jan 19 • 40
CoreML Collection Models for Apple devices. See https://github.com/FluidInference/FluidAudio for usage details • 16 items • Updated Jun 4 • 7