Harnessing CLIP and DINO: An Uncertainty-Aware Cascaded Fusion Network for Generalizable Deepfake Image Detection Paper • 2609.07670 • Published 6 days ago • 17
WorldReward: Reward Modeling for Camera-Conditioned World Models Paper • 2609.03952 • Published 10 days ago • 26
Listening Forward: Next Patch Embedding Prediction Enables Scalable Audio Learners Paper • 2608.19863 • Published 24 days ago • 6
OmniScientist: An Omni-Modal Omni-Discipline AI Scientist Paper • 2608.13558 • Published about 1 month ago • 94
SKILLER: Language-Level Reinforcement Learning for Reusable Skill Extraction in Small Language Models Paper • 2608.10538 • Published Aug 11 • 15
nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-NVFP4 Text Generation • 18B • Updated 2 days ago • 1.28M • 414
CAPEval: A Decoupled Caption Evaluation across Understanding and Generation Paper • 2608.02589 • Published Aug 3 • 25