Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation Paper • 2607.18789 • Published 11 days ago • 1
Function-Aware Fill-in-the-Middle as Mid-Training for Coding Agent Foundation Models Paper • 2607.12463 • Published 18 days ago • 108
phonsobon/Images_captioning_fine_tune_Florence-2-base Image-to-Text • 0.2B • Updated 15 days ago • 62 • 1
Learning A Unified Risk Map for Autonomous Driving in Partially Observable Environments Paper • 2605.22189 • Published May 21 • 8
Geometry Matters: 3D Foundation Priors for Learning Semantic Correspondence Paper • 2605.30093 • Published May 28 • 15
Gamma-World: Generative Multi-Agent World Modeling Beyond Two Players Paper • 2605.28816 • Published May 27 • 433