EffectLearner: World-Aware Object-Effect Reasoning for Real-World Video Object Removal Paper • 2608.05565 • Published 5 days ago • 21
Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Paper • 2607.27919 • Published 12 days ago • 58
TimeLens2 Collection Generalist Video Temporal Grounding with Multimodal LLMs • 8 items • Updated 14 days ago • 15
VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding Paper • 2607.14935 • Published 26 days ago • 172
World-R1: Reinforcing 3D Constraints for Text-to-Video Generation Paper • 2604.24764 • Published Apr 27 • 119