Mage-Flow: An Efficient Native-Resolution Foundation Model for Image Generation and Editing Paper • 2607.19064 • Published 1 day ago • 55
RESOURCE2SKILL: Distilling Executable Agent Skills from Human-Created Multimodal Resources Paper • 2606.29538 • Published 7 days ago • 134
SkillOpt-Lite: Better and Faster Agent Self-evolution via One Line of Vibe Paper • 2607.03451 • Published 20 days ago • 33
World-R1: Reinforcing 3D Constraints for Text-to-Video Generation Paper • 2604.24764 • Published Apr 27 • 119
view article Article NEO-unify: Building Native Multimodal Unified Models End to End sensenova • Mar 5 • 170
RoboPocket: Improve Robot Policies Instantly with Your Phone Paper • 2603.05504 • Published Mar 5 • 35
UniG2U-Bench: Do Unified Models Advance Multimodal Understanding? Paper • 2603.03241 • Published Mar 3 • 88