Select, Compress, Reinvest: A Controlled Study of Visual-Token Allocation in Long-Video MLLMs Paper • 2609.03820 • Published 7 days ago • 16
EnvHarness: Awakening Static Worlds for Agent Learning Paper • 2608.19880 • Published 21 days ago • 275
SemComp-Bench: Benchmarking Semantic Task Completion in Video Generation Paper • 2608.17426 • Published 23 days ago • 159
Can We Defend Against AI-Generated Video Attacks on Real-World Crisis Events? A Systematic Evaluation of Detectors, Generators and Social Dissemination Paper • 2608.14391 • Published 27 days ago • 282
DreamX-Phi 1.0: Action-Conditioned Video World Model for Robotic Manipulation Paper • 2608.13489 • Published 28 days ago • 99
From Inaudible Inputs to Model Failures: Low-Frequency Safety Risks in LALMs Paper • 2608.09158 • Published Aug 10 • 10
OpenART: Scaling Agent Red Teaming via Open-Ended Environment Evolution Paper • 2608.00677 • Published Aug 1 • 263
CalibForge: Adversarial Solver Calibration for Scaling Learnable Terminal Tasks Paper • 2608.06352 • Published Aug 6 • 23
On-Policy Delta Distillation for Multilingual Math Reasoning Paper • 2608.05802 • Published Aug 6 • 32
LuffyTheFox/Qwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V13-GGUF Image-Text-to-Text • 35B • Updated 3 days ago • 843k • 607
SwanTale: Unified Multi-Speaker Speech and Audio Generation for Instruct and Zero-Shot Tasks Paper • 2608.02023 • Published Aug 3 • 158
LeapTalk: Breaking the Latency-Quality Trade-off in Talking Head Generation Paper • 2608.00079 • Published Jul 29 • 18