BeaconKV: Key-Value Cache Compression Guided by Beacon Queries for Efficient Large Reasoning Model Inference Paper • 2609.04971 • Published 14 days ago • 40
PaperGym: Rubric-Centered Evolution for Research-Plan Generation Paper • 2608.31119 • Published 18 days ago • 44
FlowBalance: Verifier-Grounded Self-Improvement from On-Policy Reasoning Experience Paper • 2609.03241 • Published 15 days ago • 100
Lucida: Parse, Generate, and Place for Composable Real-to-Sim Scene Modeling Paper • 2608.30821 • Published 18 days ago • 122
SpatialBlock: Enhancing Spatial Intelligence in LVLMs via Synthetic Block-Stacking Problem Paper • 2609.07064 • Published 11 days ago • 141
Bilevel Coordinated Reflection: A Game-Theoretic Approach to Multi-Agent LLM Systems Paper • 2609.02750 • Published 16 days ago • 147
Random Attention: Rethinking KV Cache Eviction for Efficient Reasoning Paper • 2609.03430 • Published 15 days ago • 186
Knowing When Not to Reuse: Conditional Experience Transfer in Autonomous LLM Post-Training Paper • 2608.26730 • Published 22 days ago • 157
Does On-Policy Distillation Really Distill? From Noisy Teacher to Self-Improvement Paper • 2608.31046 • Published 18 days ago • 157
PhysBrain 1.5: From Vision-Language Models to Physical Foundation Models Paper • 2609.14973 • Published 4 days ago • 175
Dr. Claw: An AI Scientist Workspace for Vibe Research Paper • 2609.00365 • Published 18 days ago • 183
Benchmark Radar: A Living Database and Search Engine for AI Benchmarks and Evaluation Paper • 2609.11115 • Published 8 days ago • 207
Dream-RSI: Recursive Self-Improvement through Evolving Worlds Paper • 2609.14858 • Published 4 days ago • 290
HarnessDev: Can LLMs Create and Evolve Their Own Agent Harness? Paper • 2609.01437 • Published 17 days ago • 269
SenseNova-U1.5: Towards Native Unified Visual Intelligence Paper • 2609.11929 • Published 8 days ago • 261
NeoHorse-1: Towards Recursive Self-Improvement via Agentic Post-Training with Routing Harness Paper • 2609.08183 • Published 10 days ago • 420
Atria Dawn: The Dawn of Agentic Superintelligence Paper • 2609.15818 • Published 4 days ago • 433
Scaling Automatic Research Agents via World Models Paper • 2608.12564 • Published 20 days ago • 459