Confidence Comes from Experience: Experiential Confidence Estimation from Reasoning to Agents Paper • 2609.17708 • Published 4 days ago • 54
The Other Half of the Memory Wall: Serving 35B MoEs from SSD with Trained Routing Prediction Paper • 2609.18063 • Published 3 days ago • 10
Flattening Every Memory Peak in Long-Context Mixture-of-Experts Training Paper • 2609.14306 • Published 6 days ago • 10
Agora: Git as Shared Memory for Collective AutoResearch Paper • 2609.18094 • Published 3 days ago • 41
ActionPiece: Rethinking Action Tokenization for Autoregressive Vision-Language-Action Models Paper • 2609.18487 • Published 3 days ago • 39
Rethinking Critic Learning in PPO: Understanding and Mitigating Value Flattening Paper • 2609.18708 • Published 3 days ago • 62
LimiX-2: A Contextual Mechanism Network Towards General Structured-Data Intelligence Paper • 2609.17488 • Published 4 days ago • 184
ScienceIDE: Turning World's Scientific Codebase into Agent Learnable Environments Paper • 2609.19134 • Published 3 days ago • 73
Generalized Agent Iteration: One Formal Framework for Iterative Policy Improvement and Recursive Self-Improvement Paper • 2609.13406 • Published 8 days ago • 36
ModularRSI: Modular and Generalizable Recursive Harness Self-Improvement Paper • 2609.14857 • Published 5 days ago • 160
PhysStream: Streaming Physics-Grounded Video Generation with Structured Scene Memory and Fine-Grained Motion Control Paper • 2609.17521 • Published 4 days ago • 6
The Last AI Built by Humans: Toward Genuine Recursive Self-Improvement Paper • 2609.11873 • Published 9 days ago • 89
ScienceBuddy: Recursive-in-Recursive Self-Improvement for Interactive Scientific Agents Paper • 2609.17523 • Published 4 days ago • 24
Continual Learning Mechanisms Compose for Long-Horizon Memorization Paper • 2609.06986 • Published 12 days ago • 325
Root-Cause Attribution Is a Search Problem: Continual Search for Long-Horizon Agent Failures Paper • 2609.13463 • Published 8 days ago • 3