InternReviewer & InternAdvocate: Objective Reward and Evaluation for Agentic Reinforcement Learning in Peer Review and Rebuttal Paper • 2608.28612 • Published Jul 21 • 9
Blind Men and the Elephant: Probing the Epistemic Myopia of LLMs under Long-Tail Divergent Knowledge Paper • 2608.28478 • Published Aug 28 • 20
ContextPilot: Teaching Agents for Proactive Context Management via Fine-grained RL Paper • 2608.28476 • Published Aug 28 • 28
DataFlow-Harness: A Grounded Code-Agent Platform for Constructing Editable LLM Data Pipelines Paper • 2607.16617 • Published Jul 18 • 99
CausalMix: Data Mixture as Causal Inference for Language Model Training Paper • 2607.01104 • Published Jul 1 • 20
Dockerless: Environment-Free Program Verifier for Coding Agents Paper • 2606.28436 • Published Jun 26 • 86
Tmax Collection Data and models associated with "Tmax: A simple recipe for terminal agents". paper: https://arxiv.org/abs/2606.23321 • 23 items • Updated Jun 23 • 20
Tokenizing 3D Molecule Structure with Quantized Spherical Coordinates Paper • 2412.01564 • Published Dec 2, 2024
BioMatrix: Towards a Comprehensive Biological Foundation Model Spanning the Modality Matrix of Sequences, Structures, and Language Paper • 2606.22138 • Published Jun 20 • 26
BioMatrix: Towards a Comprehensive Biological Foundation Model Spanning the Modality Matrix of Sequences, Structures, and Language Paper • 2606.22138 • Published Jun 20 • 26
BioMatrix: Towards a Comprehensive Biological Foundation Model Spanning the Modality Matrix of Sequences, Structures, and Language Paper • 2606.22138 • Published Jun 20 • 26