From RLVR to RLSVR: Task Transformation Induces Self-Verifiable Rewards for Open-Ended LLM Self-Improvement Paper • 2607.23802 • Published 20 days ago • 105
Agents-A1 Collection Agents-A1 is a Long-horizon Agentic Model that reaches trillion-parameter-level performance by scaling the agent horizon. • 12 items • Updated 30 days ago • 41
DSpark: Confidence-Scheduled Speculative Decoding with Semi-Autoregressive Generation Paper • 2607.05147 • Published Jul 6 • 43
Laguna XS 2.1 Collection Designed for agentic coding and long-horizon work on a local machine. Licensed under OpenMDW-1.1. • 9 items • Updated Jul 2 • 24