DiagEvo: Diagnosis-Guided Self-Evolution via Hierarchical Error Memory Paper • 2609.00768 • Published about 1 month ago • 23
SAF-OPD: Stable Advantage Fusion for On-Policy Distillation Paper • 2607.29209 • Published Jul 31 • 34
LongCat-Flash-Prover: Advancing Native Formal Reasoning via Agentic Tool-Integrated Reinforcement Learning Paper • 2603.21065 • Published Mar 22 • 80
InteractComp: Evaluating Search Agents With Ambiguous Queries Paper • 2510.24668 • Published Oct 28, 2025 • 100
ReCode: Unify Plan and Action for Universal Granularity Control Paper • 2510.23564 • Published Oct 27, 2025 • 124