Collections
Discover the best community collections!
Collections including paper arxiv:2607.21461
-
Hierarchical Sparse Attention Done Right: Toward Infinite Context Modeling
Paper • 2607.02980 • Published • 83 -
Gemma 4 Technical Report
Paper • 2607.02770 • Published • 75 -
SkillOpt-Lite: Better and Faster Agent Self-evolution via One Line of Vibe
Paper • 2607.03451 • Published • 34 -
TurnOPD: Making On-Policy Distillation Turn-Aware for Efficient Long-Horizon Agent Training
Paper • 2607.05804 • Published • 19
-
Self-Taught Self-Correction for Small Language Models
Paper • 2503.08681 • Published • 16 -
Learning from Failures: Correction-Oriented Policy Optimization with Verifiable Rewards
Paper • 2605.14539 • Published • 8 -
SEED: Self-Evolving On-Policy Distillation for Agentic Reinforcement Learning
Paper • 2607.14777 • Published • 103 -
Self-Improvements in Modern Agentic Systems: A Survey
Paper • 2607.13104 • Published • 31
-
OmniLottie: Generating Vector Animations via Parameterized Lottie Tokens
Paper • 2603.02138 • Published • 151 -
CiteAudit: You Cited It, But Did You Read It? A Benchmark for Verifying Scientific References in the LLM Era
Paper • 2602.23452 • Published • 18 -
DARE: Aligning LLM Agents with the R Statistical Ecosystem via Distribution-Aware Retrieval
Paper • 2603.04743 • Published • 54 -
Multimodal OCR: Parse Anything from Documents
Paper • 2603.13032 • Published • 46
-
Hierarchical Sparse Attention Done Right: Toward Infinite Context Modeling
Paper • 2607.02980 • Published • 83 -
Gemma 4 Technical Report
Paper • 2607.02770 • Published • 75 -
SkillOpt-Lite: Better and Faster Agent Self-evolution via One Line of Vibe
Paper • 2607.03451 • Published • 34 -
TurnOPD: Making On-Policy Distillation Turn-Aware for Efficient Long-Horizon Agent Training
Paper • 2607.05804 • Published • 19
-
OmniLottie: Generating Vector Animations via Parameterized Lottie Tokens
Paper • 2603.02138 • Published • 151 -
CiteAudit: You Cited It, But Did You Read It? A Benchmark for Verifying Scientific References in the LLM Era
Paper • 2602.23452 • Published • 18 -
DARE: Aligning LLM Agents with the R Statistical Ecosystem via Distribution-Aware Retrieval
Paper • 2603.04743 • Published • 54 -
Multimodal OCR: Parse Anything from Documents
Paper • 2603.13032 • Published • 46
-
Self-Taught Self-Correction for Small Language Models
Paper • 2503.08681 • Published • 16 -
Learning from Failures: Correction-Oriented Policy Optimization with Verifiable Rewards
Paper • 2605.14539 • Published • 8 -
SEED: Self-Evolving On-Policy Distillation for Agentic Reinforcement Learning
Paper • 2607.14777 • Published • 103 -
Self-Improvements in Modern Agentic Systems: A Survey
Paper • 2607.13104 • Published • 31