-
Multi-Agent Computer Use
Paper • 2606.01533 • Published • 6 -
OpenSkill: Open-World Self-Evolution for LLM Agents
Paper • 2606.06741 • Published • 29 -
Socratic-SWE: Self-Evolving Coding Agents via Trace-Derived Agent Skills
Paper • 2606.07412 • Published • 13 -
Bayesian-Agent: Posterior-Guided Skill Evolution for LLM Agent Harnesses
Paper • 2606.08348 • Published • 16
Collections
Discover the best community collections!
Collections including paper arxiv:2609.02749
-
RESOURCE2SKILL: Distilling Executable Agent Skills from Human-Created Multimodal Resources
Paper • 2606.29538 • Published • 146 -
SKT: Skill-Use Training at Scale via Verified Synthetic Data Generation
Paper • 2608.02287 • Published • 31 -
Toward Skill-Native LLMs: Skill Entropy for Benchmarking and Training Long-Horizon Reasoning
Paper • 2608.05139 • Published • 27 -
SkillZip: Evaluation-Free Skill Compression for Self-Evolving Agents by Discovering Reusable Structure
Paper • 2608.11079 • Published • 16
-
AutoResearchClaw: Self-Reinforcing Autonomous Research with Human-AI Collaboration
Paper • 2605.20025 • Published • 89 -
OpenComputer: Verifiable Software Worlds for Computer-Use Agents
Paper • 2605.19769 • Published • 67 -
WildClawBench: A Benchmark for Real-World, Long-Horizon Agent Evaluation
Paper • 2605.10912 • Published • 36 -
EvolveMem:Self-Evolving Memory Architecture via AutoResearch for LLM Agents
Paper • 2605.13941 • Published • 15
-
LEGO-Anything: Coding Agents for 3D Scene Reconstruction
Paper • 2609.36380 • Published • 107 -
LimiX-2: A Contextual Mechanism Network Towards General Structured-Data Intelligence
Paper • 2609.17488 • Published • 811 -
Repo-To-Skill: Distilling GitHub Repositories Into AI4AI Skills
Paper • 2609.02749 • Published • 406
-
VibeVoice Technical Report
Paper • 2508.19205 • Published • 180 -
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Paper • 2509.22186 • Published • 179 -
SenseNova-U1: Unifying Multimodal Understanding and Generation with NEO-unify Architecture
Paper • 2605.12500 • Published • 198 -
BDH-CQ: In-Context Learning with Recurrent Latent Reasoning
Paper • 2608.09888 • Published • 795
-
MOPD: Multi-Teacher On-Policy Distillation for Capability Integration in LLM Post-Training
Paper • 2606.30406 • Published • 26 -
TurnOPD: Making On-Policy Distillation Turn-Aware for Efficient Long-Horizon Agent Training
Paper • 2607.05804 • Published • 20 -
Trust Region Policy Distillation
Paper • 2607.04751 • Published • 33 -
SEED: Self-Evolving On-Policy Distillation for Agentic Reinforcement Learning
Paper • 2607.14777 • Published • 104
-
Warp-as-History: Generalizable Camera-Controlled Video Generation from One Training Video
Paper • 2605.15182 • Published • 40 -
STALE: Can LLM Agents Know When Their Memories Are No Longer Valid?
Paper • 2605.06527 • Published • 47 -
Learning to Build the Environment: Self-Evolving Reasoning RL via Verifiable Environment Synthesis
Paper • 2605.14392 • Published • 9 -
World Action Models: The Next Frontier in Embodied AI
Paper • 2605.12090 • Published • 73
-
Multi-Agent Computer Use
Paper • 2606.01533 • Published • 6 -
OpenSkill: Open-World Self-Evolution for LLM Agents
Paper • 2606.06741 • Published • 29 -
Socratic-SWE: Self-Evolving Coding Agents via Trace-Derived Agent Skills
Paper • 2606.07412 • Published • 13 -
Bayesian-Agent: Posterior-Guided Skill Evolution for LLM Agent Harnesses
Paper • 2606.08348 • Published • 16
-
LEGO-Anything: Coding Agents for 3D Scene Reconstruction
Paper • 2609.36380 • Published • 107 -
LimiX-2: A Contextual Mechanism Network Towards General Structured-Data Intelligence
Paper • 2609.17488 • Published • 811 -
Repo-To-Skill: Distilling GitHub Repositories Into AI4AI Skills
Paper • 2609.02749 • Published • 406
-
VibeVoice Technical Report
Paper • 2508.19205 • Published • 180 -
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Paper • 2509.22186 • Published • 179 -
SenseNova-U1: Unifying Multimodal Understanding and Generation with NEO-unify Architecture
Paper • 2605.12500 • Published • 198 -
BDH-CQ: In-Context Learning with Recurrent Latent Reasoning
Paper • 2608.09888 • Published • 795
-
RESOURCE2SKILL: Distilling Executable Agent Skills from Human-Created Multimodal Resources
Paper • 2606.29538 • Published • 146 -
SKT: Skill-Use Training at Scale via Verified Synthetic Data Generation
Paper • 2608.02287 • Published • 31 -
Toward Skill-Native LLMs: Skill Entropy for Benchmarking and Training Long-Horizon Reasoning
Paper • 2608.05139 • Published • 27 -
SkillZip: Evaluation-Free Skill Compression for Self-Evolving Agents by Discovering Reusable Structure
Paper • 2608.11079 • Published • 16
-
MOPD: Multi-Teacher On-Policy Distillation for Capability Integration in LLM Post-Training
Paper • 2606.30406 • Published • 26 -
TurnOPD: Making On-Policy Distillation Turn-Aware for Efficient Long-Horizon Agent Training
Paper • 2607.05804 • Published • 20 -
Trust Region Policy Distillation
Paper • 2607.04751 • Published • 33 -
SEED: Self-Evolving On-Policy Distillation for Agentic Reinforcement Learning
Paper • 2607.14777 • Published • 104
-
AutoResearchClaw: Self-Reinforcing Autonomous Research with Human-AI Collaboration
Paper • 2605.20025 • Published • 89 -
OpenComputer: Verifiable Software Worlds for Computer-Use Agents
Paper • 2605.19769 • Published • 67 -
WildClawBench: A Benchmark for Real-World, Long-Horizon Agent Evaluation
Paper • 2605.10912 • Published • 36 -
EvolveMem:Self-Evolving Memory Architecture via AutoResearch for LLM Agents
Paper • 2605.13941 • Published • 15
-
Warp-as-History: Generalizable Camera-Controlled Video Generation from One Training Video
Paper • 2605.15182 • Published • 40 -
STALE: Can LLM Agents Know When Their Memories Are No Longer Valid?
Paper • 2605.06527 • Published • 47 -
Learning to Build the Environment: Self-Evolving Reasoning RL via Verifiable Environment Synthesis
Paper • 2605.14392 • Published • 9 -
World Action Models: The Next Frontier in Embodied AI
Paper • 2605.12090 • Published • 73