JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Paper • 2607.23588 • Published 7 days ago • 122
From Player to Master: Enhancing Test-Time Learning of LLM Agents via Reinforcement Learning over Memory Paper • 2606.08656 • Published Jun 7 • 3
VideoSeeker: Incentivizing Instance-level Video Understanding via Native Agentic Tool Invocation Paper • 2605.16079 • Published May 15 • 29
Towards Efficient Large Language Reasoning Models via Extreme-Ratio Chain-of-Thought Compression Paper • 2602.08324 • Published May 15
JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Paper • 2607.23588 • Published 7 days ago • 122
VimRAG: Navigating Massive Visual Context in Retrieval-Augmented Generation via Multimodal Memory Graph Paper • 2602.12735 • Published Feb 13 • 8
GroupGPT: A Token-efficient and Privacy-preserving Agentic Framework for Multi-User Chat Assistant Paper • 2603.01059 • Published Mar 1 • 1
Unify-Agent: A Unified Multimodal Agent for World-Grounded Image Synthesis Paper • 2603.29620 • Published Mar 31 • 49
SkillFlow:Benchmarking Lifelong Skill Discovery and Evolution for Autonomous Agents Paper • 2604.17308 • Published Apr 19 • 23
OpenSearch-VL: An Open Recipe for Frontier Multimodal Search Agents Paper • 2605.05185 • Published May 6 • 106
SCOPE: Structured Decomposition and Conditional Skill Orchestration for Complex Image Generation Paper • 2605.08043 • Published May 8 • 10