AgentGrad: Intervention-guided Prompt Optimization for Multi Agent Systems Paper • 2609.08572 • Published 11 days ago • 98
WearableQA: A Benchmark for Health Reasoning over Real-World Wearable Data Paper • 2609.05405 • Published 15 days ago • 40
Reason in the Words You Speak: Idiolectal Paraphrasing Off-Policy Traces for Reasoning Distillation in VideoLLMs Paper • 2608.26684 • Published 23 days ago • 6
Retrieve What's Missing: Coverage-Maximizing Retrieval for Consistent Long Video Generation Paper • 2606.02479 • Published Jun 1 • 24
HarnessEval-W: Agentifying the Evaluation of Visual Worlds Paper • 2608.16859 • Published Aug 17 • 121
Spark-to-Paper: End-to-End Research Paper Generation as a Composable Skill Paper • 2608.11924 • Published Aug 12 • 108
LongHorizon-Harness: Advancing Long-Horizon Agents for Real-World Tasks Paper • 2608.01964 • Published Aug 3 • 186
Qwen-UI-Agent Technical Report: Toward Next-Generation Real-World Centric Foundation GUI Agents Paper • 2607.28227 • Published Jul 30 • 157
ReDesign: Recovering Editable Design Structures from Images via Agentic Decomposition Paper • 2607.25565 • Published Jul 28 • 65
Harness Handbook: Making Evolving Agent Harnesses Readable,Navigable, and Editable Paper • 2607.13285 • Published Jul 14 • 125
MoE-GRPO: Optimizing Mixture-of-Experts via Reinforcement Learning in Vision-Language Models Paper • 2603.24984 • Published Mar 29 • 1
VideoSearch-R1: Iterative Video Retrieval and Reasoning via Soft Query Refinement Paper • 2607.00446 • Published Jul 1 • 24
Memory is Reconstructed, Not Retrieved: Graph Memory for LLM Agents Paper • 2606.06036 • Published Jun 4 • 78
UniSteer: Text-Guided Flow Matching in Activation Space for Versatile LLM Steering Paper • 2605.30076 • Published May 28 • 26
Exploring Autonomous Agentic Data Engineering for Model Specialization Paper • 2605.30407 • Published May 28 • 22
Qwen-VLA: Unifying Vision-Language-Action Modeling across Tasks, Environments, and Robot Embodiments Paper • 2605.30280 • Published May 28 • 145
On the Non-decoupling of Supervised Fine-tuning and Reinforcement Learning in Post-training Paper • 2601.07389 • Published Jan 12 • 3
Graph-Based Chain-of-Thought Pruning for Reducing Redundant Reflections in Reasoning LLMs Paper • 2604.05643 • Published Apr 7 • 13