NeoHorse-1: Towards Recursive Self-Improvement via Agentic Post-Training with Routing Harness Paper • 2609.08183 • Published 2 days ago • 374
Repo-To-Skill: Distilling GitHub Repositories Into AI4AI Skills Paper • 2609.02749 • Published 8 days ago • 538
Agentic Game Development as a Verifiable Trajectory Data Engine for Scaling World Models Paper • 2608.25518 • Published 15 days ago • 196
HarnessEval-W: Agentifying the Evaluation of Visual Worlds Paper • 2608.16859 • Published 24 days ago • 340
EnvHarness: Awakening Static Worlds for Agent Learning Paper • 2608.19880 • Published 21 days ago • 275
trl-internal-testing/tiny-Qwen2ForCausalLM-2.5 Text Generation • 2.44M • Updated 2 days ago • 17.9M • 22
Macaron-V1: Towards Open Continual Learning with Self-Improvement and Mixture-of-LoRA Paper • 2608.09819 • Published about 1 month ago • 342
Qwen-UI-Agent Technical Report: Toward Next-Generation Real-World Centric Foundation GUI Agents Paper • 2607.28227 • Published Jul 30 • 311
Beyond Euclidean Clipping: Overcoming Exploration Collapse in LLM RL via Riemannian Isometric Policy Optimization Paper • 2607.10169 • Published Jul 11 • 14
Parallelized Autoregressive Decoding for Omni-Modal Dense Video Captioning Paper • 2607.02963 • Published Jul 3 • 29