UI-Mate: Advancing Open-Weight Foundation GUI Agents with In-Context Demonstrations Paper • 2608.15930 • Published 2 days ago • 21
Recursive Synthesis for Long-Horizon Terminal Tasks Collection CC BY 4.0 datasets and SFT/RL checkpoints for recursive task synthesis; base-model and third-party terms still apply. • 6 items • Updated 7 days ago • 18
Stale but Stable: Staleness-Adaptive Trust Regions for Stabilizing Asynchronous Reinforcement Learning Paper • 2607.18722 • Published 28 days ago • 36
Harness Handbook: Making Evolving Agent Harnesses Readable,Navigable, and Editable Paper • 2607.13285 • Published Jul 14 • 233
Long-Horizon-Terminal-Bench: Testing the Limits of Agents on Long-Horizon Terminal Tasks with Dense Reward-Based Grading Paper • 2607.08964 • Published Jul 9 • 77
Graph of Skills: Dependency-Aware Structural Retrieval for Massive Agent Skills Paper • 2604.05333 • Published Apr 7 • 24