Characterizing Warp Divergence from Pascal to Blackwell Paper • 2607.23402 • Published 16 days ago • 9
MOPD: Multi-Teacher On-Policy Distillation for Capability Integration in LLM Post-Training Paper • 2606.30406 • Published Jun 29 • 23
Qwen-RobotManip Technical Report: Alignment Unlocks Scale for Robotic Manipulation Foundation Models Paper • 2606.17846 • Published Jun 17 • 31
Qwen-RobotNav Technical Report: A Scalable Navigation Model Designed for an Agentic Navigation System Paper • 2606.18112 • Published Jun 18 • 29
The Verification Horizon: No Silver Bullet for Coding Agent Rewards Paper • 2606.26300 • Published Jun 24 • 53
Qwen-Image-Agent: Bridging the Context Gap in Real-World Image Generation Paper • 2606.26907 • Published Jun 25 • 50
Deeper is Not Always Better: Mitigating the Alignment Tax via Confident Layer Decoding Paper • 2606.21906 • Published Jun 20 • 26
Self-Improvements in Modern Agentic Systems: A Survey Paper • 2607.13104 • Published 28 days ago • 33
Harness Handbook: Making Evolving Agent Harnesses Readable,Navigable, and Editable Paper • 2607.13285 • Published 28 days ago • 232
The Harness Effect: How Orchestration Design Sets the Token Economics of Enterprise Agentic AI Paper • 2607.06906 • Published Jul 8 • 2
Running 60 Don't Train the Model, Evolve the Harness 🌿 60 Evolving an agent's harness, not its model, on Harvey's LAB
ClawsBench: Evaluating Capability and Safety of LLM Productivity Agents in Simulated Workspaces Paper • 2604.05172 • Published Apr 6 • 25
SkillsBench: Benchmarking How Well Agent Skills Work Across Diverse Tasks Paper • 2602.12670 • Published Feb 13 • 65