LongHorizon-Harness: Advancing Long-Horizon Agents for Real-World Tasks Paper • 2608.01964 • Published 3 days ago • 144
From RLVR to RLSVR: Task Transformation Induces Self-Verifiable Rewards for Open-Ended LLM Self-Improvement Paper • 2607.23802 • Published 11 days ago • 97
Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Paper • 2607.27919 • Published 7 days ago • 58
Flux-OPD: On-Policy Distillation with Evolving Contexts Paper • 2607.28022 • Published 7 days ago • 43
Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills Paper • 2607.22529 • Published 13 days ago • 48
Harness Handbook: Making Evolving Agent Harnesses Readable,Navigable, and Editable Paper • 2607.13285 • Published 23 days ago • 232
SEED: Self-Evolving On-Policy Distillation for Agentic Reinforcement Learning Paper • 2607.14777 • Published 21 days ago • 104
view article Article NVIDIA brings agents to life with DGX Spark and Reachy Mini +1 jeffboudier, nader-at-nvidia, alecfong • Jan 5 • 67
Toward Generalist Autonomous Research via Hypothesis-Tree Refinement Paper • 2606.11926 • Published Jun 10 • 130
EvoArena: Tracking Memory Evolution for Robust LLM Agents in Dynamic Environments Paper • 2606.13681 • Published Jun 11 • 143
AURA: Intent-Directed Probing for Implicit-Need Surfacing in Situated LLM Agents Paper • 2606.05557 • Published Jun 4 • 1