From RLVR to RLSVR: Task Transformation Induces Self-Verifiable Rewards for Open-Ended LLM Self-Improvement Paper • 2607.23802 • Published 9 days ago • 79
Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Paper • 2607.27919 • Published 5 days ago • 53
Flux-OPD: On-Policy Distillation with Evolving Contexts Paper • 2607.28022 • Published 5 days ago • 41
Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills Paper • 2607.22529 • Published 11 days ago • 47
Harness Handbook: Making Evolving Agent Harnesses Readable,Navigable, and Editable Paper • 2607.13285 • Published 21 days ago • 232
SEED: Self-Evolving On-Policy Distillation for Agentic Reinforcement Learning Paper • 2607.14777 • Published 19 days ago • 103
view article Article NVIDIA brings agents to life with DGX Spark and Reachy Mini +1 jeffboudier, nader-at-nvidia, alecfong • Jan 5 • 67
Toward Generalist Autonomous Research via Hypothesis-Tree Refinement Paper • 2606.11926 • Published Jun 10 • 130
EvoArena: Tracking Memory Evolution for Robust LLM Agents in Dynamic Environments Paper • 2606.13681 • Published Jun 11 • 143
AURA: Intent-Directed Probing for Implicit-Need Surfacing in Situated LLM Agents Paper • 2606.05557 • Published Jun 4 • 1