LongHorizon-Harness: Advancing Long-Horizon Agents for Real-World Tasks Paper • 2608.01964 • Published 26 days ago • 183
From RLVR to RLSVR: Task Transformation Induces Self-Verifiable Rewards for Open-Ended LLM Self-Improvement Paper • 2607.23802 • Published Jul 26 • 106
Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Paper • 2607.27919 • Published about 1 month ago • 59
MemHarness: Memory Is Reconstructed, Not Replayed Paper • 2607.28272 • Published about 1 month ago • 17
Flux-OPD: On-Policy Distillation with Evolving Contexts Paper • 2607.28022 • Published about 1 month ago • 44
Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills Paper • 2607.22529 • Published Jul 24 • 47
Harness Handbook: Making Evolving Agent Harnesses Readable,Navigable, and Editable Paper • 2607.13285 • Published Jul 14 • 235
SEED: Self-Evolving On-Policy Distillation for Agentic Reinforcement Learning Paper • 2607.14777 • Published Jul 16 • 106
nvidia/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-BF16 Any-to-Any • 33B • Updated 4 days ago • 367k • 422
view article Article NVIDIA brings agents to life with DGX Spark and Reachy Mini +1 jeffboudier, nader-at-nvidia, alecfong • Jan 5 • 67
Toward Generalist Autonomous Research via Hypothesis-Tree Refinement Paper • 2606.11926 • Published Jun 10 • 130
EvoArena: Tracking Memory Evolution for Robust LLM Agents in Dynamic Environments Paper • 2606.13681 • Published Jun 11 • 143
AURA: Intent-Directed Probing for Implicit-Need Surfacing in Situated LLM Agents Paper • 2606.05557 • Published Jun 4 • 1