Echoverse: Deep, Evolving Environments for Training Computer-Use Agents at Scale Paper • 2607.28074 • Published 4 days ago • 9
Filesystem-Based Memory for LLM Agents: Organization, Evolution, and Sustainability Paper • 2607.26637 • Published 5 days ago • 8
Σ-Mem: An Online Reliability Memory for LLM-based Multi-Agent Systems Paper • 2607.27958 • Published 4 days ago • 12
Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Paper • 2607.27919 • Published 4 days ago • 49
Frontis-MA1: Training an AI4AI Model towards Recursive Self-Improvement in Machine Learning Engineering Paper • 2607.28568 • Published 4 days ago • 168
Can AI agents conduct open-ended AI research? Early evidence from two case studies Paper • 2607.27191 • Published 5 days ago • 16
CoRT: Counterfactual Replay for Token-Level Rubric-Guided Policy Optimization Paper • 2607.25659 • Published 6 days ago • 80
CAST: Game Solvers as Turn-Level Teachers for LLM Agents Paper • 2607.25308 • Published 6 days ago • 39
SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Paper • 2607.26784 • Published 5 days ago • 25
DecoEvo: Score-Decoupled Co-Evolution of Solver and Rubric-Generator Skills in Text Space Paper • 2607.25675 • Published 6 days ago • 61
Building to the Test: Coding Agents Deliver What You Check, Not What You Requested Paper • 2606.28430 • Published Jun 26 • 9
GRPO, Dr. GRPO, and DAPO Are Three Operations on One Number: The Group-Standard-Deviation Identity Paper • 2607.00152 • Published Jun 30 • 9
Are Performance-Optimization Benchmarks Reliably Measuring Coding Agents? Paper • 2607.01211 • Published Jul 1 • 9
Pass the Baton: Trajectory-Relayed On-Policy Distillation Paper • 2607.26057 • Published 6 days ago • 31
Keep It InMind: Benchmarking the Implicit-Association Blind Spot in Agent Memory Paper • 2607.24368 • Published 7 days ago • 31
CodeNib: A Multi-View Data System for Serving Repository Context to Coding Agents Paper • 2607.25431 • Published 6 days ago • 103