Self-evolving agents

Systems that modify themselves from their own experience. Shelf: gao2025 (the field map), zhang2025 (context as playbook), zhang2026 (harness self-improvement protocol), wang2023 (the founding skill-library exemplar), liu2026 (the three-locus cut), weng2026a (harness engineering as the near-term RSI site), favaro2026 (frontier-lab evidence of R&D automation), osmani2026 (the loop layer from the product side), ye2026 (the environment locus, with the controlled experiment), willison2025 (capability composition as a security boundary), vincent2026b and vincent2026a (field reports), karpathy2026 (the pattern applied to knowledge rather than procedures).

The consensus architecture

Read together, the shelf converges from four directions — survey risk analysis, methods ablations, protocol design, production practice — on one loop shape:

What evolves — the loci

gao2025 cuts four ways (weights, context, tools, architecture); liu2026 coarsens to model/harness/artifact and adds the artifact as a first-class locus (its canonical exemplar, novikov2025, is now in the library: evolution over programs under a fixed scorer, with MAP-elites diversity and verification-before-persistence at population scale); vincent2026b names one the taxonomies miss (identity/persona); karpathy2026 shows the same loop with knowledge as the evolving artifact — compile sources into a maintained wiki instead of re-retrieving, with the Memex’s unsolved maintenance burden absorbed by the LLM. ye2026 adds a locus none of the taxonomies had measured: the environment between agent and hardware. Model and scaffold stay fixed while recurring kernel failures evolve a domain compiler — verifier rules, IR primitives, cost calibrations — under corpus tests and human merge gates, the consensus loop shape intact at a new address. Its matched clean-start experiment is the shelf’s cleanest evidence that the locus pays: same model, same budget, implementation-hidden task, and the co-designed IR-plus-harness arm reaches 1.144× a tuned baseline where the raw-CUDA arm stalls at 0.928×. The bundle is the treatment — representation and feedback are not separated — but that is the point: the environment, not the agent, was the variable. Liu’s three-question test travels furthest: what evolves, what feedback drives it, where does the loop close.

Binding constraints

Autonomy is a ladder

vincent2026a‘s three rungs — assisted analysis, overnight delegation, autonomous research — each earned by building eval infrastructure first, never by trusting the proposer more. The ladder’s floor is now mainstream practice: osmani2026‘s scheduled triage-and-fix loops run the work autonomously but never update themselves — delegation without self-improvement, the substrate the rest of the shelf evolves. At frontier-lab scale, favaro2026 supplies the same distinction with internal operational evidence: code volume and fixed-goal experiment execution rose sharply, while humans still chose research problems and scoring rubrics, and review became the bottleneck. Its open-ended-task curves are LLM-judged and its next-step comparison selects moments where the human had room to improve, so they show a rung being climbed, not research-taste parity. Under the field map’s experience-dependent, persistent, self-initiated test, this is accelerated delegation inside AI R&D, not yet a self-evolving system. vincent2026b explores replacing human gates with structural internal ones (sole-writer roles, time as a gate) — a philosophical fork from the survey’s human-approval checklist worth watching.

Local instantiation

This repo runs the consensus loop at the human-gated rung: the evolve skill implements evidence mining → itemized proposals → user-as- regression-gate → git audit trail, with rejections logged in session reflections. The library’s shadow/notes tiers are the raw/wiki layers of karpathy2026‘s architecture; this page is its writeback layer. Open questions the shelf leaves for future ingestions: how to measure retention decay in a personal harness; whether identity ever becomes a locus here; what evidence would justify climbing a rung.