A note in this repository, Claude/self-improving-agents/README.md, on Anthropic’s post How Warp builds self-improving agents on Claude (from a webinar with Warp’s founder and Anthropic’s Applied AI team) and Warp’s related posts, reviewed on 2026-09-16.1
Sourcing caveat carried from the note: the primary pages were not reachable, so the note was assembled from search-result excerpts and secondary coverage; wording and figures are second-hand.1 Its “Where it breaks down” section is the author’s analysis.
Takeaways
- A scheduled outer skill reads human feedback on an agent’s work and opens a pull request editing the inner skill the agent reads.1 See Self-improving skill loop.
- Improvements arrive as pull requests, so the team, not the agent, accepts them.1 See Agents propose, people and policy accept.
- A few detailed, domain-specific comments beat many thumbs-downs.1