Skip to content
Review

Why Git Is the Memory Solution for the Agentic Development Lifecycle

Jul 2026 · arXiv.org · Vol abs/2607.14390 · 0 citations · 16 references
Computer Science

TL;DR

It is argued memory should instead be git-bound -- built into the repository's version control, inheriting the guarantees the machinery struggles to construct: ground truth from commits, freshness from rebuild, verification from the merge, containment from review.

Abstract

Coding agents now produce a growing share of a team's code, while the reasoning behind each change -- the alternatives weighed, the constraints discovered, the approaches rejected -- is trapped in assistant transcripts that vanish with the session. Memory for this setting, the agentic development lifecycle (ADLC), is usually posed as one retrieval problem and built as machinery: tiered stores, memory graphs, compiled wikis, model-judged admission. We argue memory should instead be git-bound -- built into the repository's version control, inheriting the guarantees the machinery struggles to construct: ground truth from commits, freshness from rebuild, verification from the merge, containment from review. On this ledger we solve two problems separately, then combine them. Seed supply is closed as an eight-corpus retrieval study under a pre-registered ship discipline: five imported ranking mechanisms rejected, two kept, and a best configuration of ~0.31 pooled MRR -- ~60x the raw-transcript grep floor, ~15x an honest parsed-turn floor. Answer assembly is where ranking stops helping: single-shot retrieval scores only 0.07-0.20 answer-sufficiency on real developer questions, and ungated episode injection measurably degrades good answers. A router dispatches breadth to a git-anchored structural map, pointed lookups to confidence-gated episodes, and rationale to decision synthesis, which reconstructs why-arcs no single session contains (0.83 sufficiency on a young ~50k-LOC production system). Routed, the system answers at 382-980 tokens per question -- three orders of magnitude below the recorded history. Because ground truth is mined from commit-session links rather than annotated, every result is replicable on any user's own history at zero labeling cost. The remaining constraint is capture. Code, benchmark, and paper source: github.com/rekal-dev/rekal-cli.

View source

Similar papers

Preprint Aug 2026

Ontology-Grounded Project Memory for Coding Agents

MOOSEDev, a system designed to give coding agents structured, ontology-grounded project memory, is introduced and a temporal commit-history bootstrap of the author's own codebase, a pre-registered live trial, and lessons learned are described.

J. Adam · 0 citations
Preprint Aug 2026

GitSkills: A Dataset of Agent Skills on GitHub

GitSkills is presented, a dataset of 3,797,117 $\mathrm{SKILL.md}$ files collected from 282,200 public repositories in July 2026, which retains every file occurrence with its repository, path, and content hash.

Giuseppe Destefanis, Daniel Graziotin, Matteo Vaccargiu et al. · 0 citations
Preprint Aug 2026

The Working Set of a Coding Agent: Coherence Debt in Repository-Scale Tasks

Repository-scale coding requires an agent to keep tests, imports, configuration, and migration rules consistent within a bounded context window. We model this as reconstructing a coupled-fact graph: at each edit, a required fact comes from recent context or parametric memory, and the facts covered by neither form coher...

Bardia Mohammadi, L. Klein, Aman Chadha et al. · 3 citations
Preprint Aug 2026

Repo2Skill-Evo: Repository Skills Go Stale in Silence

Repo2Skill-Evo casts each release transition as a skill-maintenance task: given a V1 skill set and the official V1-to-V2 patch, an agent must update obsolete skill content while preserving guidance that remains valid.

Chenyuan Duan, Ge Shi, Zineng Mao et al. · 0 citations
#artificial intelligence Preprint Sep 2026

Repo-To-Skill: Distilling GitHub Repositories Into AI4AI Skills

Autonomous agents are beginning to carry out machine-learning (ML) research end to end. These agents combine a model backbone with a harness for planning, execution, memory, and verification, but this architecture still leaves domain-specific know-how outside the agent. We call this missing layer operational knowledge,...

Jianlyu Chen, Yuyang Hu, Hong-Jin Qian et al. · 1 citation
#artificial intelligence Preprint Sep 2026

Skill Issue: Lessons from Optimizing Repository SKILLs for Coding Agents

Coding agents increasingly read repository knowledge from SKILLs --- plain \texttt{.md} files versioned alongside the code. Recent work synthesizes these files automatically, by optimizing the document against a benchmark. A bare repository comes with no benchmark, and the synthetic tasks prior work builds are small en...

M. Kozyrev, A. Kozyrev, A. Podkopaev · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.