GitHub is re-architecting its Git infrastructure to support the demands of "agentic software development," where AI agents generate millions of commits and interact with repositories at unprecedented scale. The new design focuses on decoupling durability from scale, minimizing coordination, and separating storage from compute to handle orders of magnitude increase in writes and concurrent operations, while maintaining reliability and existing developer workflows.
Read original on GitHub EngineeringThe advent of "agentic software development" (AI agents performing frequent, automated commits and operations) is pushing Git infrastructure to new limits. GitHub is experiencing exponential growth in Git activity, with pushes increasing 4.9x year-over-year and overall events more than doubling. This shift creates unique challenges for their existing architecture, particularly around write throughput, commit latency, and scaling reads without impacting writes.
GitHub's previous architecture, based on "Spokes" fileservers, stored full repository copies on local disks. While providing low-latency reads and redundancy, it coupled durability with scale. Adding read replicas (for more capacity) also meant adding participants to every write via a three-phase commit protocol. This design choice inherently made writes slower as read capacity increased, creating a ceiling for the busiest repositories where losing a replica reduced read capacity and losing quorum halted writes.
The Durability-Scale Trade-off
In the previous architecture, adding read replicas directly impacted write performance because every replica participated in the three-phase commit protocol. This tightly coupled durability and scalability for writes, leading to a bottleneck at high activity levels.
The redesigned infrastructure aims to separate durability from scale, minimize coordination, and decouple storage from compute, all while preserving existing developer workflows and reliability. The core tenets include:
This new approach has shown internal benchmarks of up to 35 times higher write throughput and independently scaling read capacity, providing a robust foundation for the future of automated software development.