GitFarm: how Uber eliminated local Git clones
GitFarm from Uber eliminates local Git clones in large monorepos and reduces client resources, speeding up checkout and Git workflows.
Architecture and Infra on ThecoreGrid covers the foundations of designing and operating scalable, reliable systems at BigTech level. This category brings together system design and infrastructure practices: distributed architectures, highload patterns, cloud-native platforms, and core layers such as compute, networking, and storage. We focus on real engineering decisions — how to balance reliability, performance, cost, and long-term system evolution. Topics include Infrastructure as Code, Kubernetes, multi-region deployments, traffic management, and platform design. Content is grounded in production experience: incident post-mortems, large-scale migrations, and lessons from operating infrastructure under heavy load. Instead of abstract theory, you get practical trade-offs, proven patterns, and insights drawn from real-world systems. Architecture & Infra is built for architects, backend and platform engineers, DevOps teams, and SREs responsible for complex distributed systems and mission-critical infrastructure.
GitFarm from Uber eliminates local Git clones in large monorepos and reduces client resources, speeding up checkout and Git workflows.
Netflix commerce architecture: how payments, billing, and entitlements evolved from the U.S. DVD model to a global platform with local payment options.
Predictive autoscaling for GPU workloads in Kubernetes reduces the gap between traffic spikes and GPU provisioning. In this case, the system failed because reactive scaling was always late. The system encountered not a bug, but the physics of infrastructure. A critical service crashed under load rather than simply degrading: users experienced 15–20% errors, while Kubernetes … Read more
DNS cache memory in Big Pineapple: how five storage changes cut footprint, improved locality, and reduced lookup latency at Cloudflare scale.
AI architecture for enterprise: 4 layers of the production stack, trade-offs between managed API and self-hosting, integration, serving, and operations.
Agent optimization in Microsoft Foundry begins not with reducing token cost, but with the price of a successful outcome. For an agentic system, this is more important because one result often requires multiple model requests. The main issue here is not the model itself, but that a prototype can easily become the production default. In … Read more
Signature Search in tree networks: probabilistic analysis of five strategies, exact and approximate time estimation, occupancy and synchronization overhead.
MetaRoCE for AI Infrastructure: How Offloading Intelligence to the NIC Helps Ethernet Maintain Throughput, Low Tail Latency, and Graceful Recovery.
Retained bridge share in AWS Organizations: how to preserve AWS Lake Formation permissions when migrating accounts and not lose the control plane.
Dynamic power caps for LLM serving: how POWERSLIDER distributes power across stages, maintains goodput, and withstands grid demand response.
Controls: ← → to move, ↑ to rotate, ↓ to drop.
Mobile: use buttons below.