Bitly — URL Shortener
Distributed ID generation, edge-terminated redirects under 10ms, viral-link cache skew, and analytics that never touch the read path.
Distributed ID generation, edge-terminated redirects under 10ms, viral-link cache skew, and analytics that never touch the read path.
Why a burst of one-hit-wonders evicts your most valuable keys under LRU, how Count-Min Sketch admission control fixes it, and when SLRU or 2Q are enough.
PagedAttention and KV-cache memory, continuous batching, prefix-aware routing, disaggregated prefill and decode, fairness across tenants, and why VRAM is the real database.
Consistent hashing with virtual nodes, stampede prevention, slab allocation, hot keys that sharding cannot fix, and invalidation that is actually correct.
Aggregator-leaf architecture, TAO-style graph caching, read-after-write across datacenters, ad injection, and pagination that stays stable while the world changes underneath it.
The exact trade-off between permanent and temporary redirects, why analytics and revocation force 302, and the hybrid that gets most of both.
Hybrid fan-out for celebrities, an async media pipeline, two-stage ranking under 100ms, ephemeral stories, and counters too hot for a database row.