Skip to main content

7 docs tagged with "caching"

View all tags

Bitly — URL Shortener

Distributed ID generation, edge-terminated redirects under 10ms, viral-link cache skew, and analytics that never touch the read path.

Cache eviction — LRU vs W-TinyLFU

Why a burst of one-hit-wonders evicts your most valuable keys under LRU, how Count-Min Sketch admission control fixes it, and when SLRU or 2Q are enough.

ChatGPT — LLM Serving

PagedAttention and KV-cache memory, continuous batching, prefix-aware routing, disaggregated prefill and decode, fairness across tenants, and why VRAM is the real database.

Distributed Cache

Consistent hashing with virtual nodes, stampede prevention, slab allocation, hot keys that sharding cannot fix, and invalidation that is actually correct.

Facebook News Feed

Aggregator-leaf architecture, TAO-style graph caching, read-after-write across datacenters, ad injection, and pagination that stays stable while the world changes underneath it.

HTTP 301 vs 302 for shorteners

The exact trade-off between permanent and temporary redirects, why analytics and revocation force 302, and the hybrid that gets most of both.

Instagram

Hybrid fan-out for celebrities, an async media pipeline, two-stage ranking under 100ms, ephemeral stories, and counters too hot for a database row.