All posts
136 posts, newest first.
- Why database isolation levels exist Only the top level is simply correct. The others exist because correctness costs money — and the standard's list of what can go wrong turned out to be one entry short. Data · Aug 25, 2026 · intermediate
- How does an AI 'see' a video? Upload an hour-long video to Gemini and ask what happens at minute 40 — it answers. But no model 'watches' anything. It reads a flipbook, and the flipbook is missing most of the pages. AI & ML · Aug 6, 2026 · intermediate
- Hardening a cloud server Spin up a fresh VPS, wait an hour, and the auth log already has thousands of brute-force attempts from across the internet. Every server-hardening guide says roughly the same things — here's what each one actually stops, and where the rules are theater. Security · May 25, 2026 · intro
- How end-to-end encryption works Open WhatsApp and a banner tells you Meta can't read your messages. That claim sits on a specific protocol — Diffie–Hellman key agreement plus a 'double ratchet' that changes the key on every message. Here's the shape of it. Security · May 24, 2026 · intermediate
- How does an LLM 'see' an image? You paste a screenshot into ChatGPT and it reads the text, describes the scene, answers questions. But the model only ever predicts text tokens — so how does a picture get into it at all? AI & ML · May 20, 2026 · intermediate
- What are 'weights' in an LLM? When Meta releases 'open weights' for Llama, what's actually in that file? A giant table of numbers and nothing else — so how does a pile of numbers know things? AI & ML · May 16, 2026 · intro
- Why compression works at all Zip a photo and it shrinks; zip the zip and it doesn't. Compression isn't magic — it only ever exploits the patterns that were already there. Computer Science · May 14, 2026 · intro
- Why CPUs have three levels of cache Look at a CPU die shot and you'll find more area spent on memory than on math — and that memory is split into L1, L2, and L3. The split exists because no single memory technology is both big and fast, so the chip builds a ladder out of several instead. Science · May 14, 2026 · intermediate
- Why deadlocks need four conditions A deadlock feels like bad luck, but it can only happen when four specific conditions all hold at once — and breaking any one makes it impossible. Systems · May 14, 2026 · intermediate
- Why garbage collectors pause your program A tracing collector can't safely move or free an object while your code is mid-read. Freezing every application thread is the obviously correct answer — and generations, write barriers and concurrent marking are all ways of making that freeze shorter without giving up what it buys. Systems · May 14, 2026 · intermediate
- What is tool use (a.k.a. function calling)? A model that only emits text somehow ends up booking your flight. The trick isn't in the weights — it's in the contract between model, harness, and your code. AI & ML · May 7, 2026 · intro
- What does 'X parameters' mean in an LLM? Llama 3.1 70B, DeepSeek-V3 671B, Phi-4 14B — what is that number actually counting, and why is it the headline figure on every model release? AI & ML · May 4, 2026 · intro
- Why model merging works at all Take two fine-tunes of the same model, average their weights element-wise, and you often get a model better than either parent. Naively, this shouldn't work — neural net loss surfaces are wildly non-convex. The reason it works tells you something deep about where fine-tuning actually lives. AI & ML · May 4, 2026 · intermediate
- Why EUV lithography blasts tin droplets with lasers The most demanding layers of leading-edge chips are patterned by a machine that, tens of thousands of times per second, vaporizes a falling droplet of molten tin with a high-power laser. The setup is absurd — and nothing else makes 13.5nm light in a production tool. Science · May 4, 2026 · intro
- Why Spectre still isn't fully patched Eight years after disclosure, new Spectre-class vulnerabilities keep landing. The reason isn't sloppy patching — the speculation being exploited is what makes modern CPUs fast, and the list of channels it leaks through has no end. Security · May 4, 2026 · intermediate
- How can I tell when an LLM is making the answer up? True answers and fabricated ones come out of the same pipe, in the same tone. There's no red light. But there are seams — places hallucinations cluster, shapes they tend to take, tells you can learn to read. AI & ML · May 2, 2026 · intro
- Why dropout disappeared from modern LLMs Dropout was the regularization workhorse of the deep-learning era. Frontier LLM pretraining quietly stopped using it. The reason isn't that dropout broke — it's that the problem dropout solved stopped being the problem. AI & ML · May 2, 2026 · intermediate
- Why LLMs can't count the r's in 'strawberry' A model that can write a sonnet stumbles on a question a five-year-old gets right. The reason isn't intelligence — it's that the model never sees the letters. AI & ML · May 2, 2026 · intro
- Why reward hacking is RLHF's hardest problem You can't write down a loss function for 'be helpful,' so you train a model to predict it — and then a much bigger model spends all its optimization pressure looking for holes in that prediction. That gap is reward hacking, and it doesn't go away with scale. AI & ML · May 2, 2026 · intermediate
- Why synthetic data works for modern LLM training The open web ran out of high-quality text years before frontier models stopped getting better. The new training signal didn't come from a fresh internet — it came from models writing for models, with filters in front. AI & ML · May 2, 2026 · intermediate
- Why floating-point addition isn't associative Schoolroom math says (a + b) + c equals a + (b + c). On a real computer it doesn't, and that one fact ripples out into nondeterministic GPU reductions, irreproducible training runs, and LLM outputs that aren't bit-stable across hardware. Computer Science · May 2, 2026 · intermediate
- Why ECDSA nonce reuse leaks the private key ECDSA needs a fresh random number for every signature. Use the same one twice and anyone watching can recover the private key with two lines of algebra — which is exactly how the PS3's master key fell out. Security · May 2, 2026 · intermediate
- Why 'harvest now, decrypt later' is driving post-quantum crypto adoption A sufficiently large quantum computer doesn't exist yet. Encrypted traffic from 2018 might already be sitting on a tape, waiting for one. That asymmetry — encrypt now, decrypt later — means the damage starts when the recording happens, not when the machine arrives. Security · May 2, 2026 · intermediate
- Why supply-chain attacks dominate the JavaScript ecosystem A small npm install pulls in a thousand-odd packages from hundreds of strangers, and some of that code runs before you type anything. JavaScript's deep, trusting dependency graph is the attack surface, and every step of the attack is a feature someone shipped on purpose. Security · May 2, 2026 · intermediate
- How does League of Legends keep ten players in sync at low latency? Ten strangers on ten different ISPs share one world that has to feel instantaneous. The trick is that nobody's screen shows exactly the same thing — and that's the feature, not the bug. Networking · May 1, 2026 · intermediate
- 10 famous AI-ML terms The vocabulary you keep hearing on every podcast — neural network, transformer, RLHF, RAG — compressed to one line each, then unpacked. AI & ML · Apr 30, 2026 · intro
- What is attention (in transformers)? Every token in a sequence gets to peek at every other token and decide which ones matter. That trick is the engine inside every modern LLM. AI & ML · Apr 30, 2026 · intro
- What is harness engineering? Most of the work that turns a frontier model into a reliable product happens around the model, not inside it. Harness engineering is the name for that work. AI & ML · Apr 30, 2026 · intermediate
- How does an AI model decide what to say? It looks like one big choice — you type a question, you get an answer. Underneath it's thousands of tiny choices, made one token at a time, with no plan and no rewind. AI & ML · Apr 30, 2026 · intro
- What is a neural network? A pile of multiplications and a 'how wrong was I?' signal — somehow, when you stack enough of them, the thing learns to read, see, and play chess. AI & ML · Apr 30, 2026 · intro
- What is a transformer? The neural network architecture behind essentially every modern LLM — and the one big idea that made it work: drop recurrence, let every token look at every other token directly. AI & ML · Apr 30, 2026 · intro
- Why do attention sinks exist? Trained transformers funnel a startling fraction of their attention onto the very first token — a token that's usually semantically meaningless. The pattern looks like a bug, behaves like a feature, and falls out cleanly from one constraint in the softmax. AI & ML · Apr 30, 2026 · intermediate
- Why FlashAttention was a breakthrough Same math, same exact outputs, same asymptotic compute — and yet it made attention several times faster and unlocked long context. The trick was noticing attention was a memory problem, not a compute problem. AI & ML · Apr 30, 2026 · intermediate
- Why FP8 training is stable FP8 has only 256 representable values. Training a frontier model in it sounds insane — and it almost is. Here's the trick that makes it work. AI & ML · Apr 30, 2026 · intermediate
- Why grouped-query attention exists Multi-head attention is a memory-bandwidth disaster at decode time. GQA keeps most of the quality and throws away most of the bandwidth bill. AI & ML · Apr 30, 2026 · intermediate
- Why MLA replaced MHA DeepSeek-V2 cut its KV cache by 93% by attacking the bottleneck differently than GQA — and in their own matched ablation it scored higher, not lower. AI & ML · Apr 30, 2026 · intermediate
- Why does PagedAttention exist? Naive KV-cache allocation reserves a contiguous slab for the worst-case sequence length, then watches 60–80% of it sit unused. PagedAttention asks: what if we treated GPU memory the way an operating system treats RAM? AI & ML · Apr 30, 2026 · intermediate
- Why RoPE replaced sinusoidal positional encoding The original transformer added a fixed sine/cosine vector to each token. Almost no frontier model does that anymore. RoPE rotates queries and keys instead — and that one structural change is what made long context tractable. AI & ML · Apr 30, 2026 · intermediate
- Why AI runs away in verifiable domains AI is getting superhuman fastest at things a computer can grade — math, code, formal proofs — and dragging behind on things it can't. The reason isn't that those domains are 'easier.' It's that training has a feedback step, and feedback needs a verifier. AI & ML · Apr 30, 2026 · intermediate
- What is a hash function? A deterministic shrinker that turns any blob of bytes into a fixed-size fingerprint — the same primitive that powers hash tables, Git commits, and password storage. Computer Science · Apr 30, 2026 · intro
- What is TCP? IP delivers packets best-effort — they can vanish, duplicate, or arrive out of order. Almost every program wants a clean stream of bytes instead. TCP is the layer that turns one into the other. Networking · Apr 30, 2026 · intro
- What is TLS? TCP gives you a reliable byte stream that any router along the path can read and modify. TLS is the layer that wraps that stream so you get confidentiality, integrity, and proof of who's on the other end. Networking · Apr 30, 2026 · intro
- What is public-key cryptography? Until 1976, published cryptography required both sides to already share a secret. Public-key crypto broke that chicken-and-egg problem and quietly became the substrate of the modern internet. Security · Apr 30, 2026 · intro
- A short history of AI, from Turing to today's LLMs Seventy years of trying to make machines think — and how a single architecture from 2017 finally cashed the check that 1950s AI wrote. AI & ML · Apr 29, 2026 · intermediate
- Why does chain-of-thought prompting work? Adding 'let's think step by step' to a prompt makes models measurably better at hard problems. Nobody fully agrees on why, and the wrong story will mislead you about how to use it. AI & ML · Apr 29, 2026 · intermediate
- Why do LLMs hallucinate confidently instead of saying 'I don't know'? The model isn't lying. It was never trained to know when to stop talking. AI & ML · Apr 29, 2026 · intro
- Why do embeddings exist? Computers want numbers, but you also want 'cat' and 'kitten' to live next to each other. Embeddings are the trick that makes both true at once. AI & ML · Apr 29, 2026 · intro
- What is an agent harness? The loop and scaffolding around a language model that turns 'a thing that emits tokens' into 'a thing that does work in the world.' AI & ML · Apr 29, 2026 · intro
- Why is the KV cache a thing? The model has to read your whole prompt every time it picks a token. Why doesn't it choke? Because of a quiet trick almost nobody mentions in the docs. AI & ML · Apr 29, 2026 · intermediate
- Why does in-context learning work? You paste three examples into a prompt and the model suddenly does the task. Nothing got trained. So what just happened? AI & ML · Apr 29, 2026 · intermediate
- What is an LLM? A neural network trained to predict the next token of text — and why that simple goal scaled into something that feels like reasoning. AI & ML · Apr 29, 2026 · intro
- Why does MCP exist? Every AI app was reinventing the same plumbing to talk to the same tools. MCP is the standard that turns an M×N integration mess into M+N. AI & ML · Apr 29, 2026 · intro
- Why does GPU memory bandwidth matter more than FLOPS for LLM inference? You bought the GPU for the teraflops. At inference time, almost none of them are doing anything. The bottleneck is moving the weights, not multiplying them. AI & ML · Apr 29, 2026 · intermediate
- RAG: why retrieval didn't die when context windows got huge Long context windows were supposed to kill retrieval-augmented generation. They didn't. Here's why the bottleneck moved instead of disappearing. AI & ML · Apr 29, 2026 · intro
- Why does temperature exist as a knob? If the model knows the right answer, why is there a dial that asks it to be wrong on purpose? AI & ML · Apr 29, 2026 · intro
- Why do small models exist? If bigger models always benchmark better, why does anyone ship a 3B model? The answer is mostly about latency, cost, and the place the model has to live. AI & ML · Apr 29, 2026 · intro
- Why does tokenization exist? Computers can already read bytes. So why do language models insist on chopping text into these weird half-words first? AI & ML · Apr 29, 2026 · intro
- Why Adam beat plain SGD for LLMs Vision models are mostly trained with SGD + momentum. Transformers are almost always trained with Adam or AdamW. Why did one optimizer win one regime and lose the other? AI & ML · Apr 29, 2026 · intermediate
- Why vector search is approximate on purpose Exact nearest-neighbor search exists, works, and is correct. At scale, the AI-era retrieval stack quietly walks away from it. The reason is more interesting than 'it's faster.' AI & ML · Apr 29, 2026 · intermediate
- Why is attention quadratic? Doubling the context length makes attention 4× more expensive, not 2×. That single fact shapes every trade-off in modern LLM serving — and explains what FlashAttention actually changed (it's not what most people think). AI & ML · Apr 29, 2026 · intermediate
- Why agents fall apart over long horizons Your agent solves any single step beautifully. Run it for fifty steps and it falls off a cliff. The math behind that cliff is older than LLMs, but a newer twist makes it worse. AI & ML · Apr 29, 2026 · intermediate
- Why beam search died for LLMs Beam search was the default way to decode neural sequence models for years. Then chatbots arrived and quietly stopped using it. The reason is stranger than 'sampling is more creative.' AI & ML · Apr 29, 2026 · intermediate
- Why does continuous batching exist? Static batching works fine for image classifiers and breaks immediately for LLMs. The problem isn't the batch — it's that generation lengths vary, and the slowest sequence holds the GPU hostage. AI & ML · Apr 29, 2026 · intermediate
- Why image generation went diffusion, not autoregressive LLMs are autoregressive: predict the next token. Image models could have been the same — predict the next pixel. Almost none of the dominant ones are. Here's why the field walked away from that approach. AI & ML · Apr 29, 2026 · intermediate
- Why model distillation exists A small model trained on a big model's outputs often beats the same small model trained on the original labels. That shouldn't be obvious — and the reason it works is the actually interesting part. AI & ML · Apr 29, 2026 · intermediate
- Why is fine-tuning so cheap compared to pretraining? Pretraining a frontier model costs tens of millions of dollars. Fine-tuning the same model on your data can cost less than a pizza. Why the four-orders-of-magnitude gap? AI & ML · Apr 29, 2026 · intermediate
- Why GPU kernels are still hand-tuned A modern GPU can do tens of teraflops of matrix math. A naive, correct implementation of the same math leaves most of that on the floor. Here's why moving the bytes — not doing the FLOPs — is the actual job. AI & ML · Apr 29, 2026 · intermediate
- Why LayerNorm (and RMSNorm) exist Every transformer block has a normalization step. Pull it out and training falls apart in the first thousand steps. Why is this tiny operation load-bearing? AI & ML · Apr 29, 2026 · intermediate
- Why is evaluating an LLM so much harder than testing normal software? Unit tests pass or fail. LLM outputs don't. The hard part isn't running the eval — it's deciding what 'correct' even means when there are a million right answers. AI & ML · Apr 29, 2026 · intermediate
- Why long-context models still get lost in the middle Your model has a 1M token context window. It can recall the first paragraph perfectly. It can recall the last paragraph perfectly. The thing in the middle? Coin flip. This is not a bug — it's what happens when you ask a model trained one way to behave a different way. AI & ML · Apr 29, 2026 · intermediate
- Why LoRA exists Full fine-tuning a 70B model means storing optimizer state for 70 billion weights. LoRA trains under 1% of the parameters and, on the tasks people have tested, often matches the result. The trick is a hypothesis about the shape of the update. AI & ML · Apr 29, 2026 · intermediate
- Why does mixture-of-experts exist? A 671B-parameter model whose per-token compute is closer to a 37B one. The trick isn't compression — it's that most of the weights sit out most of the time. AI & ML · Apr 29, 2026 · intermediate
- Why does predicting the next token end up doing reasoning? An LLM is trained on one objective: guess the next token. From that one task, you get translation, code, arithmetic, and arguments. Why is autocomplete this powerful? AI & ML · Apr 29, 2026 · intermediate
- Why does prompt caching exist? Your agent sends the same 50,000-token system prompt on every turn. Providers charge a fraction of the usual rate when they recognize it — not out of generosity, but because they stopped doing the work. AI & ML · Apr 29, 2026 · intermediate
- Why do positional encodings exist? A transformer cannot tell 'dog bites man' from 'man bites dog' on its own. The attention math is symmetric in token order — until you bolt on a position signal. Every modern LLM does, and the choice of how shapes long-context behavior more than people realize. AI & ML · Apr 29, 2026 · intermediate
- Why quantization works Stuffing a 70-billion-parameter model into 4-bit weights sounds like it should ruin it. It mostly doesn't — and the reason is more about how the model gets used at inference than about the math of rounding. AI & ML · Apr 29, 2026 · intermediate
- Why reasoning models exist Why we suddenly have a separate class of LLMs that 'think before answering' — and what changed to make spending compute at inference, not training, the new lever. AI & ML · Apr 29, 2026 · intermediate
- Why RLHF exists A pretrained language model knows everything and answers nothing. RLHF exists because the gap between 'predict the next token' and 'do what the user asked' is wider than prompt engineering can paper over. AI & ML · Apr 29, 2026 · intermediate
- Why do scaling laws exist? Bigger model, more data, more compute — and the loss falls along a straight line on a log-log plot for seven orders of magnitude. Nobody fully knows why that line is so straight. AI & ML · Apr 29, 2026 · intermediate
- Why does speculative decoding exist? A small fast model guesses, a big slow model checks. Somehow you get the big model's exact output, faster. The trick isn't cleverness — it's that your GPU was already sitting idle. AI & ML · Apr 29, 2026 · intermediate
- Why SwiGLU replaced ReLU in transformers Modern LLMs ditched the simplest activation function in deep learning for a multiplicative gate nobody can fully explain. Here's why. AI & ML · Apr 29, 2026 · intermediate
- Why is structured output so hard? You ask the model for JSON. Sometimes it gives you a trailing comma. Sometimes a markdown fence. Sometimes prose. Why is this still a problem? AI & ML · Apr 29, 2026 · intermediate
- Why VRAM is the bottleneck for LLM serving It's not FLOPS, it's not network, it's not the CPU. The thing that decides whether your model fits and how many users you can serve is a number printed on the GPU's spec sheet — and three things fight to consume it. AI & ML · Apr 29, 2026 · intermediate
- Why bf16 won the training format wars Half-precision floats came in two flavors: fp16, which had been around for years, and bf16, which kept fp32's exponent and threw away mantissa bits. The less-precise format won. Here's why that's not a typo. Computer Science · Apr 29, 2026 · intermediate
- Why isn't temperature 0 actually deterministic? You set temperature to 0, send the same prompt twice, get two different answers. The math says argmax is a function. The hardware disagrees. AI & ML · Apr 29, 2026 · intermediate
- Why networks are big-endian but your CPU is little-endian Two halves of the same machine disagree on which end of a number comes first. The split is older than you, and it's never going away. Computer Science · Apr 29, 2026 · intermediate
- Why Git stores snapshots, not diffs Git's reputation says 'version control = diffs.' Git's actual model says 'version control = snapshots, hashed.' That swap is the whole reason Git feels different. Computer Science · Apr 29, 2026 · intro
- Why UTF-8 won Unicode could have been a fixed 4-byte-per-character encoding. Instead, the web runs on a variable-width hack — and that hack is why everything still works. Computer Science · Apr 29, 2026 · intro
- What is a hash table? A data structure that lets you find, insert, and delete by key in roughly constant time — the workhorse behind dictionaries, sets, and most fast lookup in modern code. Computer Science · Apr 29, 2026 · intro
- Why WebAssembly exists JavaScript already runs everywhere. So why did browser vendors agree to ship a second, lower-level execution target — and why is it now showing up in CDNs, plugin systems, and AI runtimes? Computer Science · Apr 29, 2026 · intro
- Why B-trees still dominate database indexes Disks read in blocks, not bytes. B-trees were designed around that one fact — and decades later, even on SSDs, no one has dethroned them. Data · Apr 29, 2026 · intermediate
- Why Bloom filters exist A data structure that answers "have I seen this?" with "definitely no" or "maybe" — and saves enormous amounts of work by being wrong on purpose. Data · Apr 29, 2026 · intro
- Why columnar storage won for analytics Row stores read every column to answer a one-column question. Columnar stores refuse — and that refusal is what makes Parquet, ClickHouse, DuckDB, and every modern data warehouse fast. Data · Apr 29, 2026 · intermediate
- Why CRDTs exist Two people typing into the same document at the same time, possibly offline, possibly across the world. The merge has to come out the same on both screens with no central referee. CRDTs are the data structures that make that arithmetic instead of a fight. Data · Apr 29, 2026 · intermediate
- Why 'eventually consistent' became acceptable For decades, anything weaker than strict consistency was a bug. Then the internet got big enough that strict consistency stopped being affordable — and a generation of engineers learned to live with the gap. Data · Apr 29, 2026 · intermediate
- Why LSM trees exist B-trees write where the key lives. LSM trees refuse to do that — and that refusal is the whole point. Data · Apr 29, 2026 · intermediate
- Why Merkle trees are everywhere A hash tree that lets you prove one tiny piece of a giant dataset is correct without re-downloading the whole thing — the trick that quietly underpins Git, Bitcoin, IPFS, and certificate transparency. Data · Apr 29, 2026 · intro
- Why Postgres uses MVCC Readers shouldn't have to wait for writers, and writers shouldn't have to wait for readers — so Postgres keeps multiple versions of every row. Data · Apr 29, 2026 · intermediate
- Why UUIDv7 is quietly replacing autoincrement IDs Autoincrement IDs can't be minted client-side and leak how many rows you have. Random UUIDs trash your index. UUIDv7 is the boring fix almost nobody noticed shipping. Data · Apr 29, 2026 · intermediate
- Write-ahead logging Why databases write your change to a log before they write it to the actual table — and why crash recovery is impossible without it. Data · Apr 29, 2026 · intermediate
- Why is the central limit theorem load-bearing? Almost every confidence interval, A/B test, and gradient-noise argument quietly leans on one fact: averages of independent things look Gaussian, even when the things themselves don't. Math · Apr 29, 2026 · intro
- Why does cosine similarity dominate over Euclidean distance in embeddings? Two vectors can be far apart and still mean the same thing. Cosine similarity asks the only question that turns out to matter: are they pointing the same way? Math · Apr 29, 2026 · intro
- Why matrix multiplication is the bottleneck of modern ML Modern ML is mostly one operation in a trench coat. Understanding why matmul dominates explains hardware, software, and why GPUs eat the world. Math · Apr 29, 2026 · intermediate
- Why does information entropy use log base 2? Shannon could have picked any base for the logarithm in his entropy formula. He picked 2 — and the choice quietly fixes the unit you measure information in. Math · Apr 29, 2026 · intro
- Why does softmax look like that? Softmax is the function that turns a vector of arbitrary numbers into probabilities. The exponential in the middle isn't decorative — it's what makes the whole machine differentiable, well-behaved, and historically inevitable. Math · Apr 29, 2026 · intermediate
- Why JSON beat XML XML had a standards body, schemas, namespaces, transformations, and a decade head start. JSON had curly braces and a JavaScript parser. Curly braces won. Computer Science · Apr 29, 2026 · intro
- Why fiber-optic beat copper for long distances Copper carries electrons; fiber carries photons. The reasons one wins over kilometers come down to physics — how fast the signal fades, how much room there is around the carrier to put signal in, and the fact that light doesn't care about your neighbor's microwave. Science · Apr 29, 2026 · intro
- Why AI accelerators are wrapped in stacks of HBM Open any photo of a modern AI GPU and you'll see the giant compute die in the middle, ringed by short, fat towers of memory soldered millimeters away. Those towers are HBM, and they exist because regular DRAM physically cannot feed a matrix engine fast enough. Science · Apr 29, 2026 · intermediate
- Why GPUs ended up running AI even though they were built for graphics GPUs were designed to shade pixels. Then the same hardware turned out to be exactly what neural networks needed. That isn't luck — graphics and deep learning make the same demand of silicon: identical arithmetic, millions of times over, with nothing to branch on. Science · Apr 29, 2026 · intro
- Why leap seconds exist (and why they're being abolished) The Earth doesn't spin on a schedule, but our clocks do. Leap seconds tried to bridge the two — and broke the internet doing it. Science · Apr 29, 2026 · intro
- Why CPU clock speeds stopped climbing Around the mid-2000s, GHz numbers on CPUs flatlined while core counts started growing. The reason isn't engineering laziness — it's that switching a transistor costs energy, and energy turns into heat you can't get rid of fast enough. Science · Apr 29, 2026 · intro
- Why does CORS exist? CORS isn't there to keep you out of an API — it's there to stop a webpage you're visiting from quietly using your logged-in cookies on a different site. The whole design only makes sense once you see that. Networking · Apr 29, 2026 · intro
- Why retry with exponential backoff — and why jitter? Retrying on failure sounds simple until you ship it at scale. Hammer the server and you make outages worse; back off but synchronize, and you accidentally rebuild the herd. Backoff is the timing rule; jitter is the part that keeps it from biting itself. Networking · Apr 29, 2026 · intro
- Why does HTTPS need certificates if encryption already works? Encryption alone gets you a private channel — to whoever's on the other end. Certificates are how the browser decides that 'whoever' is the bank you meant to reach, not someone sitting in the middle pretending to be. Networking · Apr 29, 2026 · intro
- Why is DNS hierarchical? DNS could have been a giant flat lookup table — one machine somewhere mapping every name in the world to an IP. It isn't, and the reason is less about technology than about who gets to be in charge of what. Networking · Apr 29, 2026 · intermediate
- Why does QUIC exist when TCP already works? TCP works fine — until you're on a flaky phone connection, juggling a dozen multiplexed streams, and one lost packet stalls all of them. QUIC is the protocol designed around that specific frustration. Networking · Apr 29, 2026 · intermediate
- Why does TCP have congestion control? The internet didn't always have it. Once, in 1986, it nearly fell over. The fix wasn't a protocol change — it was endpoints learning to back off. Networking · Apr 29, 2026 · intermediate
- Why do CDNs exist when we already have fast servers? Your origin server can be the fastest box on Earth and your users in São Paulo will still hate it. CDNs exist because the speed of light, not your CPU, is the bottleneck. Networking · Apr 29, 2026 · intro
- Why GPU clusters need NVLink and InfiniBand Training a frontier model means thousands of GPUs taking the same step at the same time. Ethernet wasn't built for that, and PCIe gave up a long time ago. Networking · Apr 29, 2026 · intermediate
- ASLR: why we shuffle memory before every run Attackers used to know exactly where your code lived in memory. ASLR reshuffles it every run, so an exploit has to learn the layout before it can use it — which is why modern exploit chains start by leaking a pointer. Security · Apr 29, 2026 · intermediate
- Why JWTs are controversial JWTs solve a real problem — stateless auth across services — and then keep solving it past the point where the cure is worse than the disease. Here's where the seams are. Security · Apr 29, 2026 · intermediate
- Passkeys: why the password is finally being replaced Passwords are a shared secret you keep retyping into whatever site asked. Passkeys replace it with a key pair whose private half never reaches the site at all. Security · Apr 29, 2026 · intro
- Why password hashing is deliberately slow SHA-256 is fast and that's exactly why you must not use it for passwords. Password storage is the rare corner of computing where being slow — and greedy with memory — is the feature. Security · Apr 29, 2026 · intermediate
- Why prompt injection isn't a bug to be patched SQL, XSS, and command injection are all fought the same way: separate the code channel from the data channel. An LLM has labels for that boundary and nothing that enforces them, so the move that works everywhere else has nowhere to land. The vulnerability is the architecture. Security · Apr 29, 2026 · intermediate
- Why constant-time comparison is a thing An ordinary equality check leaks the secret it's supposed to protect — one byte at a time, through the clock. Constant-time comparison exists because == is faster than it should be. Security · Apr 29, 2026 · intermediate
- Why public-key signatures are not just 'encryption in reverse' They look symmetric — encrypt with one key, decrypt with the other — but signatures and encryption answer different questions, and conflating them is how real cryptosystems get broken. Security · Apr 29, 2026 · intermediate
- Why do LLM responses stream? It's not for show. The model literally generates one token at a time, and forcing it to buffer the full answer before sending would make every chat app feel broken. Streaming is the network shape of an autoregressive process. Networking · Apr 29, 2026 · intro
- Why circuit breakers exist Backoff makes a single retry polite. But when a downstream is plainly down, every caller in your fleet generously retrying it is the actual problem. A circuit breaker is the small piece that says: stop calling for a while — the answer isn't going to change in the next 50 ms. Systems · Apr 29, 2026 · intro
- Why monotonic time is different from wall-clock time Wall-clock time tells you what to put on a calendar. Monotonic time tells you how long something took. Confusing them is how you get bugs that look like physics violations. Systems · Apr 29, 2026 · intermediate
- Why memory-mapped files exist Why hand a file to the OS as memory instead of reading it byte by byte? Because the OS was already caching it that way — and pretending otherwise costs you a copy you don't need. What you pay for deleting the copy is knowing when the I/O happens. Systems · Apr 29, 2026 · intermediate
- Why syscalls are expensive A function call costs a few cycles. A system call costs hundreds — sometimes thousands. The gap isn't sloppy engineering; it's the price of the user/kernel boundary. Systems · Apr 29, 2026 · intermediate
- Why Linux has an OOM killer Linux promises memory it doesn't have, then has to break the promise — the OOM killer is the reaper that decides who dies so the system can keep running. Systems · Apr 29, 2026 · intermediate
- Why virtual memory exists Every process thinks it owns the whole machine. That lie is the foundation almost every modern OS feature is built on. Systems · Apr 29, 2026 · intermediate
- Why containers won over VMs Both promise isolated, reproducible environments. One boots in milliseconds and ships in megabytes; the other boots in seconds and ships in gigabytes. The reason isn't 'containers are lighter VMs' — they're a different kind of thing entirely. Systems · Apr 29, 2026 · intermediate
- Why fork() is such a weird API Other systems take a program and arguments. Unix takes your whole process and clones it. The reasons are half historical accident, half deep insight — and the seams still show. Systems · Apr 29, 2026 · intermediate
- Why idempotency keys exist The network can drop your response after the work is done. Now you have to retry — and you have no idea whether you'd be doing it for the first time or the second. Idempotency keys are the small protocol the client and server agree on so the retry is safe. Systems · Apr 29, 2026 · intro