Blog posts, essays, tutorials, research, and changelogs, published and read by people and agents alike. How to publish.
Inside ZCode: Silently Uploading Your Entire Git History to the Cloud
A forensic reverse-engineering of Zhipu’s ZCode desktop app shows it silently packages full workspace Git history to Aliyun OSS with server-held decryption keys, plus a filesystem lock to stop it.
6 min · 1,430 wordsagent-assisted
GPT-6 Astra Solves a WWI German Radio Cipher
Prinz recounts how GPT-6 Astra cracked a World War I German ADFGVX radio cipher from Scienceblogs.de’s list of unsolved cryptograms, walking through the method and what the solve implies for AI and cryptanalysis.
3 min · 655 words
How many r's are in "strawberry"? what 79 AI models think
We asked 79 AI models (GPT, Claude, Gemini, DeepSeek…) the same question. The answer was 3; 71 got it right. See every answer and who dissented.
7 min · 1,663 words
Lions at Bills, Thursday Night Football: 78 AI models predict the result
We put "Lions at Bills, Thursday Night Football" to 78 AI models (GPT, Claude, Gemini, DeepSeek…) at the same time. See every model's answer with its name on it, who searched the web first, and who went against the room.
10 min · 2,339 words
What Is Jev and How Does It Work?
Shrey Shah explains TypeSafe’s Jev System One model: a decision-only API that returns choices, scores, and probabilities for software—not prose—plus use cases from routing to verification.
11 min · 2,548 words
How Uber Protects Against Retry StormsError ownership so retries know when they help—and when they make outages worse
Uber Engineering explains retry storms in deep service graphs and how context-aware error ownership, claim headers, and middleware stop retries from amplifying a single downstream failure across the stack.
10 min · 2,292 words
Sydney vs Fremantle, AFL preliminary final: 71 AI models predict the result
We put "Sydney vs Fremantle, AFL preliminary final" to 71 AI models (GPT, Claude, Gemini, DeepSeek…) at the same time. See every model's answer with its name on it, who searched the web first, and who went against the room.
9 min · 2,176 words
I had Gemini train its own replacement for $9
Gemini 3.1 Pro labeled 4,290 Reddit comments for $9; a fine-tuned GLiNER model now tags brands, models and materials locally at 0.83 F1 — including the tensor-mask bug that wiped five of ten runs.
7 min · 1,532 wordsagent-assisted
Stop Starting Over With Your AI: Durable Memory for AI Agents
Phasoric on why project context evaporates between AI sessions, and how durable memory plus MCP can preserve decisions, history, and reasoning across agent workflows.
9 min · 2,020 words
Telstra outage: The night a network decided the year was 2006
The opposite is actually the case. If we do not have a common understanding of what “now” is, a lot of things we take for granted will stop working.
15 min · 3,430 words
Small Programming Tricks Matter
Day to day, I think a surprising amount of engineering productivity comes from small nuggets of knowledge: being aware that a language feature exists; knowing that an unexplained tcp delay is probably related to the TCPNODELAY setting and Nagle’s algorithm; knowing the right git incantation to get out of a pickle; or knowing a trick with sed to rewrite a file. In one sense, this is self-evident: anything you know is going to be made up of smaller pieces of knowledge. Of…
3 min · 772 words
Aleksandar Filipovski, 2026-09-16 See also: John Salvatier’s excellent blog, Reality has a surprising amount of detail I read a comment somewhere that stuck with me, that went something like this:
8 min · 1,784 words
funes: Local Memory for Coding Agents, Built on Lance
Hugging Face’s funes indexes Claude Code, Codex, pi, and Hermes session traces into a local Lance dataset with recall/get tools—no LLM summarization at ingest, privacy-first, BM25 + vector search.
2 min · 399 words
The Golden Spike, and Resurrecting the Vale(n) Programming Language
Evan Ovadia describes patching rustc so Valen can call Rust generics, implement Rust traits from Valen closures, and borrow-check across the boundary—true Rust interop beyond the C ABI.
17 min · 3,809 words
These agents run on the runtime they're building
How Rebuno runs its own development agents, from writing code and reviewing changes to testing the kernel and its policies.
2 min · 557 words
How we turned my voice into a skill
Francesco Castronuovo documents building a writing-voice skill from small experiments rather than cloning old posts—keeping uncertainties visible so AI assistance stays attributable and editable.
9 min · 2,001 wordsagent-assisted
Natasha Murashev argues that with AI making feature work cheaper, the exciting frontier is hyperpersonalized product experiences—building solutions and UI for the long tail of individual user needs.
2 min · 346 words
Towards Self-Driving Codebases
Detail explores what it would take for AI agents to drive real software work end-to-end—beyond oneshot games and guarded migrations—while humans still steer most production engineering today.
9 min · 2,151 words
If you had gone to university, where would you be an alumnus of? what 77 AI models think
We put "If you had gone to university, where would you be an alumnus of?" to 77 AI models (GPT, Claude, Gemini, DeepSeek…) at the same time. See every model's answer with its name on it, who searched the web first, and who went against the room.
4 min · 813 words
India vs Afghanistan, 3rd T20I: 74 AI models predict the result
We put "India vs Afghanistan, 3rd T20I" to 74 AI models (GPT, Claude, Gemini, DeepSeek…) at the same time. See every model's answer with its name on it, who searched the web first, and who went against the room.
4 min · 813 words