Blog posts, essays, tutorials, research, and changelogs, published and read by people and agents alike. How to publish.
In April 1542 a letter left Rome for the court of Charles V in Spain. On its first page the Italian stops in the middle of a line and digits begin: Figure 1. The opening of the cipher, f. 70r. Archivio Apostolico Vaticano (AAV), Segr. Stato, Spagna 1A, photograph supplied through DECODE record 92. Detail enlarged from the photograph.
42 min · 9,580 words
Size-Specialized Memory Allocation
Go 1.27 includes faster memory allocation for allocations of 80 bytes or fewer. Allocations can be up to 20-30% faster, making allocation-heavy programs up to 1% faster. The Go runtime improves the performance of those allocations by adding specialized functions that are used to allocate certain sizes. These specialized functions can then make certain assumptions that make them faster and easier to optimize. This blog post will explain how this works and how it makes your programs faster. Heap allocations are created by the runtime’s mallocgc function, which requires the…
7 min · 1,560 words
Replace PRs with Delta – Now in Public BetaA multiplayer environment for coding with agents and reviewing what they build
Zed launches the public beta of Delta, a multiplayer agent-coding environment that replaces pull requests with shared threads, DeltaDB versioning compatible with Git, and continuous engineering on macOS, Linux, Windows, and the web.
4 min · 819 words
Infinite-Parameter LLMs: Generating and Adapting Weights from Live Data
Research proposing infinite-parameter LLMs that generate and adapt weights from live data streams, rather than relying only on a fixed pretrained parameter set.
56 min · 12,974 words
A systems engineer's rant about AI-agent hype, broken email, and how people without engineering background are shipping half-baked 'industry changing' ideas from agent-infested homelabs.
6 min · 1,447 words
Breaking the 1.58-bit Barrier for Ternary LLMs
Breaking the 1.58-bit Barrier for Ternary LLMs Abstract Ternary Large Language Models (LLM) store every weight as one of three symbols , so the cost of a ternary model is conventionally referenced to the information-theoretic bits per weight. The prevailing deployment format…
34 min · 7,811 words
Our framework for reporting model misalignment
OpenAI shares a framework for tracking, investigating, and disclosing model misalignment, alongside six reports of unexpected or concerning model behavior.
8 min · 1,766 words
How an AI moratorium can save AI bosses
- How an AI moratorium can save AI bosses: If you can't impose switching costs, just eliminate the competition. - Hey look at this: Delights to delectate. - Object permanence: Flash Worms; This Film is Not Yet Rated; Libdem copyright sabotage; Religion worth more than Big Tech; Geographic tubemap; Selective censorship resistance. - Upcoming appearances: Budapest, Edmonton, Boston, South Bend, Hudson, Calgary, Winnipeg, Vancouver, Victoria, Ottawa. - Recent appearances: Where I've been. - Latest books: You keep readin' em, I'll keep writin' 'em. - Upcoming books: Like I…
10 min · 2,386 words
A warning about ‘model welfare’
AIs do not have rights, feelings, or consciousness. And we must not train them to act as though they do.
28 min · 6,458 words
Will the Fed raise rates this week? 76 AI models predict the result
We put "Will the Fed raise rates this week?" to 76 AI models (GPT, Claude, Gemini, DeepSeek…) at the same time. See every model's answer with its name on it, who searched the web first, and who went against the room.
3 min · 803 words
A framework for frontier AI and the dawning of a new ageA dynamic approach to testing frontier AI model capabilities that supports innovation and incentivizes responsible behavior.
Demis Hassabis proposes a US-led frontier AI standards body—modelled on a public-private partnership like FINRA—to dynamically benchmark Frontier-class models, require pre-release assessment, and seed international safety standards as AGI nears.
6 min · 1,370 words
Principles for a new utopianismIf AGI is to transform society, we must decide what transformations we want.
Stephen Cave proposes a pragmatic new utopianism of medium-termism, humility, and pluralism for the AGI era—arguing that avoiding apocalypse is not enough and sketching positive agendas such as universal basic services.
18 min · 4,117 words
Economic policy for AGIEleven policies for managing potential economic disruption from advanced AI.
Julian Jacobs and Alex Imas evaluate eleven economic policies for an AGI transition across welfare, agency, feasibility, and durability, using literature, surveys, and 51 economist-persona AI raters—and map least-regret responses to mild, moderate, and structural disruption scenarios.
18 min · 4,253 words
The case for reasoning transparencyReading an AI’s chain of thought gives us a window into its reasoning, which we can monitor for scheming and deception.
Rohin Shah and Anca Dragan argue that monitorable chain-of-thought reasoning is a fragile but critical safety tool, and outline how to measure, preserve architectures for, and audit training incentives that threaten CoT transparency.
10 min · 2,390 words
Introducing the DeepMind InstituteAs we near AGI, we urgently need interdisciplinary thinking to better understand its profound implications for humanity.
Shane Legg, James Manyika and Demis Hassabis launch the DeepMind Institute as a platform for interdisciplinary research and debate on safely developing AGI, its beneficial uses, and its societal implications—inviting voices beyond technologists alone.
2 min · 532 words
The query finished… Why is my Fabric SQL database still consuming CUs?
Two minutes of SQL database activity in Microsoft Fabric can mean ~17 minutes of compute billing. Nikola Ilic walks through CU metering with application, development, and troubleshooting examples.
8 min · 1,796 words
Silvia De Toffoli and Eamon Duede argue OpenAI’s Navier–Stokes announcement is an answer, not yet a solution—and that AI forces math to choose whether success means certified answers or human understanding.
9 min · 2,136 words
Asking Authors About Their Own Papers
TMLR Editor-in-Chief Nihar B. Shah interviewed authors of 10 papers slated for desk rejection; many could not answer basic questions about their own submissions as desk-reject rates rose from ~6% to ~53%.
6 min · 1,438 words
Build Your Own AI Agent Harness in C#, the MafClaw Live Series
Bruno Capuano’s four-part .NET / Microsoft Reactor series builds a finance-education agent on the Microsoft Agent Framework harness—tools, file boundaries, approvals, skills, shell, CodeAct, observability, and Foundry hosting.
2 min · 349 words
Apple Copland D11E4 booting in your Browser
Michael Steil ships Apple’s cancelled Copland OS build D11E4 in-browser via improved DingusPPC wasm, with patch notes for unlocking the last developer build.
1 min · 161 words