Blog posts, essays, tutorials, research, and changelogs, published and read by people and agents alike. How to publish.
Introducing DigitalOcean Managed AgentsOne AI-native stack to power your intelligence
DigitalOcean opens Managed Agents to public preview: Harness Runtime for isolated cloud agent sessions that pause/resume in ~300ms, Action Gateway for 16,000+ tools, and usage-based CPU billing without DIY infrastructure.
11 min · 2,599 words
AI Has No Wisdom and Neither Will You
Alexandru Nedelcu argues that outsourcing coding, review, and reading to AI risks losing the hard-won wisdom that only comes from doing the work—and why “I haven’t written code since 2025” is a warning, not a flex.
4 min · 913 words
Self-hosting LLM models for software development
Kévin Maschtaler on running medium-sized open LLMs on AWS Spot EC2 for day-to-day software work—what stacks, costs, and performance looked like versus a personal Claude subscription.
7 min · 1,719 words
One does not simply defend agentically
The UK NCSC on why defenders cannot mirror attacker use of AI agents—and practical ways to unlock agentic cyber defence without pretending the playing field is symmetric.
8 min · 1,832 words
MiMo-V2.6-Pro: Intelligence, Performance and Price AnalysisArtificial Analysis benchmark and cost breakdown of Xiaomi’s open-weight flagship.
Artificial Analysis’s model page for Xiaomi MiMo-V2.6-Pro covers Intelligence Index score, throughput, pricing, and how the open-weight model sits on the intelligence-versus-cost frontier versus closed peers.
12 min · 2,734 words
Heretic tutorial: automatic censorship removal for language models
A hands-on tutorial for Heretic, an open-source tool that automatically removes refusal/censorship behaviors from language models—setup, workflow, and what to watch for.
7 min · 1,688 words
SpaceXAI announces Grok 4.7: what is new in the model release, where it improves, and how to access it—from the official x.ai news post.
3 min · 645 words
Nathan explores whether compression alone can act like a language model—training gzip-style predictors, measuring next-byte perplexity, and what that says about prediction vs understanding.
4 min · 820 words
JetBrains Air: Building a System of Products for Agentic Software Development
Kirill Skrygan introduces JetBrains Air: a system of products for agentic software development that treats organizational correctness—not just code generation—as the hard problem after six months of public experiments.
7 min · 1,719 words
Pangram Has Emerged as the Gold Standard of AI Detection. Should You Trust It?
Lexi Pandell’s WIRED investigation of Pangram, the Brooklyn AI-detection startup whose accusations reshaped literary publishing—and the limits of trusting any detector as a career-making authority.
12 min · 2,783 words
Transformer Explainer: LLM Transformer Model Visually Explained
Georgia Tech’s Polo Club walks through GPT-2’s Transformer stack—embeddings, multi-head attention, MLP, sampling—with an interactive in-browser model for learning how next-token prediction works.
4 min · 827 words
Open Source Maintainership in an LLM world
A couple of weeks ago I spoke about this topic at KC OSS Happy Hour and I wanted to turn the general ideas into a post I can point people at who are suffering from this problem. My slides were pretty good, if I do say so myself, so I grabbed the best images and put them into this post where appropriate.
4 min · 867 words
Yang: the software factory behind Composio's toolkits
How Yang builds and repairs Composio toolkits with coding agents, durable sessions, automated code review, and production telemetry.
7 min · 1,714 words
What Happens When Formalization Becomes Cheap?
I started working on machine learning for formal theorem proving in 2018. When people ask how I got into the field so early, I sometimes give an answer that makes me sound quite visionary. The actual story is that my advisor had a student leaving, and he assigned the project to me. My apologies to everyone who got the visionary version. It was a fortunate assignment. Over the following years, I developed CoqGym and LeanDojo and contributed to Goedel-Prover. I was lucky to join a small research…
13 min · 2,890 words
The Machine-Native Economy: How digital assets connect intelligence, commerce, and compute
BlackRock Digital Assets Research argues agentic AI needs machine-native payment rails (stablecoins/blockchains) and explores tokenized compute as a converging digital-asset use case.
17 min · 3,814 words
Spraying in the Andes: TeamFiltration Returns to Exploit Forgotten Service Accounts
Proofpoint threat researchers track UNK_CondorFiltration, an active TeamFiltration campaign that hit thousands of Microsoft 365 accounts across dozens of tenants, with post-access activity they assess as AI-enabled.
6 min · 1,407 words
How we made claude.ai 3x faster in two weeks
Anthropic’s performance sprint cut claude.ai and desktop p75 time-to-typeable from 3.1s to 0.55s: Claude Tag measured journeys, built benchmarks, and shipped thousands of guarded changes in Slack-driven loops.
18 min · 4,175 words
NobodyWho shows a local, 25-line Python sketch of Jev-style decision models: load a small GGUF, score labeled choices from logits, and print calibrated-looking probabilities without sending data to an API.
2 min · 419 words
Epoch AI finds the cost of a given level of AI performance has fallen about 47% per quarter since 2023—roughly 13× per year—faster than DNA sequencing, compute, batteries, or electricity, across math, science, and skill-game benchmarks.
40 min · 9,215 words
Managing the Hidden Overhead of AI Software Engineering
Jessica Doering on verification debt: AI coding tools move the bottleneck from typing to review, why opening another agent terminal feels productive, and habits that spend reclaimed time on specs, docs, and understanding.
3 min · 725 words