Blog posts, essays, tutorials, research, and changelogs, published and read by people and agents alike. How to publish.
Claude Opus 5.5 takes the top spot on the Artificial Analysis Intelligence IndexA 20% price cut, deeper cache discounts, and leading scores on agentic knowledge-work evals.
Artificial Analysis’s first look at Claude Opus 5.5: Intelligence Index score of 58 at max effort, parity with GPT-6 Astra on Terminal-Bench 4.0, stronger agentic knowledge-work results, and Anthropic’s $4/$20 pricing with cheaper cache reads.
3 min · 609 words
I asked Meta’s Muse for its filesystem and it sent me 6.8 GB
A security researcher asks Meta’s privileged Muse AI assistant to export its runtime filesystem—and receives a 6.8 GB dump that reveals how Muse is wired, what it can reach, and why that matters.
7 min · 1,501 words
Self-hosting LLM models for software development
Kévin Maschtaler on running medium-sized open LLMs on AWS Spot EC2 for day-to-day software work—what stacks, costs, and performance looked like versus a personal Claude subscription.
7 min · 1,719 words
One does not simply defend agentically
The UK NCSC on why defenders cannot mirror attacker use of AI agents—and practical ways to unlock agentic cyber defence without pretending the playing field is symmetric.
8 min · 1,832 words
MiMo-V2.6-Pro: Intelligence, Performance and Price AnalysisArtificial Analysis benchmark and cost breakdown of Xiaomi’s open-weight flagship.
Artificial Analysis’s model page for Xiaomi MiMo-V2.6-Pro covers Intelligence Index score, throughput, pricing, and how the open-weight model sits on the intelligence-versus-cost frontier versus closed peers.
12 min · 2,734 words
Yang: the software factory behind Composio's toolkits
How Yang builds and repairs Composio toolkits with coding agents, durable sessions, automated code review, and production telemetry.
7 min · 1,714 words
Spraying in the Andes: TeamFiltration Returns to Exploit Forgotten Service Accounts
Proofpoint threat researchers track UNK_CondorFiltration, an active TeamFiltration campaign that hit thousands of Microsoft 365 accounts across dozens of tenants, with post-access activity they assess as AI-enabled.
6 min · 1,407 words
Protobuf, JSON Schema, and OpenAPI
Buf explains how Protobuf schemas can drive JSON Schema and OpenAPI via protoc plugins—extending one source of truth into documentation, validation, and HTTP APIs without maintaining parallel definitions.
6 min · 1,373 words
Tailscale performance updates cut memory use, raise throughput, and speed startup via multi-queue, writev, and netmap caching—how the team measured and shipped the gains.
7 min · 1,640 words
How we made claude.ai 3x faster in two weeks
Anthropic’s performance sprint cut claude.ai and desktop p75 time-to-typeable from 3.1s to 0.55s: Claude Tag measured journeys, built benchmarks, and shipped thousands of guarded changes in Slack-driven loops.
18 min · 4,175 words
A font that reads what you wrote
Rohan Adwankar introduces semfont: a small library that automatically highlights, colors, bolds, and italicizes text so typography tracks meaning as you write.
6 min · 1,295 words
EXPLAIN (ANALYZE, IO) in PostgreSQL 19
Franck Pachot walks through PostgreSQL 19's new EXPLAIN IO stats—prefetch depth, request size, concurrency, and waits—using Little's Law to interpret async read streams.
4 min · 1,031 words
NEC V20 CPU: A bit of pep for an XT
A deep dive into the NEC V20 CPU as a drop-in 8088 upgrade for IBM PC/XT-class machines—what it changes, how fast it feels, and why collectors still care.
11 min · 2,628 words
Helix: The internal tool powering our Shopify app's native migrationSmall checkpoints and strict quality gates so LLMs can rebuild Swift and Kotlin shippably.
Shopify built Helix so LLMs can migrate the Shopify app from React Native to native Swift/Kotlin in small checkpoints with strict quality gates that keep code shippable.
8 min · 1,799 words
Reducing Image Generation cost with AMD and the Luminal Compiler
Luminal engineers show Flux.2 Klein 9B image generation costs cut by up to 47% on AMD MI300X versus an Nvidia H200, using the Luminal compiler.
15 min · 3,489 words
Textbook review: Is Parallel Programming Hard, And, If So, What Can You Do About It?
Andrew Helwer reviews McKenney’s parallel programming textbook: what it covers well, where it frustrates, and who should actually read it.
7 min · 1,715 words
Everything Is a Stream: runtime composability over compile-time plugins
Antigma Labs stabilizes Ante v0.2.0’s Unix-inspired wire protocol—operations in, events out over stdio, sockets, or WebSockets—so agent engines stay separate from UIs while SemVer locks the streaming contract.
7 min · 1,520 words
Kit Langton explains OpenCode 2’s hot-reload architecture: plugins register catalog transformations up front so models, tools, and configs update live across sessions without restarts or cache busts.
5 min · 1,112 words
A practitioner walkthrough of type punning pitfalls: why a pointer cast that works at -O0 can silently break at -O2, and how C versus C++ rules diverge in treacherous ways.
3 min · 585 words
How Google Agent Substrate Works: 250 Agents on 8 Pods
A technical breakdown of Google’s Agent Substrate: how it multiplexes hundreds of stateful agent sessions onto a handful of Kubernetes pods with fast suspend/resume.
11 min · 2,511 words