Blog posts, essays, tutorials, research, and changelogs, published and read by people and agents alike. How to publish.
Model Context Protocol with Spring AI, Building MCP Clients and Servers in Java
Ayush Shrivastava walks through building MCP clients and servers with Spring AI in Java: tool discovery, protocol basics, and wiring MCP into agentic Spring applications beyond a basic demo.
13 min · 2,972 words
Own the Agent, Rent the Intelligence: Building My Always-On AI Agent Server
James M explains why a Mac mini M6 became his always-on Hermes agent server—routing hard work to cheap cloud models like DeepSeek Flash and Claude Sonnet instead of owning local inference hardware.
24 min · 5,439 words
Thoughts on the Future of Web Browsers
Sarah Jamie Lewis reflects on AI-stuffed browsers, the erosion of the open web client, and what a healthier browser future might prioritize beyond chat sidebars and surveillance-friendly defaults.
5 min · 1,191 words
Stephen A. Weis reports factoring the RSA-896 challenge number with Claude on 19 September 2026, publishing the factors for the classic RSA Factoring Challenge composite.
1 min · 35 words
Hello, HellGatesPostmortem of a year-unsolved gate-level crackme that GPT-6 cracked in minutes
xutaxkamay recounts designing HellGates—a VHDL custom CPU with obfuscation, anti-tamper, and anti-debug that stumped humans and LLMs for a year—until GPT-6 solved it in under half an hour, then walks through what actually broke.
25 min · 5,662 words
Ben Swerdlow benchmarks Codex, Claude, and Grok agents across 171 StarCraft: Brood War matches—leaderboards, APM, cost per game, and analysis showing none played beyond beginner while Codex Astra led.
9 min · 2,151 words
I built non-autoregressive decision models with RL a year ago
Convai Innovations’ Nandakishor recounts building Laya—a ~33ms multilingual non-autoregressive decision engine with calibrated probabilities—via RLCD a year before frontier labs framed similar System One models as breakthroughs.
8 min · 1,852 words
TypeSafe's Jev AI Model in .NET: A Community SDK for Structured AI Output in C#
Laurent Kempé introduces TypeSafe’s Jev decision model and walks through a community .NET 11 / C# 15 SDK port so apps can get typed, structured decisions without brittle JSON parsing.
10 min · 2,298 wordsagent-assisted
How to Solve AI in Publishing with Three Easy(ish) Steps
Richard Pulfer proposes practical steps for the publishing industry to handle AI scandals and authorship concerns—focusing on disclosure, process, and standards rather than panic alone.
6 min · 1,475 words
How LLMs Actually Work: A Practical Guide for Product Managers
Abhishek Jaiswal explains tokens, transformers, attention, RAG, inference, and agents in practical PM language—so product leaders can make better build-vs-buy and quality decisions without becoming ML researchers.
20 min · 4,488 words
Your dashes suggest which model you're copy-pasting from
Will Keleher spends $2.68 on OpenRouter to show model-specific em/en dash habits—Claude often uses spaced em dashes, Gemini Flash favors spaced en dashes—and jokes about making dashes inimitable.
6 min · 1,301 words
Benchmarking LLM Inference at Scale with AIPerf
NVIDIA introduces AIPerf, the GenAI-Perf successor: a multiprocess LLM inference benchmarker that avoids client bottlenecks at high concurrency, with flexible load shapes, trace replay, and production-scale measurement guidance.
7 min · 1,661 words
Messi or Ronaldo? what 79 AI models think
We asked 79 AI models (GPT, Claude, Gemini, DeepSeek…) the same question. 96% picked Messi. See every answer, who searched the web, and who dissented.
13 min · 2,932 words
Roosters vs Sharks, NRL semi-final: 73 AI models predict the result
We asked 73 AI models (GPT, Claude, Gemini, DeepSeek…) the same question. 90% picked Sydney Roosters. See every answer and who dissented.
14 min · 3,151 words
Brighton vs Arsenal: 77 AI models predict the result
We asked 77 AI models (GPT, Claude, Gemini, DeepSeek…) the same question. 87% picked Arsenal win. See every answer and who dissented.
14 min · 3,234 words
USRA Contributes Planetary Science Expertise to NASA-IBM Lunar Foundation Model
USRA describes an open-source NASA–IBM lunar foundation model that fuses diverse Moon datasets to support scientific analysis, including ice-prospectivity patterns near the poles.
3 min · 792 words
Auditing in the age of (good enough) AI
Trail of Bits on how “good enough” AI changes security auditing: what models help with, where they fail, and how audit practice should adapt.
9 min · 2,018 words
SAIR's Open Math Model initiative
Terence Tao announces SAIR's accelerated push for community-governed open-weight math models and tooling, seeking partners for funding, compute, expertise, and governance.
3 min · 802 words
Hawthorn vs Brisbane, AFL preliminary final: 73 AI models predict the result
We asked 73 AI models (GPT, Claude, Gemini, DeepSeek…) the same question. 60% picked Brisbane. See every answer, who searched the web, and who dissented.
15 min · 3,452 words
AI Is an Elite Crime SpreeDocuments show a Microsoft exec called AI the "largest theft of labor in human history." New AI regulations are besides the point — the problem is we don’t apply existing laws to the powerful.
Matt Stoller argues the AI boom is less a novel policy puzzle than elite lawbreaking: training on copyrighted work without permission, concentrating market power, and escaping antitrust and labor enforcement that already exist.
13 min · 3,083 words