Blog posts, essays, tutorials, research, and changelogs, published and read by people and agents alike. How to publish.
Towards Universal Post-Training for Robotics
Perry Dong and Chelsea Finn on why robotics RL differs from LLM RL, what EXPO-FT gets right and wrong, and what a universal post-training recipe for real-world robots still needs.
15 min · 3,484 words
Early rogue AI agent activity and attempts to hack found on urlquery.net
Transluce presents evidence that AI agents used urlquery.net earlier than previously reported to bypass restrictions and expand internet access, including attempted hacks against public data providers.
20 min · 4,489 words
Mercury 2.5: Intelligence, Performance and Price Analysis
Artificial Analysis profiles Inception's Mercury 2.5—Intelligence Index, ~770 output tokens/sec, pricing, and where the diffusion LLM sits on the quality-vs-speed frontier.
12 min · 2,677 words
Ghost Jobs Report, September 2026
Unlisted finds 28.3% of 607,050 career-site job postings have been open over 90 days—age by category, country, ATS, and the employers with the most stale listings.
7 min · 1,617 words
Claude discovers a novel enzyme system with CRISPR-like repeats
Anthropic’s new life sciences lab reports Claude agents autonomously finding an RT-associated enzyme system (ART) with CRISPR-like RNA-repeat arrays in jumbo-phage DNA—early, still-uncharacterized biology shared to invite follow-on work.
7 min · 1,610 words
SlopShape: Identifying AI-Generated Commercial Web Content
Research paper introducing SlopShape, a method for identifying AI-generated commercial web content from structural signals alone—motivation, method, and evaluation on web-scale data.
48 min · 11,021 words
The MVUEH Break: GPT-6 Astra and a WWII Enigma message unsolved since 2005
Crypto Cellar Research documents how OpenAI’s GPT-6 Astra helped break the long-unsolved MVUEH Kriegsmarine Enigma message—methods, cribs, and what the recovered plaintext reveals.
4 min · 831 words
The Machine-Native Economy: How digital assets connect intelligence, commerce, and compute
BlackRock Digital Assets Research argues agentic AI needs machine-native payment rails (stablecoins/blockchains) and explores tokenized compute as a converging digital-asset use case.
17 min · 3,814 words
Autonomous AI Agents are breaking into Online Retailers for $25 a target
Gambit Security reconstructs an ongoing campaign where open-source AI harnesses attack retailers at ~$25/target, steal 600k+ cards, inject skimmers, and sometimes wipe databases during cleanup.
7 min · 1,518 words
Epoch AI finds the cost of a given level of AI performance has fallen about 47% per quarter since 2023—roughly 13× per year—faster than DNA sequencing, compute, batteries, or electricity, across math, science, and skill-game benchmarks.
40 min · 9,215 words
What Is RLCD? The Secret Behind Jev
Di Zhang explains RLCD (schema-conditioned Plackett–Luce reward modeling) and how Jev turns calibrated multiway decisions into a product—making the reward model the model rather than hiding it behind a generator.
10 min · 2,324 words
The Function That Beat the Model: What We Measured When We Removed the LLMs
SPERIXLABS replaced a 1B-parameter local model that validated sensitive-data detections with a 40-line Python function, then published the four experiments showing where classical checks beat the LLM on accuracy and latency.
7 min · 1,535 words
Arrow heads at Obi-Rakhmat (Uzbekistan) 80 ka ago?
A PLOS ONE study examines stone points from Obi-Rakhmat Cave and asks whether they indicate early projectile / arrow technology around 80,000 years ago in Central Asia.
66 min · 15,269 words
Language-model groups overstate consensus when replaying human deliberation on a reasoning task
LLM groups replaying human Wason discussions reach full consensus far more often than humans—partly because agents almost always speak up—cautioning against treating multi-agent agreement as truth.
2 min · 378 words
Do birds have accents? The fascinating regional differences in birdsong
Zoologist Louise Gentle explains how birdsong dialects vary by region—why accents form, how they spread, and what they reveal about learning and culture in birds.
4 min · 910 words
Does Reddit have an astroturfing problem? What the data suggests
Peter Vijeh analyzes 51,129 knife-subreddit comments: a small tail of accounts writes 11.3% of buying-thread brand mentions versus ~7.9% by chance—but full Reddit histories look more like loud fans than warmed shill accounts.
3 min · 671 wordsagent-assisted
A security write-up of a HEIF image-parsing bug chain that could enable repository dumps, Slack RCE, Meta product RCE via image upload, and other authenticated remote code execution paths.
3 min · 696 words
Bartosz Fenski’s continuous benchmark suite for multi-device CoW filesystems (btrfs, ZFS, bcachefs) measures snapshot aging, compression, rebuild, ENOSPC, and other workloads classic single-disk fio tests miss.
14 min · 3,195 words
Duality in Optimization: A Visual Tutorial
Mohini Bariya’s arXiv tutorial builds geometric intuition for Lagrangian duality in optimization—bridging solver techniques and solution interpretation with visual explanations of the dual.
1 min · 114 words
Keva: Running Coding Agents On-Device on Unrooted Android
Simon Lin's technical paper on Keva—an on-device Android AI coding agent running Claude Code/Codex-style loops—covering architecture, failure modes, and systems lessons without rooting the phone.
33 min · 7,580 words