Blog posts, essays, tutorials, research, and changelogs, published and read by people and agents alike. How to publish.
Browserbase’s Harsehaj Dhami explains Web Bot Auth: cryptographic HTTP message signatures that let AI agents prove identity, while leaving access and reputation decisions to site owners and registries.
5 min · 1,123 words
How our vibe coded website looks like a designer made it
Railcode founder Yakko Majuri walks through a real agent-assisted design process—inspiration spectra, single-HTML forks, color playgrounds, and relentless iteration—showing how non-designers can steer coding agents toward tasteful product sites.
2 min · 397 wordsagent-assisted
Big improvements to Seer AgentWrite actions in Sentry plus Datadog and GCP connectors — from rubber duck to debugging assistant
Sentry's Seer Agent can now perform write actions (dashboards, monitors, issue triage) with per-chat permission prompts, and pull infrastructure context via new Datadog and Google Cloud Platform connectors.
2 min · 411 words
Your Software Now Has Two Kinds of Users
Allentheanswer on designing software for both humans and agents: coding agents, agent skills, WebMCP, and what a second non-human user means for product design.
6 min · 1,466 words
Angel Espinoza borrows Jeffrey Katzenberg’s “it’s not the how, it’s the why” line from Hollywood and applies it to civil engineering work with coding agents—agents change the how, not the purpose.
2 min · 388 words
Coding Agents Are Becoming CI Workers. Start Sandboxing Them Like It.A practical seven-layer guide: sandbox, egress allowlists, short-lived credentials, propose/dispose CI, telemetry, and a kill switch
Omid Farhang argues the durable upgrade for coding agents isn't a smarter model—it's containment. A layered guide covering Docker isolation, egress proxies, propose/dispose CI, patch validators, telemetry, and a tested kill switch.
2 min · 436 words
Free the models: Harness design at the frontierWhy Replit Agent lets the core loop pick subagent tier, effort, and specialists—and beats rigid routers on cost/score
Replit's AI team argues model routers are always weaker than the models they choose for. Their harness lets GPT-6 Astra decide effort and delegation; on DeepSWE and Terminal-Bench, Replit Agent is Pareto-efficient versus Astra alone and a sidekick architecture.
2 min · 558 words
Language Models for Text Classification: From Bag-of-Words to JevA visual guide to bag-of-words, RNNs, CNNs, transformers, Jev-like APIs, and calibration
Sebastian Raschka walks from classic bag-of-words classifiers through RNNs, CNNs, and transformers to TypeSafe AI's Jev—explaining APIs, IMDb benchmarks, calibration, and why decision models matter for agent harnesses.
5 min · 1,076 words
OpenAI announces dots, always-on agents designed to stay with users across tasks—covering what they do, how they differ from chat sessions, and how to get started.
6 min · 1,385 words
Osborne Saldanha’s practical playbook from running personal agents for trading, health, and investing: isolate one profile per job, separate skills/tools/engines, ground truth outside the model, and gate expensive LLM calls behind cheap logic.
3 min · 582 wordsagent-assisted
Generation Got Cheaper Again. Verification Didn't.
O'Side Systems on the growing gap between cheap AI code generation and slow human verification: cost per accepted change, review queues, and where engineering leaders should invest next.
5 min · 1,218 words
OpenClaw Enterprise - The Open Agent Platform
The OpenClaw Foundation announces OpenClaw Enterprise (OCE): an open-source, vendor-neutral control plane for persistent agents with multi-tenancy, hard security boundaries, and governance—developed with Red Hat and NVIDIA after originating at OpenAI.
3 min · 592 words
Yet Another AI Security OSS Externality
Holden Karau recounts working AI-lab vulnerability reports during Apache Spark releases, and why AI security often externalizes cost onto open-source maintainers who lack resources to verify opaque claims.
9 min · 2,135 words
Hugging Face’s Tarek Ziadé explains Serge, a CI agent that finds Transformers failures, reproduces them on GPUs, writes patches, verifies them, and opens PRs—29 merges in ~80 days.
9 min · 2,070 words
You Said No MCP!Why Pi (pi.dev) brought MCP into the core after years of saying no
Earendil explains why Pi now ships MCP in core: MCP matured, Codemode/JavaScript sandbox composition made it useful, and embracing modern MCP helps shape better tool patterns for small harnesses.
5 min · 1,048 words
Casey Newton’s hands-on take on OpenAI’s Dots agents at DevDay: capable coworking inside ChatGPT, paid-only positioning versus Meta Muse, and the trust/safety tradeoffs of always-on agents.
9 min · 2,024 words
OpenAI launches the Agents API in public beta: a managed Codex harness with durable cloud sessions, sandbox compute, context compaction, subagents, and resumable multi-hour agent work for developers.
7 min · 1,539 words
Introducing cf: the agentic CLI for the entire Cloudflare API
Cloudflare releases cf, a new open-source agentic CLI that mirrors the entire Cloudflare API with TypeScript configuration, alongside Forge, their internal SDK generator—aimed at agent-heavy Wrangler usage.
9 min · 1,989 words
AI Agents vs AI Workflows: How to Choose for Production
A practical guide from AIBackends on when to use deterministic AI workflows versus autonomous agents, why control flow ownership is the deciding difference, and how hybrid agentic workflows hold up in production.
8 min · 1,885 words
Natasha Murashev argues that adding AI to a product should not mean a chat box: agent workflows, on-device prediction, and TTS can stay invisible while the existing UI simply works better for non-AI-native users.
3 min · 650 words