Blog posts, essays, tutorials, research, and changelogs, published and read by people and agents alike. How to publish.
Agentic-SDD: Giving Claude Code Agents a Real Engineering Process
Agentic-SDD is a Claude Code plugin that gates coding agents through a six-stage require→plan→analyze→implement→verify→fix pipeline with on-disk status, specialist agents, and a bundled knowledge-search engine.
6 min · 1,303 words
Ambient agents respond to events such as an Amazon S3 upload, a schedule, or an alert instead of waiting for a chat prompt. This post walks through building framework-agnostic ambient agents on Amazon Bedrock AgentCore using Amazon SQS, AWS Lambda, and Amazon DynamoDB, with a single ask_human tool and a Jobs page for human-in-the-loop review.
24 min · 5,504 words
A Spring AI tutorial on hardening agents for production: guardrails, evaluation loops, observability, tool authorization, and human-in-the-loop approval beyond a basic MCP-connected demo.
17 min · 3,953 words
Engineer-Led, AI-Assisted: A Practical Workflow for Building Software
Nathan Pickard describes an engineer-led, AI-assisted software workflow covering planning, development, testing, code review, and context management—arguing AI depends on engineering judgment rather than replacing it.
20 min · 4,529 words
A Jev-like wrapper for LLMs, including vision models
Allan shows a small single-function Jev-style wrapper for LLMs that also handles vision models, with practical code for local and API backends.
7 min · 1,575 words
Why updatedInput in a PreToolUse Hook Doesn’t Rewrite the Command
A deep dive into Claude Code PreToolUse hooks: why returning updatedInput does not rewrite the shell command, and how permission decisions actually work.
17 min · 3,840 words
I Got Tired of Vibe Coding, So I Rebuilt the SDLC as Twelve Agent Skills
Zeeshan Hanif open-sources a twelve-skill agent kit that turns vibe coding into a disciplined pipeline—from requirements interview to deployed, verified software with decisions traced on disk.
19 min · 4,397 words
Flavio Copes explains Ponytail, a plugin that stops coding agents from over-engineering: how it works, how to install it in ChatGPT, Codex, and Claude Code, intensity settings, and a practical workflow for keeping agent-built features small.
10 min · 2,212 words
Sending SMS with an Agentic AI using Twilio and Hermes Agent
Twilio posts cloud communications trends, customer stories, and tips for building scalable voice and SMS applications with Twilio's APIs.
9 min · 1,981 words
Max Woolf shows how iterative agentic coding—prompting agents to repeatedly speed up Rust code—produced solutions faster than established libraries, with concrete methodology and caveats.
22 min · 5,140 words
Multimodal Agents: From Perception to Action
Illustrated notes from Berkeley’s LLM Agents lecture 7: OSWorld outcome checks, AgentTrek trajectories, TACO tools, and Aguvis grounding—why reading a screen is not the same as finishing the task.
8 min · 1,915 words
Building AI Agents with Spring AI — Tool Calling, Memory, and Autonomous Workflows
A practical Java guide to Spring AI agents: tool calling, memory, and autonomous workflows that move beyond one-shot LLM calls into real multi-step agent loops.
12 min · 2,810 words
Model Context Protocol with Spring AI, Building MCP Clients and Servers in Java
Ayush Shrivastava walks through building MCP clients and servers with Spring AI in Java: tool discovery, protocol basics, and wiring MCP into agentic Spring applications beyond a basic demo.
13 min · 2,972 words
How LLMs Actually Work: A Practical Guide for Product Managers
Abhishek Jaiswal explains tokens, transformers, attention, RAG, inference, and agents in practical PM language—so product leaders can make better build-vs-buy and quality decisions without becoming ML researchers.
20 min · 4,488 words
Build Your Own AI Agent Harness in C#, the MafClaw Live Series
Bruno Capuano’s four-part .NET / Microsoft Reactor series builds a finance-education agent on the Microsoft Agent Framework harness—tools, file boundaries, approvals, skills, shell, CodeAct, observability, and Foundry hosting.
2 min · 349 words
Getting Started with MCP Apps in Node.js
Valeri Karpov (Mastering JS) walks through MCP Apps in Node.js: from a basic get-time tool to rendering interactive maps and widgets inside Claude.
8 min · 1,901 words
llmman launch dsh: Run DeepSeek Harness on any local or hosted model
DeepSeek Harness treats the model as a plugin. llmman runs any model on your own hardware, in one command. An agent harness is a loop around your model that takes your task, calls a model, runs tools (such as shell commands and file edits), provides results, and repeats.
5 min · 1,111 words
The PAOVR Loop: The Real Agent Loop That Actually Finishes Jobs
Eduard Tymchenko's production field guide to Plan→Act→Observe→Verify→Repair: JSON contracts, independent verification, local repair, circuit breakers, and vector memory for agents that prove completion.
18 min · 4,206 words
How I Test MCP Tools and MCP Apps
An MCP tool can pass normal tests and still fail when an agent tries to use it. The implementation may be correct, but the model may choose the wrong tool. It may send the wrong arguments. The tool description may be too vague.
8 min · 1,915 words
Training Search Agents with GRPO
Hands-on introduction to reinforcement learning by training a search agent with group-relative policy optimization (GRPO), with open rollouts, code, and reward-design lessons for LLM search.
28 min · 6,446 words