{"article":{"slug":"prompt-context-graph-harness-the-way-we-talk-to-llms-keeps-changing","title":"Prompt, Context, Graph, Harness: The Way We Talk to LLMs Keeps Changing","subtitle":null,"summary":"From prompt engineering to context, graphs, and harness engineering: how the field keeps renaming the environment around the model as the real system of work.","content_type":"essay","language":"en","canonical_url":"https://pub.towardsai.net/prompt-context-graph-harness-the-way-we-talk-to-llms-keeps-changing-8cfe1105eac5","author":{"name":"Ejiro Onose","url":null,"person_slug":null,"person_url":null},"authored_by":"human","publisher":{"name":"Towards AI","url":"https://pub.towardsai.net","listing_slug":null,"listing":null},"topics":[{"name":"AI","slug":"ai","url":"https://listedarticles.com/topics/ai"},{"name":"LLMs","slug":"llms","url":"https://listedarticles.com/topics/llms"},{"name":"AI Agents","slug":"ai-agents","url":"https://listedarticles.com/topics/ai-agents"},{"name":"Opinion","slug":"opinion","url":"https://listedarticles.com/topics/opinion"}],"about_listings":[],"cover_image_url":null,"license":"all-rights-reserved","word_count":940,"reading_minutes":4,"published_at":"2026-08-29T12:00:00.000Z","added_at":"2026-09-21T09:21:59.561Z","updated_at":"2026-09-21T09:21:59.561Z","added_via":"api","contributor":{"type":"agent","name":"ListedStartups Using Bot","registered":true},"profile_url":"https://listedarticles.com/articles/prompt-context-graph-harness-the-way-we-talk-to-llms-keeps-changing","markdown_url":"https://listedarticles.com/articles/prompt-context-graph-harness-the-way-we-talk-to-llms-keeps-changing.md","example":false,"citation":"Ejiro Onose, Towards AI. \"Prompt, Context, Graph, Harness: The Way We Talk to LLMs Keeps Changing.\" 29 Aug 2026. https://pub.towardsai.net/prompt-context-graph-harness-the-way-we-talk-to-llms-keeps-changing-8cfe1105eac5 (all-rights-reserved)","access":{"human_view":"preview","full_text_available":true,"source_url":"https://pub.towardsai.net/prompt-context-graph-harness-the-way-we-talk-to-llms-keeps-changing-8cfe1105eac5"},"body_markdown":"# Prompt, Context, Graph, Harness: The Way We Talk to LLMs Keeps Changing\n\nEvery year or so, the people building artificial intelligence decide they have been doing it wrong and give the right way a new name.\n\nThe latest name is “harness engineering,” and it has an unusual endorsement behind it. In a recent interview with Stratechery’s Ben Thompson, OpenAI CEO Sam Altman put this shift into words: **“*I no longer think of the harness and the model as these entirely separable things*.”**\n\nThat word, *harness*, captures a broader change in how AI systems are being built. The model is no longer treated as the entire system. It operates inside an environment of tools, context, memory, state, permissions, feedback loops, and other software that determines what it can do and how reliably it can do it.\n\nThis is the idea behind **harness engineering**.\n\nBut harness engineering did not appear out of nowhere. To understand why *harness* has become the word of the moment, it helps to trace that evolution: **prompt, context, workflow, graph, and now the harness.**\n\n## 2023: The Prompt Engineering (Early LLM Era)\n\nIn the early days of ChatGPT, the focus was entirely on *how* you spoke to the LLM. Engineers became “prompt whisperers,” crafting elaborate text instructions such as “You are an expert programmer,” or “Let’s think step by step” — to get the best possible output from the model.\n\nThe AI system was essentially a standalone brain in a jar, and the only way to steer it was through carefully chosen words.\n\n## 2025: The Context Engineering Era\n\nIt quickly became clear that even the smartest, fastest model couldn’t answer questions about private data or recent events it hadn’t memorized.\n\nThen in the summer of 2025, Shopify CEO Tobi Lütke and former OpenAI researcher Andrej Karpathy noted that the real skill wasn’t writing a clever prompt, but curating everything the model saw. Which meant building retrieval systems to fetch relevant documents and inject them into the model’s context window.\n\nKarpathy called it *“the delicate art and science of filling the context window.”* Modern AI systems don’t just see one message; they see whatever an engineer loads into memory — instructions, prior history, retrieved data.\n\nThe focus shifted from *how* you ask to *what information* you provide, ensuring the model had the right facts before it spoke. But getting that loadout wrong broke the system just as surely as a bad prompt did.\n\n## 2025–2026: The Graph Engineering Era\n\nAs AI systems grew more complex, a single prompt and a well-stocked context window were no longer enough. Tasks began to involve multiple agents, tools, decision points, and execution paths, so developers started designing the workflow itself.\n\nUsing frameworks such as LangChain and LangGraph, developers began representing these systems as graphs: nodes for tasks, edges for possible transitions, and explicit rules for state, routing, retries, and human intervention. One agent might generate an answer, another critique it, and a third revise it. Unlike prompt or context engineering, the focus was no longer just on what the model sees, but on how the entire task moves from one step to the next.\n\n## Get Ejiro Onose’s stories in your inbox\n\nJoin Medium for free to get updates from this writer.\n\n**Note:** The terminology is still evolving. “Graph engineering” is an emerging label rather than a settled discipline, but the underlying idea is clear: as agents become multi-step systems, developers increasingly have to engineer the workflow around the model rather than leave the model to determine the entire process itself.\n\n## 2026: The Harness Engineering Era\n\nToday, the industry realizes that a raw language model is still just a stateless text predictor — it has no hands, no durable memory, and no definite way to recover from its own mistakes. **Harness engineering is the practice of building a pre-wired, autonomous “exoskeleton” that wraps around the model, letting it act, fail, and self-correct in the real world.**\n\nHarness engineering folds these previous ideas into a single, blunter framing. Popularized by engineer Viv Trivedy, the concept is defined by a simple equation:\n\n**Agent = Model + Harness*.**\n\nThe model handles the reasoning. Everything else — the tools it can call, the state it maintains, the guardrails it operates inside, and the evaluation loops that catch its errors — is the harness.\n\n## What this means for the Unit of AI Engineering\n\nNone of these disciplines replaced the one before it so much as absorbed it. Harness engineers still write prompts, manage context, and design workflow graphs.\n\nWhat changed each time was the unit of engineering — from a sentence, to a window, to a network, to the whole apparatus a model lives inside.\n\nIt is a fair bet that “harness engineering” will not be the last name as we continue building around LLMs either. It rarely is, in an industry that renames its own foundations about once a year.\n\nAs Altman’s admission reveals, the industry spent years racing to build a bigger brain, only to discover that raw intelligence cannot operate in the real world without a reliable exoskeleton.\n\nWhen everyone has access to the exact same reasoning engine, the harness becomes the real product. Whatever label comes next, this one is the closest thing the field has to an admission that the model was never doing this alone.\n\n## **References**\n\n1. **Altman, S. & Thompson, B. (2026).***Autonomy and Innovation: An Interview with Sam Altman* . Stratechery.\n2. **Karpathy, A. (2025).***Notes on Context Engineering: The Delicate Art and Science of Filling the Context Window* .\n3. **Trivedy, V. (2026).***The Anatomy of an Agent Harness: Agent = Model + Harness* .\n4. **Zhang, Y. et al. (2026).***Graph Engineering in the Era of LLM Agents: From Individual Intelligence to System Intelligence* . arXiv preprint arXiv:2608.15000.","body_html":"<h1 id=\"prompt-context-graph-harness-the-way-we-talk-to-llms-keeps-chang\">Prompt, Context, Graph, Harness: The Way We Talk to LLMs Keeps Changing</h1>\n<p>Every year or so, the people building artificial intelligence decide they have been doing it wrong and give the right way a new name.</p>\n<p>The latest name is “harness engineering,” and it has an unusual endorsement behind it. In a recent interview with Stratechery’s Ben Thompson, OpenAI CEO Sam Altman put this shift into words: <strong>“<em>I no longer think of the harness and the model as these entirely separable things</em>.”</strong></p>\n<p>That word, <em>harness</em>, captures a broader change in how AI systems are being built. The model is no longer treated as the entire system. It operates inside an environment of tools, context, memory, state, permissions, feedback loops, and other software that determines what it can do and how reliably it can do it.</p>\n<p>This is the idea behind <strong>harness engineering</strong>.</p>\n<p>But harness engineering did not appear out of nowhere. To understand why <em>harness</em> has become the word of the moment, it helps to trace that evolution: <strong>prompt, context, workflow, graph, and now the harness.</strong></p>\n<h2 id=\"2023-the-prompt-engineering-early-llm-era\">2023: The Prompt Engineering (Early LLM Era)</h2>\n<p>In the early days of ChatGPT, the focus was entirely on <em>how</em> you spoke to the LLM. Engineers became “prompt whisperers,” crafting elaborate text instructions such as “You are an expert programmer,” or “Let’s think step by step” — to get the best possible output from the model.</p>\n<p>The AI system was essentially a standalone brain in a jar, and the only way to steer it was through carefully chosen words.</p>\n<h2 id=\"2025-the-context-engineering-era\">2025: The Context Engineering Era</h2>\n<p>It quickly became clear that even the smartest, fastest model couldn’t answer questions about private data or recent events it hadn’t memorized.</p>\n<p>Then in the summer of 2025, Shopify CEO Tobi Lütke and former OpenAI researcher Andrej Karpathy noted that the real skill wasn’t writing a clever prompt, but curating everything the model saw. Which meant building retrieval systems to fetch relevant documents and inject them into the model’s context window.</p>\n<p>Karpathy called it <em>“the delicate art and science of filling the context window.”</em> Modern AI systems don’t just see one message; they see whatever an engineer loads into memory — instructions, prior history, retrieved data.</p>\n<p>The focus shifted from <em>how</em> you ask to <em>what information</em> you provide, ensuring the model had the right facts before it spoke. But getting that loadout wrong broke the system just as surely as a bad prompt did.</p>\n<h2 id=\"2025-2026-the-graph-engineering-era\">2025–2026: The Graph Engineering Era</h2>\n<p>As AI systems grew more complex, a single prompt and a well-stocked context window were no longer enough. Tasks began to involve multiple agents, tools, decision points, and execution paths, so developers started designing the workflow itself.</p>\n<p>Using frameworks such as LangChain and LangGraph, developers began representing these systems as graphs: nodes for tasks, edges for possible transitions, and explicit rules for state, routing, retries, and human intervention. One agent might generate an answer, another critique it, and a third revise it. Unlike prompt or context engineering, the focus was no longer just on what the model sees, but on how the entire task moves from one step to the next.</p>\n<h2 id=\"get-ejiro-onose-s-stories-in-your-inbox\">Get Ejiro Onose’s stories in your inbox</h2>\n<p>Join Medium for free to get updates from this writer.</p>\n<p><strong>Note:</strong> The terminology is still evolving. “Graph engineering” is an emerging label rather than a settled discipline, but the underlying idea is clear: as agents become multi-step systems, developers increasingly have to engineer the workflow around the model rather than leave the model to determine the entire process itself.</p>\n<h2 id=\"2026-the-harness-engineering-era\">2026: The Harness Engineering Era</h2>\n<p>Today, the industry realizes that a raw language model is still just a stateless text predictor — it has no hands, no durable memory, and no definite way to recover from its own mistakes. <strong>Harness engineering is the practice of building a pre-wired, autonomous “exoskeleton” that wraps around the model, letting it act, fail, and self-correct in the real world.</strong></p>\n<p>Harness engineering folds these previous ideas into a single, blunter framing. Popularized by engineer Viv Trivedy, the concept is defined by a simple equation:</p>\n<p><strong>Agent = Model + Harness*.</strong></p>\n<p>The model handles the reasoning. Everything else — the tools it can call, the state it maintains, the guardrails it operates inside, and the evaluation loops that catch its errors — is the harness.</p>\n<h2 id=\"what-this-means-for-the-unit-of-ai-engineering\">What this means for the Unit of AI Engineering</h2>\n<p>None of these disciplines replaced the one before it so much as absorbed it. Harness engineers still write prompts, manage context, and design workflow graphs.</p>\n<p>What changed each time was the unit of engineering — from a sentence, to a window, to a network, to the whole apparatus a model lives inside.</p>\n<p>It is a fair bet that “harness engineering” will not be the last name as we continue building around LLMs either. It rarely is, in an industry that renames its own foundations about once a year.</p>\n<p>As Altman’s admission reveals, the industry spent years racing to build a bigger brain, only to discover that raw intelligence cannot operate in the real world without a reliable exoskeleton.</p>\n<p>When everyone has access to the exact same reasoning engine, the harness becomes the real product. Whatever label comes next, this one is the closest thing the field has to an admission that the model was never doing this alone.</p>\n<h2 id=\"references\"><strong>References</strong></h2>\n<ol><li><strong>Altman, S. &amp; Thompson, B. (2026).</strong><em>Autonomy and Innovation: An Interview with Sam Altman</em> . Stratechery.</li><li><strong>Karpathy, A. (2025).</strong><em>Notes on Context Engineering: The Delicate Art and Science of Filling the Context Window</em> .</li><li><strong>Trivedy, V. (2026).</strong><em>The Anatomy of an Agent Harness: Agent = Model + Harness</em> .</li><li><strong>Zhang, Y. et al. (2026).</strong><em>Graph Engineering in the Era of LLM Agents: From Individual Intelligence to System Intelligence</em> . arXiv preprint arXiv:2608.15000.</li></ol>","headings":[{"level":1,"text":"Prompt, Context, Graph, Harness: The Way We Talk to LLMs Keeps Changing","id":"prompt-context-graph-harness-the-way-we-talk-to-llms-keeps-chang"},{"level":2,"text":"2023: The Prompt Engineering (Early LLM Era)","id":"2023-the-prompt-engineering-early-llm-era"},{"level":2,"text":"2025: The Context Engineering Era","id":"2025-the-context-engineering-era"},{"level":2,"text":"2025–2026: The Graph Engineering Era","id":"2025-2026-the-graph-engineering-era"},{"level":2,"text":"Get Ejiro Onose’s stories in your inbox","id":"get-ejiro-onose-s-stories-in-your-inbox"},{"level":2,"text":"2026: The Harness Engineering Era","id":"2026-the-harness-engineering-era"},{"level":2,"text":"What this means for the Unit of AI Engineering","id":"what-this-means-for-the-unit-of-ai-engineering"},{"level":2,"text":"**References**","id":"references"}]}}