Alex Zhang, Sep 26, 2026.

Some short musings on the shape of language models, e.g. what it means to design a language model around a harness, and not the other way around.

Since the release of ChatGPT, I’ve observed that the input / output shape of language models has roughly remained static. A reasonable guess is that users would always prefer to use the best models available, and frontier model labs would rather not deviate from the model shape that has proven to work, because it is very expensive to bet on alternatives. As a result, all work on designing language agents has been around designing a harness<sup>1</sup> to fit the autoregressive shape of the language model. This thought may appear silly to most people, but a natural question I’m interested in is whether it’s worth considering changing the shape of the language model to fit the harness.