One month coding with GLM 5.3 Flash

Setting a challenge to spend the whole of September on only one efficient open model felt like a great idea at the time. Turns out not so much in practice. 2B tokens later, here’s how it went.

Where tokens went this month

Here’s the tokens distribution according to AgentsView, one of our Agentic engineering recommendations to keep tabs on AI usage:

Zooming in on the models split specifically:

The goal was to spend the whole month on GLM 5.3 Flash pictured in teal. Here’s what went well:

  • Successfully spent the first half of the month on just that model.
  • That model’s usage was well within our budget ($68, about 4kWh of energy use / 365 grams of carbon emissions).