Tech article
One month coding with GLM 5.3 Flash
No preview is available. Read the original article for the full story.
Hacker News | Oct 2, 2026 | ThibWeb
Automated excerpt
I chose the 'wrong' model for the prototype, and we spent 450M tokens / $150 / 5kWh of energy use almost overnight. This is particularly essential as we start to benchmark models’ performance on Wagtail tasks, where we need data across a wide range of models. A viable target is probably that the majority of AI inference work should be done with such efficient models, measured in cost or energy use rather than meaningless tokens.
Selected automatically from source text; not independently written or fact-checked. Read the original for full context.