GLM-5.3-Flash is now the cheapest way to reach Artificial Analysis Intelligence Index 52.3 to 57.5, at 8.7¢ per task, and displaced Muse Spark 1.2, Grok 4.5, Gemini 3.7 Flash, and DeepSeek V4 Pro from the Pareto frontier.
Author here. I made these plots because I had been searching for them for months and never found quite what I wanted: how the cheapest way to reach a fixed capability level has moved over time. Artificial Analysis publishes enough data to reconstruct it. If someone knows of a source that already tracks this, with historical prices, please share.
It blows my mind how fast models are getting better and this is the first article I’ve seen that shows just that and leaves almost no room for disagreement. Well done.
I wonder if it might drive the point even further if the graph scales were linear? Or maybe the progress has been so great that this would make the graphs unreadable?
advance card: https://catalystneuro.com/llm-cost-frontier/images/advances/...
tracker: https://catalystneuro.com/llm-cost-frontier/