Found this announcement interesting since allegedly OpenAI is retiring their Terra tier. I think for everyday work, two models with various thinking efforts seem enough, plus some frontier level model like Fable or Astra to coordinate.
At introduction, Terra was a good mid-tier model. Terra [high] was on the pareto frontier of DeepSWE's score over cost, if only ever so slightly.
When I had Sol orchestrate Luna and Terra as implementation agents, Sol was a lot happier with what Terra produced and would find far fewer issues than what was implemented by Luna.
But a few weeks after introduction, OpenAI slashed Luna's cost by 80% and Terra's only by 20%. Only then did it become uneconomical to run Terra and its reason to exist stopped.
I'd be really curious to see benchmarks of haiku vs lower effort on bigger models. My own evals found Fable 5.1 at low to be better than Opus 5 on high.
Nice. I was starting to think that Haiku got abandoned.