$10/$50 is incredibly expensive compared to Chinese models which are cents.
I think they’re really going to struggle selling these models long-term. My company is already massively cutting down on access because they’ve realised most people don’t actually produce any value using it. All the tokenmaxers have ruined it for the rest of us now that accounting have seen the costs.
Assuming what GP said is accurate (tokenmaxxers not delivering more value for the usage), the analogy is more like realizing the seniors you hired aren't actually producing any better results than juniors you could get for a quarter of the cost.
An official motto of a president of one of the biggest software companies in Poland (Comarch) was "you can replace any experienced engineer with a finite amount of sutdents... for half of the cost".
Not really comparable IMO. Astra and Fable are not the every day workhorse you reach for to do basic tasks (unless your company has fuck you-money), they’re the tool you break out when you need the absolute strongest performance. There are plenty of tasks where finding and fixing one or two extra edge cases saves the business a lot of money, even if the cost is high. The best example would be scanning for vulnerabilities, if these models weren’t kneecapped in that area.
We have ChatGPT Pro at work and I usually use Terra medium/high and only bring out Sol High when the big or feature actually requires "thinking"/complex behavior. This has worked pretty well for me and it's very token efficient
How much code are you shipping in a day? I find I can pretty comfortable use Sol high most of the day and stay within the 5 hour limit. I’ve got too many meetings to allow me time to write code continuously for an entire day. I usually finish about one ticket a day and then review 1-3 tickets for my colleagues.
We have a very lean development process and are in early stages of releasing the software, so almost no meetings and 90% of my time is development time (obv I talk with colleagues etc in that time). For example working with yocto eats through tokens like crazy since codex has to work with a pretty big codebase and look up tons of stuff.
I think they’re really going to struggle selling these models long-term. My company is already massively cutting down on access because they’ve realised most people don’t actually produce any value using it. All the tokenmaxers have ruined it for the rest of us now that accounting have seen the costs.