But I don't want to use your CLI. I already have my own harnesses and workflows. The friction is too high to "just try out" a new model like this. It would be preferable if I can evaluate it over, say, open router like all the other models and then decide from there if it's worth downloading a bespoke tool chain for only 1 lab's models
It’s preferable to keep inference capacity available for users using main Devin products than openrouter atm. Might change in the future. Even OAI is cutting off new plan signups to keep up with demand.
> SWE-2 is free to use for users like yourself for the next month, and almost all usage should be supported via our CLI (https://docs.devin.ai/cli)
Thanks for pointing that out; I wouldn't have found out about it otherwise. I bought the 20$ subscription and have been using SWE-2 for a few days now, and it's actually good. I don't think it's Fable-class, but SWE-2 Max feels substantially better than K3 and on par with Opus 5 Max.
I just gave it a try and it doesn't appear to be free, it used up some of my on demand usage. It does say 75% off though. Seems like for Pro subscribers SWE-1.7 is free, maybe SWE-2 is free for them?
That list ordering drives me nuts. What's up with 1024?!
And how come much larger numbers have already been solved? Based on that information one cannot strictly assume that the current solution required improvements to the strategy or hardware, no?
The ordering is stupid because some of them are named by the number of bits and some of them are named by the number of decimal digits. They are in order of size and this is the largest one so far, despite the confusing names.
reply