Friend of mine works for a corp that is one of the top spenders on Claude models. He complained about these nerfs during peak demand.
Their Anthropic contact changed something and it did not happen since.
Artificial intelligence is only as useful as its intelligence. A dumb model provides no value to anybody.
Flock to a dumb model? - doubt it. Forced by the EU? - i can see it happen, sadly.
" a judge agent then attempts each task to verify that it is actually solvable "
I understand you need to verify the goal is achievable. But if the judge agent has the same goal as the training agent (solve), and both are of the same model, then aren't the judge and the training agent doing the exact same thing? What is the point then? Can someone explain this to me.
reply