A year ago I was trying to have these models write css for me to support mobile styles for an ERP system. There were so many quirks that you typically have with css, and it couldn't figure out how to properly get it done. I'd ask it to fix something and it'd break something else - pretty much as good as any average developer is with css.
Out of curiosity, I tried the same project yesterday with OpenAI's newest model, and it had 0 issues. It seems to have a much deeper understanding of how global styles and local styles work across different modules.
It's fun seeing how even a year ago these models seemed so capable to us, and yet they're still improving greatly. I'm still not using Anthropic. I used to change back and fourth periodically as the models would surpass each other, and get the most expensive plans, but I've settled and am happy with the capabilities of the lowest tier subscriptions now because they've progressed much faster than I've had use for them. Maybe I've gotten more efficient with language and instructing the model since a year ago. Still, it's been a fun ride and I'm excited to see what else we can do with these models as time goes on.
The easiest solution is to make an empty alias account on twitter just for the purpose of accessing links. I use an email address for this. Spam@myemaildomain.com. it forwards all emails to the address right to spam, it's great! No cease and desist letters from Mr. musk , yet !
My only problem is that by viewing tweets with your alias account is that you're still contributing to their DAU / MAU (daily/monthly active user) count. I want the service to die; their usage numbers should go to zero.
Benchmarks and cost don't really help me understand how good a model actually is. Anyone have hands on experience working with the current newest models? What work did you do, and how did the model performance ?
I have been using it since it was available in opencode Go plan.
It replaced all other models for me.
It's much less chatty than DeepSeek V4. And it feels noticeably smarter. I also use Claude at work and GLM 5.3 feels like Opus 4.8 (which I consider better than Opus 5).
It tends to be more proactive with suggestions after a task is done also.
DS4 Flash is awesome but GLM 5.3 is better despite being a bit more expensive.
Out of curiosity, I tried the same project yesterday with OpenAI's newest model, and it had 0 issues. It seems to have a much deeper understanding of how global styles and local styles work across different modules.
It's fun seeing how even a year ago these models seemed so capable to us, and yet they're still improving greatly. I'm still not using Anthropic. I used to change back and fourth periodically as the models would surpass each other, and get the most expensive plans, but I've settled and am happy with the capabilities of the lowest tier subscriptions now because they've progressed much faster than I've had use for them. Maybe I've gotten more efficient with language and instructing the model since a year ago. Still, it's been a fun ride and I'm excited to see what else we can do with these models as time goes on.
reply