When these came out, I thought, what is the average PC user going to get out of this? Faster image touchups, and...? You can't run an LLM chatbot of any comparison to free cloud models.
> Could’ve added more gpu or cpu or cache and people would’ve been happier.
I actually wonder about this - my regular dev workstation is a Qualcomm ARM laptop - by-and-large my bottleneck for CPU-bound tasks is thermal. Would adding more CPU or cache instead of the NPU help with performance/watt?
Mine is used for background retouches, and possibly speech recognition when dictating (Firefox had to download a model the last time I think...).
The potential is there, software just needs to catch up.
As usual Microsoft messes up marketing and abandons the possibly usable tech before it spreads. Then competitors will be successful with it and Microsoft will lag behind. Good strategy, bad execution, as standard MS practice.
Yes, you can. Of course relatively limited, but you absolutely can run local models.
Based on my own testing, the Windows drivers are a bit restricted to do some more advanced stuff, and it may be due to some security issues in the past? at least that is what I gathered from exploring this issue with GLM 5.3
35-50 low power TOPS taking 1/5 of the die area and not being used by anyone...