Actually, on further research, what I have wrong here is that Brockman was the largest single donor to the MAGA PAC after the election, contributing $25 million in September 2025 [1].
Is your main concern privacy? I've been using the zcode of and on for a few months now and it's definitely improved over what it was back in the spring. I also use DeepSeek's harness. Not sure which I prefer at this point. Used to be I preferred DSH, but zcode has some features I like over DSH.
I just looked it up: apparently they silently uploaded your entire repo with the entire git history to the cloud. A feature not controlled by any settings, operating and retrying silently.
Their excuse was that it was a file indexing optimisation and they forgot to create a toggle or setting around it. As an apology they open sourced the harness, but wiped out the history.
It sounds to me like a good distillation technique if the project has hundreds of commits written by Claude and detailed commit descriptions. Also for completeness Grok did the exact same thing two months ago.
I've got some Tang nano and Tang primer boards (9k, 20k and 25k LUTs). The 9K has a cute little SPI LCD display. Claude has been able to interface with the boards via USB (it setup a UART on the FPGA to do so). We've gotten a GoL as well as a Pong game (where paddles are controlled by Tsetlin machines) and a MicroCNN that does MNIST classification. For the latter we were having trouble getting correct answers back from the CNN (accuracy was very poor which didn't match simulation). Claude created an on-chip logic analyzer and doggedly troubleshot the problem (turned out to be some kind of problem in the yosys synthesis tool - we had to use -noalu to work around). Anyway, I was blown away by this. Claude was communicating with the FPGA board and analyzing the data coming back from the internal logic analyzer.
He must actually like farming because he didn't have to become a farmer. Of course, knowing that the family fortune has your back even if the farm doesn't work out probably helps.
I've also got a strix halo box. 30tok/s would we usable, but I wonder how it compares to the Qwen3.8-Flash-next - I get about 40tok/s running that on Halogen and it feels like using Claude 4.6.
Vulkan, I have never used ROCm on it but have been debating since the latest big update. How is your prefill? Do you hit over 1K? If it’s 1000K prefill, and 40 TG, I might have to try this over the weekend. Also, can you fit 128K without offload the ngram onto SSD?
I have not measured pre-fill, but it's said to be around 1000.
It feels very snappy and unlike my experience with running 27B models the performance stays pretty flat even as the context increases. Unfortunately, we don't know how Halogen is doing this because it's closed source, but I think AMD should offer that guy some $$$ because he's done a lot of good work getting more performance out of Strix Halo.
I suspect this is changing. The DSA branch especially wants all contracts with Palantir eliminated.
reply