> It's avoiding two distinct vowel sounds in succession
That implies that English added a consonant there to avoid double vowels, but it's the other way around: it used to have "n" always, same as other Germanic languages, and then it was dropped to avoid two consonants in a row
> So the (PCI-E) bandwidth strongly affects time to first token
On dedicated inference hardware I'd expect model weights to never leave the RAM, and you'd probably load them on startup before even starting to serve requests
That's probably a bit pedantic. Whether you draw a magic line with dinosaurs on one side and (avian) dinosaurs/birds on the other or just say that birds continued on from the avian dinosaur clade is pretty definitional.
I do. But the difference is: do I go around bragging that my agent worked for 88h to solve a problem? Where is the credit coming from? Is it coming from the: I was the first one to think about throwing a prompt: "Solve Riemann Hypothesis" and it turns out that by luck of the non-deterministic behavior of the agent, it got right? Wow, that is a lot of merit really. Congrats...
Sorry the sarcasm, but really your point makes absolutely no sense.
It is completely different to design something with AI and then execute, validate, evolve vs just prompt it machine-g-brrrr style and get a result.
This brings an important question. Nowadays I don't write code, I review code, I review systems behavior and get paid for it. Will that be the same for math researchers? Their prompt/problem is already well-posed out there. Ours, in the day-to-day, are not. Will the first one to verify AI work get the credit? or is it going to be the dumdum that types a simple prompt and has the compute to run it for 21321 hours? I absolutely don't get your point here. Or you are just rage baiting
How many humans can you name who really did that? Like, in the whole history of our species
reply