That seems impossible (unless maybe you meant to write "predates", and even then it seems to only be about 3 years).
The comic was published on 9 February 2007 [0].
The PS3 was first released in November 2006. I haven't watched the video yet, but its description says "2010 saw the first hacks for the Playstation 3".
Even though it's called "T9 Speed Test", it seems to default to T9 being off, and you have to actually click on the T9 link at the top right to enable it. Really confusing at first!
It's not real Python. It's essentially a made-up language that resembles a subset of Python.
That said, if I was building a small ugly toy Python implementation, I would probably also not implement long integers. Or the power operator. At least not in the first iteration.
> But there are a handful of designated platforms that aren’t paying anything in the first round as they reported a loss during the preceding financial year — including Amazon, Pinterest, Snapchat and Wikimedia.
B) A mathematician working for Anthropic solved a problem mathematicians have been working on for more than a century, and then credited it to Claude for PR purposes
If you believe B is more likely, why would you then believe a proof in the form of a chat log, when said chat log could itself have been faked by Anthropic way more easily than solving the mathematical problem in the first place?
While I agree that we need the inputs to properly evaluate what this means for LLM capabilities, I don't really believe that the amount of knowledge input matters much for the overall significance of the result.
These kinds of results are interesting for LLMs because mathematicians have been working on them for decades. If the result doesn't already exist, there's no way it's in the training data, and if mathematicians have been unsuccessfully tackling the problem for decades, it is believable that the use of a new tool made the result possible, even if guided by a great mathematician.
Yes, of course, that would be interesting info to have. I just mean that from what we have already we can reasonably infer that the LLM played a role in the result being obtained.
If I am not mistaken, all of the flurry of novel results has come from existing mathematicians. This makes me suspect that the models aren't at the level where just any layman can get results. They require a skilled human in the loop to keep them on the rails and to properly explore the solution space.
reply