Looks like there might be a bit of an openai vs anthropic dick waving competition over mathematical proofs. I haven’t had time to digest all of this yet because it is a bit hard for me to read, but there are a few interesting nuggets in there such as making it clear that the ai firms treat this as a marketing exercise (no surprises there). They’re also apparently trying to obscure the amount of human assistance involved (again, no surprises).
Levent had been told by Sebastien “very little human input” had been used. This turned out not to be true. Over the course of the call, as members of their team sent Sebastien corrections and details over their internal chat, it emerged that an entire team had been working on the problem, that this was one of a number of things that was tried, that work had started on the unforced problem, that the team first set the model on easier problems, including Euler, that even the prompt that had been shown to me had been written by prompting Codex, and that an insane amount of compute had been used.
The main statement is here: https://cims.nyu.edu/~tristanb/statement.pdf
Associated stuff linked from here: https://mastodon.social/@tristanbuckmaster/117233413705701198
Do note that the author made heavy use of llm assistance for their work, and does consider it to be the future of developing mathematical proofs… this isn’t human vs ai.
all 21 comments