Hong: DeepMind's Formal Math Slowdown Post-AlphaProof Was Non-Technical
“After AlphaProof, kind of like, we didn't see a lot of the formal math you know, results or kind of progress from Google DeepMind, and that's actually because of reasons that are not necessarily technical.”
Tenev: Harmonic surpassed Google AlphaProof's capabilities within one year
“Alpha Proof was the first AI model to get a silver medal, to achieve silver medal performance at the IMO last year.
But Alpha Proof they did not announce a gold this year.
So, yeah, we were we were, the Harmonic team was excited about that, that, you know, i…”
DeepMind Abandoned AlphaProof to Run Gemini End-to-End for IMO Math
“We wanted to try to, like, use, actually use Gemini as an end-to-end model. Basically, no, no second system with alpha proof. No second system. In, text out.”
DeepMind math AI achieves Olympiad silver medals despite making basic errors
“You have systems some systems that we work on, like alpha proof, alpha geometry, that are getting, you know, silver medals in maths olympiads, which is fantastic, but on the other hand, Some of our systems, those same systems are still making some fairly basic…”
Guo: DeepMind's AlphaProof solved four of six 2024 IMO problems
“Alphaproof had this really amazing results of solving four of the six problems this year.”
Hubert: Formal math proofs enable self-improving reinforcement learning loops
“The advantage of that is that once kind of the proof is complete then you know, the machine would give you a signal back to say, yes, your proof is correct or not. And so we could search for kind of correct proofs. Once we find a correct proof, we can learn fr…”
Mehta: AlphaProof Is Strongest in IMO Algebra and Number Theory
“So the IMO problems have come in four categories. So there's algebra number theory, combinatrix, and geometry. The two that it's strongest at are algebra and number theory. It's relatively weaker at combinatrix, although it can do quite good at some IMO combin…”
Mehta: AlphaProof Uses Test-Time RL on Problem Variations to Find Proofs
“One of the ways in which it navigates this massive search space is via an idea that we came up with, which we call test time RL. So this is an idea where, like, let's say you're confronted with a problem that you don't know how to solve. And you know, you can …”
Sartran: AlphaProof cannot perform theory building or invent new theories
“Maybe the main thing that Alphaproof doesn't do is theory building. It doesn't invent its own theories.”
Mehta: AlphaProof's RL scaling and test-time compute generalize across domains
“Some of the sort of tech we developed here of like, you know, like scaling RL and like figuring out how to spend a lot of inference time compute stuff like this feels like it's Quite generally applicable to many other problems.”
Sartran: AlphaProof develops an alien style by learning from self-discovered proofs
“The way alpha proof currency operates there is that it discovers its own proofs, and when they are valid, it learns from them and develop its own style which has been commented upon as looking, yeah, quite, quite alien.”
Sartran: Translating all known human proofs would save years of compute
“With more supervised data, we can avoid the exploration problem, and we could translate all the proofs that are known to man, and that certainly would make the agent much better, and that would save us years of compute for sure.”
Hubert: AlphaProof reached high school level but cannot rival Terence Tao
“And to be honest with you know, like kind of we, at the moment we can't rival at all with someone like Terry Tao. We, I think we demonstrated that what we've demonstrated is that we can learn general mathematics almost from scratch and arrive at kind of an imp…”
Mehta: Fields Medalist Tim Gowers failed to find AlphaProof's IMO construction
“Tim Gowers, who was one of our judges and is also a fields medalist. Tried this question for a couple hours and he couldn't find the construction for function that had this property”