{"id":2681,"date":"2025-07-22T07:27:07","date_gmt":"2025-07-22T07:27:07","guid":{"rendered":"https:\/\/violethoward.com\/new\/google-deepmind-makes-ai-history-with-gold-medal-win-at-worlds-toughest-math-competition\/"},"modified":"2025-07-22T07:27:07","modified_gmt":"2025-07-22T07:27:07","slug":"google-deepmind-makes-ai-history-with-gold-medal-win-at-worlds-toughest-math-competition","status":"publish","type":"post","link":"https:\/\/violethoward.com\/new\/google-deepmind-makes-ai-history-with-gold-medal-win-at-worlds-toughest-math-competition\/","title":{"rendered":"Google DeepMind makes AI history with gold medal win at world’s toughest math competition"},"content":{"rendered":" \r\n
Want smarter insights in your inbox? Sign up for our weekly newsletters to get only what matters to enterprise AI, data, and security leaders.<\/em> Subscribe Now<\/em><\/p>\n\n\n\n Google DeepMind announced Monday that an advanced version of its Gemini artificial intelligence model has officially achieved gold medal-level performance at the International Mathematical Olympiad, solving five of six exceptionally difficult problems and earning recognition as the first AI system to receive official gold-level grading from competition organizers.<\/p>\n\n\n\n The victory advances the field of AI reasoning and puts Google ahead in the intensifying battle between tech giants building next-generation artificial intelligence. More importantly, it demonstrates that AI can now tackle complex mathematical problems using natural language understanding rather than requiring specialized programming languages.<\/p>\n\n\n\n \u201cOfficial results are in \u2014 Gemini achieved gold-medal level in the International Mathematical Olympiad!\u201d Demis Hassabis, CEO of Google DeepMind, wrote on social media platform X Monday morning. \u201cAn advanced version was able to solve 5 out of 6 problems. Incredible progress.\u201d<\/p>\n\n\n\n Official results are in \u2013 Gemini achieved gold-medal level in the International Mathematical Olympiad! ? An advanced version was able to solve 5 out of 6 problems. Incredible progress \u2013 huge congrats to @lmthang<\/a> and the team! https:\/\/t.co\/pp9bXF7rVj<\/p>\u2014 Demis Hassabis (@demishassabis) July 21, 2025<\/a><\/blockquote> \n\n\n\n The International Mathematical Olympiad, held annually since 1959, is widely considered the world\u2019s most prestigious mathematics competition for pre-university students. Each participating country sends six elite young mathematicians to compete in solving six exceptionally challenging problems spanning algebra, combinatorics, geometry, and number theory. Only about 8% of human participants typically earn gold medals.<\/p>\n\n\n\n The AI Impact Series Returns to San Francisco – August 5<\/strong><\/p>\n\n\n\n The next phase of AI is here – are you ready? Join leaders from Block, GSK, and SAP for an exclusive look at how autonomous agents are reshaping enterprise workflows – from real-time decision-making to end-to-end automation.<\/p>\n\n\n\n Secure your spot now – space is limited: https:\/\/bit.ly\/3GuuPLF<\/p>\n\n\n\n Google\u2019s latest success far exceeds its 2024 performance, when the company\u2019s combined AlphaProof and AlphaGeometry systems earned silver medal status by solving four of six problems. That earlier system required human experts to first translate natural language problems into domain-specific programming languages and then interpret the AI\u2019s mathematical output.<\/p>\n\n\n\n This year\u2019s breakthrough came through Gemini Deep Think, an enhanced reasoning system that employs what researchers call \u201cparallel thinking.\u201d Unlike traditional AI models that follow a single chain of reasoning, Deep Think simultaneously explores multiple possible solutions before arriving at a final answer.<\/p>\n\n\n\n \u201cOur model operated end-to-end in natural language, producing rigorous mathematical proofs directly from the official problem descriptions,\u201d Hassabis explained in a follow-up post on the social media site X, emphasizing that the system completed its work within the competition\u2019s standard 4.5-hour time limit.<\/p>\n\n\n\n We achieved this year\u2019s impressive result using an advanced version of Gemini Deep Think (an enhanced reasoning mode for complex problems). Our model operated end-to-end in natural language, producing rigorous mathematical proofs directly from the official problem descriptions \u2013\u2026<\/p>\u2014 Demis Hassabis (@demishassabis) July 21, 2025<\/a><\/blockquote> \n\n\n\n The model achieved 35 out of a possible 42 points, comfortably exceeding the gold medal threshold. According to IMO President Prof. Dr. Gregor Dolinar, the solutions were \u201castonishing in many respects\u201d and found to be \u201cclear, precise and most of them easy to follow\u201d by competition graders.<\/p>\n\n\n\n The announcement comes amid growing tension in the AI industry over competitive practices and transparency. Google DeepMind\u2019s measured approach to releasing its results has drawn praise from the AI community, particularly in contrast to rival OpenAI\u2019s handling of similar achievements.<\/p>\n\n\n\n \u201cWe didn\u2019t announce on Friday because we respected the IMO Board\u2019s original request that all AI labs share their results only after the official results had been verified by independent experts & the students had rightly received the acclamation they deserved,\u201d Hassabis wrote, appearing to reference OpenAI\u2019s earlier announcement of its own olympiad performance.<\/p>\n\n\n\n Btw as an aside, we didn\u2019t announce on Friday because we respected the IMO Board’s original request that all AI labs share their results only after the official results had been verified by independent experts & the students had rightly received the acclamation they deserved<\/p>\u2014 Demis Hassabis (@demishassabis) July 21, 2025<\/a><\/blockquote> \n\n\n\n Social media users were quick to note the distinction. \u201cYou see? OpenAI ignored the IMO request. Shame. No class. Straight up disrespect,\u201d wrote one user. \u201cGoogle DeepMind acted with integrity, aligned with humanity.\u201d<\/p>\n\n\n\n The criticism stems from OpenAI\u2019s decision to announce its own mathematical olympiad results without participating in the official IMO evaluation process. Instead, OpenAI had a panel of former IMO participants grade its AI\u2019s performance, a approach that some in the community view as lacking credibility.<\/p>\n\n\n\n \u201cOpenAI is quite possibly the worst company on the planet right now,\u201d wrote one critic, while others suggested the company needs to \u201ctake things seriously\u201d and \u201cbe more credible.\u201d<\/p>\n\n\n\n You see?<\/p> OpenAI ignored the IMO request. Shame. No class. Straight up disrespect. <\/p> Google DeepMind acted with integrity, aligned with humanity. <\/p> TRVTHNUKE pic.twitter.com\/8LAOak6XUE<\/a><\/p>\u2014 NIK (@ns123abc) July 21, 2025<\/a><\/blockquote> \n\n\n\n Google DeepMind\u2019s success appears to stem from novel training techniques that go beyond traditional approaches. The team used advanced reinforcement learning methods designed to leverage multi-step reasoning, problem-solving, and theorem-proving data. The model was also provided access to a curated collection of high-quality mathematical solutions and received specific guidance on approaching IMO-style problems.<\/p>\n\n\n\n The technical achievement impressed AI researchers who noted its broader implications. \u201cNot just solving math\u2026 but understanding language-described problems and applying abstract logic to novel cases,\u201d wrote AI observer Elyss Wren. \u201cThis isn\u2019t rote memory \u2014 this is emergent cognition in motion.\u201d<\/p>\n\n\n\n Ethan Mollick, a professor at the Wharton School who studies AI, emphasized the significance of using a general-purpose model rather than specialized tools. \u201cIncreasing evidence of the ability of LLMs to generalize to novel problem solving,\u201d he wrote, highlighting how this differs from previous approaches that required specialized mathematical software.<\/p>\n\n\n\n It wasn’t just OpenAI.<\/p> Google also used a general purpose model to solve the very hard math problems of the International Math Olympiad in plain language. Last year they used specialized tool use<\/p>
\n<\/div>
\n\n\n\n
\n<\/div>How Google DeepMind\u2019s Gemini Deep Think cracked math\u2019s toughest problems<\/h2>\n\n\n\n
OpenAI faces backlash for bypassing official competition rules<\/h2>\n\n\n\n
Inside the training methods that powered Gemini\u2019s mathematical mastery<\/h2>\n\n\n\n