This repository, maintained by Google DeepMind's Superhuman Reasoning team led by Thang Luong, collects projects and datasets for advanced mathematical reasoning in AI. The main components include IMO Bench, a suite of benchmarks with short-answer problems, proof-based problems, and human grading datasets designed to evaluate robust mathematical reasoning; Aletheia, a math research agent powered by Gemini Deep Think that generates, verifies, and revises solutions on research-level mathematics problems; and LEAP, an agentic framework built in Lean 4 that decomposes complex mathematical problems into subgoals and refines formal proofs using compiler feedback and LLM reviews. All software is licensed under Apache 2.0, while other materials are licensed under Creative Commons Attribution 4.0 International.