2016Program paper
Research
Our research focuses on machine reasoning and artificial general intelligence, with mathematics as our primary test case for abstract thinking capabilities.
- Autonomous mathematics
- Frontier-model benchmarking
- Lean formalization
- Research-level reasoning
Autoresearch
Browse AI-generated research papers with their abstracts, publication dates, and transparent verification records.
- Lean verification
- Human review
- LLM verification
27 listed works
Public papers, datasets, repositories, samples, and linked Lean formalizations.
01
Research program
AI for Mathematics
We treat mathematics as the next game at which machines would excel, serving as a test case for machine reasoning capabilities.
2025Paper
2026Benchmark paper
2026System paper
2026Paper
2026Benchmark paper
02
Research program
AI-powered Mathematics
Research level mathematics we have done while testing reasoning capabilities of various LLMs. 16 open Erdos problems claimed (and some formally verified in Lean) with many partial results
Mar 2026Full solution
Erdos Problem 1148
formalization in Lean with UlamAI Prover.
Apr 2026Full solution
Apr 2026Full solution
Apr 2026Full solution
Apr 2026Full solution
Apr 2026Full solution
Apr 2026Full solution
Apr 2026Full solution
Apr 2026Full solution
Apr 2026Full solution
Apr 2026Full solution
Apr 2026Full solution
Apr 2026Full solution
Apr 2026Full solution
Apr 2026Full solution
Apr 2026Full solution
May 2026Full solution
Jul 2026Counterexamples
03
Research program
Machine Reasoning Theory
Research-level reasoning of large language models.
Try a broader search or select a different research program.
From program to proof.
The archive brings together the long-running DeepAlgebra program, AI-generated research mathematics, benchmarks, formalization work, and reasoning datasets in one coherent index.
