Jump to content
Wikipedia The Free Encyclopedia

FrontierMath

From Wikipedia, the free encyclopedia
AI benchmark testbed for mathematical problem solving
Part of a series on
Artificial intelligence (AI)
Glossary

FrontierMath is a test bed to benchmark [1] various artificial intelligence systems in their attempts to solve 14 bespoke[2] heretofore unexamined mathematical problems[3] (none of which are on the scale of the Millennium Problems). It was established by the non-profit research organization Epoch AI in November 2024.[4] The first such open problemof the "moderately interesting" rankto be solved was in hypergraph theory: "A Constant-Factor Lower Bound For H (n)" by GPT-5.4.[5]

See also

[edit ]

References

[edit ]
  1. Glazer, Elliot; Erdil, Ege; Besiroglu, Tamay; Chicharro, Diego; Chen, Evan; Gunning, Alex; Olsson, Caroline Falkman; Denain, Jean-Stanislas; Ho, Anson (2025年12月23日). "FrontierMath: A Benchmark for Evaluating Advanced Mathematical Reasoning in AI". arXiv:2411.04872 [cs.AI].
  2. Team, MindStudio (April 7, 2026). "What Is the Frontier Math Benchmark? Why Open Research Problems Expose True AI Reasoning". MindStudio.
  3. "FrontierMath: Open Problems - Unsolved Mathematical Challenges". Epoch AI.
  4. "AI Math Benchmarks: AI's Growing Capabilities - IEEE Spectrum". spectrum.ieee.org.
  5. Johnson, Olivia (March 14, 2026). "GPT-5.4 solves its first open math problem from FrontierMath benchmark". remio.

AltStyle によって変換されたページ (->オリジナル) /