College Hard Math Problems

AI’s math problem: FrontierMath benchmark shows how far technology still has to go

Artificial intelligence systems may be good at generating text, recognizing images, and even solving basic math problems—but when it comes to advanced mathematical reasoning, they are hitting a wall.

Hosted on MSN

Top AI models are failing hard at solving fresh math problems

Top artificial intelligence systems now ace many textbook-style math questions, yet they still fall apart on genuinely new problems. The gap between polished performance on familiar benchmarks and ...

Phys.org

Testing AI systems on hard math problems shows they still perform very poorly

A team of AI researchers and mathematicians affiliated with several institutions in the U.S. and the U.K. has developed a math benchmark that allows scientists to test the ability of AI systems to ...

Some results have been hidden because they may be inaccessible to you

Show inaccessible results

AI’s math problem: FrontierMath benchmark shows how far technology still has to go

Top AI models are failing hard at solving fresh math problems

Testing AI systems on hard math problems shows they still perform very poorly

Trending now