Global Big Data Conference

OpenAI’s o3 AI model scores lower on a benchmark than the company initially implied Posted on : Apr 21 - 2025

A discrepancy between first- and third-party benchmark results for OpenAI’s o3 AI model is raising questions about the company’s transparency and model testing practices.

When OpenAI unveiled o3 in December, the company claimed the model could answer just over a fourth of questions on FrontierMath, a challenging set of math problems. That score blew the competition away — the next-best model managed to answer only around 2% of FrontierMath problems correctly. View More

Get the

Global Big Data Conference

Newsletter

Acknowledgement

Global Big Data Conference

Industry News Details

OpenAI’s o3 AI model scores lower on a benchmark than the company initially implied Posted on : Apr 21 - 2025