Skip to main content
Close-up of a futuristic AI interface displaying GPT-5.2 performance metrics, showing it outperforming professionals on 70.9%

Editorial illustration for GPT-5.2 Outperforms Professionals on 70.9% of Tasks, OpenAI Analysis Reveals

GPT-5.2 Beats Pros on 70.9% of Tasks, OpenAI Study Shows

Analysis overhauls AI Index; GPT-5.2 beats professionals on 70.9% of tasks

Updated: 2 min read

The AI industry has stopped bragging about test scores and started clocking in. A major research group has scrapped its old benchmarks for measuring intelligence. They now look at whether models can do a job, not pass a test.

The results from this new AI Intelligence Index are blunt. According to OpenAI, GPT-5.2 now beats or ties top professionals on 70.9% of well-specified tasks across 44 occupations. Companies like Notion, Box, Shopify, Harvey, and Zoom call its performance state-of-the-art.

The question is no longer about solving a math puzzle. It is about completing the work.

On the original GDPval evaluation, GPT-5.2 beat or tied top industry professionals on 70.9% of well-specified tasks, according to OpenAI.

This pivot matters. The industry is finally valuing economic output over academic achievement. But a new evaluation called CritPT draws a hard line.

It shows these models still fail at graduate-level physics. They cannot reason scientifically.

So we have a split. AI can now reliably perform many defined jobs. It remains hopeless at the open-ended inquiry that drives real discovery.

The tool is getting very good. The scientist is not inside it.

Common Questions Answered

How did GPT-5.2 perform across professional tasks in OpenAI's evaluation?

According to OpenAI's research, GPT-5.2 beat or tied top industry professionals on 70.9% of well-specified tasks. The system demonstrated exceptional performance across 44 different occupations, showcasing advanced long-horizon reasoning and tool-calling capabilities.

Which major tech companies have observed GPT-5.2's performance?

Companies including Notion, Box, Shopify, Harvey, and Zoom have directly observed GPT-5.2's state-of-the-art performance. These organizations have noted the AI's remarkable ability to handle complex knowledge work tasks with unprecedented efficiency.

What makes GPT-5.2's performance significant for workplace dynamics?

GPT-5.2's ability to outperform professionals on nearly 71% of tasks suggests a potential transformative shift in how knowledge work is conducted. The research challenges traditional assumptions about professional competence and indicates that AI could fundamentally reshape workplace productivity and task execution.

LIVE12:53AI Agents Face Identity Challenge Before Gateway Integration