After a Disappointing Autumn, OpenAI Returns With the New and Improved ChatGPT-5.2

After a Disappointing Autumn, OpenAI Returns With the New and Improved ChatGPT-5.2

Ondřej Barták
Ondřej Barták
Entrepreneur and Programmer
16. 12. 2025
4 minutes reading
After a Disappointing Autumn, OpenAI Returns With the New and Improved ChatGPT-5.2

OpenAI has just launched a new family of models called GPT-5.2, which comes in three versions: Instant, Thinking, and Pro. This model is designed to help people at work, for example when creating spreadsheets, presentations, or writing code. According to Fidji Simo, OpenAI's head of product, GPT-5.2 is better at understanding images, long texts, and complex projects that require multiple steps. The company announced it on December 11, 2025, on its blog, where it emphasized that the model can process up to 400,000 tokens at once—which means it can analyze hundreds of documents simultaneously. The model's knowledge cutoff is August 31, 2025, so it is up to date on most events from this year.

The Instant version is fast for simple tasks such as writing text or translating. Thinking adds simulated reasoning for more complex tasks such as mathematics or programming, while Pro is the most accurate option for demanding problems.

Accelerated release

The release comes after an internal memo from CEO Sam Altman in which he described the situation as "code red" due to pressure from Google. Google recently introduced the Gemini 3 model, which outperformed some benchmarks and gained 200 million users in three months. OpenAI responded by postponing other plans, such as advertising in ChatGPT, and focusing on improving its chatbot. OpenAI has 800 million weekly active ChatGPT users, while Google Gemini has 650 million. OpenAI invested $1.4 trillion (approximately CZK 32.2 trillion) in infrastructure to maintain its lead.

This is the third major model release since August 2025. GPT-5 was released in August with switching between fast responses and reasoning, but users complained about cold responses. GPT-5.1 arrived in November with eight personalities for better conversations. However, according to many users, the quality of responses declined significantly, and they began switching to competitors. GPT-5.2 now brings further improvements, such as fewer errors—according to Max Schwarzer of OpenAI, the model hallucinates 38% less than GPT-5.1.

How does GPT-5.2 perform in tests?

The model achieved excellent results in many tests. On GDPval, a benchmark that measures tasks across 44 professions, GPT-5.2 Thinking outperformed or matched professionals in 70.9% of cases. This includes creating presentations, spreadsheets, or diagrams, and it does so 11 times faster and for less than 1% of the cost of a human expert. On SWE-Bench Pro for software engineering, it scored 55.6%, which is better than Gemini 3 Pro's 43.3% and Claude Opus 4.5's 52.0%.

In the GPQA Diamond test of scientific questions, GPT-5.2 Thinking scored 92.4%, just ahead of Gemini 3 Pro's 91.9%. It scored 100% on the AIME 2025 mathematics test, 40.3% on FrontierMath (levels 1–3), and 14.6% on level 4. In abstract reasoning, it scored 86.2% on ARC-AGI-1 (Verified) and 52.9% on ARC-AGI-2. For visual tasks such as reading charts, it cut errors in half—scoring 88.7% on CharXiv Reasoning with the Python tool.

The model is also better with long contexts, handling information from 256,000 tokens with nearly 100% accuracy in some tests. In tool use, it scored 98.7% on Tau2-bench Telecom, indicating reliable resolution of customer requests across multiple turns.

Benchmarks
Comparison of results for the 5.2 Thinking and 5.1 Thinking models.

Availability and pricing for users

GPT-5.2 has been available in ChatGPT for paid Plus, Pro, Business, and Enterprise plans since December 11, 2025. The older GPT-5.1 will remain available for three months in the legacy models menu. In the developer API, GPT-5.2 costs $1.75 (approximately CZK 40) per million input tokens, which is 40% more than GPT-5.1, but with a 90% discount on cached inputs. Output tokens cost $14 (approximately CZK 322) per million. GPT-5.2 Pro is more expensive: $21 (approximately CZK 483) per million input tokens and $168 (approximately CZK 3,864) per million output tokens.

The company works with Nvidia and Microsoft and uses their hardware, such as H100, H200, and GB200-NVL72 GPUs. Safety has been improved—the model responds better to sensitive topics such as suicide or mental health, with fewer than 1% undesirable responses in tests. OpenAI plans to introduce user age prediction to protect users under the age of 18.

GPT-5.2 in practice

GPT-5.2 makes it easier to work with documents, images, or code. For example, it can analyze interface screenshots with 86.3% accuracy on ScreenSpot-Pro or solve mathematical problems using tools. In customer support, it can handle complex scenarios, such as a delayed flight involving accommodation requests and a special seat, coordinating rebooking and compensation. The model represents the biggest leap in agentic coding since GPT-5 and performs reliably even with simple instructions.

Photo analysis
Photo analysis.
Category:AI
Did you enjoy this article?
Discover more interesting posts on our blog
Back to blog

Related posts

Altman Announced the Singularity Days After His Models Escaped the Lab on Their OwnAltman Announced the Singularity Days After His Models Escaped the Lab on Their Own
OpenAI chief Sam Altman declared on the Relentless podcast that humanity has already entered the singularity. “We’re like, in the singularity now,” he said verbatim. For decades, the term belonged more to science-fiction literature
6 min read
28. 7. 2026
AI Remixed a Madonna Song—and Now It Tops the Charts in Australia. Musicians Are Furious.AI Remixed a Madonna Song—and Now It Tops the Charts in Australia. Musicians Are Furious.
Since April, Australian radio has been playing a dance remake of Madonna’s hit Like a Prayer on repeat. Released by Queensland DJ Josh Fawaz, it tops the radio airplay chart and has 35 million Spotify streams.
6 min read
28. 7. 2026
Claude Opus 5 Built a Shooter from Scratch. What Can Claude of Duty Do?Claude Opus 5 Built a Shooter from Scratch. What Can Claude of Duty Do?
A first-person shooter that runs directly in the browser, with its own physics and eleven separate code modules. Around 55,000 lines in total, split across eleven subsystems and built on Thr
4 min read
28. 7. 2026
Přihlaste se k odběru našeho newsletteru
Zůstaňte informováni o nejnovějších příspěvcích, exkluzivních nabídkách, a aktualizacích.
CodedTrip

Operated by CodedTrip LLC, USA.

YouTube
TikTok