• About Us
  • Privacy Policy
  • Disclaimer
  • Contact Us
AimactGrow
  • Home
  • Technology
  • AI
  • SEO
  • Coding
  • Gaming
  • Cybersecurity
  • Digital marketing
No Result
View All Result
  • Home
  • Technology
  • AI
  • SEO
  • Coding
  • Gaming
  • Cybersecurity
  • Digital marketing
No Result
View All Result
AimactGrow
No Result
View All Result

OpenAI releases GPT-5.2 after “code crimson” Google menace alert

Admin by Admin
December 14, 2025
Home Technology
Share on FacebookShare on Twitter


In making an attempt to maintain up with (or forward of) the competitors, mannequin releases proceed at a gentle clip: GPT-5.2 represents OpenAI’s third main mannequin launch since August. GPT-5 launched that month with a brand new routing system that toggles between instant-response and simulated reasoning modes, although customers complained about responses that felt chilly and medical. November’s GPT-5.1 replace added eight preset “persona” choices and centered on making the system extra conversational.

Numbers go up

Oddly, although the GPT-5.2 mannequin launch is ostensibly a response to Gemini 3’s efficiency, OpenAI selected to not checklist any benchmarks on its promotional web site evaluating the 2 fashions. As a substitute, the official weblog put up focuses on GPT-5.2’s enhancements over its predecessors and its efficiency on OpenAI’s new GDPval benchmark, which makes an attempt to measure skilled information work duties throughout 44 occupations.

Through the press briefing, OpenAI did share some competitors comparability benchmarks that included Gemini 3 Professional and Claude Opus 4.5 however pushed again on the narrative that GPT-5.2 was rushed to market in response to Google. “You will need to word this has been within the works for a lot of, many months,” Simo instructed reporters, though selecting when to launch it, we’ll word, is a strategic choice.

In line with the shared numbers, GPT-5.2 Pondering scored 55.6 p.c on SWE-Bench Professional, a software program engineering benchmark, in comparison with 43.3 p.c for Gemini 3 Professional and 52.0 p.c for Claude Opus 4.5. On GPQA Diamond, a graduate-level science benchmark, GPT-5.2 scored 92.4 p.c versus Gemini 3 Professional’s 91.9 p.c.

GPT-5.2 benchmarks that OpenAI shared with the press.
GPT-5.2 benchmarks that OpenAI shared with the press.


Credit score:

OpenAI / Venturebeat


OpenAI says GPT-5.2 Pondering beats or ties “human professionals” on 70.9 p.c of duties within the GDPval benchmark (in comparison with 53.3 p.c for Gemini 3 Professional). The corporate additionally claims the mannequin completes these duties at greater than 11 occasions the velocity and fewer than 1 p.c of the price of human specialists.

GPT-5.2 Pondering additionally reportedly generates responses with 38 p.c fewer confabulations than GPT-5.1, in response to Max Schwarzer, OpenAI’s post-training lead, who instructed VentureBeat that the mannequin “hallucinates considerably much less” than its predecessor.

Nevertheless, we at all times take benchmarks with a grain of salt as a result of it’s straightforward to current them in a approach that’s constructive to an organization, particularly when the science of measuring AI efficiency objectively hasn’t fairly caught up with company gross sales pitches for humanlike AI capabilities.

Impartial benchmark outcomes from researchers exterior OpenAI will take time to reach. Within the meantime, if you happen to use ChatGPT for work duties, anticipate competent fashions with incremental enhancements and a few higher coding efficiency thrown in for good measure.

Tags: alertCodeGoogleGPT5.2OpenAIRedReleasesThreat
Admin

Admin

Next Post
Diablo 4’s subsequent growth is bringing two new lessons to the sport, one in all which you’ll be able to play proper now

Diablo 4's subsequent growth is bringing two new lessons to the sport, one in all which you'll be able to play proper now

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Recommended.

TurboLearn AI Evaluate: The Final Examine Hack for College students

TurboLearn AI Evaluate: The Final Examine Hack for College students

June 4, 2025
The GPT-5 rollout has been a giant mess

The GPT-5 rollout has been a giant mess

August 20, 2025

Trending.

Backrooms director Kane Parsons explains the birds, the portals, and his sensible results

Backrooms director Kane Parsons explains the birds, the portals, and his sensible results

May 31, 2026
100 Most Costly Key phrases for Google Advertisements in 2026

100 Most Costly Key phrases for Google Advertisements in 2026

January 13, 2026
Resident Evil followers have adopted a Love & Deepspace character because the son of Leon S. Kennedy and one in every of his potential spouses

Resident Evil followers have adopted a Love & Deepspace character because the son of Leon S. Kennedy and one in every of his potential spouses

April 4, 2026
AI & data-driven Starbucks – Deep Brew

AI & data-driven Starbucks – Deep Brew

May 18, 2026
Parental Lock Code Puzzle Defined

Parental Lock Code Puzzle Defined

July 27, 2025

AimactGrow

Welcome to AimactGrow, your ultimate source for all things technology! Our mission is to provide insightful, up-to-date content on the latest advancements in technology, coding, gaming, digital marketing, SEO, cybersecurity, and artificial intelligence (AI).

Categories

  • AI
  • Coding
  • Cybersecurity
  • Digital marketing
  • Gaming
  • SEO
  • Technology

Recent News

Anthropic and DOD Set to Face Off Thursday Over Blacklisting

Anthropic and DOD Set to Face Off Thursday Over Blacklisting

July 27, 2026
Combating Soul’s Most Fashionable Characters in Open Beta

Combating Soul’s Most Fashionable Characters in Open Beta

July 27, 2026
  • About Us
  • Privacy Policy
  • Disclaimer
  • Contact Us

© 2025 https://blog.aimactgrow.com/ - All Rights Reserved

No Result
View All Result
  • Home
  • Technology
  • AI
  • SEO
  • Coding
  • Gaming
  • Cybersecurity
  • Digital marketing

© 2025 https://blog.aimactgrow.com/ - All Rights Reserved