Close Menu
    Facebook LinkedIn YouTube WhatsApp X (Twitter) Pinterest
    Trending
    • These Were My Favorite Things Samsung Unpacked During Its 2026 Galaxy Event
    • AI minister role boosted but tech department axed in Burnham shake-up
    • Loop Engineering for RAG Question Parsing: The Small Loop That Runs Before Retrieval
    • The risk of weather data sabotage is rising
    • Hand-E Now Reaches 100 mm Without Giving Up an Ounce of Precision
    • Weight loss drug effectiveness and long term maintenance
    • Here’s what Albo’s ‘Office of AI’ means for Australian tech
    • YouTube and X Have Become ‘Gateways’ to Nudify Apps
    Facebook LinkedIn WhatsApp
    Times FeaturedTimes Featured
    Thursday, July 23
    • Home
    • Founders
    • Startups
    • Technology
    • Profiles
    • Entrepreneurs
    • Leaders
    • Students
    • VC Funds
    • More
      • AI
      • Robotics
      • Industries
      • Global
    Times FeaturedTimes Featured
    Home»News»OpenAI releases GPT-5.2 after “code red” Google threat alert
    News

    OpenAI releases GPT-5.2 after “code red” Google threat alert

    Editor Times FeaturedBy Editor Times FeaturedDecember 14, 2025No Comments3 Mins Read
    Facebook Twitter Pinterest Telegram LinkedIn Tumblr WhatsApp Email
    Share
    Facebook Twitter LinkedIn Pinterest Telegram Email WhatsApp Copy Link


    In making an attempt to maintain up with (or forward of) the competitors, mannequin releases proceed at a gradual clip: GPT-5.2 represents OpenAI’s third main mannequin launch since August. GPT-5 launched that month with a brand new routing system that toggles between instant-response and simulated reasoning modes, although customers complained about responses that felt chilly and scientific. November’s GPT-5.1 replace added eight preset “persona” choices and centered on making the system extra conversational.

    Numbers go up

    Oddly, though the GPT-5.2 mannequin launch is ostensibly a response to Gemini 3’s efficiency, OpenAI selected to not record any benchmarks on its promotional web site evaluating the 2 fashions. As an alternative, the official blog post focuses on GPT-5.2’s enhancements over its predecessors and its efficiency on OpenAI’s new GDPval benchmark, which makes an attempt to measure skilled data work duties throughout 44 occupations.

    Throughout the press briefing, OpenAI did share some competitors comparability benchmarks that included Gemini 3 Professional and Claude Opus 4.5 however pushed again on the narrative that GPT-5.2 was rushed to market in response to Google. “It is very important be aware this has been within the works for a lot of, many months,” Simo told reporters, though selecting when to launch it, we’ll be aware, is a strategic resolution.

    Based on the shared numbers, GPT-5.2 Pondering scored 55.6 p.c on SWE-Bench Pro, a software program engineering benchmark, in comparison with 43.3 p.c for Gemini 3 Professional and 52.0 p.c for Claude Opus 4.5. On GPQA Diamond, a graduate-level science benchmark, GPT-5.2 scored 92.4 p.c versus Gemini 3 Professional’s 91.9 p.c.

    GPT-5.2 benchmarks that OpenAI shared with the press.


    Credit score:

    OpenAI / Venturebeat


    OpenAI says GPT-5.2 Pondering beats or ties “human professionals” on 70.9 p.c of duties within the GDPval benchmark (in comparison with 53.3 p.c for Gemini 3 Professional). The corporate additionally claims the mannequin completes these duties at greater than 11 instances the velocity and fewer than 1 p.c of the price of human consultants.

    GPT-5.2 Pondering additionally reportedly generates responses with 38 p.c fewer confabulations than GPT-5.1, in line with Max Schwarzer, OpenAI’s post-training lead, who told VentureBeat that the mannequin “hallucinates considerably much less” than its predecessor.

    Nonetheless, we at all times take benchmarks with a grain of salt as a result of it’s simple to current them in a means that’s optimistic to an organization, particularly when the science of measuring AI efficiency objectively hasn’t fairly caught up with company gross sales pitches for humanlike AI capabilities.

    Unbiased benchmark outcomes from researchers outdoors OpenAI will take time to reach. Within the meantime, should you use ChatGPT for work duties, count on competent fashions with incremental enhancements and a few higher coding efficiency thrown in for good measure.



    Source link

    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
    Editor Times Featured
    • Website

    Related Posts

    Kalshi lawsuits dominate prediction market news today

    July 13, 2026

    Catawba Tribe Plans Two More North Carolina Casinos

    July 3, 2026

    Polymarket scrutiny, Schwab entry – latest prediction market news

    June 23, 2026

    Honolulu gambling raid in Waimakua Place nets machines

    June 13, 2026

    New Mexico lawsuit targets Kalshi sports contracts

    June 6, 2026

    Rhode Island Senate approves sports betting market expansion

    June 5, 2026

    Comments are closed.

    Editors Picks

    These Were My Favorite Things Samsung Unpacked During Its 2026 Galaxy Event

    July 22, 2026

    AI minister role boosted but tech department axed in Burnham shake-up

    July 21, 2026

    Loop Engineering for RAG Question Parsing: The Small Loop That Runs Before Retrieval

    July 19, 2026

    The risk of weather data sabotage is rising

    July 18, 2026
    Categories
    • Founders
    • Startups
    • Technology
    • Profiles
    • Entrepreneurs
    • Leaders
    • Students
    • VC Funds
    About Us
    About Us

    Welcome to Times Featured, an AI-driven entrepreneurship growth engine that is transforming the future of work, bridging the digital divide and encouraging younger community inclusion in the 4th Industrial Revolution, and nurturing new market leaders.

    Empowering the growth of profiles, leaders, entrepreneurs businesses, and startups on international landscape.

    Asia-Middle East-Europe-North America-Australia-Africa

    Facebook LinkedIn WhatsApp
    Featured Picks

    These 8 iPhone 17 Pro Max Feature Rumors Have Me Questioning My Earlier Phone Choices

    September 3, 2025

    Roam Rider twin-slide pop-up pickup camper

    April 30, 2026

    A retro British comeback with 20 hp

    September 20, 2025
    Categories
    • Founders
    • Startups
    • Technology
    • Profiles
    • Entrepreneurs
    • Leaders
    • Students
    • VC Funds
    Copyright © 2024 Timesfeatured.com IP Limited. All Rights.
    • Privacy Policy
    • Disclaimer
    • Terms and Conditions
    • About us
    • Contact us

    Type above and press Enter to search. Press Esc to cancel.