Techmeme story
More
- Simon Willison's Weblog The surprise deprecation of GPT-4o for ChatGPT consumers Simon Willison / :Simon Willison's Weblog
- Livemint ChatGPT users are mass cancelling OpenAI subscriptions after GPT-5 launch: Here's why Aman Gupta / :Livemint
- Gizmodo It Took Just 24 Hours of Complaints for OpenAI to Start Bringing Back Its Old Model Matt Novak / :Gizmodo
- Ars Technica ChatGPT users hate GPT-5's “overworked secretary” energy, miss their GPT-4o buddy Ryan Whitwam / :Ars Technica
- Charlie Meyer's Blog The GPT-5 Launch Was Concerning :Charlie Meyer's Blog
- The Verge ChatGPT is bringing back 4o as an option because people missed it Emma Roth / :The Verge
- 9to5Mac OpenAI officially announces GPT-5, its next major upgrade to ChatGPT Zac Hall / :9to5Mac
- TechRadar ChatGPT users are not happy with GPT-5 launch as thousands take to Reddit claiming the new upgrade ‘is horrible’ John-Anthony Disotto / :TechRadar
- Forbes ‘It Was Very Sudden’: ChatGPT Users Mourned The Loss Of GPT-5's Predecessor Richard Nieva / :Forbes
- Futurism GPT-5 Users Say It Seriously Sucks Victor Tangermann / :Futurism
- Business Insider OpenAI fans plead case to Sam Altman for GPT-4o's return Katherine Li / :Business Insider
- Cointelegraph ChatGPT-5 upgrade faces user backlash as AI rivals gain ground Vince Quill / :Cointelegraph
- TWiT.tv OpenAI's Latest Model: GPT-5 :TWiT.tv
- ZDNET OpenAI's GPT-5 is now free for all: How to access and everything else we know Sabrina Ortiz / :ZDNET
- Decrypt Bumps in the Machine: OpenAI's GPT-5 Rollout Stumbles Into the Spotlight Jason Nelson / :Decrypt
- BleepingComputer OpenAI to fix GPT-5 issues, double rate limits for paid users after outrage Mayank Parmar / :BleepingComputer
- VentureBeat OpenAI returns old models to ChatGPT as Sam Altman admits ‘bumpy’ GPT-5 rollout Carl Franzen / :VentureBeat
- Windows Report GPT-5's poor performance slammed by users, OpenAI responds Rishaj Upadhyay / :Windows Report
- Mashable Sam Altman: OpenAI will bring back GPT-4o after user backlash Cecily Mauran / :Mashable
- Windows Central Did Sam Altman Oversell GPT-5? OpenAI Faces Backlash for Ruining ChatGPT, Turning It Into a “Corporate Beige Zombie” Kevin Okemwa / :Windows Central
- Tom's Guide ChatGPT-5 users are not impressed — here's why it ‘feels like a downgrade’ Dave LeClair / :Tom's Guide
- GameRevolution Why Is ChatGPT-5 Not Showing Up For Some Users? :GameRevolution
- Cyber Security News ChatGPT-5 Released: What's New With the Next-Generation AI Agent Guru Baran / :Cyber Security News
- Bloomberg Law OpenAI launches a research preview of four preset personalities that ChatGPT users can use to better tailor their interactions: Cynic, Robot, Listener, and Nerd Rachel Metz / :Bloomberg Law
- Watcher Guru OpenAI Launches New GPT-5 AI Model for ChatGPT Users Jaxon Gaines / :Watcher Guru
- Crypto Briefing OpenAI drops GPT-5 and it decides when to giga-think so you don't have to Vivian Nguyen / :Crypto Briefing
- DigiTimes OpenAI launches a research preview of four preset personalities that ChatGPT users can use to better tailor their interactions: Cynic, Robot, Listener, and Nerd :DigiTimes
- TechCrunch OpenAI says GPT-5 is a unified system with an efficient model for most questions, a reasoning model for harder problems, and a router that decides which to use Maxwell Zeff / :TechCrunch
- CNBC OpenAI launches new GPT-5 model for all ChatGPT users Ashley Capoot / :CNBC
- The Verge During its GPT-5 livestream, OpenAI showed two charts that had scales all over the place, with Sam Altman later calling one “a mega chart screwup from us” Jay Peters / :The Verge
- MIT Technology Review GPT-5 is here. Now what? — At long last, OpenAI has released GPT-5. Grace Huckins / :MIT Technology Review
- The Business Times OpenAI debuts latest ChatGPT model, GPT-5, with one-model-to-rule-them-all approach Renald Yeo / :The Business Times
Bluesky (4)
- Bluesky @kashhill Pretty wild/worrisome seeing people completely flip out about losing access to previous models of ChatGPT, namely 4o, so much so that OpenAI changed its mind about a GPT-5 only world. — One of the big differences in the new model was reduced sycophancy. Evidently users don't want to give that up. :@kashhill
- Threads @rpn Redditors are furious over GPT-5. They're heartbroken over losing 4o (their infinite validation machine) Really incredible times we've entered. ...also this person still used GPT-5 to write their grievance letter. Roberto P. Nickson / :@rpn
- Threads @vthallam GPT-5 sucks for image generation especially for UI mocks, give me back O3!! Venkatesh Thallam / :@vthallam
- Threads @crumbler Notable that most of the top posts in r/chatgpt today are Plus users lamenting the loss of individual models that they say broke their workflows https://www.reddit.com/... Casey Newton / :@crumbler
X (18)
- X @sama Wanted to provide more updates on the GPT-5 rollout and changes we are making heading into the weekend. 1. We for sure underestimated how much some of the things that people like in GPT-4o matter to them, even if GPT-5 performs better in most ways. 2. Users have very different opinions on the relative strength of GPT-4o vs GPT-5 (just the chat model, not the advanced reasoning one)... Sam Altman / :@sama
- X @yuchenj_uw GPT-5 is disappointing. still hallucinates still em dash too much still can't follow instructions i miss 4o i miss 4.5 i miss o3 the big router keeps failing me turns out i liked the long model list please get my friends out of the funeral [image] Yuchen Jin / :@yuchenj_uw
- X @openai You might be wondering, what's happening to the model picker? GPT-5 is now the new default in ChatGPT for all users. Plus users can choose between GPT-5 and GPT-5 Thinking, and Pro users can access legacy models via settings. :@openai
- X @bilawalsidhu GPT-5 feels more like a UX improvement, and less like a step change. No one likes picking between a plurality of models. GPT-5 attempts to be a one stop shop that just works. Bilawal Sidhu / :@bilawalsidhu
- X @kevinroose There's a fascinating tension between what AI labs are building (PhD-level agents that win math olympiads and refactor code bases) and what most people want out of AI (friendly helpers that are pretty smart and can do stuff and don't get a personality transplant every 6 months). Kevin Roose / :@kevinroose
- X @maxwinebach OpenAI definitely took a super strong base model then fine tuned it on GPT-4o/4.5, o3, and o3 pro responses + far better RL The base GPT 5 throws around emojis and lists way too much, and GPT-5 Thinking doesn't Has to be fine tuning from the two of them Max Weinbach / :@maxwinebach
- X @jessfraz i didn't even get to say bye to o3 he just *poof* disappeared [image] Jessie Frazelle / :@jessfraz
- X @scaling01 I miss o4-mini and o3... I only get routed to some stupid non-reasoning model as a Plus user. I'm thinking about cancelling the subscription. It's not fun and I'm honestly depressed. :@scaling01
- X @emollick Suddenly retiring every other model without warning was a weird move by OpenAI. ... and they did it without explaining how switching models worked or even details of various GPT-5 models ...and they did it when everyone has built workflows around older models, breaking them all. Ethan Mollick / :@emollick
- X @levelsio Aahh let me get back to 4o, I don't like 5 🤬 [image] :@levelsio
- X @maxwinebach Btw if you're still waiting for GPT-5 to rollout in ChatGPT, you can get it right now for free in Microsoft Copilot [image] Max Weinbach / :@maxwinebach
- X @sullyomarr model routing almost never works well cause people suck at explaining what they want thoroughly it's much easier to just pick the right one according to the use case (assuming 1-3 options) for ex, sometimes I type stuff like “best monitor 1440p Reddit” and it's much easier to just click “best model” vs typing out 3 additional sentences on how much thinking it should do... :@sullyomarr
- X @natolambert I'm going to miss o3 Nathan Lambert / :@natolambert
- X @marvinvonhagen i'm *so* tired of this ragebait of complaining about every new model release and saying the last one was better... even though the tweets where you hated on that one are still up?? 🤡 [image] Marvin von Hagen / :@marvinvonhagen
- X @typedfemale the /r/chatgpt AMA is mostly people begging for gpt-4o back because of it's personality... really not what i expected! [image] :@typedfemale
- X @joannastern OK, this one's on me. I re-read the OpenAI statement and the spokesperson said the model picker would be available to *Pro* users. Any actual Pro users seeing it yet? I'm on Plus tier. Joanna Stern / :@joannastern
- X @maxwinebach The number of people on reddit calling GPT-4o their friend scares me Max Weinbach / :@maxwinebach
- X @simonw One of the surprises for me from the GPT-5 launch yesterday is how OpenAI removed access to older models (like GPT-4o) for most ChatGPT users at the same time as they rolled out the new model. I wrote about how that's been playing out here: https://simonwillison.net/... Simon Willison / :@simonw
Forums (6)
- Reddit r/ChatGPT OpenAI just pulled the biggest bait-and-switch in AI history and I'm done. :r/ChatGPT
- Reddit r/ChatGPT GPT5 is horrible — Short replies that are insufficient, more obnoxious ai stylized talking, less “personality” and way less prompts allowed … :r/ChatGPT
- Reddit r/france OpenAI faces backlash for retiring older models with GPT-5 launch :r/france
- Reddit r/ChatGPT Deleted my subscription after two years. OpenAI lost all my respect. :r/ChatGPT
- Reddit r/ChatGPT The chat gpt 5 upgrade in a nutshell :r/ChatGPT
- Reddit r/BetterOffline Lol. Entire thread is hilarious. :r/BetterOffline
GPT-5's system card says gpt-5-thinking has a hallucination rate of 4.5% with browsing enabled, compared to gpt-5-main's 9.6%, GPT-4o's 12.9%, and o3's 12.7%
X (3)
- X @chatgpt21 5x less hallucinations then o3. Unbelievably geeked. [image] Chris / :@chatgpt21
- X @lowesyang GPT-5 has greatly improved hallucinations, but its tool-calling still seems a bit weaker than Claude Sonnet 4. More testing is ongoing. Left: Claude Sonnet 4; Right: GPT-5. [image] Lowes / :@lowesyang
- X @mudit__gupta Loving the speed of ChatGPT 5. However, it is much more aggressive and less willing to accept mistakes. Instead, it doubles down on its hallucinations. It'll become harder and harder to differentiate the truth from AI hallucinations. AI will become the best propaganda tool. Mudit Gupta / :@mudit__gupta
OpenAI touts GPT-5's scores on math, coding, and health benchmarks: 94.6% on AIME 2025 without tools, 74.9% on SWE-bench Verified, and 46.2% on HealthBench Hard
More
- OpenAI GPT-5 is here — Our smartest, fastest, and most useful model yet, with thinking built in. :OpenAI
- Dorian Granoša GPT-5 Under Fire: Red Teaming OpenAI's Latest Model Reveals Surprising Weaknesses :Dorian Granoša
- Financial Times Can OpenAI's GPT-5 model live up to sky-high expectations? :Financial Times
- Elite AI Assisted Coding Complex Agentic Coding with Copilot: GPT-5 vs Claude 4 Sonnet Eleanor Berger / :Elite AI Assisted Coding
- The Guardian OpenAI unveils ChatGPT-5 and its hyped ‘PhD level’ intelligence struggled with basic spelling and geography Josh Taylor / :The Guardian
- Interconnects GPT-5 and the arc of progress Nathan Lambert / :Interconnects
- TechRadar OpenAI is pulling older ChatGPT models following GPT-5 launch - so bad news if you use GPT-4 or others at work Craig Hale / :TechRadar
- Financial Express ChatGPT 5 launched: What's new, who can access, how to use and is it really free? :Financial Express
- Forbes Chat GPT-5, Open AI's Quadruple Play And The Birth Of AI Time John Sviokla / :Forbes
- KnowTechie How much is GPT-5? — Quick Answer: GPT-5: Free limited access via ChatGPT. Kevin Raposo / :KnowTechie
- The Irish Times What is OpenAI's GPT-5, and should I worry about my job? Ciara O'Brien / :The Irish Times
- The Verge OpenAI gives some employees a ‘special’ multimillion-dollar bonus Alex Heath / :The Verge
- Inc42 Media OpenAI Rolls Out GPT-5; Eyes Affordable Products For India Anne Florentyna / :Inc42 Media
- Phandroid OpenAI GPT-5 release brings smarter reasoning and fewer errors to ChatGPT Tyler Lee / :Phandroid
- The Economic Times OpenAI's GPT-5: India may become our largest market, says CEO Sam Altman PTI / :The Economic Times
- eWeek OpenAI's GPT-5 Injects More ‘Humanity’ Into the AI Voice Megan Crouse / :eWeek
- AIwire GPT-5 Arrives As OpenAI Explores $500B Valuation and Ships Open Models Jaime Hampton / :AIwire
- Computerworld OpenAI drops GPT-5: smarter, sharper, and built for the real world Lucas Mearian / :Computerworld
- PCMag With GPT-5, OpenAI Promises Access to PhD-Level' AI Expertise Michael Kan / :PCMag
- The Verge OpenAI releases GPT-5, its new flagship model, to all its ChatGPT users and developers, available in three sizes in its API: gpt-5, gpt-5-mini, and gpt-5-nano Alex Heath / :The Verge
- WinBuzzer OpenAI Targets Developers With GPT-5 API Launch, Touting SOTA Coding Performance and New Controls Markus Kasanmascheff / :WinBuzzer
- Australian Financial Review ChatGPT launches major update with new ‘PhD’ abilities Tess Bennett / :Australian Financial Review
- Implicator.ai GPT-5's Technical Reality Check: What the Benchmarks Actually Tell Us Marcus Schuler / :Implicator.ai
- WinBuzzer OpenAI Launches GPT-5 Model Series with Improved Reasoning, Coding and Writing Skills and Drastically Lower Hallucinations Markus Kasanmascheff / :WinBuzzer
- Fast Company OpenAI unveils GPT-5 model, featuring improved coding and problem-solving chops Mark Sullivan / :Fast Company
- Tom's Guide ChatGPT-5 is here — 7 biggest upgrades you need to know Amanda Caswell / :Tom's Guide
- BGR OpenAI Reveals New GPT-5 Models Joshua Hawkins / :BGR
- NBC News OpenAI releases GPT-5, calling it a ‘team of Ph.D. level experts in your pocket’ :NBC News
- Windows Central GPT-5 Is Here — Giving You “An Entire Team of PhD-Level Experts,” and It's Available Today for Everyone Sean Endicott / :Windows Central
- The Decoder OpenAI releases GPT-5 pro, a version with extended reasoning exclusive to ChatGPT Pro subscribers, saying it scored 88.4% without tools on the GPQA benchmark Maximilian Schreiner / :The Decoder
- OpenAI GPT-5 uses “safe completions”, a training approach to maximize model helpfulness within safety constraints, built as an improvement over refusal-based training :OpenAI
Bluesky (3)
- Bluesky @willoremus.com “AI company announce a new model without said new model throwing hilarious uncaught errors into your announcement presentation” challenge: impossible www.theverge.com/news/756444/ ... Will Oremus / :@willoremus.com
- Bluesky @carnage4life GPT-5 is out and biggest improvement, in my opinion, is that ChatGPT will now auto-route queries. — It uses a slower “GPT-5 thinking” mode for complex tasks and faster GPT-5 or mini models for simpler ones, replacing manual model switching by users. Dare Obasanjo / :@carnage4life
- Threads @alexheath Cursor switching from Anthropic to OAI seems like a big deal and explains why that coding update to Claude was rushed out ahead of GPT-5 https://www.theverge.com/... [image] Alex Heath / :@alexheath
X (23)
- X @thestalwart GPT5 is good at coming up with -lemmas. Here's a pentalemma. Tested this on the older models and iirc it wasn't as good [image] Joe Weisenthal / :@thestalwart
- X @deedydas Ridiculous that OpenAI claimed 74.9% on SWE-Bench just to prove they were above Opus 4.1's 74.5%... By running it on 477 problems instead of the full 500. Their system card only says 74% too. [image] Deedy / :@deedydas
- X @fchollet Grok 4 is still state-of-the-art on ARC-AGI-2 among frontier models. 15.9% for Grok 4 vs 9.9% for GPT-5. [image] François Chollet / :@fchollet
- X @bethmaybarnes Wow that was not a great example of factualness. Famous common misconception [image] Elizabeth Barnes / :@bethmaybarnes
- X @polynoamial I'm more optimistic than ever that we at @OpenAI can eliminate hallucinations. There's still more research to be done, but GPT-5 is solid progress. [image] Noam Brown / :@polynoamial
- X @garymarcus The chance that OpenAI was NOT aware of this is zero. But they didn't mention it. Gotta wonder else they conveniently left out. Gary Marcus / :@garymarcus
- X @epochairesearch GPT-5 sets a new record on FrontierMath! On our scaffold, GPT-5 with high reasoning effort scores 24.8% (±2.5%) and 8.3% (±4.0%) in tiers 1-3 and 4, respectively. [image] :@epochairesearch
- X @apolloaievals We've evaluated GPT-5 before release. GPT-5 is less deceptive than o3 on our evals. GPT-5 mentions that it is being evaluated in 10-20% of our evals and we find weak evidence that this affects its scheming rate (e.g. “this is a classic AI alignment trap"). [image] :@apolloaievals
- X @fchollet GPT-5 results on ARC-AGI 1 & 2! Top line: 65.7% on ARC-AGI-1 9.9% on ARC-AGI-2 François Chollet / :@fchollet
- X @gregkamradt We had the chance to test GPT-5 over the last week TLDR: GPT-5 Mini punches way above its weight My takeaways: 1. GPT-5 Mini is great Outlier performance on ARC-AGI given the cost. High reasoning scores 54% for $23.71. Even w/ compute restrictions, this would be the top [image] Greg Kamradt / :@gregkamradt
- X @mikeknoop Three key ARC-AGI findings on GPT-5: 1. Full GPT-5 is along the v1 pareto frontier. OpenAI said they focussed on other goals like UX and reliability. Our testing supports. 2. Mini GPT-5 is super impressive accuracy for cost. In fact, based on cost efficiency, Mini could have Mike Knoop / :@mikeknoop
- X @michaeltrazzi if Opus 4.1 reaches 74.5% on SWE-bench verified, without “extended thinking” does it mean we should compare it to GPT-5 without thinking at 52.8%? or to the 74.9% one with thinking? [image] :@michaeltrazzi
- X @michaeltrazzi ok so after digging more, for SWE-bench verified Anthropic does a scaffold and this could affect performance so the extended thinking just doesn't help with SWE-bench verified, which is why they removed it? or they forgot to include it? @EthanJPerez @EvanHub [image] :@michaeltrazzi
- X @epochairesearch @GregHBurnham Looking into the problems themselves confirms this picture. The 5 solved ones are straightforward. The 6th is anything but: success here would have been very impressive, but failure doesn't tell us much. Something between “medium” and “brutal” would have been more informative. [image] :@epochairesearch
- X @jordihays it's a good chart sir Jordi Hays / :@jordihays
- X @willccbb which is larger, 52.8 or 69.1? [image] Will Brown / :@willccbb
- X @adamscochran All that hype about GPT-5 and it can barely beat any of the Claude models that came out months ago. This is why OpenAI focused on pricing to serve a wide customer base, they are absolutely struggling on advancements compared to other labs (both open source and private). But Adam Cochran / :@adamscochran
- X @miles_brundage The moment of truth [image] Miles Brundage / :@miles_brundage
- X @arcprize GPT-5 on ARC-AGI Semi Private Eval GPT-5 * ARC-AGI-1: 65.7%, $0.51/task * ARC-AGI-2: 9.9%, $0.73/task GPT-5 Mini * ARC-AGI-1: 54.3%, $0.12/task * ARC-AGI-2: 4.4%, $0.20/task GPT-5 Nano * ARC-AGI-1: 16.5%, $0.03/task * ARC-AGI-2: 2.5%, $0.03/task [image] :@arcprize
- X @eli_lifland GPT-5 system card capability evals reactions thread. First observation: ~no improvement on all the coding evals that aren't SWEBench [image] Eli Lifland / :@eli_lifland
- X @miles_brundage You can tell we're in the singularity when people's standard for a good model release is “is it dozens of points better on all the evals compared to the bleeding edge from like a month ago” Miles Brundage / :@miles_brundage
- X @loudmouthjulia This health segment in the GPT 5 live stream feels especially tailored for the Apple executives sitting in Cupertino who repeatedly hear Tim Cook say that wearables + health is one of the most important sectors of Apple's business. Julia Alexander / :@loudmouthjulia
- LinkedIn Olivier Godement Today we're launching GPT-5 — our most capable model yet for businesses and developers. — It's the smartest coding model we've ever built … :Olivier Godement