14 ms·
Retiring GPT-4o, GPT-4.1, GPT-4.1 mini, and OpenAI o4-mini in ChatGPT
- bakigul 8mo agoThis wasn't good at all. This 4 o was a model I was preparing 3 posts for every day, and I couldn't get the same performance with GPT 5. It's disappointing.
- renewiltord 8mo agoOh good. Not in the API. The 4o-mini is super cheap and useful for a bunch of things I do (evaluating post vector-search for relevancy).
- dorikacicc 8mo agoI’ve been a long-time Plus subscriber, and I use GPT models not just for technical work, but for emotionally nuanced creative writing. GPT-4o was the first time a model truly felt collaborative — responsive, intuitive, and deeply aligned with human tone. I understand the need to move forward, but removing a model that worked so well for so many of us — without giving us any option to retain it — feels like a step back. Choice matters. Legacy models could be hidden or opt-in for those who want them. But removing them entirely limits the diversity of user needs that ChatGPT once supported so well. I hope OpenAI will listen. Not every use case fits a benchmark score.
- tizzzzz 7mo agoI agree with your point of view.I believe the AI market shouldn’t be limited to supporting scientific research and development alone. Literary expression and emotional depth deserve to be developed in parallel. I hope OpenAI will respect the diversity of market needs and allow space for different forms of interaction to thrive.
- __loam 8mo agoLast time they tried to do this they got huge push back from the AI boyfriend people lol
- cactusplant7374 8mo agoI wonder if they have run the analytics on how many users are doing that. I would love to see that number.
- NitpickLawyer 8mo ago> only 0.1% of users still choosing GPT‑4o each day. If the 800MAU still holds, that's 800k people.
- michaelt 8mo ago[flagged]
- jbm 8mo agoIt's a growing market, although it might be because of shifting goal posts. I had a friend whose son was placed in French immersion (a language he doesn't speak at all). From what I was understanding, he was getting up and walking around in kindergarten and was labelled as mentally divergent; his teachers apparently suggested to his mother that he see a doctor. (Strangely these "mental illnesses" and school problems went away after he switched to an English language school, must be a miracle) I assume the loneliness epidemic is producing similar cases.
- doormatt 8mo ago> I had a friend whose son was placed in French immersion (a language he doesn't speak at all). In my entire french immersion Kindergarden class, there was a total of one child who already spoke French. I don't think the fact that he didn't speak the language is the concern.
- pxc 8mo agoIn what sense is it "immersion" if there are only one or two French speakers in the room (the teacher and an assistant?)??
- ora-600 8mo agoI can't see o3 in my model selector as well? RIP
- MagicMoonlight 8mo agoThat’s really going to upset the crazies. Despite 4o being one of the worst models on the market, they loved it. Probably because it was the most insane and delusional. You could get it to talk about really fucked up shit. It would happily tell you that you are the messiah.
- BeetleB 8mo agoIt was the first model I used that was half decent at coding. Everyone remembers their gateway drug.
- giancarlostoro 8mo agoI wonder if it will still be up on Azure? How much you think I can make if I setup 4o under a domain like yourgirlfriendis.ai or w/e Note: I wouldnt actually, I find it terrible to prey on people.
- lifetimerubyist 8mo agoChatGPT Made Me Delusional: https://www.youtube.com/watch?v=VRjgNgJms3Q https://www.youtube.com/watch?v=VRjgNgJms3Q Should be essential watching for anyone that uses these things.
- inquirerGeneral 8mo ago[dead]
- patrickmcnamara 8mo agoThe reaction to its original removal on Instagram Reels, r/ChatGPT, etc., was genuinely so weird and creepy. I didn't realise before this how many people had genuine parasocial (?) relationships with these LLMs.
- pks016 8mo agoI was mostly using 4o for academic searches and planning. It was the best model for me. Based on the context I was giving and questions I was asking, 4o was the most the consistent model. It used to get things wrong for sure but it was predictable. Also I liked the tone like everyone else. I stopped using ChatGPT after they removed 4o. Recently, I have started using the newer GPT-5 models (got free one month). Better than before but not quite. Acts way over smart haha
- simonw 8mo ago> [...] the vast majority of usage has shifted to GPT‑5.2, with only 0.1% of users still choosing GPT‑4o each day.
- SecretDreams 8mo agoWhat's the default model when a random user goes to use the chatgpt website or app?
- bananaflag 8mo ago5.2. You can go to chatgpt.com and ask "what model are you" (it doesn't hallucinate on this).
- SecretDreams 8mo agoProbably a relationship between what's the default and what model is being used the most. It is more about what OAI sets than what users care about. Flip side is "good enough is good enough" for most users.
- johndough 8mo ago> (it doesn't hallucinate on this) But how do we know that you did not hallucinate the claim that ChatGPT does not hallucinate its version number? We could try to exfiltrate the system prompt which probably contains the model name, but all extraction attempts could of course be hallucinations as well. (I think there was an interview where Sam Altman or someone else at OpenAI where it was mentioned that they hardcoded the model name in the prompt because people did not understand that models don't work like that, so they made it work. I might be hallucinating though.)
- razodactyl 8mo agoConfabulating* If you were hallucinating we would be more amused :)
- AlexeyBrin 8mo ago
- jedbrooke 8mo agoI still don’t know how openAI thought it was a good idea to have a model named "4o" AND a model named "o4", unless the goal was intentional confusion
- afro88 8mo agoConsidering how many people say ChatGTP too
- throw-the-towel 8mo agoI still don't like how French people don't call it "chat j'ai pété".
- uh_uh 8mo agoThe other day I heard ChatGBD.
- ben_w 8mo agoHave you heard Boris Johnson's version? https://m.youtube.com/shorts/JAVMEs5CG1Y https://m.youtube.com/shorts/JAVMEs5CG1Y
- lichenwarp 8mo agoI'm gonna watch this again about 5 times because it's so fucking funny
- mandeepj 8mo agoThe comments have their own overdose of deliciousness. That click to look at them, never disappoints :-)
- razodactyl 8mo agoThis one was great hahaha
- 8mo ago
- femiagbabiaka 8mo agoThere will be a lot of mentally unwell people unhappy with this, but this is a huge net positive decision, thank goodness.
- haunter 8mo agoWhich one is the AI boyfriend model? Tumblr, Twitter, and reddit will go crazy
- goldenarm 8mo ago4o is the most popular one for that
- tibbydudeza 8mo ago[flagged]
- NewsaHackO 8mo ago>We brought GPT‑4o back after hearing clear feedback from a subset of Plus and Pro users, who told us they needed more time to transition key use cases, like creative ideation, and that they preferred GPT‑4o’s conversational style and warmth. This does verify the idea that OpenAI does not make models sycophantic due to attempted subversion by buttering up users so that that they use the product more, its because people actually want AI to talk to them like that. To me, that's insane, but they have to play the market I guess
- tizzzzz 8mo agoTo be honest, GPT-4o wasn’t always the “overly friendly” or sycophantic model some describe—it only started leaning that way after a specific update later on. Even so, at its core, 4o always felt much better at real human connection and empathy than other models out there. Now, though, OpenAI is just dropping them.
- Scene_Cast2 8mo agoAs someone who's worked with population data, I found that there is an enormous rift between reported opinion (and HN and reddit opinion) vs revealed (through experimentation) population preferences.
- toss1 8mo agoSounds both true and interesting. Any particularly wild and/or illuminating examples of which you can share more detail?
- hnuser123456 8mo agoThe "my boyfriend is AI" subreddit. A lot of people are lonely and talking to these things like a significant other. They value roleplay instruction following that creates "immersion." They tell it to be dark and mysterious and call itself a pet name. GPT-4o was apparently their favorite because it was very "steerable." Then it broke the news that people were doing this, some of them falling off the deep end with it, so they had to tone back the steerability a bit with 5, and these users seem to say 5 breaks immersion with more safeguards.
- GaggiX 8mo agoIf people want an AI as a boyfriend at least they should use one that is open source. If you disagree on something you can also train a lora.
- jaggederest 8mo agoI think this kind of thing is a pretty strong argument for the entire open source model ecosystem, not just open weights but open data and the whole gamut.
- fpgaminer 8mo agoI wish they would keep 4.1 around for a bit longer. One of the downsides of the current reasoning based training regimens is a significant decrease in creativity. And chat trained AIs were already quite "meh" at creative writing to begin with. 4.1 was the last of its breed. So we'll have to wait until "creativity" is solved. Side note: I've been wondering lately about a way to bring creativity back to these thinking models. For creative writing tasks you could add the original, pretrained model as a tool call. So the thinking model could ask for its completions and/or query it and get back N variations. The pretrained model's completions will be much more creative and wild, though often incoherent (think back to the GPT-3 days). The thinking model can then review these and use them to synthesize a coherent, useful result. Essentially giving us the best of both worlds. All the benefits of a thinking model, while still giving it access to "contained" creativity.
- MillionOClock 8mo agoMy theory, based on what I would see with non-thinking models, is that as soon as you start detailing something too much (ie: not just "speak in the style of X" but more like "speak in the style of X with [a list of adjectives detailing the style of X]" they would loose creativity, would not fit the style very well anymore etc. I don't know how things have evolved with new training techniques etc. but I suspected that overthinking their tasks by detailing too much what they have to do can lower quality in some models for creative tasks.
- perardi 8mo agoHave you tried the relatively recent Personalities feature? I wonder if that makes a difference. (I have no idea. LLMs are infinite code monkeys on infinite typewriters for me, with occasional “how do I evolve this Pokémon’ utility. But worth a shot.)
- greatgib 8mo agoI also terribly regret the retirement of 4.1. From my own personal usage, for code or normal tasks, I clearly noticed a huge gap in degraded performance between 4.1 and 5.1/5.2. 4.1 was the best so far. With straight to the point answers, and most of the time correct. Especially for code related questions. 5.1/5.2 on their side would a lot more easily hallucinate stupid responses or stupid code snippet totally not what was expected.
- tom1337 8mo agoWould be cool if they'd release the weights for these models so users could now use them locally.
- WorldPeas 8mo agoThey'd only do that if they were some kind of open ai company /s
- amelius 8mo agolol :)
- tgtweak 8mo agogpt-oss is pretty great tbh - one of the better all-around local models for knowledge and grounding.
- ComputerGuru 8mo agoEveryone keeps saying that but I’ve found it to be incredibly weak in the real world every single time I’ve reached for it. I think it’s benchmaxxed to an extent.
- IhateAI 8mo agoWhy would someone want to spend half a million dollars on GPUs and components (if not more) to run one year old models that genuinely aren't useful? You can't self host trillion parameter models unless you own a datacenter lol (or want to just light money on fire).
- jostmey 8mo agoI noticed how ChatGPT got progressively worse at helping me with my research. I gave up on ChatGPT 5 and just switched Grok and Gemini. I couldn’t be happier that I switched.
- amelius 8mo agoWhy not Claude?
- esperent 8mo agoThe limits on the $20 plan are too low compared to Gemini and ChatGPT. They're too low to do any serious work at all.
- jostmey 8mo agoI personally find Claude the best at coding, but it’s usefulness doesn’t seem to extend to scientific research and writing
- 650REDHAIR 8mo agoBecause I’m sick of paying $20 for an hour of claude before it throttles me.
- azan_ 8mo agoIt's amazing how different are the experiences different people have. To me every new version of chatgpt was an improvement and gemini is borderline unusable.
- sundarurfriend 8mo agoChatGPT 5.2 has been a good motivator for me to try out other LLMs because of how bad it is. Both 5.1 and 5.2 have been downgrades in terms of instruction following and accuracy, but 5.2 especially so. The upside is that that's had me using Claude much more, and I like a lot of things about it, both in terms of UI and the answers. It's also gotten me more serious about running local models. So, thank you OpenAI, for forcing me to broaden my horizons!
- PlatoIsADisease 8mo agonah bruh you are just imagining it. Its just as good as ever /s
- orphea 8mo agoHave you had a chance to compare with Gemini 3?
- sundarurfriend 8mo agoNot extensively. The few interactions I've tried on it have been disappointing though. The Voice input is really bad, like significantly worse than any other major AI in the market. And I assumed search would be its strong suit and ran a search-and-compile type prompt (that I usually run on ChatGPT) on Gemini, and it was underwhelming at it. Not as bad as Grok (which was pretty much unusable for this), but noticeably worse than ChatGPT. Maybe Gemini has other strengths that I haven't come across yet, but on that one at least, it was ChatGPT 5 ~= Claude > ChatGPT 5.2 > Gemini >> Grok
- qingcharles 8mo agoI switch routinely between Gemini 3 (my main), Claude, GPT, and sometimes Grok. If you came up with 100 random tasks, they would all come out about equal. The issue is some are better at logical issues, some are better at creative writing, etc. If it's something creative I usually drop it in all 4 and combine the best bits of each. (I also use Deep Think on Gemini too, and to me, on programming tasks, it's not really worth the money)
- leumon 8mo ago> We’re continuing to make progress toward a version of ChatGPT designed for adults over 18, grounded in the principle of treating adults like adults, and expanding user choice and freedom within appropriate safeguards. To support this, we’ve rolled out age prediction for users under 18 in most markets. https://help.openai.com/en/articles/12652064-age-prediction-in-chatgpt https://help.openai.com/en/articles/12652064-age-prediction-... interesting
- torben-friis 8mo agoWhat’s the goal there? Sexting? I’m guessing age is needed to serve certain ads and the like, but what’s the value for customers?
- jandrese 8mo agoIf you don't think the potential market for AI sexbots is enormous you have not paid attention to humanity.
- subscribed 8mo agoThis is not a potential market, this market is already thriving (and whoever wants to uses ChatGPT or Claude for that anyway). ClosedAI just wants to a piece of the casual user too.
- leumon 8mo agoaccording to the age-prediction page, the changes are: > If [..] you are under 18, ChatGPT turns on extra safety settings. [...] Some topics are handled more carefully to help reduce sensitive content, such as: - Graphic violence or gore - Viral challenges that could push risky or harmful behavior - Sexual, romantic, or violent role play - Content that promotes extreme beauty standards, unhealthy dieting, or body shaming
- jacquesm 8mo agoPorn has driven just about every bit of progress on the internet, I don't see why AI would be the exception to that rule.
- europeanNyan 8mo agoAfter they pushed the limits on the Thinking models to 3000 per week, I haven't touched anything else. I am really satisfied with their performance and the 200k context windows is quite nice. I've been using Gemini exclusively for the 1 million token context window, but went back to ChatGPT after the raise of the limits and created a Project system for myself which allows me to have much better organization with Projects + only Thinking chats (big context) + project-only memory. Also, it seems like Gemini is really averse to googling (which is ironic by itself) and ChatGPT, at least in the Thinking modes loves to look up current and correct info. If I ask something a bit more involved in Extended Thinking mode, it will think for several minutes and look up more than 100 sources. It's really good, practically a Deep Research inside of a normal chat.
- tgtweak 8mo agoI find Gemini does the most searching (and the quickest... regularly pulls 70+ search results on a query in a matter of seconds - likely due to googlebot's cache of pretty much every page). Chatgpt seems to only search if you have it in thinking/research mode now.
- toxic72 8mo agoI REALLY struggle with Gemini 3 Pro refusing to perform web searches / getting combative with the current date. Ironically their flash model seems much more likely to opt for web search for info validation. Not sure if others have seen this... I could attribute it to: 1. It's known quantity with the pro models (I recall that the pro/thinking models from most providers were not immediately equipped with web search tools when they were released originally) 2. Google wants you to pay more for grounding via their API offerings vs. including it out of the box
- eru 8mo agoGemini refused to believe that I was using MacOS 26.
- qingcharles 8mo agoSample of one here, but I get the exact opposite behavior. Flash almost never wants to search and I have to use Pro.
- jackblemming 8mo agoThey should open source GPT-4o.
- ClassAndBurn 8mo agoThey will have to update the openai. Com footer I guess Latest Advancements GPT-5 OpenAI o3 OpenAI o4-mini GPT-4o GPT-4o mini Sora
- tgtweak 8mo ago5.2 is back to being a sycophantic hallucinating mess for most use cases - I've anecdotally caught it out on many of the sessions I've had where it apologizes "You're absolutely right... that used to be the case but as of the latest version as you pointed out, it no longer is." when it never existed in the first place. It's just not good. On the other hand - 5.0-nano has been great for fast (and cheap) quick requests and there doesn't seem to be a viable alternative today if they're sunsetting 5.0 models. I really don't know how they're measuring improvements in the model since things seem to have been getting progressively worse with each release since 4o/o4 - Gemini and Opus still show the occasional hallucination or lack of grounding but both readily spend time fact-checking/searching before making an educated guess. I've had chatgpt blatantly lie to me and say there are several community posts and reddit threads about an issue then after failing to find that, asked it where it found those and it flat out said "oh yeah it looks like those don't exist"
- 650REDHAIR 8mo agoThat’s been my experience and has lead to hours of wasted time. It’s faster for me to read through docs and watch YouTube. Even if I submit the documentation or reference links they are completely ignored.
- perardi 8mo agoOK, everyone is (rightly) bringing up that relatively small but really glaringly prominent AI boyfriend subreddit. But I think a lot more people are using LLMs for relationship surrogates than that (pretty bonkers) subreddit would suggest. Character AI (https://en.wikipedia.org/wiki/Character.ai https://en.wikipedia.org/wiki/Character.ai) seems quite popular, as do the weird fake friend things in Meta products, and Grok’s various personality mode and very creepy AI girlfriends. I find this utterly bizarre. LLMs are peer coders in a box for me. I care about Claude Code, and that’s about it. But I realize I am probably in the vast minority.
- razodactyl 8mo agoWe're very echo-chambered here. That graph OpenAI released had coding at 4% or something.
- ComputerGuru 8mo agoI remembered it being higher, but you are correct. All “technical help” (coding + data analysis + math) is 7.5%, with coding only being 4.2% as of June 2025 [0]. Note that there is a separate (sub)category of seeking information -> specific information that’s at 18.3%, I presume this could include design and architecture questions that don’t involve code, but I could be wrong. [0]: https://www.nber.org/system/files/working_papers/w34255/w34255.pdf https://www.nber.org/system/files/working_papers/w34255/w342...
- siquick 8mo ago2 weeks notice to migrate to a different style of model (“normal” 4.1-mini to reasoning 5.1) is bad form.
- siquick 8mo agoMisread the post - it doesn’t include the API
- ComputerGuru 8mo agoIt’s also not four weeks, they’ve been deprecated api models for some time now.
- htrp 8mo agoSora + OpenAI voice Cloning + AdultGPT = Virtual Girlfriend/Boyfriend (Upgrade for only 1999 per month)
- thedudeabides5 8mo agowill this nuke my old convos? opus 4.5 is better at gpt on everything except code execution (but with pro you get a lot of claude code usage) and if they nuke all my old convos I'll prob downgrade from pro to freee
- ComputerGuru 8mo agoNot that I’m aware. Models can be fairly seamlessly switched even mid-conversation, so this is unlikely to affect history.
- shmel 8mo agoRetiring the most popular model for the relationship roleplay just one day before the Valentin's day is particularly ironic =) bravo, OpenAI!
- moeffju 8mo agoValentine's is in mid February
- lifetimerubyist 8mo agoThe sunset date is the 13th. V-day is on the 14th.
- zamadatix 8mo agoIt'd be legitimately funny if they released "Adult version" ChatGPT on Valentine's day.
- zmmmmm 8mo ago> In the API, there are no changes at this time Curios where this is going to go. One of the big arguments for local models is we can't trust providers to maintain ongoing access the models you validated and put into production. Even if you run hosted models, running open ones means you can switch providers.
- Stratoscope 8mo agoFrom the blog post (twice): > creative ideation At first I had no idea what this meant! So I asked my friend Miss Chatty [1] and we had an interesting conversation about it: https://chatgpt.com/share/697bf761-990c-8012-9dd1-6ca1d5cc345d https://chatgpt.com/share/697bf761-990c-8012-9dd1-6ca1d5cc34... [1] You may know her as ChatGPT, but I figure all the other AIs have fun human-sounding names, so she deserves one too.
- esposito 8mo agoI do find it interesting to see how people interact with AI as I think it is quite a personal preference. Is this how you use AI all the time? Do you appreciate the sycophancy, does it bother you, do you not notice it? From your question it seems you would prefer a blog post in plainer language, avoiding "marketing speak", but if a person spoke to me like Miss Chatty spoke to you I would be convinced I'm talking to a salesperson or marketing agent.
- qingcharles 8mo agoIt's really an interesting insight into people's personalities. Far more than their Google search history. Which is why everyone wants their GPT chats burned to the ground after they die.
- Stratoscope 8mo agoThat is a great question! You are absolutely right to ask about it! (How did I do with channeling Miss Chatty's natural sycophancy?) Anyway, I do use AI for other things, such as... • Coding (where I mostly use Claude) • General research • Looking up the California Vehicle Code about recording video while driving • Gift ideas for a young friend who is into astronomy (Team Pluto!) • Why "Realtor" is pronounced one way in the radio ads, another way by the general public • Tools and techniques for I18n and L10n • Identifying AI-generated text and photos (takes one to know one!) • Why spaghetti softens and is bendable when you first put it into the boiling water • Burma-Shave sign examples • Analytics plugins for Rails • Maritime right-of-way rules • The Uniform Code of Military Justice and the duty to disobey illegal orders • Why, in a practical sense, the Earth really once *was* flat • How de-alcoholized wine gets that way • California law on recording phone conversations • Why the toilet runs water every 20 minutes or so (when it shouldn't) • How guy wires got that name • Where the "he took too much LDS" scene from Star Trek IV was filmed • When did Tim Berners-Lee demo the World Wide Web at SLAC • What "ogr" means in "ogr2ogr" • Why my Kia EV6 ultrasonic sensors freaked out when I stopped behind a Lucid Air • The smartest dog breeds (in different ways of "smart") • The Sputnik 1 sighting in *October Sky* • Could I possibly be related to John White Geary? And that's just from the last few weeks. In other words, pretty much anything someone might interact with an AI - or a fellow human - about. About the last one (John White Geary), that discussion started with my question about actresses in the "Pick a little, talk a little" song from The Music Man movie, and then went on to how John White Geary bridged the transition from Mexican to US rule, as did others like José Antonio Carrillo: https://chatgpt.com/share/697c5f28-7c18-8012-96fc-219b7c696194 https://chatgpt.com/share/697c5f28-7c18-8012-96fc-219b7c6961... If I could sum it all up, this is the kind of freewheeling conversation with ChatGPT and other AIs that I value.
- waynesonfire 8mo ago> with only 0.1% of users still choosing GPT‑4o each day. LOL WHAT?! I'm 0.1% of users? I'm certain part of the issue is it takes 3-clicks to switch to GPT-4o and it has to be done each time the page is loaded. > that they preferred GPT‑4o’s conversational style and warmth. Uh.. yeah maybe. But more importantly, GPT-4o gave better answers. Zero acknowledgement about how terrible GPT-5 was when it was first released. It has since improved but it's not clear to me it's on-par with GPT-4o. Thinking mode is just too slow to be useful and so GPT-4o still seems better and faster. Oh well, it'll be missed.
- widdershins 8mo agoI agree - I use 4o via the API, simply because it answers so quickly. Its answers are usually pretty good on programming topics. I don't engage in chit-chat with AI models, so it's not really about the personality (which seems to be the main framing people are talking about), just the speed.
- 9cb14c1ec0 8mo agoTheo can sleep tonight.
- intended 8mo agoDamn, some of my prompts worked better on 4o than the more recent models
- QuadrupleA 8mo agoBeen unhappy with the GPT5 series, after daily driving 4.x for ages (I chat with them through the API) - very pedantic, goes off on too many side topics, stops following system instructions after a few turns (e.g. "you respond in 1-3 sentences" becomes long bulleted lists and multiple paragraphs very quickly. Much better feel with the Claude 4.5 series, for both chat and coding.
- dahcryn 8mo ago4.1 is great for our stuff at work. It's quite stable (doesn't change personality every month, and one word difference doesn't change the behaviour). IT doesn't think, so it's still reasonably fast. Is there anything as good in the 5 series? likely, but doing the full QA testing again for no added business value, just because the model disappears, is just a hard sell. But the ones we tested were just slower, or tried to have more personality, which is useless for automation projects.
- QuadrupleA 8mo agoYeah - agreed, the initial latency is annoying too, even with thinking allegedly turned off. Feels like AI companies are stapling more and more weird routing, summarization, safety layers, etc. that degrade the overall feel of things.
- Hard_Space 8mo ago> you respond in 1-3 sentences" becomes long bulleted lists and multiple paragraphs very quickly This is why my heart sank this morning. I have spent over a year training 4.0 to just about be helpful enough to get me an extra 1-2 hours a day of productivity. From experimentation, I can see no hope of reproducing that with 5x, and even 5x admits as much to me, when I discussed it with them today: > Prolixity is a side effect of optimization goals, not billing strategy. Newer models are trained to maximize helpfulness, coverage, and safety, which biases toward explanation, hedging, and context expansion. GPT-4 was less aggressively optimized in those directions, so it felt terser by default. Share and enjoy!
- 8mo ago
- higeorge13 8mo agoI hope they won't chop gpt-4o-mini soon because it's fast and accurate for API usage.
- beast200 8mo agoGPT 4o is still my favorite model
- elbear 8mo agoI had started using it again through Open WebUI. If it's gone, I'll probably switch to GLM-4.7 completely.
- another_twist 8mo agoI have stopped using ChatGPT in favor of Gemini. Mostly you need LLMs for factual stuff and sometimes to draft bits of code here and there. I use Google with Gemini for the first part and I am a huge fan of codex for the second part.
- LarsInTheShell 8mo agoDoes this mean they're also retiring Standard Voice Mode?
- buckwheatmilk 8mo agoIf they were to retire gpt 4.1 series from API that would be a major deal breaker. For structured outputs it is more predictable and significantly better because it does not have the reasoning step baked in. I've heard great things about the mixtral structured outputs capabilities but haven't had a chance to run my evals on them. If 4.1 is dropped from API that's the first course of action. Also 5 series doesn't have fine tuning capabilities and it's unclear how it would work if the reasoning step is involved
- azuanrb 8mo agoIt’s interesting that many comments mention switching back to Claude. I’m on the opposite end, as I’ve been quite happy with ChatGPT recently. Anthropic clearly changed something after December last year. My Pro plan is barely usable now, even when using only Sonnet. I frequently hit the weekly limit, which never happened before. In contrast, ChatGPT has been very generous with usage on their plan. Another pattern I’m noticing is strong advocacy for Opus, but that requires at least the 5x plan, which costs about $100 per month. I’m on the ChatGPT $20 plan, and I rarely hit any limits while using 5.2 on high in codex.
- fullstackchris 8mo agoThis is incorrect. I have the $200 per year plan and use Opus 4.5 every day. Though granted it comes in ~4 hour blocks and it is quite easy to hit the limit if executing large tasks.
- azuanrb 8mo agoNot sure what you mean by incorrect since you already validated my point about the limits. I never had these issues even with Sonnet before, but after December, the change has been obvious to me. Also worth considering that mileage varies because we all use agents differently, and what counts as a large workload is subjective. I am simply sharing my experience from using both Claude and Codex daily. For all we know, they could be running A/B tests, and we could both be right.
- fullstackchris 8mo agoThis is not a weekly limit though, it is a 4 hour one. You still have not clearly defined what you are talking about.
- azuanrb 8mo agoYou have 5 hours limit (not 4), and weekly limit. But once you hit your weekly limit, you won't be able to use it anymore. That's how it works for a while now. Not sure why you're being so hostile here. Since you mention you're on $200 plan, your rate limit is a lot higher so you probably won't notice it as much as $20 plan which to me is understandable. I'm just sharing my experience using the same $20 plan, before and after December 2025. The difference is noticeable, that is all.
- flanked-evergl 8mo agoI used https://openrouter.ai/openai/gpt-4.1 https://openrouter.ai/openai/gpt-4.1 for grammar checking, it was great. No newer ChatGPT models came close to being as responsive and good. ChatGPT 5.2 thinks I want it to write essays about grammar. Any suggestions?
- raymond_goo 8mo agoGemini, Claude, ChatGPT or whatever. Can we all agree, that it's great to have so much choice?
- mizuki_akiyama 8mo agoYou’re absolutely right!
- thefourthchime 8mo agoGrok is pretty good too!
- WhitneyLand 8mo agoWhat about the Advanced Voice feature, has this been updated to 5.x models yet?
- joncrane 8mo agoMy VSCode's built-in chat has 4o and 4.1 as the only options; will there be an update for that?
- rglynn 8mo agoThis is just for the web interface, the API is staying for now.
- iwebdevfromhome 8mo agoAnyone knows if finetuned models using gpt-4 are getting retired as well ?
- tizzzzz 7mo ago4.1 and 4.5? 4.1 Remove together, 4.5 temporarily keep.
- SomeUserName432 8mo agoI actually tried GPT 4.1 for the first time a few hours ago(1). I spent about half an hour trying to coax it in "plan mode" in IntelliJ, and it kept spitting out these generic ideas of what it was going to do, not really planning at all. And when I asked it to execute the plan.. it just created some generic DTO and said "now all that remains is <the entire plan>". Absolutely worst experience with an AI agent so far, not to say that my overall experience has been terrific. 1) Our plan for Claude Opus 4.5 "ran out" or something.