The title is misleading. The only thing they seem to have done was add a $100 plan identical to Claude's, which gives 5x usage of ChatGPT Plus. There is still a $200 plan that gives 20x usage.
It's not worse, Anthropic simply has no equivalent model (if you don't consider Mythos) of GPT 5.4 Pro. Google does though: Gemini 3.1 Deep Think.
GPT 5.4 Pro is extremely slow but thorough, so it's not meant for the usual agentic work, rather for research or solving hard bugs/math problems when you provide it all the context.
I'm genuinely asking, when you say Gemini 3.1 DT is an equivalent model of GPT 5.4 Pro, is there a specific benchmark/comparison you're referring to or is this more anecdotal?
And do you mean to say that you don't really use GPT 5.4 Pro unless it's for a hard bug? Curious which models you use for system design/architecture/planning vs execution of a plan/design.
TIA! I'm still trying to figure out an optimal system for leveraging all of the LLMs available to us as I've just been throwing 100% of my work at Claude Code in recent months but would like to branch out.
So, reading the tea leaves, they're either losing subscribers for the $200 plan, or they're not following the same hockey stick path of growth they thought they were... maybe?
Edit: I wonder if this is actually compute-bound as the impetus
Nope, it's just that a lot of people (especially those using Codex) asked us for a medium-sized $100 plan. $20 felt too restrictive and $200 felt like a big jump.
Pricing strategy is always a bit of an art, without a perfect optimum for everyone:
- pay-per-token makes every query feel stressful
- a single plan overcharges light users and annoyingly blocks heavy users
- a zillion plans are confusing / annoying to navigate and change
This change mostly just adds a medium-sized plan for people doing medium-sized amounts of work. People were asking for this, and we're happy to deliver.
Did you modify the Plus plans usage recently or as part of this introduction? Given that Pro plans usage are multiples of it (5x/20x) and given reports of less Plus usage, clarification would be appreciated?
Transparency on this sort of thing is the best way to address negative company sentiment.
I'm honestly not sure, as I don't work on it. My understanding from afar is:
- There was a 2x promotion in March that ended on April 2, so limits have felt tighter since then
- We sometimes reset rate limits after bugs or milestones or because Tibo feels generous, which can make some days feel different than others (they are typically announced here: https://x.com/thsottiaux)
- Recently Plus was tweaked to have a smaller 5h limit but an increased weekly limit
- Lastly, as part of the new Pro launch, the $100 & $200 Pro tiers are getting a 2x promotion, meaning they are temporarily 10x/40x instead of 5x/20x
I've asked our team to clarify the pricing page. Agree it's not clear.
All good, I interpreted it as postulation and not accusation. :)
I do like the job! Much more organic than yanking tickets, though I'm on the model training side of things, rather than product side. Always a balance between short-term sprints patching bad behaviors for the next model vs long-term investments in infra and science that make future work easier. Sometimes the negative press gets to me a bit (it's a very different feeling than 2022 or 2023), but my goal is just to make the most useful product I can for people. It's been wild how much Codex has already changed my day-to-day work, I'm so curious to see what it looks like in 2030 or 2040.
What kind of bad behaviors? How is the whole SDLC lifecycle there? I imagine, given that this tech is kind of redefining how software is being written, it's not your standard workflow pipeline? Are there code reviews at all? Have you been in any particularly interesting meetings about how you're trying to "shape" the models?
I won't misrepresent myself, I've never spent a penny on any of these services. I am just super curious what it's like to work at one of these frontrunner companies. I bet it's pretty neat.
If you want it to deeply research something pro is great. I had a problem I just couldn’t find with my oven so I gave it a lot of information and it went off on its own for about 2 hours and then gave me what I needed to fix the problem (fan was turning off too quickly which was causing the panel to overheat). I have no idea how it figured it out and I couldn’t find anything after hours of googling so it was very impressive. I even went and googled for it once I knew what the problem was and I still couldn’t find the solution that it came up with.
Thanks for sharing this experience. Does it cost a lot of token in the deep analysis - which will make the $100 plan much quicker to drain all budgets.
I think it’s going to be very hard to blow through your tokens just using chat. I mostly bought the plan so I could use Codex and on the $200 a month plan I’ve basically been using it 15 hours a day almost nonstop and I don’t run out of tokens for the week.
Notably, up until now Pro had 6x usage of Plus. So the title is only slightly misleading.
On the other hand, the benchmark of Plus usage seems to be to be all over the place, so it’s difficult to say now how does the usage compare to the old Pro.
r/codex is reporting that $20 (Plus) seems to have had its usage limit reduced (some people are saying it feels like 1/3 the previous limit now). The theory[1] is that reducing $20's limit lets them claim $200 has 20x $20's limit (and $100 has 10x).
If that's true, then the value comparison is not so positive for Codex any more