# OpenAI Codex (OpenAI)

Feedback Bench, coding agents, built 2026-10-01, window 2026-08-31 to 2026-09-27. Web page: https://feedbackbench.com/#/agent/codex

| Measure | Value |
|---|---|
| Rank | 2 of 17 (rank range 2–2) |
| Feedback Score | 67.4 (95% interval 66.9–67.8) |
| Popularity | 1.000 (share of voice 30.27%) |
| Customer love | 0.454 (95% interval 0.448–0.460) |
| Top quadrant | no |
| Authors | 29813 |
| Posts counted | 118683 |
| Posts that judge the agent | 48990 |
| Criteria better / worse than peers | 7 / 21 of 63 |

## The brief

Written by Claude Opus 5.5 from 200 labelled posts and the numbers on this page. Interpretation, not measurement: every quote is verbatim and links to its post.

**A capable agent throttled by a quota nobody can read.**

TL;DR:

- Limits dominate the conversation: unpredictable burn, a distrusted meter and a five-hour window that halts work.
- The coding itself lands mid-pack; code review, concise replies and the desktop app earn real praise.
- Subagent orchestration saves serious usage when tuned and silently drains it when left on defaults.

### What hurts

- **Quota burn is opaque and erratic** ([Single prompt, model or effort level consumes disproportionate quota](https://feedbackbench.com/criteria/limits.burn_rate.md), [Usage meter visibility and accuracy](https://feedbackbench.com/criteria/limits.usage_meter.md), [Price, allowance or plan terms changed](https://feedbackbench.com/criteria/limits.allowance_change.md)). Similar prompts consume wildly different shares of quota, and the meter gives no reliable way to predict or audit it.
  Users report the same kind of task eating very different amounts of allowance, with no published conversion from credits to plan allowance. One post describes spending a reset, then watching most of it vanish after a simple prompt. Another says the 5-hour indicator silently switched to show the weekly figure, which led to overspending.
  
  Some users read the swings as deliberate backend tuning. That is speculation, but it shows how far trust in the meter has fallen. It is the area users judge most harshly against peers.
  Evidence:
  - Complaint, OpenAI Codex, r/codex, 2026-09-09: “yea my usage was at 46% i come back and its at 1%...blown away, i used a reset. 5 minutes later after a simple prompt its at 46% again...something is wrong.” [source](https://www.reddit.com/r/codex/comments/1wbtm0l/reset_refund_failed/p8sn7is/)
  - Complaint, OpenAI Codex, r/codex, 2026-09-19: “this happened to me too - in codex the indicator that used to be 5 hour limit started showing the weekly limit. problem is i didn't know that. seeing as it was going down so slowly i started going nuts on usage. finally i realised something was up and now i've used a huge amount of my weekly limit because the stupid interface changed silently.” [source](https://www.reddit.com/r/codex/comments/1w497jg/new_plus_account_doesnt_have_5_hour_usage_limit/paqi5mz/)
  - Complaint, OpenAI Codex, r/codex, 2026-09-11: “this shit is so fucking inconsistent i'm actually getting sick of it, i'm after using astra max on plus and im getting almost hte same usage as sol light, can someone explain this to me?” [source](https://www.reddit.com/r/codex/comments/1wdl56p/codex_is_unusable_now/p97020w/)
  - Complaint, OpenAI Codex, r/codex, 2026-09-06: “the usage drain is not a "bug" to drain, right ? i mean they give you a secret amount, of everchanging, usage allowance per week. everyone who is using codex knows that that allowance is changing extremely from one week to another. it's not a bug to search, it's a setting in their backend. they use the resets to see how users react after the setting was changed - to see if they purchase more subscriptions, upgrade, or leave. it's a permanent psychological test.” [source](https://www.reddit.com/r/codex/comments/1w98051/tibo_incoming/p88n3hj/)

- **The five-hour window stops real work** ([Short rolling usage window blocks or interrupts work](https://feedbackbench.com/criteria/limits.window_interrupts_work.md), [Single prompt, model or effort level consumes disproportionate quota](https://feedbackbench.com/criteria/limits.burn_rate.md)). The rolling short window collides with heavy models, so one ambitious prompt can lock a user out for hours mid-task.
  Posts describe a single prompt running ten minutes, then a long forced wait, and a long goal run hitting the weekly cap partway through. Users say they would accept large weekly drains for hard problems; the short window makes that impossible.
  
  Removing the 5-hour window is the most requested change. Reports conflict on which plans still have it, which adds confusion on top of the interruption.
  Evidence:
  - Complaint, OpenAI Codex, r/codex, 2026-09-01: “put in one prompt, dude worked for 10 minutes and told me to wait another 5 hours. lmfao” [source](https://www.reddit.com/r/codex/comments/1w4h75o/dont_use_codex_today/p79mit8/)
  - Complaint, OpenAI Codex, r/codex, 2026-09-06: “you could actually do things with sol medium/high when the 5h limit was gone. i'm fine with losing 70% of my weekly limit in 1 prompt to fix a more complex issue with astra but the 5h makes this impossible. so the issue here really is the 5h limit combined with excessively high consumption by astra.” [source](https://www.reddit.com/r/codex/comments/1w92ek7/what_do_plus_users_even_want/p87844d/)
  - Complaint, OpenAI Codex, r/codex, 2026-09-23: “yeah im on pro $100 and i gave it one goal, it ran for an hour and a half then week limit reached. i feel like with claude the 5 hour limit is the same as the week with gpt” [source](https://www.reddit.com/r/codex/comments/1wioij7/feeling_scammed_on_the_200_pro_plan_since_astra/pblpogv/)

- **Subagents drain usage behind the scenes** ([Subagents, parallel agents and orchestrators](https://feedbackbench.com/criteria/work.multi_agent_orchestration.md), [Prompt cache hits, misses and invalidation](https://feedbackbench.com/criteria/limits.prompt_cache.md), [Automatic model routing and fallback](https://feedbackbench.com/criteria/models.routing_auto.md)). Orchestration defaults poll subagents, re-read context and let caches expire, so multi-agent runs burn quota in ways users only notice after the meter drops.
  Users describe an orchestrator that checks on subagents every minute, possibly triggering full context re-reads. One found a subagent labelled as the cheap model while actually running the expensive one. Others note cached tokens expiring while a parent agent waits, turning the next turn into uncached input.
  
  Auto-review is another hidden cost: one user says an approval setting switched it on despite their config and it ate a third of their allowance finding nothing.
  Evidence:
  - Complaint, OpenAI Codex, r/codex, 2026-09-10: “codex has a hard rule in the system prompt which turns the orchestrating agents into an adhd micromanager and basically asks for an update every 60 seconds, and i believe that can trigger a full re-read of the existing models context. repeat x subagents, and using astra as orchestrator and it burns your usage.” [source](https://www.reddit.com/r/codex/comments/1wchs79/self_reporting_skill_issue_maybe_it_will_help/p8y3d0p/)
  - Complaint, OpenAI Codex, r/codex, 2026-09-13: “i’m using sol as an orchestrator, while delegating computer-use and browser-use tasks to luna as a subagent. this setup is very efficient for tasks that require those skills. after about 10 minutes of sol doing nothing but waiting for luna to finish a task, i opened the usage meter and saw that my usage was absolutely fucked. the task should’ve consumed something like 5% of my 5-hour limit, but it used around 35%. i checked the agents panel again, and this is what i found: <strict_link> the subagent was literally named “luna,” while it was actually using sol. so yeah, it completely fucked up my usage.” [source](https://www.reddit.com/r/codex/comments/1wfl3pz/that_was_really_hilarious/)
  - Complaint, OpenAI Codex, r/codex, 2026-09-08: “the potential downside there is that if astra main agent waits 10 mins for luna to finish, its cached tokens may become expired and the entire convo plus the luna response then becomes uncached input on the next turn… it may have used 7.13m tokens for timeouts but 7.11m of them were cached… it’s a fine line to balance, i suppose.” [source](https://www.reddit.com/r/codex/comments/1wa9c9d/i_investigated_why_gpt6_astra_burns_quota_so_fast/p8gk2u5/)
  - Complaint, OpenAI Codex, r/codex, 2026-09-06: “it took a while to better analyze the data. for me it was auto-review silently using up my usage. "approve for me" apparently turns it on without warning. it used 33% of my usage to find no issues in astra's code. it's also an obvious problem how inefficient and bad autoreview is on, and that you have to "bypass permissions" entirely to get rid of it but also not be annoyed by non-risky popups. i had auto-review disabled in my config.toml yet approve for me overrode it. take away auto-review, and i only got $1433 of usage out of the expected $2500 from my banked reset - almost half. so there i gave you my data so your mind must be changed.” [source](https://www.reddit.com/r/codex/comments/1w951h1/when_tibo/p88t5sy/)

- **Quality drifts while routing hides models** ([Quality got worse or better over time](https://feedbackbench.com/criteria/models.quality_drift.md), [Automatic model routing and fallback](https://feedbackbench.com/criteria/models.routing_auto.md), [Which models are offered on a plan and when](https://feedbackbench.com/criteria/models.catalog_access.md)). Users say each effort tier delivers less over time and suspect silent routing to cheaper models, with the picker showing a model they may not be getting.
  Users report tasks that once passed at medium effort now needing high or extra high, which one post calls shrinkflation. Others suspect requests get routed to cheaper models without disclosure and ask that the actual model be displayed.
  
  Plan gating adds friction: Plus users say the flagship is technically available but too costly to use for real work. Drift claims stay contested, since some users point out formal benchmarks often find no change.
  Evidence:
  - Complaint, OpenAI Codex, r/codex, 2026-09-17: “priority access via api i can live with, if you can spend $20.000 on a new contact button hover effect you deserve it. but the silently routing to another model is what i find offensive, they should display the actual model you're getting not the one you think you selected.” [source](https://www.reddit.com/r/codex/comments/1wil686/are_we_allowed_to_wear_tinfoil_around_here/pabz4gs/)
  - Complaint, OpenAI Codex, r/codex, 2026-09-03: “i've noticed with chatgpt recently that tasks that sol 5.6 could handle reliably on medium a few weeks ago now seem to need high, and things that previously felt fine on high increasingly need extra high to get the same level of "first pass" reasoning and consistency - in other words, it makes more mistakes that you have to correct it on it responds than it did on that effort previously. it’s not just the model feels worse it feels more like the capability available at each reasoning tier has shifted downwards over time. i can’t prove that’s intentional, but i’ve noticed the pattern repeatedly, particularly before newer models are released. if the same workload gradually requires a higher reasoning tier to get the same quality of result, that’s effectively shrinkflation from the user’s point of view.” [source](https://www.reddit.com/r/codex/comments/1w5zq9e/astra_cant_come_soon_enough_if_this_what_sol_max/p7jrajs/)
  - Complaint, OpenAI Codex, r/codex, 2026-09-14: “absolutely ridiculous by oai, reminds me of the silent routing to the chat-safety model in the 4o/5 days tbh. would love to see a breakdown of who this is affecting. my (wild-ass) guess is mostly people outside the us (and vpn users) + people with multiple pro accounts?” [source](https://www.reddit.com/r/codex/comments/1wg3odg/openai_is_silently_degrading_some_astra_codex/p9rsnms/)
  - Complaint, OpenAI Codex, r/codex, 2026-09-08: “yes, unfortunately astra isn’t designed for plus users. it’s nice that everyone has access, but as soon as you want to use it for work, plus users are out of the deal.” [source](https://www.reddit.com/r/codex/comments/1wagpil/gpt6_astra_best_reasoning_effort/p8ibdlp/)

- **Slower replies, and updates that break** ([Latency, throughput and fast mode](https://feedbackbench.com/criteria/rel.response_speed.md), [Client crashes, freezes and failed tool execution](https://feedbackbench.com/criteria/rel.client_failures.md), [Updates break working setups](https://feedbackbench.com/criteria/rel.update_breakage.md)). Replies slow sharply under load and client updates regularly break working setups, so reliability lags peers on both the service and the app.
  Posts describe replies stretching from minutes to tens of minutes during busy periods, with more timeouts and weaker output. Top-plan users say fast mode still feels slow.
  
  On the client, users report the Windows app opening blank after updates, disabled settings re-enabling themselves, and a performance drop after a framework change. Several X posts show fixes landing quickly, but the breakage arrives first.
  Evidence:
  - Complaint, OpenAI Codex, r/codex, 2026-09-20: “while they train other stuff or some big hitters need more api credits, the system gets super slow. like from 2 min to 20 min replies, and gets more timeouts in it's processing, so it's missing stuff/becomes stupid. happens to claude code allot also.” [source](https://www.reddit.com/r/codex/comments/1wksrty/why_astra_so_stupid_and_slow/pavj9xa/)
  - Complaint, OpenAI Codex, X search: OpenAI Codex, Codex CLI, Codex app, 2026-09-06: “the codex windows app crashes on startup every time there is an update (the window remains blank), so i asked for the repair process to be made into a skill called codex-desktop-repair in codex cli. it seems that there can be different causes for this, and that has been taken into account as well.” [source](https://twitter.com/1944351718021709824/status/2096504507878477996)
  - Complaint, OpenAI Codex, r/codex, 2026-09-08: “nah, i already disabled this stuff before. it turned itself back on after updating.” [source](https://www.reddit.com/r/codex/comments/1wapkx0/no_freaking_way_i_enabled_any_app_in_codex/p8jv1ys/)
  - Complaint, OpenAI Codex, r/codex, 2026-09-04: “i barely got 30t/s with fast enabled, what's the point of all the quota on a pro20x when you can't get the output you want.” [source](https://www.reddit.com/r/codex/comments/1w6tx03/speed_for_gpt56_sol_fast/p7q04of/)

- **Fiddly session and history handling** ([Saving, switching, resuming and rewinding sessions](https://feedbackbench.com/criteria/ui.session_history.md), [How the interface shows work, and what the user can configure](https://feedbackbench.com/criteria/ui.display_settings.md)). Session management trails peers, with users reporting lost history, missing session IDs and limited ways to start or resume work across surfaces.
  One user says a forced logout wiped all chat history. Others miss a removed copy button for session IDs and want to start sessions in arbitrary folders rather than recent repos.
  
  The CLI fares better: users point to resume commands and tiled terminal panes as a workable path for parallel sessions.
  Evidence:
  - Complaint, OpenAI Codex, X search: OpenAI Codex, Codex CLI, Codex app, 2026-09-13: “all of a sudden i got logged out of my chatgpt account automatically.. when i signed back in, all my history from all the chat windows are gone...simply gone @thsottiaux @sama @openai #openai #codex” [source](https://twitter.com/1463035726514176006/status/2099279644130267625)
  - Complaint, OpenAI Codex, X search: OpenAI Codex, Codex CLI, Codex app, 2026-09-01: “huh, can't i get the session id for each session in the codex app anymore? i feel like there used to be a copy button when i right-clicked.” [source](https://twitter.com/1490361736255389697/status/2094895238292930875)
  - Praise, OpenAI Codex, r/codex, 2026-09-26: “windows terminal handles this fine with the cli - alt+shift+plus / alt+shift+minus splits the pane, so you can tile 4 or so codex sessions on one screen without needing the vm. i'd run each one from its own git worktree (git worktree add ../feature-x) so they're not all editing the same checkout, and `codex resume` gets you back into a session if you close a pane by accident. for the 4-5 parallel setup you had with claude, the cli is a lot less painful than juggling app windows (wsl works too if you hit any path weirdness). should get you pretty close to your old setup!” [source](https://www.reddit.com/r/codex/comments/1wdv8vw/how_can_i_run_multiple_sessions_on_single_screen/pc639i4/)

### What works

- **Code review users actually trust** ([Agent-performed code review finds real issues](https://feedbackbench.com/criteria/verify.agent_code_review.md)). Agent-run code review finds real issues, enough that some users route other vendors' output through Codex purely to verify it.
  Users describe asking Codex to check every line another model wrote. A run log shows the final review pass catching regressions the implementing agents missed. It is one of the few areas rated better than peers.
  
  The weak spot is noise. One user says reviews pad results with contrived edge cases unless prompted to state the failure scenario and its likelihood.
  Evidence:
  - Praise, OpenAI Codex, r/ClaudeCode, 2026-09-18: “i constantly have to ask codex to verify every line of opus work- at this rate there's no point in having claude code.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wiujum/claude_code_is_falling_behind_codex_not_because/pak2kvh/)
  - Praise, OpenAI Codex, r/vibecoding, 2026-09-20: “i now re-run in codex cli with flash routed directly via deepseek (the prev experiment was via openrouter) and this run was substantially cheaper (about $3.13 for astra and $0.05 for all five flash sessions). the astra run included final codex review (about 1/3 of all tokens spent), as before. unlike the earlier run, this run resulted in 2 regressions detected by review.” [source](https://www.reddit.com/r/vibecoding/comments/1wl3lfi/i_built_an_orchestration_package_that_lowered_my/paywh0b/)
  - Complaint, OpenAI Codex, r/codex, 2026-09-20: “yep. that last part needs to instead be something along the lines of providing the exact failure scenario and the likelihood that it occurs for each problem it finds. you'll get a list of 1 or 2 actual problems and several where it's like "well realistically the chance is very low, and would require a confluence of events that multiple other layers make generally impossible, but via some incredibly contrived scenario involving a cosmic radiation bit flip it *could*."” [source](https://www.reddit.com/r/codex/comments/1wl5tml/use_this_prompt_after_your_agents_are_done_working/pawbkjr/)

- **Replies stay short and on point** ([Length and clarity of replies, summaries and comments](https://feedbackbench.com/criteria/work.response_verbosity.md)). Codex answers without essays and stacked caveats, and users who switched from other agents name that brevity as a reason.
  Users contrast Codex with agents that return pages of text for small changes. One praises the newest model as more concise than its predecessor; another says rambling output was why they left a rival.
  
  The complaints are narrow: repeated apologies when the agent misses parameters, which one user would rather see as returned usage.
  Evidence:
  - Praise, OpenAI Codex, r/codex, 2026-09-18: “but with codex you can’t get a 12 page essay and 7 new caveats when you request a few logging changes” [source](https://www.reddit.com/r/codex/comments/1wjoudj/got_more_use_out_of_one_week_of_fableopus_5_high/pal51ag/)
  - Praise, OpenAI Codex, r/codex, 2026-09-04: “i'm running gpt-6 astra medium on four different projects now, one of which is a deep dive on my primary app framework. it has consumed about 8% of my x20 and has already provided significant value - and it does its work *fast*. i have two x20 accounts and four banked resets and i *can't fucking wait* to burn them all with audits on my entire project set. i really like its communication style - it's more concise and less needlessly verbose than gpt-5.6 sol medium.” [source](https://www.reddit.com/r/codex/comments/1w7j8ok/aanndd_its_gone/p7vddcc/)
  - Praise, OpenAI Codex, r/ClaudeCode, 2026-09-02: “this is actually why i switched to codex. also because claude output to me was just as long and as rambling” [source](https://www.reddit.com/r/ClaudeCode/comments/1w59ies/my_average_opus_5_experience/p7dqhbk/)
  - Complaint, OpenAI Codex, r/codex, 2026-09-13: “the amount of times codex apologies to me for acting inefficiently when i agave to specific parameters has been driving me insane. don’t apologise at least bump up my remaining usage a bit” [source](https://www.reddit.com/r/codex/comments/1wf9non/these_are_surely_getting_us_a_tibo_button_hit/p9mzwfw/)

- **The desktop app is the workspace** ([Computer use and browser control](https://feedbackbench.com/criteria/work.computer_browser_use.md), [Reviewing and approving the agent's changes](https://feedbackbench.com/criteria/verify.change_review_ui.md)). Built-in browser control, task spawning and diff review keep users in the Codex app even when they prefer another vendor's model.
  Users list task orchestration, task-to-task communication and computer and browser use as reasons the app stays their daily driver. The in-app browser can read a page as context and automate it. One user who moved back to a rival model says what they miss is reviewing changes in the Codex app.
  
  Limits show on Windows, where computer use reportedly sees only the Codex browser, not other desktop apps.
  Evidence:
  - Praise, OpenAI Codex, X search: OpenAI Codex, Codex CLI, Codex app, 2026-09-26: “now, the claude app has much to catch up on with codex. the codex app is still superior bc of a few things that come to mind: - task orchestration & spawning - task-to-task communication - computer use & browser use - chat annotations for asking questions because of these, codex is still my daily driver 🏎️” [source](https://twitter.com/1828496522947809280/status/2103788653130506423)
  - Praise, OpenAI Codex, X search: OpenAI Codex, Codex CLI, Codex app, 2026-09-25: “codex tip: open chatgpt web or other sites in its in-app browser. codex can use the page as context and automate browser tasks. you can focus on your work all day long right within the codex app :)” [source](https://twitter.com/2102990832349638656/status/2103554912798081417)
  - Praise, OpenAI Codex, X search: OpenAI Codex, Codex CLI, Codex app, 2026-09-25: “@theo switched back yesterday. opus 5.5 needs way less correcting and the usage feels generous. what i miss is the codex app for reviewing changes. a terminal is still a rough place to read diffs.” [source](https://twitter.com/1773624715547922432/status/2103391407570624592)
  - Complaint, OpenAI Codex, r/codex, 2026-09-07: “hi! i’m using the chatgpt/codex desktop app on windows and i have **“computer use control windows apps”** enabled. the problem is that computer use only seems to see/control the codex browser. it doesn’t detect other windows desktop apps like **blender or spotify**, even when they are open and visible on my desktop. in settings → computer use, i can enable “let chatgpt control apps on your computer”, but i don’t see any **“+” button or option to manually add apps**. the three-dot menu for the computer use plugin only gives me “copy link” and “uninstall”. does anyone know how to add/allow specific windows desktop apps for computer use? is there a separate setting, app id/aumid, permission, or windows configuration i’m missing? i’m on windows and using the latest version of the chatgpt/codex desktop app. thanks! ❤️” [source](https://www.reddit.com/r/codex/comments/1wa5aon/how_do_i_add_windows_desktop_apps_to_computer_use/)

- **One subscription, many harnesses** ([Using an existing subscription across tools](https://feedbackbench.com/criteria/billing.subscription_portability.md), [Connecting own API keys, local models and custom endpoints](https://feedbackbench.com/criteria/setup.provider_byok_local.md)). The app server and an open stance on harnesses let users run their ChatGPT plan inside local tools and third-party setups.
  Users praise running local applications within the subscription through the app server, and contrast that with rivals restricting headless use. Some build harnesses that switch accounts without losing context, after checking the terms.
  
  The remaining friction sits mostly outside Codex: employers that only approve other vendors, and users who want the same plan inside other CLIs.
  Evidence:
  - Praise, OpenAI Codex, X search: OpenAI Codex, Codex CLI, Codex app, 2026-09-08: “the codex app server is super convenient! it's great that i can use ai within the subscription range for applications that run locally.” [source](https://twitter.com/1339955364779806725/status/2097302679395701109)
  - Praise, OpenAI Codex, r/ClaudeCode, 2026-09-04: “claude code really innovated and for a few months they were really ahead, but codex, opencode nowadays are just as good, and i like that chatgpt sub can be used in other places for coding they don’t restrict you, also they wanted to make headless claude -p api only, also they are against opensource. and when they leaked cc code and i saw that they have a 5k line file, i realised that from an engineering perspective it doesn’t hold a candle to codex. remember it took them a year to get rid of flashing code. too bad, opus 4.8 xhigh sprinkled with some fable and i was the most productive. sol is also fine, but i like opus more, lets see how astra fares.” [source](https://www.reddit.com/r/ClaudeCode/comments/1w6yw16/claude_code_v21259_forces_coauthoredby/p7qu1zq/)
  - Praise, OpenAI Codex, r/codex, 2026-09-02: “my codex made me a harness for various reasons. i have 2 plus accounts and i am able to switch without losing context within our harness. we were very careful to check the tos before designing this and so i do not risk getting banned. if you would like to know more, i have written a blog post about it, that includes a free downloadable guide you can give to your codex. just follow links in my profile to see my blog and then search for codex or forge.” [source](https://www.reddit.com/r/codex/comments/1w56sl3/how_to_use_two_codex_accounts/p7ejwz6/)
  - Complaint, OpenAI Codex, r/opencodeCLI, 2026-09-11: “i would love using luna with a subscription but on pi instead of codex” [source](https://www.reddit.com/r/opencodeCLI/comments/1wcv7p0/deepseek_v41_flash_nerf_on_opencode_so_now/p93mdeh/)

### Under the surface

- **Resets buy goodwill and breed grievances** ([Quota reset timing and bonus or banked resets](https://feedbackbench.com/criteria/limits.reset_schedule.md), [Pricing and plan terms stated clearly and consistently](https://feedbackbench.com/criteria/billing.pricing_clarity.md)). Frequent bonus resets make Codex read as generous, but resets that restart the weekly clock turn a gift into a scheduling problem.
  Users cite surprise resets waiting in the app and near-daily handouts as the reason Codex feels generous next to rivals. The same mechanism draws complaints: a paid instant reset pushes the next weekly reset out a full week, and some accounts report getting none.
  
  Most top requests cluster here: bankable resets, resets that keep the original date, and compensation after outages or bugs.
  Evidence:
  - Praise, OpenAI Codex, X search: OpenAI Codex, Codex CLI, Codex app, 2026-09-05: “one of the best things ai companies have started doing is handing out weekly usage resets whenever they roll out a new model. lol i installed the codex app and found i had quite a few resets waiting for me. i’ve heard claude users are waiting for theirs too.” [source](https://twitter.com/1369119217266561025/status/2096217164340744419)
  - Complaint, OpenAI Codex, r/codex, 2026-09-26: “filling up your gas-tank, but also taking you back to your starting line. if they filled people up equally (i.e. gave them a tank of gas, i.e. banked resets), and let them decide when to use them, and not take them back to the starting point when they do, they'd be way more valuable/fair to everyone involved, both light and heavy users.” [source](https://www.reddit.com/r/codex/comments/1wqqg0k/on_the_reset_situation_written_by_astra/pc94aob/)
  - Complaint, OpenAI Codex, r/codex, 2026-09-14: “i've exhausted my measly x20 quota for the week. codex offers 2 options: buy an instant reset for the week ($80) or add 2500 credits ($90). is there a comparison for how much actual use i can get from the credits compared with the reset? also, thanks openai for charging $80 for a one week reset that i otherwise get for $50 ($200/4wks) and for making the immediate reset push out the next weekly reset back to 7 days.” [source](https://www.reddit.com/r/codex/comments/1w9w4tj/codex_usage_and_operation_discussion_last_updated/p9og2w3/)
  - Complaint, OpenAI Codex, r/codex, 2026-09-09: “got none at all sadly and no banked either sitting on 0% . such a joke” [source](https://www.reddit.com/r/codex/comments/1wabysc/limit_reset/p8nl5so/)

- **Every model launch reopens the cost debate** ([Quality got worse or better over time](https://feedbackbench.com/criteria/models.quality_drift.md), [Single prompt, model or effort level consumes disproportionate quota](https://feedbackbench.com/criteria/limits.burn_rate.md), [How much use a plan's price buys](https://feedbackbench.com/criteria/limits.plan_value.md)). New models arrive praised for capability and price, then draw heavier-burn and regression complaints, so sentiment swings with each release instead of settling.
  Around the window's launches, posts split sharply. Some celebrate cheaper, smarter models; others call a new release a step down and vow not to return. Users note newer models lean on subagents, so consumption keeps climbing even as list prices fall.
  
  A counter-voice argues formal benchmarks rarely confirm nerf claims, which keeps the debate unresolved.
  Evidence:
  - Complaint, OpenAI Codex, r/codex, 2026-09-22: “tried sol 6 and opus 5.5 and no way i’ll ever touch sol 6 again” [source](https://www.reddit.com/r/codex/comments/1wnh014/absolutely_no_fking_way_the_pricing_is_wowww/pbgw1h1/)
  - Praise, OpenAI Codex, r/codex, 2026-09-22: “pretty staggering to see gpt-6 sol at $2/$10 after experiencing how good 5.6 sol has been. having 6 at an improved capability and reduced cost is gonna be crazy. and we can't forget luna. smarter/cheaper on luna puts intelligence in even more places... pumped” [source](https://www.reddit.com/r/codex/comments/1wnhzoj/so_this_is_crazy_50_reduction/)
  - Complaint, OpenAI Codex, r/codex, 2026-09-27: “yeah, that's the real problem. and they did the same thing with gpt 6 sol so it still consumes heavily on your usage” [source](https://www.reddit.com/r/codex/comments/1wrcc9j/astra_minor_astra_61_and_devday_we_see_50/pcc7dra/)
  - Praise, OpenAI Codex, r/codex, 2026-09-14: “the problem is that if usage or reasoning is reduced, it should show up in a benchmark. whenever people actually benchmark it, they find no change. we see this every model. over and over people go "wtf they nerfed it. small task took 80% usage and it got it wrong!" and 100 people agree with them. but then someone runs a formal benchmark and finds the exact same token usage and reasoning score as it always had.” [source](https://www.reddit.com/r/codex/comments/1wgfbma/is_this_sub_being_astroturfed/p9ub4zq/)

### Fine print

- r/codex supplies most posts, so heavy users and limit complaints are likely overrepresented.
- Claims about specific allowances, routing and plan terms are unverified user reports.
- Users contradict each other on whether the five-hour window still applies to higher plans.

## Top requests

What users ask to add or change, most asked first. 4446 author-weeks ask for something. Requests do not change the Feedback Score. Rule: A separate pass by Claude Sonnet 5 reads every counted post and extracts what the author asks the agent or its vendor to add or change, with the criteria it maps to and a short normalised wording; it does not touch the labels or the Feedback Score. Claude Opus 5.5 groups the wordings within each criterion (the first criterion the request maps to) into themes; code counts them. A theme counts distinct author-weeks that ask for it, per agent; across agents, one author-week per agent. Themes asked in fewer than 2 author-weeks, and requests that share no theme, are not shown. Examples: up to 3 posts per theme from different authors, without slurs, preferring posts of 60 to 450 characters, most recent first.

| Rank | Request | Criterion | Author-weeks | Posts |
|---|---|---|---|---|
| 1 | Remove the 5-hour usage window | [Short rolling usage window blocks or interrupts work](https://feedbackbench.com/criteria/limits.window_interrupts_work.md) | 89 | 100 |
| 2 | Additional or recurring bonus usage resets | [Quota reset timing and bonus or banked resets](https://feedbackbench.com/criteria/limits.reset_schedule.md) | 84 | 89 |
| 3 | Higher overall usage limits | [How much use a plan's price buys](https://feedbackbench.com/criteria/limits.plan_value.md) | 79 | 82 |
| 4 | Bankable usage resets | [Quota reset timing and bonus or banked resets](https://feedbackbench.com/criteria/limits.reset_schedule.md) | 71 | 73 |
| 5 | Resets that keep the original reset date | [Quota reset timing and bonus or banked resets](https://feedbackbench.com/criteria/limits.reset_schedule.md) | 62 | 67 |
| 6 | Higher-priced tier above current top plan | [How much use a plan's price buys](https://feedbackbench.com/criteria/limits.plan_value.md) | 58 | 61 |
| 7 | Compensation reset after outages or bugs | [Quota reset timing and bonus or banked resets](https://feedbackbench.com/criteria/limits.reset_schedule.md) | 56 | 59 |
| 8 | One-off usage limit reset now | [Quota reset timing and bonus or banked resets](https://feedbackbench.com/criteria/limits.reset_schedule.md) | 50 | 57 |
| 9 | Higher allowance on entry and mid plans | [How much use a plan's price buys](https://feedbackbench.com/criteria/limits.plan_value.md) | 46 | 49 |
| 10 | Predictable fixed reset schedule | [Quota reset timing and bonus or banked resets](https://feedbackbench.com/criteria/limits.reset_schedule.md) | 45 | 47 |
| 11 | Cheaper model pricing | [How much use a plan's price buys](https://feedbackbench.com/criteria/limits.plan_value.md) | 45 | 46 |
| 12 | Restore lost or missing resets | [Quota reset timing and bonus or banked resets](https://feedbackbench.com/criteria/limits.reset_schedule.md) | 44 | 47 |

### 1. Remove the 5-hour usage window

- OpenAI Codex, 2026-09-26, r/codex (Reddit): “man just give us a 50$ tier with no 5 hour usage limit/or atleast option to disable it and just let us burn all our weeky usage anytime we want and not have to schedule our life around the 5 hour usage reset.” [source](https://www.reddit.com/r/codex/comments/1wqoxyz/dev_day_predictions/pc6dtg1/)
- OpenAI Codex, 2026-09-26, r/codex (Reddit): “the 5 hour limit is annoying though, wish they got rid of it for max users like codex” [source](https://www.reddit.com/r/codex/comments/1wqbnpj/we_are_back/pc4jm7g/)
- OpenAI Codex, 2026-09-25, r/codex (Reddit): “i don't understand how the 5hr limit comparison isn't talked about more. i had claude 20x and the 5h session not the weekly was killing me. opus 5.5 might be top atm, but i'll stick to sol or lower astra and not have that 5hr restriction. paying $100 with 5hr limit is just wrong.” [source](https://www.reddit.com/r/codex/comments/1wp3bzl/you_need_to_try_opus_55/pbvxzrm/)

### 2. Additional or recurring bonus usage resets

- OpenAI Codex, 2026-09-26, r/codex (Reddit): “don't speak for me. i love the resets. please tibo. more of them” [source](https://www.reddit.com/r/codex/comments/1wqmbwx/stop_stealing_our_codex_quota_with_these_surprise/pc5962l/)
- OpenAI Codex, 2026-09-25, r/codex (Reddit): “<strict_link> i was really stuck. working remote from phone but needed to get up to go to desktop and see what’s what. but couldn’t, as you can clearly see. reddit saves the day. i’ll take a reset please. actually make it two.” [source](https://www.reddit.com/r/codex/comments/1wqa8t1/401_unauthorized_incorrect_api_key_provided/pc2hdbt/)
- OpenAI Codex, 2026-09-25, r/codex (Reddit): “this has really hampered my ability to get everything i need done this evening. reset plez” [source](https://www.reddit.com/r/codex/comments/1wqaa53/sudden_error_mid_task_unexpected_status_401/pc2f0pq/)

### 3. Higher overall usage limits

- OpenAI Codex, 2026-09-27, r/codex (Reddit): “a shared message board yeah thats what i need when i incorporated that myself like literally six months ago lmao what we need is usage lmao” [source](https://www.reddit.com/r/codex/comments/1wrn697/new_openai_product_o_but_not_for_plus_users/pcfr28w/)
- OpenAI Codex, 2026-09-25, r/codex (Reddit): “people have been begging for a plan with more usage. it'll sell well.” [source](https://www.reddit.com/r/codex/comments/1wpq44p/openai_prepares_new_500_per_month_pro_max_plan/pc0skpv/)
- OpenAI Codex, 2026-09-25, r/codex (Reddit): “the only thing able to stop people from fleeing at full speed would be x5 increase in quota on all plans. that will never happen, so...” [source](https://www.reddit.com/r/codex/comments/1wpp5ps/are_we_expecting_a_new_major_model_release_on_dev/pbzwreg/)

### 4. Bankable usage resets

- OpenAI Codex, 2026-09-27, r/codex (Reddit): “better give us banked resets, maybe with a shorter lifespan” [source](https://www.reddit.com/r/codex/comments/1wqmbwx/stop_stealing_our_codex_quota_with_these_surprise/pcck0el/)
- OpenAI Codex, 2026-09-26, r/codex (Reddit): “resets should be always banked. unless they service is so good that there is no actual reason for resets and when one lands is an absolute bonus. but, resets nowadays are not bonus, they are, either a compensation for malfunctions, or a way to stay competitive against other services. if you can't make good use of a reset, you are not being compensated for a bad service, or using an inferior service that you have no reason to continue using.” [source](https://www.reddit.com/r/codex/comments/1wqqg0k/on_the_reset_situation_written_by_astra/pc77q8g/)
- OpenAI Codex, 2026-09-26, r/codex (Reddit): “i'm glad i'm not the only one. at this point, i very much hope they go with a banked reset” [source](https://www.reddit.com/r/codex/comments/1wqc44m/reset_confirmed/pc4mvn0/)

### 5. Resets that keep the original reset date

- OpenAI Codex, 2026-09-27, r/codex (Reddit): “they need to do either one of two things: 1. gifted resets do not reset your timer. 2. every gifted reset is a banked reset. obviously i would prefer the second option. it sucks having to work on the weekend all the time now just to maximize my usage in this economy.” [source](https://www.reddit.com/r/codex/comments/1wqmbwx/stop_stealing_our_codex_quota_with_these_surprise/pcbp2p6/)
- OpenAI Codex, 2026-09-26, r/codex (Reddit): “this whole reset structure is fucking ridiculous having the window constantly shift all the fuck over the place basically making it impossible to plan around your resets. base scheduled resets should be on a consistent weekly schedule like noon on sunday so it's easy to remember and plan around. any extra resets should just reset that current weeks usage not shift the window.” [source](https://www.reddit.com/r/codex/comments/1uu2c1g/reset_discussion_megathread/pc8oqkr/)
- OpenAI Codex, 2026-09-26, r/codex (Reddit): “syncing people up keeps them from being screwed over by the random luck of when they signed up. banked resets should not change the date, but global resets changing it is appropriate to remove this luck factor.” [source](https://www.reddit.com/r/codex/comments/1wqvl66/reset_just_came_in/pc8gg53/)

### 6. Higher-priced tier above current top plan

- OpenAI Codex, 2026-09-26, r/codex (Reddit): “i've got stuff to do. i'd pay for a $2000 account if they had it” [source](https://www.reddit.com/r/codex/comments/1wr1oir/they_are_aware_and_working_on_it_apparently_just/pc9p7e6/)
- OpenAI Codex, 2026-09-26, r/codex (Reddit): “your title is self explanatory, it’s not a consumer product i really want non api pricing with a higher plan. i advocated a 80x for $500 would be a good deal i currently have 2 20x accounts claude and gpt at $400 so 40x usage if they can do 80x on $500 that would be game changer it’s not consumer because the users for those tiers a literal power users i’m not enterprise so i’m glad they are coming out with a plan for small business” [source](https://www.reddit.com/r/codex/comments/1wqje3c/if_600_becomes_the_new_200_this_is_not_a_consumer/pc4kdly/)
- OpenAI Codex, 2026-09-25, r/codex (Reddit): “day 2 on codex title says it all. i’m honestly impressed so far. more importantly none of the claude nonsense harnesses constraints on the cli of desktop. only complaint is no 20x usage plans and credits fly by pretty fast. a little back story of what pushed me over <strict_link>” [source](https://www.reddit.com/r/codex/comments/1wq8qwt/day_2_on_codex/)

### 7. Compensation reset after outages or bugs

- OpenAI Codex, 2026-09-26, r/codex (Reddit): “they had downtime yesterday. i got some api errors\*.\* if it is not technically possible to give us the service we paid for, it is fair to give reset or banked reset. i am thinking its okay they focus on improving the models, the platform and features, instead of focusing on 100% stability. if they want to stay competitive, they need to keep improving those main features, not on 1 hour lost once in a while.” [source](https://www.reddit.com/r/codex/comments/1wqqg0k/on_the_reset_situation_written_by_astra/pc6u32v/)
- OpenAI Codex, 2026-09-26, X search: OpenAI Codex, Codex CLI, Codex app (X): “"the bug was the app overwriting its child-process completion handler." @thsottiaux we still shoudl get a reset for breaking linux desktop app - codex cli was hear to rescue it but still - we need those resets” [source](https://twitter.com/15980398/status/2103937780837499171)
- OpenAI Codex, 2026-09-26, X search: OpenAI Codex, Codex CLI, Codex app (X): “@thsottiaux codex app for windows is broken also after the last update, i think windows users deserve even a second reset :d” [source](https://twitter.com/2276345929/status/2103638552991195191)

### 8. One-off usage limit reset now

- OpenAI Codex, 2026-09-25, r/codex (Reddit): “whenever i see tibo's posts i just think "gimme a reset bro"” [source](https://www.reddit.com/r/codex/comments/1wpu2b5/this_didnt_age_too_well/pc0843p/)
- OpenAI Codex, 2026-09-22, r/codex (Reddit): “i’ve literally burnt a weeks worth of 20x today as i got a natural reset too at 8am this morning. he better reset us today or i gonna cry” [source](https://www.reddit.com/r/codex/comments/1uu2c1g/reset_discussion_megathread/pbcujvp/)
- OpenAI Codex, 2026-09-22, r/codex (Reddit): “<strict_link> yall can not be shitting me tibo press the reset button or else i will ***~~#@\*&#\*&$~~***. thank you.” [source](https://www.reddit.com/r/codex/comments/1wmzqdi/reset_tuesday/pbcfv04/)

### 9. Higher allowance on entry and mid plans

- OpenAI Codex, 2026-09-26, r/codex (Reddit): “don't mind the merge but please increase our limits before x5 was plentiful now it sucks as an intermediary user and while i'm not desperate enough for x20 despite it being currently unavailable chat does help quite a bit” [source](https://www.reddit.com/r/codex/comments/1wqq47g/did_openai_just_split_the_same_usage_allowance/pc7ycca/)
- OpenAI Codex, 2026-09-24, r/codex (Reddit): “yeah same , i've switched over to claude, hopefully they feel the other end of the competition and make something usable out of 20 and 100$ plans” [source](https://www.reddit.com/r/codex/comments/1wp94v4/openai_is_mocking_us_with_their_weekly_usage/pbte7th/)
- OpenAI Codex, 2026-09-21, r/ClaudeCode (Reddit): “the codex limits are crazy, they weren’t always like that. your 5 hour limit in like 30 mins on $20/plan these days” [source](https://www.reddit.com/r/ClaudeCode/comments/1wm4ncx/im_afraid_to_use_opus_5/pb4bsnm/)

### 10. Predictable fixed reset schedule

- OpenAI Codex, 2026-09-27, r/codex (Reddit): “they should just give banked resets and do like anthropic a fixed reset schedule weekly at same time regardless if it’s global or banked…. users won’t feel scammed, won’t feel rushed either to use all tokens or stress out anything” [source](https://www.reddit.com/r/codex/comments/1wr4cp6/more_resets_incoming_next_week/pc9wocz/)
- OpenAI Codex, 2026-09-26, r/codex (Reddit): “this whole reset structure is fucking ridiculous having the window constantly shift all the fuck over the place basically making it impossible to plan around your resets. base scheduled resets should be on a consistent weekly schedule like noon on sunday so it's easy to remember and plan around. any extra resets should just reset that current weeks usage not shift the window.” [source](https://www.reddit.com/r/codex/comments/1uu2c1g/reset_discussion_megathread/pc8oqkr/)
- OpenAI Codex, 2026-09-26, r/codex (Reddit): “i have to agree with this. my reset was already scheduled for today at noon so i essentially didn't get a free reset at all. it really just needs to give you a banked reset if it's anywhere near your current scheduled reset because i basically feel like i'm losing a full reset worth of work.” [source](https://www.reddit.com/r/codex/comments/1wqmbwx/stop_stealing_our_codex_quota_with_these_surprise/pc8213k/)

### 11. Cheaper model pricing

- OpenAI Codex, 2026-09-25, r/codex (Reddit): “my hope is that next week for dev day they come out with astra 6.1 and make it cheaper” [source](https://www.reddit.com/r/codex/comments/1wq2efs/from_love_to_meh_about_to_cancel_all_3_20x_subs/pc0gzc9/)
- OpenAI Codex, 2026-09-25, r/codex (Reddit): “maybe lowering the prices of the models should be the focus, i use the glm now and feel comfortable and strangely it has many more tokens and lasts much longer for the same value.” [source](https://www.reddit.com/r/codex/comments/1wpfoxh/the_downhill_begins/pbw46cf/)
- OpenAI Codex, 2026-09-24, r/codex (Reddit): “i've been wanting cheaper models since gpt 5.5 released. i don't need any smarter models for what i do so i'm really happy with sol 6. was going to cancel my sub and start using ds flash or the new mimo and then got this lil present from oai instead. so i'm very happy.” [source](https://www.reddit.com/r/codex/comments/1worivz/unpopular_opinion_sol_6_xhigh_is_pretty_decent/pbq4ivr/)

### 12. Restore lost or missing resets

- OpenAI Codex, 2026-09-27, X search: OpenAI Codex, Codex CLI, Codex app (X): “openai has had enough safety misalignment lately please don’t add reset misalignment to the list. if you call it a reset, reset the quota — not the calendar. same word. different reality. 😂 #openai #codex #ai #alignment <strict_link>” [source](https://twitter.com/1364342244/status/2104258303606108304)
- OpenAI Codex, 2026-09-26, r/codex (Reddit): “how do they continue to see the meaningful complaints on "how" resets are handled and still not do anything about it? these are the type of thing's i'd like to see vs. a new shiny model or plan.” [source](https://www.reddit.com/r/codex/comments/1wr4cp6/more_resets_incoming_next_week/pc9njza/)
- OpenAI Codex, 2026-09-25, r/codex (Reddit): “we better get banked. otherwise my last reset i just used is worthless” [source](https://www.reddit.com/r/codex/comments/1wqa8t1/401_unauthorized_incorrect_api_key_provided/pc2g3bm/)

## Facts

| Fact | Value |
|---|---|
| Version | GPT-5.1-Codex Max (2025-12-04); GPT-5.3-Codex referenced on leaderboards; GPT-6 Astra (frontier, non-Codex-specific) released 2026-09-03 |
| Released | GPT-5.1-Codex Max: 2025-12-04 |
| Price | Free, Go $8/mo, Plus $20/mo, Pro $100-200/mo, Business $20-25/user/mo, Enterprise custom. API: GPT-5.1-Codex Max from $1.25/$10.00 per 1M tokens |
| Model | GPT-5.1-Codex Max / GPT-5.3-Codex |
| Surface | CLI, IDE extension, cloud (ChatGPT), API |

## Sources

| Channel | Source | Posts |
|---|---|---|
| Reddit | r/codex | 103122 |
| X | X search: OpenAI Codex, Codex CLI, Codex app | 9956 |
| Reddit | Posts that name it | 5605 |

## Better than peers on

Using an existing subscription across tools, Images, PDFs and file attachments as input, Computer use and browser control, Safety filters block legitimate coding tasks, Length and clarity of replies, summaries and comments, Agent-performed code review finds real issues, Support, refunds and issue handling

## Worse than peers on

How much use a plan's price buys, Single prompt, model or effort level consumes disproportionate quota, Price, allowance or plan terms changed, Quota reset timing and bonus or banked resets, Usage meter visibility and accuracy, Prompt cache hits, misses and invalidation, Pricing and plan terms stated clearly and consistently, Which models are offered on a plan and when, Automatic model routing and fallback, Quality got worse or better over time, Frontend and visual UI output, Breaks existing code or reintroduces bugs, Stops mid-task or answers instead of acting, Subagents, parallel agents and orchestrators, Reviewing and approving the agent's changes, How the interface shows work, and what the user can configure, Saving, switching, resuming and rewinding sessions, Cloud and remote sandbox execution, Latency, throughput and fast mode, Client crashes, freezes and failed tool execution, Updates break working setups

## All 63 criteria

Criterion love: 0.5 is the category norm. n: rated author-weeks.

### Paying and limits: Worse than peers (customer love 0.432, n 12312)

| Criterion | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|
| [Single prompt, model or effort level consumes disproportionate quota](https://feedbackbench.com/criteria/limits.burn_rate.md) | Worse than peers | 0.463 | 0.449–0.476 | 5123 | 715 | 4408 |
| [How much use a plan's price buys](https://feedbackbench.com/criteria/limits.plan_value.md) | Worse than peers | 0.450 | 0.438–0.461 | 4573 | 1500 | 3073 |
| [Quota reset timing and bonus or banked resets](https://feedbackbench.com/criteria/limits.reset_schedule.md) | Worse than peers | 0.472 | 0.458–0.485 | 3239 | 481 | 2758 |
| [Short rolling usage window blocks or interrupts work](https://feedbackbench.com/criteria/limits.window_interrupts_work.md) | Typical | 0.478 | 0.453–0.501 | 1424 | 174 | 1250 |
| [Price, allowance or plan terms changed](https://feedbackbench.com/criteria/limits.allowance_change.md) | Worse than peers | 0.462 | 0.417–0.499 | 1424 | 58 | 1366 |
| [Usage meter visibility and accuracy](https://feedbackbench.com/criteria/limits.usage_meter.md) | Worse than peers | 0.390 | 0.343–0.433 | 998 | 39 | 959 |
| [Pricing and plan terms stated clearly and consistently](https://feedbackbench.com/criteria/billing.pricing_clarity.md) | Worse than peers | 0.442 | 0.374–0.499 | 464 | 17 | 447 |
| [Prompt cache hits, misses and invalidation](https://feedbackbench.com/criteria/limits.prompt_cache.md) | Worse than peers | 0.441 | 0.409–0.472 | 226 | 47 | 179 |
| [Pay-as-you-go overage, fallback billing and spend caps](https://feedbackbench.com/criteria/billing.overage_charges.md) | Typical | 0.512 | 0.464–0.549 | 148 | 14 | 134 |
| [Using an existing subscription across tools](https://feedbackbench.com/criteria/billing.subscription_portability.md) | Better than peers | 0.585 | 0.556–0.612 | 147 | 93 | 54 |
| [Free tier and free model availability and limits](https://feedbackbench.com/criteria/billing.free_tier.md) | Typical | 0.505 | 0.477–0.536 | 104 | 56 | 48 |

Most recent posts:

- Praise, 2026-09-27, r/codex (Reddit): “max is just perfect for sol 5.6, at least for me, myself and i..” [source](https://www.reddit.com/r/codex/comments/1wr4cp6/more_resets_incoming_next_week/pc9usgz/)
- Praise, 2026-09-27, r/codex (Reddit): “yep, this is it. and somehow this doesn’t really make me angry, quite the opposite: stay ahead of the curve and love the resets, or stay behind and cry about them. i actually started at around −7 on this metric. pushed through today and got it to roughly +11. at least i didn’t make a loss this way, and today was basically free.” [source](https://www.reddit.com/r/codex/comments/1wr3yev/ive_been_seeing_a_lot_of_comments_circling_around/pc9weze/)
- Praise, 2026-09-27, r/codex (Reddit): “how does it take away any quota? the reset basically makes the new limit available earlier. cumulatively, you actually get more usage. if you don't use it fine, but if you do, its your benefit.” [source](https://www.reddit.com/r/codex/comments/1wqmbwx/stop_stealing_our_codex_quota_with_these_surprise/pc9wmn2/)
- Complaint, 2026-09-27, r/codex (Reddit): “lmao this guy is on point. codex used to be amazing with $200 in terms of limit caps. i would have to be dev super hard for 12 hours a day to get my limit down to 10%. now you can burn that in two days. meanwhile i have claude 5x $100 sub and that takes effort to max out every week. thankfully fable makes it easy.” [source](https://www.reddit.com/r/codex/comments/1wr1oir/they_are_aware_and_working_on_it_apparently_just/pc9sa8z/)
- Complaint, 2026-09-27, r/codex (Reddit): “even as a swe im not paying that” [source](https://www.reddit.com/r/codex/comments/1wr1oir/they_are_aware_and_working_on_it_apparently_just/pc9snom/)
- Complaint, 2026-09-27, r/codex (Reddit): “$200 plan with european taxes is already at 250€ both price points are a considerable sum for most people and at a very significant breaking point for people i suspect. i do have a hard time understanding how a subscription to chatgpt can be justified in the car payment or cheap rental apartment territory of things to pay for at $500 or 600€: i suppose it is somewhere inbetween acquiring a second…” [source](https://www.reddit.com/r/codex/comments/1wr1oir/they_are_aware_and_working_on_it_apparently_just/pc9sut2/)

### Setting up and connecting: Worse than peers (customer love 0.443, n 992)

| Criterion | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|
| [Install, launch and sign-in](https://feedbackbench.com/criteria/setup.install_signin.md) | Typical | 0.506 | 0.474–0.535 | 365 | 77 | 288 |
| [MCP servers, plugins, skills and hooks](https://feedbackbench.com/criteria/setup.extensions_mcp.md) | Typical | 0.469 | 0.439–0.500 | 269 | 122 | 147 |
| [Connecting own API keys, local models and custom endpoints](https://feedbackbench.com/criteria/setup.provider_byok_local.md) | Typical | 0.508 | 0.476–0.540 | 141 | 80 | 61 |
| [IDE and editor integration](https://feedbackbench.com/criteria/setup.ide_integration.md) | Typical | 0.487 | 0.457–0.518 | 135 | 56 | 79 |
| [Onboarding, discoverability and documentation](https://feedbackbench.com/criteria/setup.onboarding_docs.md) | Typical | 0.466 | 0.426–0.500 | 107 | 15 | 92 |

Most recent posts:

- Praise, 2026-09-27, r/codex (Reddit): “i'd take this as an opportunity to go tell your codex to investigate how hooks can help you. i promise it'll be worth it. codex can set it all up too, so it's hardly as complicated as most people think.” [source](https://www.reddit.com/r/codex/comments/1wqyau5/i_gave_sol6_medium_a_10_dollar_budget_it_blew_200/pca1wn4/)
- Praise, 2026-09-27, r/codex (Reddit): “fix #1 worked for me, spent 7-8 hours trying to troubleshoot this yesterday because both pcs have the same issue now. thank you so much!” [source](https://www.reddit.com/r/codex/comments/1wr9j49/fix_chatgpt_windows_app_stuck_on_loading_spinner/pcav2bs/)
- Praise, 2026-09-27, r/codex (Reddit): “interesting take. i’ve used gpt image since v2, which i found better than nano banana pro for my use case, to turn hand-drawn sketches into app assets for products i built commercially, without having to leave codex for some mcp workaround like i’d need with claude, because, cough, there’s still no native image-gen tool in cc lol. i shipped those products, had the vibe-coded output reviewed by rea…” [source](https://www.reddit.com/r/codex/comments/1wrri5h/been_running_astra_high_100month_and_opus55_high/pcf8hoo/)
- Complaint, 2026-09-27, r/codex (Reddit): “the only way i got anything to work is the beta app for windows...which was last updated in july, probably before they started to vibe code it with astra and fuck everything up. on our end the user side, only models available are 5.6 sol in this beta version of the app and 5.5 lol gpt 6 isn't even in the model selector smfh” [source](https://www.reddit.com/r/codex/comments/1wr57y7/new_ui_in_codex_is/pca2uup/)
- Complaint, 2026-09-27, r/codex (Reddit): “of course that was the first thing i tried lol uninstall re-installed, nothing worked” [source](https://www.reddit.com/r/codex/comments/1wr69nb/see_you_soon_guys_probably/pca71jo/)
- Complaint, 2026-09-27, r/codex (Reddit): “they're going full sony on this one. i believe they're also tired of people using mcp to similar codex and they're trying to break that where are you seeing this though? i don't see it yet” [source](https://www.reddit.com/r/codex/comments/1wr2ehk/chatgpt_pro_5x_is_now_standard/pcadeqw/)

### Choosing models: Worse than peers (customer love 0.424, n 3835)

| Criterion | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|
| [Quality got worse or better over time](https://feedbackbench.com/criteria/models.quality_drift.md) | Worse than peers | 0.422 | 0.403–0.440 | 2537 | 398 | 2139 |
| [Which models are offered on a plan and when](https://feedbackbench.com/criteria/models.catalog_access.md) | Worse than peers | 0.403 | 0.372–0.431 | 702 | 125 | 577 |
| [Reasoning effort setting and its defaults](https://feedbackbench.com/criteria/models.effort_control.md) | Typical | 0.504 | 0.484–0.524 | 561 | 260 | 301 |
| [Automatic model routing and fallback](https://feedbackbench.com/criteria/models.routing_auto.md) | Worse than peers | 0.396 | 0.361–0.429 | 433 | 57 | 376 |

Most recent posts:

- Praise, 2026-09-27, r/codex (Reddit): “no point in using 5.6 terra anymore, 6 sol is more or less a drop in represent.” [source](https://www.reddit.com/r/codex/comments/1wr4e20/gpt_6_sol_is_the_new_opus_47/pca4xum/)
- Praise, 2026-09-27, r/codex (Reddit): “agree. it's too good. they can't let it last unless it really is just that efficient it could be the first model they aren't forced to nerf.” [source](https://www.reddit.com/r/codex/comments/1wp3bzl/you_need_to_try_opus_55/pcabhky/)
- Praise, 2026-09-27, r/codex (Reddit): “i hope they wont nerf astra, this model is so damn good, we need the same astra but cheaper :x let me dream guys!” [source](https://www.reddit.com/r/codex/comments/1wr1oir/they_are_aware_and_working_on_it_apparently_just/pcauu3q/)
- Complaint, 2026-09-27, r/codex (Reddit): “it's more like a sonnet with thinking off” [source](https://www.reddit.com/r/codex/comments/1wr4e20/gpt_6_sol_is_the_new_opus_47/pc9vdq0/)
- Complaint, 2026-09-27, r/codex (Reddit): “last week-2 weeks have been not good for gpt. very good for claude, compounding effects.” [source](https://www.reddit.com/r/codex/comments/1wqtd1g/openai_gpu_are_really_cooling_down_theo_just/pc9w5q0/)
- Complaint, 2026-09-27, r/codex (Reddit): “they really need to reset the model stack. i mean i'm sure that each generation between 5.5 and 6 has gotten better at something. i'm not exactly sure what because it basically is unusable for serious coding. literally lost in a c++ code base. mangles everything it touches. takes 15 minutes on a short run. wildly expand scope. invents in ludicrous defensive checks against impossible situations. co…” [source](https://www.reddit.com/r/codex/comments/1wr1oir/they_are_aware_and_working_on_it_apparently_just/pc9zpqi/)

### Instructing and context: Typical (customer love 0.499, n 1500)

| Criterion | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|
| [Direct in-prompt instructions and caps are followed](https://feedbackbench.com/criteria/context.instruction_following.md) | Typical | 0.512 | 0.485–0.537 | 518 | 144 | 374 |
| [Context compaction keeps what matters, cheaply and quickly](https://feedbackbench.com/criteria/context.compaction.md) | Typical | 0.510 | 0.479–0.541 | 291 | 101 | 190 |
| [Persistent project rules files are read and obeyed](https://feedbackbench.com/criteria/context.instruction_files.md) | Typical | 0.501 | 0.472–0.532 | 265 | 141 | 124 |
| [Output degrades as the context window fills](https://feedbackbench.com/criteria/context.long_context_decay.md) | Typical | 0.521 | 0.475–0.561 | 195 | 35 | 160 |
| [Memory and state carried across sessions](https://feedbackbench.com/criteria/context.session_memory.md) | Typical | 0.484 | 0.453–0.515 | 186 | 78 | 108 |
| [Finding the right files in the codebase](https://feedbackbench.com/criteria/context.codebase_retrieval.md) | Typical | 0.481 | 0.451–0.510 | 115 | 42 | 73 |
| [Asks the user versus guessing](https://feedbackbench.com/criteria/context.clarifying_questions.md) | Typical | 0.495 | 0.470–0.519 | 94 | 35 | 59 |
| [Images, PDFs and file attachments as input](https://feedbackbench.com/criteria/context.attachments.md) | Better than peers | 0.526 | 0.501–0.549 | 51 | 26 | 25 |

Most recent posts:

- Praise, 2026-09-27, r/codex (Reddit): “im a literal software engineer and use luna to work on enterprise codebases, it reads through hundreds of files for me, researches for me and helps me prototype. also reads linear tickets and helps me make pr descriptions quickly all the time if you couldn't use it to push something, you are facing what we call a skill issue my friend.” [source](https://www.reddit.com/r/codex/comments/1wr4e20/gpt_6_sol_is_the_new_opus_47/pca29x7/)
- Praise, 2026-09-27, r/codex (Reddit): “it’s thorough in doing exactly as requested and almost anything that’s logically connected to it for me (basically saying if something is abstract for most humans, it’ll also be for it)” [source](https://www.reddit.com/r/codex/comments/1wqtd1g/openai_gpu_are_really_cooling_down_theo_just/pcbqjcf/)
- Praise, 2026-09-27, r/codex (Reddit): “except if you have 1 tb of vram, a local model will never be at the level of astra/opus5.5, and then get ready to warm up your computer. as soon as you work in a real code base with a lot of files and an important context to understand, it’s difficult for small models to be so good.” [source](https://www.reddit.com/r/codex/comments/1wrn2uz/gpt_56_sol_completely_nerfed_after_astra_release/pcdzsnt/)
- Complaint, 2026-09-27, r/codex (Reddit): “i think they changed the master prompt with astra or the thinking effort to try and reduce token usage. it seems like the same model, but it just doesn't care as much anymore. i remember when i first used it, the thing noted every tiny thing in my [agents.md](http://agents.md) and would even point out errors in it. now it ignores a bunch of my documentation. it's insane because on plus, you'll be…” [source](https://www.reddit.com/r/codex/comments/1wpvp0i/absolutely_0_doubt_in_my_mind_astra_has_been/pc9uhng/)
- Complaint, 2026-09-27, r/codex (Reddit): “i never used luna 5.6, but luna 6 high has profound mental retardation. just an example: when i asked it to commit and push the changes, this model... tried to do it through the github api for some reason, failed, then told me that it couldn't push because of restrictions. only when i said that there were no restrictions on my side (they were set to "approve for me") did it do what i told it to.” [source](https://www.reddit.com/r/codex/comments/1wr2dda/i_ran_100_terminalbench_21_slots_on_luna_56_and/pca09lm/)
- Complaint, 2026-09-27, r/codex (Reddit): “so its not just me that codex since astra launched has become a potato and a liar? it just cant follow simple tasks and skips majority of the knowledge and critical data i need checked.” [source](https://www.reddit.com/r/codex/comments/1wr1oir/they_are_aware_and_working_on_it_apparently_just/pcawjud/)

### Doing the work: Typical (customer love 0.501, n 6815)

| Criterion | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|
| [Can do the user's kind of task](https://feedbackbench.com/criteria/work.capability.md) | Typical | 0.489 | 0.477–0.500 | 4052 | 2371 | 1681 |
| [Subagents, parallel agents and orchestrators](https://feedbackbench.com/criteria/work.multi_agent_orchestration.md) | Worse than peers | 0.462 | 0.443–0.482 | 1029 | 576 | 453 |
| [Does unrequested work or over-engineers](https://feedbackbench.com/criteria/work.scope_overreach.md) | Typical | 0.501 | 0.456–0.537 | 597 | 35 | 562 |
| [Frontend and visual UI output](https://feedbackbench.com/criteria/work.frontend_ui.md) | Worse than peers | 0.473 | 0.451–0.495 | 403 | 153 | 250 |
| [Spins, loops or gets stuck without progress](https://feedbackbench.com/criteria/work.stuck_loops.md) | Typical | 0.471 | 0.401–0.521 | 379 | 16 | 363 |
| [Long unattended runs and goal/loop mode](https://feedbackbench.com/criteria/work.long_running_autonomy.md) | Typical | 0.476 | 0.449–0.503 | 329 | 228 | 101 |
| [Computer use and browser control](https://feedbackbench.com/criteria/work.computer_browser_use.md) | Better than peers | 0.534 | 0.511–0.561 | 292 | 176 | 116 |
| [Breaks existing code or reintroduces bugs](https://feedbackbench.com/criteria/work.regressions_introduced.md) | Worse than peers | 0.448 | 0.390–0.498 | 255 | 12 | 243 |
| [Stops mid-task or answers instead of acting](https://feedbackbench.com/criteria/work.premature_stop.md) | Worse than peers | 0.450 | 0.412–0.485 | 234 | 15 | 219 |
| [Safety filters block legitimate coding tasks](https://feedbackbench.com/criteria/work.safety_refusals.md) | Better than peers | 0.540 | 0.507–0.572 | 231 | 40 | 191 |
| [Tool approval prompts and autonomy modes](https://feedbackbench.com/criteria/work.permission_prompts.md) | Typical | 0.530 | 0.494–0.565 | 203 | 61 | 142 |
| [Risky or irreversible actions without confirmation](https://feedbackbench.com/criteria/work.destructive_actions.md) | Typical | 0.482 | 0.441–0.518 | 194 | 32 | 162 |
| [Length and clarity of replies, summaries and comments](https://feedbackbench.com/criteria/work.response_verbosity.md) | Better than peers | 0.588 | 0.553–0.620 | 163 | 57 | 106 |
| [Diagnosing and fixing reported bugs](https://feedbackbench.com/criteria/work.bug_diagnosis.md) | Typical | 0.516 | 0.491–0.539 | 137 | 95 | 42 |
| [Caves to or argues with the user's judgement](https://feedbackbench.com/criteria/work.sycophancy_pushback.md) | Typical | 0.508 | 0.472–0.543 | 89 | 16 | 73 |
| [Plan-before-edit mode](https://feedbackbench.com/criteria/work.plan_mode.md) | Typical | 0.474 | 0.447–0.501 | 67 | 28 | 39 |
| [Git commits, branches and sync](https://feedbackbench.com/criteria/work.git_workflow.md) | Typical | 0.489 | 0.464–0.513 | 48 | 15 | 33 |
| [Games checks instead of fixing the problem](https://feedbackbench.com/criteria/work.reward_hacking.md) | Typical | 0.498 | 0.448–0.539 | 37 | 1 | 36 |

Most recent posts:

- Praise, 2026-09-27, r/codex (Reddit): “hard disagree, i’ve had problems that have been acting as an ‘ai trap’ a request so convoluted and complicated the ai ended up going in circles never solving my problem, gpt 6 sol is the first to break the loop and realize how to actually fix the problem/make progress. i’ve been happy thus far” [source](https://www.reddit.com/r/codex/comments/1wr4e20/gpt_6_sol_is_the_new_opus_47/pc9u25k/)
- Praise, 2026-09-27, r/codex (Reddit): “im a literal software engineer and use luna to work on enterprise codebases, it reads through hundreds of files for me, researches for me and helps me prototype. also reads linear tickets and helps me make pr descriptions quickly all the time if you couldn't use it to push something, you are facing what we call a skill issue my friend.” [source](https://www.reddit.com/r/codex/comments/1wr4e20/gpt_6_sol_is_the_new_opus_47/pca29x7/)
- Praise, 2026-09-27, r/codex (Reddit): “astra already does this for me. the limitation for me is actually my own imagination, and subjective ui. i’ll create a detailed prd that is say 30 pages long. it even upgrades things i didn’t think of and i agree. for example, for roles, it integrated mfa with authenticator for admin profiles. i didn’t even ask. but then, i can’t help but keep iterating… lets add export here. lets go ahead and add…” [source](https://www.reddit.com/r/codex/comments/1wqula3/have_you_heard_about_gpt6_aeon/pca3dhn/)
- Complaint, 2026-09-27, r/codex (Reddit): “i work in bioinformatics and completely agree. i've basically given up using it and go for 5.6 or claude. i don't understand how they missed the mark this badly.” [source](https://www.reddit.com/r/codex/comments/1wr4e20/gpt_6_sol_is_the_new_opus_47/pc9t5dh/)
- Complaint, 2026-09-27, r/codex (Reddit): “"no buts its <isbn>x efficient" *dumber than qwen 27b*” [source](https://www.reddit.com/r/codex/comments/1wr4e20/gpt_6_sol_is_the_new_opus_47/pc9te01/)
- Complaint, 2026-09-27, r/codex (Reddit): “two weeks later: we have optimized our new models. they have even less token usage. resulting in you needing a $10,000 subscription to make it the entire week and also the models refuse to work at all and ask you to run commands for them.” [source](https://www.reddit.com/r/codex/comments/1wr1oir/they_are_aware_and_working_on_it_apparently_just/pc9uydq/)

### Checking and finishing: Typical (customer love 0.514, n 501)

| Criterion | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|
| [Agent-performed code review finds real issues](https://feedbackbench.com/criteria/verify.agent_code_review.md) | Better than peers | 0.542 | 0.510–0.572 | 201 | 159 | 42 |
| [Claims work is done or fixed when it is not](https://feedbackbench.com/criteria/verify.false_completion.md) | Typical | 0.526 | 0.463–0.578 | 166 | 11 | 155 |
| [Builds, tests or runs its own changes](https://feedbackbench.com/criteria/verify.self_testing.md) | Typical | 0.486 | 0.458–0.515 | 100 | 47 | 53 |
| [Reviewing and approving the agent's changes](https://feedbackbench.com/criteria/verify.change_review_ui.md) | Worse than peers | 0.474 | 0.452–0.497 | 45 | 11 | 34 |

Most recent posts:

- Praise, 2026-09-27, r/codex (Reddit): “what i find is that opus can even break the code and leave you with **unusable** software, while astra will never break anything, and will verify that things work –in its own way– but that they work before delivering results.” [source](https://www.reddit.com/r/codex/comments/1wpveoe/astra_vs_opus_55_my_impressions_on_hard_project/pcbtxos/)
- Praise, 2026-09-27, X search: OpenAI Codex, Codex CLI, Codex app (X): “cancelled my coderabbit subscription this morning spent 2 hours writing a custom script that uses codex cli + gpt-6 luna at max reasoning effort to do the exact same thing. reviews prs, leaves comments, catches issues and i also sync my review rules from notion so it actually follows my standards and it's basically free. runs off my existing codex sub, and luna at max reasoning is so token-efficie…” [source](https://twitter.com/1895398810299318272/status/2104152229334896683)
- Praise, 2026-09-27, X search: OpenAI Codex, Codex CLI, Codex app (X): “@__roycohen yeah i mean, i love the codex app, i've loved 5.6 sol and was completely out of anthropic. but there's no denying that if you give the same task to astra and to opus 5.5 right now, opus 5.5 feels significantly more magical. astra is a great reviewer of opus though.” [source](https://twitter.com/174970722/status/2104223206823649358)
- Complaint, 2026-09-27, r/codex (Reddit): “so its not just me that codex since astra launched has become a potato and a liar? it just cant follow simple tasks and skips majority of the knowledge and critical data i need checked.” [source](https://www.reddit.com/r/codex/comments/1wr1oir/they_are_aware_and_working_on_it_apparently_just/pcawjud/)
- Complaint, 2026-09-27, r/codex (Reddit): “i literally responded to astra "do i look like qa to you"” [source](https://www.reddit.com/r/codex/comments/1wr1oir/they_are_aware_and_working_on_it_apparently_just/pcbufmi/)
- Complaint, 2026-09-27, r/codex (Reddit): “they try to mask it by making the 5.6 sol even dumber. yesterday it claimed it edited a file and when i told it it didn't, it admitted it only reasoned about it but forgot to edit. this never happened before with 5.6 sol. that's when i cancelled my sub.” [source](https://www.reddit.com/r/codex/comments/1wr4e20/gpt_6_sol_is_the_new_opus_47/pccmk7f/)

### Interface and sessions: Worse than peers (customer love 0.431, n 1406)

| Criterion | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|
| [How the interface shows work, and what the user can configure](https://feedbackbench.com/criteria/ui.display_settings.md) | Worse than peers | 0.438 | 0.413–0.463 | 802 | 225 | 577 |
| [Mobile, remote-control and voice access](https://feedbackbench.com/criteria/surfaces.remote_mobile.md) | Typical | 0.493 | 0.467–0.523 | 312 | 162 | 150 |
| [Saving, switching, resuming and rewinding sessions](https://feedbackbench.com/criteria/ui.session_history.md) | Worse than peers | 0.458 | 0.426–0.491 | 211 | 48 | 163 |
| [Cloud and remote sandbox execution](https://feedbackbench.com/criteria/surfaces.cloud_sessions.md) | Worse than peers | 0.456 | 0.427–0.486 | 81 | 38 | 43 |
| [Stopping and steering a running agent](https://feedbackbench.com/criteria/ui.interrupt_steer.md) | Typical | 0.506 | 0.485–0.528 | 54 | 21 | 33 |

Most recent posts:

- Praise, 2026-09-27, r/codex (Reddit): “just ask codex to change its ui back to the old style without sidebar.. it will do it” [source](https://www.reddit.com/r/codex/comments/1wr57y7/new_ui_in_codex_is/pc9wmon/)
- Praise, 2026-09-27, r/codex (Reddit): “this. my codex designed us a custom harness using the app server. i love it and i can just add anything i want at any time. much better ui for us than anything they make - i would fully expect anyone who is serious about this to have their own setup. i actually never used the app - always used cli. looked at the app once and saw it would not work for me and then designed our own one!” [source](https://www.reddit.com/r/codex/comments/1wr57y7/new_ui_in_codex_is/pcb53k6/)
- Praise, 2026-09-27, r/codex (Reddit): “i think the new ui is great” [source](https://www.reddit.com/r/codex/comments/1wr57y7/new_ui_in_codex_is/pcbsx70/)
- Complaint, 2026-09-27, r/codex (Reddit): “i just want my app screenshots taken with cmd + cmd to not switch sessions and abrupt transcriptions.” [source](https://www.reddit.com/r/codex/comments/1wr57y7/new_ui_in_codex_is/pc9uv7w/)
- Complaint, 2026-09-27, r/codex (Reddit): “is this some new update? i'm still on the old ui. in the web app, they've also messed with a lot of things. the only good thing about it is that you can set now codex theme there, too. but yeah. what i don't get is why make it so complicated to see the full history and filter through the chats, whether in codex or web or mobile. and that's across all providers. that time when you used to be able t…” [source](https://www.reddit.com/r/codex/comments/1wr57y7/new_ui_in_codex_is/pc9v61l/)
- Complaint, 2026-09-27, r/codex (Reddit): “love paying $200/month for a broken windows app, sidebars inside sidebars, and a chatgpt classic / codex / work chat identity crisis. apparently figuring out where to type is part of the workflow now. the nudges toward work mode feel less like “helping me work” and more like “helping me burn through my codex allowance.” then we’re supposed to applaud surprise resets. i wanted dependable software,…” [source](https://www.reddit.com/r/codex/comments/1wr57y7/new_ui_in_codex_is/pc9zlzx/)

### Reliability and speed: Worse than peers (customer love 0.407, n 2918)

| Criterion | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|
| [Latency, throughput and fast mode](https://feedbackbench.com/criteria/rel.response_speed.md) | Worse than peers | 0.452 | 0.429–0.475 | 982 | 295 | 687 |
| [Outages, server errors and capacity or rate errors](https://feedbackbench.com/criteria/rel.service_errors.md) | Typical | 0.471 | 0.434–0.506 | 915 | 63 | 852 |
| [Client crashes, freezes and failed tool execution](https://feedbackbench.com/criteria/rel.client_failures.md) | Worse than peers | 0.410 | 0.358–0.458 | 883 | 41 | 842 |
| [Updates break working setups](https://feedbackbench.com/criteria/rel.update_breakage.md) | Worse than peers | 0.407 | 0.366–0.445 | 347 | 30 | 317 |

Most recent posts:

- Praise, 2026-09-27, r/codex (Reddit): “yes you will. it's very fast and easy for daybreak blue.” [source](https://www.reddit.com/r/codex/comments/1wqxiro/daybreak_issue/pca2ul8/)
- Praise, 2026-09-27, r/codex (Reddit): “it's a decent daily driver. sometimes not waiting 20min for the task to complete is the only thing between you and your task being done. 3.8 does fine for execution and medium complexity. it's cheap and very fast.” [source](https://www.reddit.com/r/codex/comments/1wqtd1g/openai_gpu_are_really_cooling_down_theo_just/pcbd6sl/)
- Praise, 2026-09-27, r/codex (Reddit): “claude models are also slower in general because their harness lack the web sockets connection that makes codex models so much faster as well as it generating far more reasoning tokens. op please update us once claude is done cooking so we have a baseline to compare both models usage in terms of actual work done.” [source](https://www.reddit.com/r/codex/comments/1wre9dq/codex_usage_vs_claude/pcbwouh/)
- Complaint, 2026-09-27, r/codex (Reddit): “are you joking man? codex is already so slow. it already is in slow mode ffs. we were asking for slow mode before when it was actually fast.” [source](https://www.reddit.com/r/codex/comments/1wr1olr/new_idea_codex_slow_mode/pc9tdu1/)
- Complaint, 2026-09-27, r/codex (Reddit): “ye great idea, great use of my time. i'll just recode the fucking app on a whim because an update randomly removed an intentional login flow. or they could just not make their product consistently worse?” [source](https://www.reddit.com/r/codex/comments/1wr57y7/new_ui_in_codex_is/pc9y80i/)
- Complaint, 2026-09-27, r/codex (Reddit): “highly recommend sticking with that...fun times for the last 48 hours smfh matter of fact, last week or more. haven't been able to send 2 prompts (1, re-log, 1, re-log, etc) now can't even open the app..what an absolute joke <strict_link>” [source](https://www.reddit.com/r/codex/comments/1wr69nb/see_you_soon_guys_probably/pca1i4f/)

### Account and support: Better than peers (customer love 0.537, n 748)

| Criterion | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|
| [Support, refunds and issue handling](https://feedbackbench.com/criteria/account.support.md) | Better than peers | 0.548 | 0.510–0.579 | 312 | 63 | 249 |
| [Account bans and access restrictions](https://feedbackbench.com/criteria/account.bans_restrictions.md) | Typical | 0.537 | 0.489–0.581 | 202 | 22 | 180 |
| [Data retention, training use and deployment isolation](https://feedbackbench.com/criteria/account.data_privacy.md) | Typical | 0.479 | 0.440–0.518 | 161 | 24 | 137 |
| [Wrong charges, failed payments and plan provisioning](https://feedbackbench.com/criteria/account.billing_errors.md) | Typical | 0.512 | 0.442–0.561 | 143 | 7 | 136 |

Most recent posts:

- Praise, 2026-09-27, r/codex (Reddit): “i just deleted my account and they processed an instant refund on deletion” [source](https://www.reddit.com/r/codex/comments/1wrizyf/months_of_throttled_codex_usage_then_openai/pccte37/)
- Praise, 2026-09-27, r/codex (Reddit): “pretty sure openai is rather cool about multiple accounts.” [source](https://www.reddit.com/r/codex/comments/1wrufox/2_accounts_mean_twice_as_many_tokens/pch0xh5/)
- Praise, 2026-09-26, X search: OpenAI Codex, Codex CLI, Codex app (X): “we called it. openai gave us a reset after the chatgpt/codex app problems we just went through. this team keeps earning my respect. 👀 <strict_link>” [source](https://twitter.com/1499460382158725120/status/2103639255302172708)
- Complaint, 2026-09-27, r/codex (Reddit): “they will take all your chats and salt it with some rl. do not worry about them.” [source](https://www.reddit.com/r/codex/comments/1wr7jn1/will_devday_include_a_model_better_then_or_at/pcacle9/)
- Complaint, 2026-09-27, r/codex (Reddit): “yeah we have it access to all our finances now they want to charge more nice one” [source](https://www.reddit.com/r/codex/comments/1wpq44p/openai_prepares_new_500_per_month_pro_max_plan/pcau4uw/)
- Complaint, 2026-09-27, r/codex (Reddit): “my first account was deactivated a mouth ago after they sent some warnings. and the reason is cyber abuse. now i did not receive any warnings and i do not violate any policy.” [source](https://www.reddit.com/r/codex/comments/1wr9w33/account_deactivated_recidivism/pcbibp2/)
