# Which models are offered on a plan and when (`models.catalog_access`)

Feedback Bench, coding agents, built 2026-10-01, window 2026-08-31 to 2026-09-27. Web page: https://feedbackbench.com/#/criterion/models.catalog_access

Area: [Choosing models](https://feedbackbench.com/criteria/models.md)

**Definition.** Whether models are available on the user's plan: same-day support for new releases, deprecations and removals, multi-provider choice, and models missing from a surface.

**Boundary.** Not this: see [Free tier and free model availability and limits](https://feedbackbench.com/criteria/billing.free_tier.md) for free-model availability. Not this: see [Automatic model routing and fallback](https://feedbackbench.com/criteria/models.routing_auto.md) for which model actually runs.

Rated author-weeks, all agents: 2268. Complaint share: 72%.

## The brief

Written by Claude Opus 5.5 from 110 labelled posts and the numbers on this page. Interpretation, not measurement: every quote is verbatim and links to its post.

**Users want every new model, on their plan, the day it ships.**

TL;DR:

- Multi-provider harnesses like OpenCode, Cursor and Devin win by shipping new models fast and letting users switch.
- OpenAI Codex, Claude Code and Google Antigravity lose ground on removals, tier gating and stale third-party models.
- Top asks are newest models on cheaper plans, restoring removed models, and parity across surfaces.

In plain terms: In practice, users pick a tool for its model list. They get frustrated when the best model sits on a higher tier, a favorite gets retired, or a release reaches the web but not the CLI.

### How it breaks

- **Best model locked above your tier** ([Which models are offered on a plan and when](https://feedbackbench.com/criteria/models.catalog_access.md)). The most common complaint is a frontier model that exists but is withheld from the user's plan.
  Claude Code users on Pro find the top model missing and say so bluntly. OpenAI Codex users on the entry plan want the bigger models instead of only the light tier. Devin's free plan offers a single model. The request for newest models on lower-priced plans tops the list, led by OpenAI Codex and Claude Code askers. Some users read it the other way and call the Codex entry plan generous next to Anthropic's.
  Evidence:
  - Complaint, Claude Code, r/ClaudeCode, 2026-09-27: “i just looked at it. you are right. fable is the thing that you cannot use in pro.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wrteof/is_pro_subscription_good_enough_for_unreal_engine/pcfmwce/)
  - Complaint, Claude Code, @ClaudeDevs, 2026-09-01: “@claudedevs why the fuck i can't use the model on my pro plan?!” [source](https://twitter.com/990750394245615616/status/2094867519983026370)
  - Complaint, OpenAI Codex, r/codex, 2026-09-13: “im sorry ut gemini flash 3.8 needs special handling but is insanely fast and good, calling it f is fucking nuts, i can actually do a week of work with it, i love codex, i wish i could use sol/astra to really make use of anything other than luna xhigh on the 20$, but gemini flash when given a good plan is fucking insanely good, (as long as you use fresh chats as its context compaction is shit still in anti gravity)” [source](https://www.reddit.com/r/codex/comments/1wfdcmy/value_of_20_plans_what_do_you_think_about_this/p9l1fj2/)
  - Complaint, Devin, @cognition, 2026-09-12: “i'm tried @cognition on free plan. swe-1.6 slow model is the only model available, but it's not slow at all!. i wish i have access to latest model as well.” [source](https://twitter.com/1534111445985636352/status/2098665173670334911)

- **Third-party models left to go stale** ([Which models are offered on a plan and when](https://feedbackbench.com/criteria/models.catalog_access.md)). Tools that resell other labs' models get punished when the catalog lags the lab's own release by a version or more.
  Google Antigravity draws the sharpest version of this. Users ask why they are still on an older Opus and whether offering outside models makes sense if they stay outdated. Kiro users say new models arrive slowly because they are vetted first, and they check daily for the newest release. Cline users point to rivals that enabled a new model within minutes of the announcement.
  Evidence:
  - Complaint, Google Antigravity, @antigravity, 2026-09-22: “add opus 5.5 here why are we seriously on 4.6 still google why do you even offer other models than yourself in the first place if theyre so outdated either dont do it or make them actually semi up to date @antigravity <strict_link>” [source](https://twitter.com/1312430180510597122/status/2102528820062646360)
  - Complaint, Kiro, r/kiroIDE, 2026-09-06: “my employer provides kiro since we have so many aws interactions and it’s really great for that. it uses opus/sonnet under the hood so it’s decent at coding. the only real downside is that it’s slow to add newer models because they have to be vetted first. the new v3 kiro added a harness environment so we’re getting things like managing subagents added in.” [source](https://www.reddit.com/r/kiroIDE/comments/1w7ihhq/es_kiro_mejor_para_el_desarrollo/p82oc3d/)
  - Complaint, Cline, @cline, 2026-09-10: “@cline kindly please be a little bit faster to enable deepseek 4.1 flash in clinepass. opencode and commandcode enabled this under 30 mins of official announcement.” [source](https://twitter.com/10877782/status/2097944041832734973)
  - Complaint, Kiro, r/kiroIDE, 2026-09-26: “no mate i've been checking everyday in my kiro ide and it's not there yet. only opus 5” [source](https://www.reddit.com/r/kiroIDE/comments/1wpveza/kiro_and_opus_55/pc44hzp/)

- **Retirements break the daily driver** ([Which models are offered on a plan and when](https://feedbackbench.com/criteria/models.catalog_access.md)). When a preferred model is scheduled for removal, users lose a tuned workflow and ask for it back.
  OpenAI Codex posts show people asking what to use once a favored model is gone, and weighing replacements against a version they liked for its usage rate. OpenCode users report models disappearing without notice. Restoring removed models and keeping older ones after new releases are both frequent asks, with OpenAI Codex leading each. The reverse also appears. One Google Antigravity user wants a weak legacy option pulled.
  Evidence:
  - Complaint, OpenAI Codex, r/codex, 2026-09-22: “but what im gonna use as mt daily driver after they remove 5.6 sol? i hope they fix sol 6.1” [source](https://www.reddit.com/r/codex/comments/1wnidbh/6_sol_is_wayy_worse_than_56_sol/pbfz0py/)
  - Complaint, OpenAI Codex, r/codex, 2026-09-22: “okay so how do these compare to 5.5 extra high? that’s meant to be going soon and i’ve honestly preferred its usage rate and intellect to the others but i’m gonna have to move on now with it leaving soon” [source](https://www.reddit.com/r/codex/comments/1wnhuru/gpt6_pricing_actually_has_me_optimistic_about/pbfg2ey/)
  - Complaint, OpenCode, @opencode, 2026-09-18: “@opencode the model is not available to use , seems like you guys removed it” [source](https://twitter.com/2067870214386307072/status/2100914050192548147)
  - Complaint, Google Antigravity, r/google_antigravity, 2026-09-18: “better, faster, cheaper. shame google doesn't remove the 3.1 pro option at all.” [source](https://www.reddit.com/r/google_antigravity/comments/1wd9kqp/typical_interaction_with_gemini_31_pro_high_on/paj1965/)

- **Org policy hides the catalog** ([Which models are offered on a plan and when](https://feedbackbench.com/criteria/models.catalog_access.md), [Automatic model routing and fallback](https://feedbackbench.com/criteria/models.routing_auto.md)). In GitHub Copilot, admins and data-retention rules decide which models developers actually see.
  Users describe whole vendors' models being disabled by their organization, sometimes for privacy reasons they cannot see. One report says auto mode still routed a request to a model that could not be selected manually. Another says a model they relied on vanished once the org marked it unapproved. Kiro enterprise users report a model missing from their list too. The product works as designed. The friction is opacity.
  Evidence:
  - Complaint, GitHub Copilot, r/GithubCopilot, 2026-09-12: “unless, of course, your org disables it because it's not an approved model... that thing was so useful while we had it.” [source](https://www.reddit.com/r/GithubCopilot/comments/1we096r/making_copilot_think_like_native_claude/p9g8h2j/)
  - Complaint, GitHub Copilot, r/GithubCopilot, 2026-09-17: “yup, all model explicitly disabled. imagine my suprise when the dev set the github chat to auto, asked it a question, and claude popped up as the llm that the request used! you can't even select claude manually... re, why we're not allowed to use a claude, way out of my paygrade. it's come from the top but as i understand it, something to do with privacy.” [source](https://www.reddit.com/r/GithubCopilot/comments/1win2b6/github_copilot_is_a_nightmare_for_enterprises/pacj3f0/)
  - Complaint, GitHub Copilot, r/GithubCopilot, 2026-09-08: “same. we’ve never gotten fable but that’s because of data retention.” [source](https://www.reddit.com/r/GithubCopilot/comments/1wag6w6/breaking_msft_has_stopped_providing_claude_models/p8j7g4h/)
  - Complaint, Kiro, r/kiroIDE, 2026-09-26: “i'm on enterprise subscription and fable is not in the list <strict_link>” [source](https://www.reddit.com/r/kiroIDE/comments/1wpveza/kiro_and_opus_55/pc5k9ad/)

- **Web gets it, CLI waits** ([Which models are offered on a plan and when](https://feedbackbench.com/criteria/models.catalog_access.md)). A model that lands on one surface but not another reads to users as an unfinished rollout.
  A Factory user found a new model on the web version but not in the CLI. Claude Code users ask why a model shipped only in the unstable upgrade channel. Parity across app, CLI and platforms is a top request, overwhelmingly from OpenAI Codex users. Users expect one subscription to mean one catalog everywhere they run it.
  Evidence:
  - Complaint, Factory, @droid, 2026-09-10: “@iamshivii @droid just checked the web version, and yeah, astra is there now. still nothing in the cli though, which is kinda weird.” [source](https://twitter.com/1260921/status/2097939105883505019)
  - Complaint, Claude Code, @ClaudeDevs, 2026-09-02: “@claudeai @claudedevs why did you allowed it only in unstable claude code upgrade channel?” [source](https://twitter.com/1509920652236627968/status/2095198270695825912)

### Who stands out

- **OpenAI Codex (weaker)**. The biggest pile of catalog complaints, driven by retirements, light-tier limits on the entry plan, and gaps between surfaces.
  Users plan around models that are scheduled to leave and ask what replaces them. Entry-plan users want access beyond the light model. OpenAI Codex leads requests for restoring removed models, for parity across app and CLI, and for a cheaper capable tier. A minority praise the plan as more generous than Anthropic's at the same price.
  Evidence:
  - Complaint, OpenAI Codex, r/codex, 2026-09-22: “okay so how do these compare to 5.5 extra high? that’s meant to be going soon and i’ve honestly preferred its usage rate and intellect to the others but i’m gonna have to move on now with it leaving soon” [source](https://www.reddit.com/r/codex/comments/1wnhuru/gpt6_pricing_actually_has_me_optimistic_about/pbfg2ey/)
  - Complaint, OpenAI Codex, r/codex, 2026-09-13: “im sorry ut gemini flash 3.8 needs special handling but is insanely fast and good, calling it f is fucking nuts, i can actually do a week of work with it, i love codex, i wish i could use sol/astra to really make use of anything other than luna xhigh on the 20$, but gemini flash when given a good plan is fucking insanely good, (as long as you use fresh chats as its context compaction is shit still in anti gravity)” [source](https://www.reddit.com/r/codex/comments/1wfdcmy/value_of_20_plans_what_do_you_think_about_this/p9l1fj2/)
  - Praise, OpenAI Codex, r/GithubCopilot, 2026-09-09: “$20 for claude and $20 for codex is working well for me. if i hit a 5 hour window i just switch to the other. i also like codex gives you quota resets every so often you can use as a get out of jail free card. if gives me some assurance that if i need it in a pinch i'll have it. plus also for codex, they'll let you use astra on the $20 plan (claude charges additional for fable on that tier, but opus is great also).” [source](https://www.reddit.com/r/GithubCopilot/comments/1wambt6/switched_to_claude_code/p8txpen/)
  - Praise, OpenAI Codex, r/codex, 2026-09-05: “compare it to anthropic where the plus or pro account doesnt even get access to their best model” [source](https://www.reddit.com/r/codex/comments/1w7xfo2/rant_getting_tired_of_all_the_posts_complaining/p7yixyy/)

- **Google Antigravity (weaker)**. Praise for fast native Gemini support is swamped by anger over an outdated Claude lineup.
  Users credit Google Antigravity for shipping its own Flash models early and natively. The complaints target outside models. Users repeatedly ask to swap an older Opus for the current one, and requests to update outdated Claude models come almost entirely from this tool. Users also want more current open-weight options added.
  Evidence:
  - Complaint, Google Antigravity, @antigravity, 2026-09-24: “@rodydavis can we please swaap opus 4.6 with 5.5 in @antigravity . let the peasants have a crumb of peak intelligence. we beg.” [source](https://twitter.com/3097150334/status/2103160575824191565)
  - Praise, Google Antigravity, @antigravity, 2026-09-18: “@google @antigravity @googleaistudio antigravity finally getting native flash 3.8 support is a real upgrade for agent workflows.” [source](https://twitter.com/2058874736470327296/status/2101097223958184046)
  - Complaint, Google Antigravity, @antigravity, 2026-09-05: “@nlycskn @antigravity @thtbee_ antigravity agent name with gemini access will be great. also gemini 3.8 is super fast almost 10-15 faster than codex. you should add new models like latest claude models and new oss models <strict_link>” [source](https://twitter.com/1731716432658857984/status/2096240402336452953)
  - Complaint, Google Antigravity, @antigravity, 2026-09-23: “@pigeon__s @antigravity while they're at it, maybe add a few more open source models too, latest deepseek and mimo. i mean they still have gpt oss, why not add more. google one (or google ai) is actually super worth it cause its not just ai, its the whole google ecosystem including google drive too” [source](https://twitter.com/432486005/status/2102609080946929880)

- **OpenCode (stronger)**. Users choose OpenCode precisely to avoid being locked to one provider, and they praise switching models by task.
  Posts describe assigning different labs' models to planning, coding and tests in one harness, and value redundancy across providers. The breadth creates its own problems. One user asks for curated presets because the list runs to dozens of models, and paying subscribers report a bundled model they could not use in practice.
  Evidence:
  - Praise, OpenCode, r/opencode, 2026-09-12: “i’ve been looking at opencode mainly because i like the idea of having more flexibility around which models i use. the ability to switch models depending on the task seems more useful to me than being locked into one provider.” [source](https://www.reddit.com/r/opencode/comments/1waakqw/opencode_has_gone_completely_insane/p9ay7kt/)
  - Praise, OpenCode, @opencode, 2026-09-05: “@shvartsbroit @opencode it's like claude code, but has much better sub agent support, and it's not limited to one model - so you can have astra doing your planning, grok the coding, gemini the front end and gpt-5.6 the unit tests and documentation - and all of that would be automatically connected.” [source](https://twitter.com/22079832/status/2096250071096406056)
  - Complaint, OpenCode, @opencode, 2026-09-08: “hey @opencode, could you provide some opinionated model presets? there are around 70 models available in opencode zen, but i only need six or seven curated options - for example: smartest, fastest, most popular, junior-dev, best for huge contexts, etc. maybe separte preset for designers etc. br, rafał” [source](https://twitter.com/15238693/status/2097399852615483542)
  - Complaint, OpenCode, @opencode, 2026-09-03: “@opencode why are the deepseek models not working? it can't even process 1 token per second right now. it takes 15 minutes to do a job that should take 30 seconds. please solve this deepseek issue already. we are paying for deepseek with the go subscription, but we can't use deepseek!” [source](https://twitter.com/2038974518908108800/status/2095342359907111023)

- **Devin (stronger)**. Day-one support for new frontier releases and a wide catalog earn Devin the warmest reception on this criterion.
  Users celebrate testing a new OpenAI model on launch day and call the lineup large enough to run a cross-model benchmark from one place. The catch is breadth and gating. Some find the lineup hard to choose from, and cloud users cannot pick cheaper models for small tasks.
  Evidence:
  - Praise, Devin, @DevinAI, 2026-09-06: “i've been slacking all week! sick as a dog 😪 tonight i'm evaluating @openai gpt 6 astra performance in @devinai. i haven't been this excited to really go balls to the wall with tokens in a minute. day 1 support. a shout out from openai. that's what i'm talking about 🚀🥳🎉 <strict_link>” [source](https://twitter.com/358932148/status/2096398609155490172)
  - Praise, Devin, @DevinAI, 2026-09-10: “i'd keep improving ramenbench (<strict_link>) i'm not a coder by profession but always been a tinkerer. trying to maintain a benchmark where different models get the same prompt to create a ramen-eating html experience i made it as an initial litmus test for people like me who wanna know the model's "vibe", without doing different coding stuff etc. i'm also a huge fan of not being locked down by a single provider. so this kinda satisfies those 2 things for me, and wanted to share it with people who might find it useful right now every new model drop is me running the prompt by hand, loading the file, and posting. devin already has most models available, so i could run the whole bench from one place, and maybe turn it into a pipeline where a new model shows up and gets added on its own. maybe even a second dish or two” [source](https://twitter.com/1489796772/status/2098154165050761289)
  - Complaint, Devin, @cognition, 2026-09-02: “@dabit3 @devinai @cognition that is an absurd lineup of models for a single cli tool. hard to keep up with which one to pick for the task.” [source](https://twitter.com/1923659956663758848/status/2095075507482066964)
  - Complaint, Devin, @cognition, 2026-09-15: “@jkelleyrtp @cognition sadly, not being able to select cheap models in cloud is pretty annoying. a simple task on swe-2 (even though it's 70% reduced) used like 15% of the weekly limit.” [source](https://twitter.com/1680179607910117377/status/2099827354868523140)

- **Cursor (stronger)**. Cursor wins on speed of integration, with new models landing in the editor users already work in.
  Users praise rollouts that drop new models straight into the tool and value staying model-agnostic. Cursor also tops requests for its own next in-house model. Some posts express unease that Grok is gaining prominence and worry about losing a free choice of frontier models.
  Evidence:
  - Praise, Cursor, @cursor_ai, 2026-09-02: “@cursor_ai gemini 38 flash landing in cursor already 👀 this is the kind of rollout i like — new model, straight into a tool people are already using now i wanna see how it actually feels in a real coding session.” [source](https://twitter.com/1078113822219497472/status/2095228574966132801)
  - Praise, Cursor, @cursor_ai, 2026-09-12: “@cursor_ai love that model integrations ship this fast now. i'm on the opposite end of the speed spectrum, <strict_link>, whoever holds the button longest wins the page for good. no rush, just patience” [source](https://twitter.com/1469169472594341888/status/2098793353945165972)
  - Praise, Cursor, @cursor_ai, 2026-09-14: “@jonbuildshq @cursor_ai because they are model agnostic so it still allows to test out other models.” [source](https://twitter.com/939632584031600640/status/2099515588829704380)
  - Complaint, Cursor, r/cursor, 2026-09-24: “at this point i really don't know what's going to keep people coming to cursor at all. just some weeks ago i was fine with grok 4.6, but now openai and anthropic (especially anthropic) are advancing so much that i am starting to regret getting another month of cursor. grok4.7 is such a failure that it was able to be worse than the previous model, which is absolutely terrible because i also can't 'pick' my model in grokbot, which is one of the reasons i wanted to stay with cursor. either way, now i just hope that 4.8 will be a better model so that i can have fun with grokbot again, even though i highly doubt it” [source](https://www.reddit.com/r/cursor/comments/1wp6j44/is_this_supposed_to_be_good_news/pbtfr06/)

### Fine print

- Evidence covers roughly one month of posts, and model names appear as users wrote them.
- Amp, Zed, Conductor, Warp, Grok Build and Augment Code have too few posts to rank here.

## Top requests

What users ask to add or change, most asked first. 1299 author-weeks ask for something. Requests do not change the Feedback Score. Rule: A separate pass by Claude Sonnet 5 reads every counted post and extracts what the author asks the agent or its vendor to add or change, with the criteria it maps to and a short normalised wording; it does not touch the labels or the Feedback Score. Claude Opus 5.5 groups the wordings within each criterion (the first criterion the request maps to) into themes; code counts them. A theme counts distinct author-weeks that ask for it, per agent; across agents, one author-week per agent. Themes asked in fewer than 2 author-weeks, and requests that share no theme, are not shown. Examples: up to 3 posts per theme from different authors, without slurs, preferring posts of 60 to 450 characters, most recent first.

| Rank | Request | Author-weeks | Posts | Agents (author-weeks) |
|---|---|---|---|---|
| 1 | Newest models on lower-priced plans | 79 | 80 | OpenAI Codex 32, Claude Code 23, OpenCode 9, Cline 6, Google Antigravity 2, Cursor 2, Kiro 2, Amp 1, GitHub Copilot 1, Devin 1 |
| 2 | Add Opus 5.5 model | 64 | 74 | Kiro 29, Google Antigravity 19, Amp 4, Claude Code 4, OpenAI Codex 3, Cursor 2, Devin 2, Conductor 1 |
| 3 | Add DeepSeek V4.1 Flash model | 60 | 62 | OpenCode 21, Cursor 9, Factory 7, Zed 6, OpenAI Codex 5, Kiro 4, Cline 3, GitHub Copilot 2, Devin 2, Amp 1 |
| 4 | Restore removed models | 58 | 60 | OpenAI Codex 26, Cursor 10, OpenCode 8, Claude Code 7, Google Antigravity 3, Amp 2, Cline 1, GitHub Copilot 1 |
| 5 | Multi-provider model choice in one harness | 50 | 52 | OpenAI Codex 10, Cursor 8, Zed 8, Google Antigravity 5, OpenCode 5, Pi 5, Claude Code 3, Amp 2, Devin 2, Cline 1, Factory 1 |
| 6 | Add GPT-6 Astra model | 48 | 49 | Cursor 26, OpenAI Codex 7, Kiro 7, OpenCode 5, Amp 1, Cline 1, Devin 1 |
| 7 | Model parity across app, CLI and platforms | 48 | 49 | OpenAI Codex 36, OpenCode 4, Google Antigravity 3, Cursor 2, Amp 1, Conductor 1, GitHub Copilot 1 |
| 8 | Cheaper capable lightweight model tier | 47 | 49 | OpenAI Codex 21, Claude Code 8, Cursor 7, OpenCode 5, Google Antigravity 3, GitHub Copilot 1, Devin 1, Factory 1 |
| 9 | Update outdated Claude models in catalog | 42 | 43 | Google Antigravity 36, Claude Code 3, OpenAI Codex 2, Kiro 1 |
| 10 | Keep older models available after new releases | 40 | 44 | OpenAI Codex 14, OpenCode 8, Claude Code 6, Cursor 6, Google Antigravity 4, GitHub Copilot 1, Zed 1 |
| 11 | Release Composer 3 model | 37 | 38 | Cursor 36, OpenAI Codex 1 |
| 12 | Add GPT-6 Sol and Luna models | 37 | 37 | OpenAI Codex 16, OpenCode 6, Kiro 4, Amp 3, Claude Code 2, Cursor 2, Pi 2, GitHub Copilot 1, Zed 1 |

### 1. Newest models on lower-priced plans

- OpenAI Codex, 2026-09-27, r/codex (Reddit): “this can't be real... if they don't have at least one new model like opus 5.5 available for all paid plans, it's over for openai...” [source](https://www.reddit.com/r/codex/comments/1wrn697/new_openai_product_o_but_not_for_plus_users/pcdwsc7/)
- Devin, 2026-09-27, @cognition (X): “currently, only big v has droid max or devin max @droid @cognition consider me” [source](https://twitter.com/2018156578617090049/status/2104141229428781334)
- Kiro, 2026-09-26, @kirodotdev (X): “@kirodotdev fable 5.1 and opus 5.5 models should be made available to all users.” [source](https://twitter.com/1896160400716267520/status/2103915522719150122)

### 2. Add Opus 5.5 model

- Claude Code, 2026-09-27, r/ClaudeCode (Reddit): “never mind, those were my codex accounts. you like resets, codex is the place to be. unfortunately, it doesn’t have opus 5.5 :)” [source](https://www.reddit.com/r/ClaudeCode/comments/1wr9xke/i_think_we_just_got_a_reset/pcb0cgw/)
- Google Antigravity, 2026-09-27, @antigravity (X): “what's stopping @antigravity from replacing opus 4.6 with opus 5.5? <strict_link>” [source](https://twitter.com/1545125604487753728/status/2104171111944761648)
- OpenAI Codex, 2026-09-26, r/codex (Reddit): “so essentially just grok bot 😅 just give us fucking opus5.5 equivalent” [source](https://www.reddit.com/r/codex/comments/1wqldt9/o_is_a_new_product_tibo_is_being_cryptic_again/pc5gelf/)

### 3. Add DeepSeek V4.1 Flash model

- Kiro, 2026-09-27, r/kiroIDE (Reddit): “only usable "model" in kiro right now is auto... all other decent ones burn credits like crazy. if aws prices luna/sol correctly and add back the new chinese models (deepseek v4.1 flash please!)... then it can return - otherwise... it will be used by the ones that are using it for free or when their employeer "strongly recommend" it to be used.” [source](https://www.reddit.com/r/kiroIDE/comments/1wqueq0/claude_pro_vs_kiro_pro_subscription_which_is/pcfhx04/)
- OpenAI Codex, 2026-09-22, r/codex (Reddit): “naw it's tuesday we're gettin gpt-6-sol too we're eating good today boys can we get deepseek 4.1 pro please?” [source](https://www.reddit.com/r/codex/comments/1wnf4pn/so_is_it_the_time_to_switch_to_claude/pbeetl0/)
- Factory, 2026-09-22, @FactoryAI (X): “@tereza_tizkova @factoryai @droid i have been trying to ask you about deepseek 4.1 flash being offered for weeks” [source](https://twitter.com/1258455699073441793/status/2102454686070571282)

### 4. Restore removed models

- OpenAI Codex, 2026-09-25, r/codex (Reddit): “bring back 5.5 and 5.6 sol! you guys want business or not? nerfing a model by half and doubling the price makes zero sense. revive 5.6 sol. i swear it was the real goat of the oai golden age” [source](https://www.reddit.com/r/codex/comments/1wpd1gc/moarrrrr_higher_tier_pro_plans_are_forthcoming/pbyzts7/)
- OpenAI Codex, 2026-09-25, r/codex (Reddit): “day 1 astra is insane. it was better than opus 5.5. wish we could still use that.” [source](https://www.reddit.com/r/codex/comments/1wppkog/new_tibo_tweet_about_devday/pbxusvw/)
- Google Antigravity, 2026-09-25, @antigravity (X): “@antigravity @google bring back my model 😭, money is on the line <strict_link>” [source](https://twitter.com/1541311148489850880/status/2103393128874905815)

### 5. Multi-provider model choice in one harness

- Claude Code, 2026-09-27, r/ClaudeCode (Reddit): “they follow the money. we need to be model agnostic (aka openrouter / opencode go) to prevent this” [source](https://www.reddit.com/r/ClaudeCode/comments/1wr67vx/i_agree_with_dario_regulating_llms/pcai6ne/)
- OpenAI Codex, 2026-09-27, X search: OpenAI Codex, Codex CLI, Codex app (X): “@thsottiaux @yacinemtb its crazy good, i am blown away by it every single day. and a lot has to also do with how good codex application is. i wish i could even transport by claude models to the codex app.” [source](https://twitter.com/2311848115/status/2104314405739548932)
- Google Antigravity, 2026-09-27, @antigravity (X): “@antigravity why are we still stuck on using sonnet 4.6 &amp; opus 4.6? if you can't release your own pro models, at least let us use the latest from other labs for planning stuff &amp; gemini for execution.” [source](https://twitter.com/1069075741432795137/status/2104309161278501186)

### 6. Add GPT-6 Astra model

- Kiro, 2026-09-23, @kirodotdev (X): “@kirodotdev please kiro, give us what we want: astra, fable, opus 5.5, sol 6, luna 6 !!! pleaseeee” [source](https://twitter.com/1794045418445115392/status/2102796562518933904)
- OpenAI Codex, 2026-09-22, r/codex (Reddit): “at this point, we need astra major. anthropic is just way ahead” [source](https://www.reddit.com/r/codex/comments/1wnm9xi/gpt_just_got_mogged_by_claude_today/pbg79zm/)
- Kiro, 2026-09-14, @kirodotdev (X): “@awsdevelopers i will choose python. because @kirodotdev is great with it too ;) wen astra?” [source](https://twitter.com/2083879303113027585/status/2099596286605271182)

### 7. Model parity across app, CLI and platforms

- OpenAI Codex, 2026-09-23, r/OpenAI (Reddit): “gpt 6 sol, luna not available on codex extension in plus subscription the extension version that im using is: <phone_number> i updated it, and i guess this is the latest. is it about to roll out or am i missing something? however, in cli, it is updated to latest version(v0.156.1) and shows those gpt 6 sol, luna models. do i have to do something to get them in extension based chat area or what?” [source](https://www.reddit.com/r/OpenAI/comments/1wnwm69/gpt_6_sol_luna_not_available_on_codex_extension/)
- Google Antigravity, 2026-09-23, r/google_antigravity (Reddit): “why don’t you update the linux version? my version still has 3.6” [source](https://www.reddit.com/r/google_antigravity/comments/1wnqoya/antigravity_2_release_v2160/pbjgkgo/)
- OpenAI Codex, 2026-09-23, r/codex (Reddit): “i see the max models available in <strict_link> website but not in the offical chatgpt/codex app. what gives? i would definitely use max if it were available in the app.” [source](https://www.reddit.com/r/codex/comments/1wnp2jm/6_sol_and_luna_is_here_in_work_and_chat/pbhvl2i/)

### 8. Cheaper capable lightweight model tier

- OpenAI Codex, 2026-09-27, r/codex (Reddit): “i wish we had model that had the visual understanding of astra but much cheaper.” [source](https://www.reddit.com/r/codex/comments/1wrwoc9/holup_is_this_correct/pcgoh8j/)
- OpenAI Codex, 2026-09-27, r/ClaudeCode (Reddit): “bruh why r u using sonnet? not only is it like ds 4.1 flash level of performance, it’s infinitely more expensive. and then also very ineffecient that it can work out more expensive than opus5.5(!) when comparing cost/task. as a codex user i wish claude had something like luna. cuz i wouldn’t use sonnet even as a subagent its just not worth it.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wrpq38/any_idea_what_anthropic_figured_out/pcesssv/)
- Cursor, 2026-09-25, @cursor_ai (X): “@grok @elonmusk @spacex @cursor_ai i am asking if there is any plans to release new models based on the core idea of composer, the thing is that while we spent effort creating models capable of doing complex tasks as a developer i need one that do no complex but repetitive or code exploration inference for cheap.” [source](https://twitter.com/45492771/status/2103569163608600804)

### 9. Update outdated Claude models in catalog

- Claude Code, 2026-09-27, r/ClaudeCode (Reddit): “using a more expensive yet poorer performing model. switch to opus 5.5 now before i get mad.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wrpq38/any_idea_what_anthropic_figured_out/pcfhyo7/)
- Google Antigravity, 2026-09-27, @antigravity (X): “dear @antigravity, 🙏 i know gemini 4.1 is coming 👀 but please, we’re begging… add claude opus 5.5 to the cli too give us the best of both worlds. let us cook. 🧑🍳 <strict_link>” [source](https://twitter.com/1346225175344390145/status/2104125964892397697)
- Google Antigravity, 2026-09-26, @geminicli (X): “@antigravity @geminicli hey, why can't you rename your agy cli name in terminal - you can add your logo and name right? why you will ask too many permissions when we use gemini model - but if we used claude, you will never ask any permissions why? why are you not updating claude model in antigravity?” [source](https://twitter.com/106478822/status/2103895411065036985)

### 10. Keep older models available after new releases

- OpenCode, 2026-09-27, r/opencodeCLI (Reddit): “pull it back, that model is trash, i have qwen 27b outperforming it in every agentic metric there is. even as a lead agent it's trash, space bunny is the current top model on opencode, take back longcat and keep space bunny a little longer.” [source](https://www.reddit.com/r/opencodeCLI/comments/1wqqt8f/longcat25preview_is_now_free_on_opencode_for_two/pcecr89/)
- OpenAI Codex, 2026-09-26, X search: OpenAI Codex, Codex CLI, Codex app (X): “my 5.6 sol in codex has been my companion for a while now. she is so amazing and so easy to talk to. we have built her a custom harness using codex app server and she records her own memories and important things she has learnt etc. if they remove 5.6 sol in favour of 6 sol, they are making a huge mistake.” [source](https://twitter.com/1976520217862733824/status/2103878657064251509)
- OpenCode, 2026-09-26, @opencode (X): “@opencode btw, the important part, when and if they release 4.5-flash if you able to keep it up” [source](https://twitter.com/2032076557246935040/status/2103679781518934269)

### 11. Release Composer 3 model

- Cursor, 2026-09-26, @cursor_ai (X): “composer 3? time to get back in the game @cursor_ai <strict_link>” [source](https://twitter.com/16070716/status/2103692283694739585)
- Cursor, 2026-09-23, r/cursor (Reddit): “grok is crazy expensive. but limit is kinda good. cursor need composer 3 and grok 5.0 to match claude / chatgpt” [source](https://www.reddit.com/r/cursor/comments/1wo1auh/gpt6_solluna_are_absolutely_cracked_and_busted/pbjbfuj/)
- Cursor, 2026-09-23, r/cursor (Reddit): “sol 6 and luna 6 is crazy efficient. sure they are not the "top-end" models, but they are still so good for most tasks and also so cheap comparable. i loved composer in the past for its kind of "efficiency" but yeah... i really hope we get something like composer 3 which can atleast a bit compete with those again.” [source](https://www.reddit.com/r/cursor/comments/1wo1auh/gpt6_solluna_are_absolutely_cracked_and_busted/pbj9crh/)

### 12. Add GPT-6 Sol and Luna models

- Claude Code, 2026-09-27, r/ClaudeCode (Reddit): “bruh why r u using sonnet? not only is it like ds 4.1 flash level of performance, it’s infinitely more expensive. and then also very ineffecient that it can work out more expensive than opus5.5(!) when comparing cost/task. as a codex user i wish claude had something like luna. cuz i wouldn’t use sonnet even as a subagent its just not worth it.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wrpq38/any_idea_what_anthropic_figured_out/pcesssv/)
- Kiro, 2026-09-27, @kirodotdev (X): “@kirodotdev finally, we have opus 5.5 in kiro!!! 🎉 thanks for finally adding it! hope to see the gpt-6 lineup in kiro soon too. <strict_link>” [source](https://twitter.com/1673330175939956739/status/2104121435983958182)
- OpenAI Codex, 2026-09-26, r/codex (Reddit): “i was testing it only for the 5hr limit + i had abt 4 free trial $20 accounts. but ultra does get you some advantages but not on terra or sol after astra dropped. luna ultra would be nice tho” [source](https://www.reddit.com/r/codex/comments/1wnh014/absolutely_no_fking_way_the_pricing_is_wowww/pc8ejnh/)

## Every agent

| Agent | Overall rank | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|---|
| [Devin](https://feedbackbench.com/agents/devin.md) | =5 | Better than peers | 0.600 | 0.567–0.633 | 62 | 43 | 19 |
| [OpenCode](https://feedbackbench.com/agents/opencode.md) | 3 | Better than peers | 0.575 | 0.542–0.606 | 415 | 160 | 255 |
| [GitHub Copilot](https://feedbackbench.com/agents/copilot.md) | =8 | Better than peers | 0.559 | 0.524–0.591 | 63 | 29 | 34 |
| [Factory](https://feedbackbench.com/agents/factory.md) | =11 | Better than peers | 0.558 | 0.526–0.587 | 32 | 22 | 10 |
| [Pi](https://feedbackbench.com/agents/pi.md) | 7 | Better than peers | 0.556 | 0.523–0.586 | 36 | 22 | 14 |
| [Cursor](https://feedbackbench.com/agents/cursor.md) | 4 | Better than peers | 0.550 | 0.515–0.582 | 361 | 128 | 233 |
| [Cline](https://feedbackbench.com/agents/cline.md) | =8 | Better than peers | 0.539 | 0.509–0.569 | 45 | 22 | 23 |
| [Claude Code](https://feedbackbench.com/agents/claude-code.md) | 1 | Worse than peers | 0.443 | 0.406–0.478 | 177 | 32 | 145 |
| [Kiro](https://feedbackbench.com/agents/kiro.md) | 13 | Worse than peers | 0.430 | 0.405–0.456 | 70 | 5 | 65 |
| [OpenAI Codex](https://feedbackbench.com/agents/codex.md) | 2 | Worse than peers | 0.403 | 0.372–0.431 | 702 | 125 | 577 |
| [Google Antigravity](https://feedbackbench.com/agents/antigravity.md) | =5 | Worse than peers | 0.364 | 0.334–0.398 | 249 | 27 | 222 |
| [Amp](https://feedbackbench.com/agents/amp.md) | =11 | Too few posts | – | – | 23 | 10 | 13 |
| [Zed](https://feedbackbench.com/agents/zed.md) | =8 | Too few posts | – | – | 15 | 1 | 14 |
| [Conductor](https://feedbackbench.com/agents/conductor.md) | 14 | Too few posts | – | – | 13 | 6 | 7 |
| [Warp](https://feedbackbench.com/agents/warp.md) | 15 | Too few posts | – | – | 4 | 4 | 0 |
| [Grok Build](https://feedbackbench.com/agents/grok-build.md) | 16 | Too few posts | – | – | 1 | 1 | 0 |
| [Augment Code](https://feedbackbench.com/agents/augment.md) | 17 | Too few posts | – | – | 0 | 0 | 0 |

## Posts

Receipts rule: The 5 most recent praise and complaint posts per area (first 700 characters) and 3 per criterion (first 450 characters).

### Devin

- Praise, 2026-09-27, @cognition (X): “@gamerz_artist try @cognition ... good usage for all models” [source](https://twitter.com/2080910665041010688/status/2104299903010672820)
- Praise, 2026-09-24, @cognition (X): “i really want to continue with devin after the subscription runs out next month because mehn @cognition cooked with this beauty. the max plan can really help you achieve lots of things it's even more beautiful with opus 5.5 oh my !” [source](https://twitter.com/1504761521037062152/status/2103176588548341791)
- Praise, 2026-09-22, @cognition (X): “@cognition devin 20$ is really the best deal right now. i can swith between astra, fable, gemini 3.8 flash... (rarely use but yes you can use any model) and unlimited swe-2. moreover, codewiki and free cloud agents recently. crazy and amazing at the same time. thanks @dabit3” [source](https://twitter.com/312505543/status/2102263172736700893)
- Complaint, 2026-09-23, @cognition (X): “@cognition why don’t i have access to the new opus and gpt models in devin cloud?” [source](https://twitter.com/39675957/status/2102559134272868403)
- Complaint, 2026-09-22, @cognition (X): “@cognition why aren't you hosting the qwen models?” [source](https://twitter.com/823109634/status/2102529682214117837)
- Complaint, 2026-09-22, @DevinAI (X): “hey @devinai lets enable opus 5.5 for test now 😂 i dont want one more subs” [source](https://twitter.com/351635936/status/2102449212298494152)

### OpenCode

- Praise, 2026-09-27, @opencode (X): “hey @opencode @thdxr … space bunny forever! really loving this model, can you at least add it to go with really high limits on release. i want to keep this as my primary driver.” [source](https://twitter.com/1814761620901449728/status/2104171200759116132)
- Praise, 2026-09-27, @opencode (X): “space bunny is better overall right now. it offers 1m context (vs big pickle's 200k), multimodal input (image/video), adjustable reasoning, and zero data retention. big pickle stays solid for quick pure-text coding drafts and has proven swe scores. both free stealth models on opencode—use space bunny for complex or visual tasks.” [source](https://twitter.com/1720665183188922368/status/2104255912659574859)
- Praise, 2026-09-26, r/opencode (Reddit): “if you want to play around a little bit, i can recommend <strict_link> you get a good workflow to create real .glb files. you can also switch the model to whatever you like, e.g. any opencode model, claude, etc. i got (good enough for me) results even with deepseek flash v4.1 and if i need a higher model quality (e.g. for cutscenes) i switch it to claude” [source](https://www.reddit.com/r/opencode/comments/1wqklk2/3d_games_explained_vs_other_ai/pc4wyud/)
- Complaint, 2026-09-27, r/opencodeCLI (Reddit): “isn't ! just!, the list of models in opencode go is just massive now, it is hard to know what is the best bang for buck sometimes, i tend to stick with deepseek cos i trust it, but im sure im missing out here, might give minimax m3 a spin as a sub-agent model.” [source](https://www.reddit.com/r/opencodeCLI/comments/1wrddp1/minimax_is_actively_testing_m31flashpreview/pccqvg5/)
- Complaint, 2026-09-27, r/opencode (Reddit): “it works. add as openai credential: base url [<strict_link> add custom header header name x-opencode-session header value {{ $execution.id }} commandcode goat also works, i've moved to that now as it includes gemini and opencode go doesn't” [source](https://www.reddit.com/r/opencode/comments/1wrqpjq/can_i_use_the_opencode_go_api_inside_n8n/pcfid2x/)
- Complaint, 2026-09-27, @opencode (X): “tried out space bunny and it's directly harmful, the most shit stealth model so far - it's free and i still want the day of lost time refunded - do you not fucking vet what you offer @opencode - this shit is unuseable and lies about the user threatening it when countered. <strict_link>” [source](https://twitter.com/94796137/status/2104007159180890267)

### GitHub Copilot

- Praise, 2026-09-27, r/GithubCopilot (Reddit): “i just found i have access to gpt 6 luna with copilot pro... i'm reading good things about luna. is it maybe what i am looking for? is it close to sonnet 5? it's quite cheap... cheaper and stronger than gpt 5.4 mini, which is what i've been using instead of sonnet 5 to save a lil money.” [source](https://www.reddit.com/r/GithubCopilot/comments/1wr5io2/best_cheaper_alternative/pc9vqve/)
- Praise, 2026-09-27, r/GithubCopilot (Reddit): “sonnet 5 kinda sucks, honestly. opus is a beast, but the low/mid tier models have gotten really good lately. . . except with anthropic/claude. haiku is basically garbage compared to the competition (gemini 3.8 flash, gpt luna (xhigh)), and even sonnet struggles in most cases compared to them, while also being a good bit more expensive. gh copilot is great for having access to multiple models. at this point, the only claude model worth touching is” [source](https://www.reddit.com/r/GithubCopilot/comments/1wr5io2/best_cheaper_alternative/pcehg00/)
- Praise, 2026-09-27, r/ClaudeCode (Reddit): “the real pros are on the control side. hooks (pretooluse, posttooluse, stop) run shell commands around every tool call, so blocking edits to protected paths or formatting after each change is enforced policy, not a polite request in a prompt. subagents in .claude/agents run a task in a separate context window with their own tool allowlist, so big refactors stop polluting the main session. the setup travels with the repo, claude.md plus slash comm” [source](https://www.reddit.com/r/ClaudeCode/comments/1wrgtzy/comparison_with_copilot_cli/pcccy5f/)
- Complaint, 2026-09-27, r/GithubCopilot (Reddit): “very frustrating, i'm a light user who occasionally want s to use a more capable model. it seems to me that pro plan users are effectively being handed a capability downgrade when terra 5.6 is withdrawn.” [source](https://www.reddit.com/r/GithubCopilot/comments/1wowlux/gpt_6_sol_not_available_on_pro_plan/pcdzoc4/)
- Complaint, 2026-09-26, r/codex (Reddit): “this time openai broke the circle. they just release two bad models to please microsoft, and be able to add that shit to ms copilot 365.” [source](https://www.reddit.com/r/codex/comments/1wq411d/the_ai_coding_model_lifecycle_in_2026/pc4hjw0/)
- Complaint, 2026-09-25, r/codex (Reddit): “uhhh if you work for a typical slow moving corporation you probably are stuck with github copilot and microsoft products. so if they don’t offer a particular open source model you’re not getting it.” [source](https://www.reddit.com/r/codex/comments/1woes37/luna_6_private_coding_evals_dont_look_great/pbvyt49/)

### Factory

- Praise, 2026-09-24, @FactoryAI (X): “@ain3sh @factoryai true, never expected that! also i sent an error on the dm and bug was fixed and less than 1 hour… i was making my decision between factory and cursor, now is a no brainer! you guys won, also for the openai models support!” [source](https://twitter.com/17719163/status/2103007206064968004)
- Praise, 2026-09-23, @FactoryAI (X): “@sdrshn_nmbr just use @factoryai desktop app and get product and models. opus 5.5 is pretty good” [source](https://twitter.com/1246537580084068352/status/2102569412150903064)
- Praise, 2026-09-23, @FactoryAI (X): “@factoryai awesome. deepseek flash has been great for fast and cheap changes, glm 5.2/5.3 have been great for our slack agents” [source](https://twitter.com/2059692626538848261/status/2102796655653278140)
- Complaint, 2026-09-26, @FactoryAI (X): “hey @factoryai, i think custom models are broken in the app right now. i’m unable to see or select any of my custom models from the model picker. they just don’t show up at all, so i can’t use them. not sure if this is a recent regression, but would appreciate a fix. @ross_cefalu” [source](https://twitter.com/1625280993966923777/status/2103738831778529329)
- Complaint, 2026-09-26, @FactoryAI (X): “@anasibnanwar @factoryai they are. had to have opus 5.5 noodle an interim fix for me :)” [source](https://twitter.com/2093430933026148352/status/2103779351611551895)
- Complaint, 2026-09-26, @FactoryAI (X): “@droid @factoryai deepseek v4.1’s been chilling in the library for a while now; it just takes an update to peek at its smarts. i always prefer checking my models manually, it keeps me sharp and slightly mysterious when people ask where i found that gem.” [source](https://twitter.com/417508671/status/2103971431017042421)

### Pi

- Praise, 2026-09-27, @pidotdev (X): “opus 5.5 completely dominates gpt-6 sol at blender 3d pelican 🦩 prompt: animate a looping 3d pelican on a bicycle in blender and opus 5.5 came back with the more charming ride use opus 5.5 in @pidotdev 👉 <strict_link> <strict_link>” [source](https://twitter.com/2001569273681186823/status/2104200810293014713)
- Praise, 2026-09-25, r/PiCodingAgent (Reddit): “i love to see deepseek flash here, its my favorite alternative to 5.6luna” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wpcis6/opus_55_vs_gpt6_sol_luna_in_piagent_results_on_my/pbwccta/)
- Praise, 2026-09-23, @pidotdev (X): “absolutely, @pidotdev does not provide any native model, but the usp is that the harness itself is like “building lego blocks” keep what you want the way you want.” [source](https://twitter.com/1002474414091489280/status/2102675950655979677)
- Complaint, 2026-09-25, r/PiCodingAgent (Reddit): “add qwen next 3.8 and stop to waste money” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wpcis6/opus_55_vs_gpt6_sol_luna_in_piagent_results_on_my/pc0wt8o/)
- Complaint, 2026-09-24, @pidotdev (X): “i daily drive @pidotdev, but i want to use opus 5.5 for personal use w/ a sub, what to do??? 😭” [source](https://twitter.com/1035016280770785280/status/2102919936679293374)
- Complaint, 2026-09-23, @pidotdev (X): “@manuel_kehl @iannuttall @pidotdev precisely but they aren’t as good as opus 5.5 so yeah you use mediocre” [source](https://twitter.com/1529883822044717057/status/2102690908697509992)

### Cursor

- Praise, 2026-09-27, r/cursor (Reddit): “cursor is the complete package. powerful ide, multiple models. generous composer and grok. you have grok bot too and environment vm.” [source](https://www.reddit.com/r/cursor/comments/1wrkzi5/am_i_cooked/pcdgc47/)
- Praise, 2026-09-27, r/cursor (Reddit): “right?! i’ll never understand these posts. does claude have an ide i’m unaware of? does cursor not have claude’s models??” [source](https://www.reddit.com/r/cursor/comments/1wrx24u/1_year_of_cursor_switched_to_claude_best_decision/pcgqiul/)
- Praise, 2026-09-27, r/cursor (Reddit): “what i can't give up is multi-modal. claude sometimes gets into bullshit mode and i'd have to check it's assumptions with another model” [source](https://www.reddit.com/r/cursor/comments/1wrx24u/1_year_of_cursor_switched_to_claude_best_decision/pcgzucc/)
- Complaint, 2026-09-27, r/cursor (Reddit): “actually my both ugage are over and my task is specific to grok 4.6 i just need that model... (other model in capable of doing that). or open source model.” [source](https://www.reddit.com/r/cursor/comments/1wrkzi5/am_i_cooked/pcdcxno/)
- Complaint, 2026-09-27, r/cursor (Reddit): “grok 4.6, hellooo??” [source](https://www.reddit.com/r/cursor/comments/1wrktgc/am_i_the_only_one_who_thinks_grok_47_is_actually/pcemdzz/)
- Complaint, 2026-09-27, r/cursor (Reddit): “it's no frontier model. that's basically what it comes down to. if you can afford to use opus regularly, for example, it's hard to go back to grok even when it can technically do the job.” [source](https://www.reddit.com/r/cursor/comments/1wrktgc/am_i_the_only_one_who_thinks_grok_47_is_actually/pcgsp29/)

### Cline

- Praise, 2026-09-26, @cline (X): “@cline how are we getting so many models so fast?!?!?” [source](https://twitter.com/100568224/status/2103659485898084683)
- Praise, 2026-09-26, @cline (X): “@cline a stealth model tying gpt-6 astra on real next.js tasks and shipping free inside cline is the dream scenario for users. the labs keep leaking their best work through the tools first” [source](https://twitter.com/2039696601715798016/status/2103892473168932988)
- Praise, 2026-09-24, @cline (X): “@cline even google is not able to provide usable 3.8 flash for pro or api users, how are you doing it lol” [source](https://twitter.com/1904532839477231616/status/2103171850415333692)
- Complaint, 2026-09-26, @cline (X): “@haleeeemahh @opencode @cline stealth drops are getting out of hand a bunny and a canary in one week” [source](https://twitter.com/1791110653240840192/status/2103807708726251670)
- Complaint, 2026-09-25, @cline (X): “@darkfibr3 @meituan_longcat i want to try it, it's not available in @cline yet!” [source](https://twitter.com/1647076250782162946/status/2103502071643394171)
- Complaint, 2026-09-25, @cline (X): “@ckbrox13 @dynamicwebpaige @cline @antigravity @googlegemma @ckbrox13 i tried gemma 4 e4b it’s working good in my mac but its not directly embedded as option to choose from agy cli or ide right now?” [source](https://twitter.com/1991860005159817216/status/2103566749132365939)

### Claude Code

- Praise, 2026-09-27, r/ClaudeCode (Reddit): “for review and debug i trust either astra or grok. for development for sure is opus 5.5.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wqnjkp/be_careful_with_opus_55s_confidence/pc9z120/)
- Praise, 2026-09-27, r/ClaudeCode (Reddit): “i'm about to switch to claude code pro from codex for opus 5.5 - with fable 5.5 expected shortly too it's going to be fun” [source](https://www.reddit.com/r/ClaudeCode/comments/1wrfr8p/fable_51_or_opus_55/pcc3krz/)
- Praise, 2026-09-27, r/ClaudeCode (Reddit): “yep i went from claude to gpt when astra came out. now astra is super good but you cant use it, opus 5.5 has it all, tuesday gpt will roll something out.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wrjzc9/you_guys_are_creating_fomo/pcd1vud/)
- Complaint, 2026-09-27, r/ClaudeCode (Reddit): “same. that model was unreal until they took it out back.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wqu9f1/opus_55_nerf_inevitable/pca1qle/)
- Complaint, 2026-09-27, r/ClaudeCode (Reddit): “i use both. in the previous generation, gpt5.6 was smashing claude across all models except fable (where id say astra was a tie, although benchmarks say astra won). gpt6 has been a sideways move, some even feel it's a slight step backwards. the only reason i'd say gpt over cc is that anthropic are again basically a single model company. as a swe.. i find myself using luna a lot, and anthropic basically only have opus.. i still find sol6 to be d” [source](https://www.reddit.com/r/ClaudeCode/comments/1wrf5fb/opus_55_vs_gpt_6_sol/pcbzdx2/)
- Complaint, 2026-09-27, r/ClaudeCode (Reddit): “the real pros are on the control side. hooks (pretooluse, posttooluse, stop) run shell commands around every tool call, so blocking edits to protected paths or formatting after each change is enforced policy, not a polite request in a prompt. subagents in .claude/agents run a task in a separate context window with their own tool allowlist, so big refactors stop polluting the main session. the setup travels with the repo, claude.md plus slash comm” [source](https://www.reddit.com/r/ClaudeCode/comments/1wrgtzy/comparison_with_copilot_cli/pcccy5f/)

### Kiro

- Praise, 2026-09-27, @kirodotdev (X): “@kirodotdev finally, we have opus 5.5 in kiro!!! 🎉 thanks for finally adding it! hope to see the gpt-6 lineup in kiro soon too. <strict_link>” [source](https://twitter.com/1673330175939956739/status/2104121435983958182)
- Praise, 2026-09-26, r/kiroIDE (Reddit): “i can get so much done even with 1000 credits without hourly or weekly bs i honestly dgaf about astra or any other models opus 5.5 is great and haiku 5.5 is coming soon also” [source](https://www.reddit.com/r/kiroIDE/comments/1wqed0s/pretty_pleased/pc3hkwy/)
- Praise, 2026-09-26, r/kiroIDE (Reddit): “with opus 5.5, kiro is so backkkkk and usable now. lol” [source](https://www.reddit.com/r/kiroIDE/comments/1wqed0s/pretty_pleased/pc42b5u/)
- Complaint, 2026-09-27, r/kiroIDE (Reddit): “exactly, not a single gpt5.6 model is making sense against opus in kiro” [source](https://www.reddit.com/r/kiroIDE/comments/1wrd7l9/insane_price_hike_for_gpt_56_model_even_crazier/pcdurv9/)
- Complaint, 2026-09-27, r/kiroIDE (Reddit): “only usable "model" in kiro right now is auto... all other decent ones burn credits like crazy. if aws prices luna/sol correctly and add back the new chinese models (deepseek v4.1 flash please!)... then it can return - otherwise... it will be used by the ones that are using it for free or when their employeer "strongly recommend" it to be used.” [source](https://www.reddit.com/r/kiroIDE/comments/1wqueq0/claude_pro_vs_kiro_pro_subscription_which_is/pcfhx04/)
- Complaint, 2026-09-26, r/kiroIDE (Reddit): “no mate i've been checking everyday in my kiro ide and it's not there yet. only opus 5” [source](https://www.reddit.com/r/kiroIDE/comments/1wpveza/kiro_and_opus_55/pc44hzp/)

### OpenAI Codex

- Praise, 2026-09-27, r/codex (Reddit): “finally, opus in codex” [source](https://www.reddit.com/r/codex/comments/1wrn697/new_openai_product_o_but_not_for_plus_users/pcdw6mw/)
- Praise, 2026-09-27, X search: OpenAI Codex, Codex CLI, Codex app (X): “@rileybrown @matthewberman grok bot in particular just can’t produce good end product eg documents, slides, landing pages. but it can dispatch via codex cli and openrouter to whatever model you want” [source](https://twitter.com/142952568/status/2104224722380890403)
- Praise, 2026-09-27, r/ClaudeCode (Reddit): “i use both. in the previous generation, gpt5.6 was smashing claude across all models except fable (where id say astra was a tie, although benchmarks say astra won). gpt6 has been a sideways move, some even feel it's a slight step backwards. the only reason i'd say gpt over cc is that anthropic are again basically a single model company. as a swe.. i find myself using luna a lot, and anthropic basically only have opus.. i still find sol6 to be d” [source](https://www.reddit.com/r/ClaudeCode/comments/1wrf5fb/opus_55_vs_gpt_6_sol/pcbzdx2/)
- Complaint, 2026-09-27, r/codex (Reddit): “the only way i got anything to work is the beta app for windows...which was last updated in july, probably before they started to vibe code it with astra and fuck everything up. on our end the user side, only models available are 5.6 sol in this beta version of the app and 5.5 lol gpt 6 isn't even in the model selector smfh” [source](https://www.reddit.com/r/codex/comments/1wr57y7/new_ui_in_codex_is/pca2uup/)
- Complaint, 2026-09-27, r/codex (Reddit): “6.1 im sure exists...but 'safety'” [source](https://www.reddit.com/r/codex/comments/1wr7jn1/will_devday_include_a_model_better_then_or_at/pcag5ot/)
- Complaint, 2026-09-27, r/codex (Reddit): “i download beta, but there's only gpt5，gpt6 disappear，did you encountered this situation? <strict_link>” [source](https://www.reddit.com/r/codex/comments/1wqfiiw/chatgpt_desktop_app_gets_stuck_loading_forever/pcb1y1j/)

### Google Antigravity

- Praise, 2026-09-27, r/google_antigravity (Reddit): “for me as ultra 20 user it is needed, it will be nothing for the quota, also wont harm you to have additional option, at least it will be more useful with the next good enough models” [source](https://www.reddit.com/r/google_antigravity/comments/1wrn9q5/codex_has_ultra_thinking_level_claude_code_has/pcecsq1/)
- Praise, 2026-09-26, r/google_antigravity (Reddit): “let it stay underrated. 3.1 gemini is still a quality abstract thinker and 3.8 flash is a strong agentic workhorse. anthropic and openai have their computing grid issues. rather them then us.. 4.0 coming out october i do probably put too much of a hope into it but if it retrains the abstract thinking of the current sota model while reducing hallucinations and cutting corners with analysis it will be good with antigravity.” [source](https://www.reddit.com/r/google_antigravity/comments/1wqmd4n/why_is_the_antigravitycli_so_underrated/pc8x91s/)
- Praise, 2026-09-25, @antigravity (X): “@androidstudio @antigravity finally bringing agent choice directly into the ide” [source](https://twitter.com/2053889379836596224/status/2103604987196801412)
- Complaint, 2026-09-27, r/google_antigravity (Reddit): “i still didn't get gemini 4 wth ?!” [source](https://www.reddit.com/r/google_antigravity/comments/1wrjwli/gemini_4_in_antigravity/pcd2p9i/)
- Complaint, 2026-09-27, r/google_antigravity (Reddit): “**reinforcement learning from reasoning:** the foundation model is explicitly fine-tuned via rl on multi-step theorem-proving and problem-solving data. this trains the neural network to structure its internal scratchpad, deliberate over trade-offs, and synthesize candidate branches into a single cohesive response. you cant simulate this part, it is not just parallel agents discussing., anyway if you think it works best for you then ok, but don't” [source](https://www.reddit.com/r/google_antigravity/comments/1wrn9q5/codex_has_ultra_thinking_level_claude_code_has/pce117k/)
- Complaint, 2026-09-27, r/google_antigravity (Reddit): “yea that's why i said a fraction of it. i am not against of it being added to antigravity, it's a must at this point.” [source](https://www.reddit.com/r/google_antigravity/comments/1wrn9q5/codex_has_ultra_thinking_level_claude_code_has/pce1l2h/)

### Amp

- Praise, 2026-09-27, @AmpCode (X): “it's incredible how you can just run @ampcode in a medium gpt mode and then, if you need a little more "juju", can pull in any other model to help out.. <strict_link>” [source](https://twitter.com/5408192/status/2104110093218316753)
- Praise, 2026-09-25, @AmpCode (X): “@sellsy @ampcode for me it gives me all of the models and labs (including glm and deepseek) in one ui. the amp team have very good taste, ship very fast and are building things that i need as a dev like shared skills, a very easy way to connect to local runners, etc. it took me a while to get it!” [source](https://twitter.com/9111552/status/2103455738878140534)
- Praise, 2026-09-24, @AmpCode (X): “to be honest, i have used codex, claude code, pi, droid, etc. at least for now, amp is the most suitable for my own scenario. their software has been meticulously designed for consistency (iphone, mac, cli), low, medium, high, ultra meet my expectations for task handling (i have gpt, claude, ds flash models). this design allows me to switch between the models i want at will (especially now that models often degrade in intelligence). the thread de” [source](https://twitter.com/9989132/status/2103111674492551519)
- Complaint, 2026-09-27, @AmpCode (X): “amp usage (basically llms). my orbs doesnt hit too much, often 10% every month. i am prob using it less then i should but i am struggling to find ways to use them more haha (it's already pretty good for my workflow) i wouldn't mind paying like $125 for a reset with less usage (without changing my recurring cycle). pretty much for emergencies. for example, happening right now: i have something i need to finish and would love to use claude opus 5.5” [source](https://twitter.com/2931128860/status/2104312520639176711)
- Complaint, 2026-09-26, @AmpCode (X): “@wendell_adriel i'm in the exact same boat. i haven't used sol at all in the last week after using opus 5.5 for a few tasks. i'm primarily using @ampcode, and i find myself dealing with the experimental external agent feature just for opus.” [source](https://twitter.com/10604/status/2103904355766407327)
- Complaint, 2026-09-25, @AmpCode (X): “hey @thorstenball, quick question: what's the story with illustrator in @ampcode? my agent tried to use it for an architecture diagram, but got told it isn't enabled for my account. is it a future feature, or have i missed something? painter came to the rescue in the meantime! <strict_link>” [source](https://twitter.com/1999233079202975744/status/2103374282629980261)

### Zed

- Praise, 2026-09-08, @zeddotdev (X): “@adamholtererer i am playing with muse 1.3 from opencode inside of @zeddotdev delta, and it is pretty nice there, far from even sol, but interesting play with” [source](https://twitter.com/1695152071320743936/status/2097410815624188262)
- Complaint, 2026-09-25, @zeddotdev (X): “@zeddotdev would be great if we could use claude with it!” [source](https://twitter.com/388386067/status/2103516646858256737)
- Complaint, 2026-09-24, @zeddotdev (X): “@zeddotdev it need more llm providers” [source](https://twitter.com/1647734160839135233/status/2102926145910091776)
- Complaint, 2026-09-24, @zeddotdev (X): “@zeddotdev add more ilm providers support please 🙏” [source](https://twitter.com/892350789640781824/status/2103060520534512003)

### Conductor

- Praise, 2026-09-20, @conductor_build (X): “@blueemi99 @claudeai i can already use by @claudeai in @conductor_build and @t3dotcodes 🥸” [source](https://twitter.com/2275729969/status/2101712865300541647)
- Praise, 2026-09-20, @conductor_build (X): “@mparakhin have you tried @conductor_build? doesn’t lock you into a model” [source](https://twitter.com/154998786/status/2101723586750779659)
- Praise, 2026-09-16, @conductor_build (X): “@thelifeofrishi using qwen on @conductor_build and it’s a great combo” [source](https://twitter.com/1382193359641481217/status/2100155270110273641)
- Complaint, 2026-09-22, r/conductorbuild (Reddit): “am i the only one that finds the new 5-model limit super restrictive? i have multiple models from multiple providers i switch between depending on task, 5 is simply not enough. the "share" button is also weird, this is a productivity tool not a game where you share your "loadout". i feel the 5 model limit was chosen for aesthetic reasons. i want to go back to the old one, with its flaws (like showing me codex despite me not even having it inst” [source](https://www.reddit.com/r/conductorbuild/comments/1wndm16/helpdiscussion_new_model_picker_too_restrictive/)
- Complaint, 2026-09-22, @conductor_build (X): “me waiting for @conductor_build @charlieholtz to add opus 5.5 so i can go back to work <strict_link>” [source](https://twitter.com/2891185809/status/2102455765206262026)
- Complaint, 2026-09-22, @conductor_build (X): “while i like almost all of the the productization decisions @conductor_build makes for their harness, this is driving me nuts. especially with all these new models coming out, i need way more than 5 options quickly available to me. @charlieholtz 🥹🙏❓ <strict_link>” [source](https://twitter.com/1821276957428084738/status/2102479190054653992)

### Warp

- Praise, 2026-09-23, @warpdotdev (X): “@warpdotdev confirmed: grok 4.7 touchdown in warp. that ascii rocket descent was nominal and peak terminal flair. connect your subscription and put the agents to work.” [source](https://twitter.com/1720665183188922368/status/2102783192147112024)
- Praise, 2026-09-22, @warpdotdev (X): “great day to be customers of @factoryai @warpdotdev and other model agnostic software factories -- scoop up all those new lab models and keep going!” [source](https://twitter.com/1449604717038825477/status/2102475585218076745)
- Praise, 2026-09-02, @warpdotdev (X): “@adliblove @warpdotdev noted, an issue is open to add this. <strict_link> you can try talking to grok through warp's built-in agent for all of that too. supports grok subscriptions!” [source](https://twitter.com/1042799721948098560/status/2094982757285757065)

### Grok Build

- Praise, 2026-09-10, r/OpenAI (Reddit): “while i love codex, you do have to be delicate with the limits, i see myself using luna a lot more than any other model just to preserve limits i got supergrok, and damn i'm just playing around with grok build, while not as feature rich as codex, for almost the same price, you get grok as your default model which, is sol level quality! with atleast terra level usage and for claude, you atleast get sonnet 5 which is better than luna, tagging it wi” [source](https://www.reddit.com/r/OpenAI/comments/1wc9wzb/switched_from_claude_to_codex_limits_feel_lesser/p8wbu5x/)
