# Latency, throughput and fast mode (`rel.response_speed`)

Feedback Bench, coding agents, built 2026-10-01, window 2026-08-31 to 2026-09-27. Web page: https://feedbackbench.com/#/criterion/rel.response_speed

Area: [Reliability and speed](https://feedbackbench.com/criteria/reliability.md)

**Definition.** How quickly the agent responds and completes work, including paid or fast modes and whether they actually deliver speed.

**Boundary.** Not this: see [Outages, server errors and capacity or rate errors](https://feedbackbench.com/criteria/rel.service_errors.md) for failures. Not this: see rel.client_crash_resources for local slowness from resource use.

Rated author-weeks, all agents: 2906. Complaint share: 63%.

## The brief

Written by Claude Opus 5.5 from 110 labelled posts and the numbers on this page. Interpretation, not measurement: every quote is verbatim and links to its post.

**Speed rides on the model and the week, not the agent.**

TL;DR:

- Most complaints describe sudden slowdowns lasting days, not a steady baseline that users can plan around.
- Fast models get praised on launch, then users report the speed quietly fading.
- Zed and Claude Code earn speed praise; OpenAI Codex draws the heaviest volume of latency complaints.

In plain terms: Users judge speed in waves. A tool feels fast at launch, then a model update or capacity crunch makes simple replies take minutes. The client matters too. The same model can feel quick in one harness and sluggish in another.

### How it breaks

- **Sudden slowdowns that last for days** ([Latency, throughput and fast mode](https://feedbackbench.com/criteria/rel.response_speed.md)). The most common complaint is a tool that was fine last week and now takes minutes to start a simple reply.
  Posts describe multi-day windows where responses crawl, then recover without explanation. Users report minutes before the agent starts, sessions that stall for hours, and speed that varies by time of day. The pattern hits several agents in the same month, which points at serving capacity more than client code. Faster model response is the top request on this page across almost every agent.
  Evidence:
  - Complaint, OpenAI Codex, r/codex, 2026-09-08: “in the past days it sometimes took 4 minutes to just start doing something on sol medium. really unusable.” [source](https://www.reddit.com/r/codex/comments/1wai111/codex_so_slow_its_unusable/p8ialsl/)
  - Complaint, OpenAI Codex, r/ClaudeCode, 2026-09-17: “i've been smashing out my very complex saas, along side 2 other mobile apps for the last 2 months. normally going between codex, and claude code and the 3 programs i'd be flat out trying to keep up with all 4 sessions. but this afternoon i feel like it's all 20x slower. the time i takes to do a simple response i've caught up with all the sessions, played a game, took a dump and mowed the lawn... and come back to it still not complete. am i cooked? ts making kimi k3 seem fast now.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wips2u/turtle_mode/)
  - Complaint, Cursor, r/cursor, 2026-09-08: “same for me, but it depends on the time of the day it seems. but overall, much slower in the last 2-3 days.” [source](https://www.reddit.com/r/cursor/comments/1wam12r/composer_and_grok_slow_down_and_stalling/p8lkyjw/)
  - Complaint, Google Antigravity, r/google_antigravity, 2026-09-19: “today it is unfeasible! 09/19/2026, i have only been using antigravity 2.0, it is very slow.” [source](https://www.reddit.com/r/google_antigravity/comments/1wiotl5/antigravvity_ide_is_running_better/pau2lzu/)

- **Fast models that stop being fast** ([Latency, throughput and fast mode](https://feedbackbench.com/criteria/rel.response_speed.md)). Users praise a model's speed at launch, then notice it slowing as models think longer or launch boosts disappear.
  One user suspects a launch badge marked temporary extra resources, and that pulling them made everything feel slow. Others report updated models reading more files and thinking longer before answering, which costs both time and tokens. A model named for speed taking a minute to answer a greeting lands hard. Newer versions of the same family are also called slower than their predecessors.
  Evidence:
  - Complaint, Google Antigravity, @antigravity, 2026-09-15: “just typed “hello” gemini flash took a full minute to reply. name says flash. reality says buffering @geminiapp @antigravity @officiallogank @_philschmid <strict_link>” [source](https://twitter.com/1551589251766108160/status/2099850501588427168)
  - Complaint, Google Antigravity, r/google_antigravity, 2026-09-15: “when 3.7 and 3.8 came out, there was a little badge that said (i think) "fast" next to those models. i assume those meant that they were running faster than they usually would, for instance, google diverting additional resources to these models so you judge them not on their speed, but rather on their content and quality. that seems like a good strategy on their part but then you get used to how fast it is and then when they pull that, everything seems slow.” [source](https://www.reddit.com/r/google_antigravity/comments/1wh1uyl/whats_your_antigravity_workflow_heres_mine/p9zqvzd/)
  - Complaint, Google Antigravity, r/google_antigravity, 2026-09-25: “looks like they updated the models to read more files and think more before responding. i have also been getting slower responses + more token costs” [source](https://www.reddit.com/r/google_antigravity/comments/1wpzp7k/ag_20_quota_issue/pbzy7bw/)
  - Complaint, Claude Code, @ClaudeDevs, 2026-09-06: “good lord. fable 5.1 is excruciatingly slow. i might swap back to fable 5 as a main. pls @anthropicai @claudedevs do something <strict_link>” [source](https://twitter.com/1518127723335790592/status/2096396255169761465)

- **Startup and history loading drag** ([Latency, throughput and fast mode](https://feedbackbench.com/criteria/rel.response_speed.md)). Before any model runs, users lose time to slow app boots, slow CLI startup and conversation lists that take ages to load.
  Users report remote conversation lists taking 30 to 50 seconds to appear, CLIs with painful startup, and editors that boot slowly enough to send people back to plain editors. Faster app and CLI startup and faster history loading both show up as explicit requests. This is local friction in the shell around the model, and it shapes first impressions every session.
  Evidence:
  - Complaint, Cursor, @cursor_ai, 2026-09-26: “@cursor_ai can you guys fix fucking fix your cli, the startup time is absolutely horrendous just let me use your sub in grok build if its trouble” [source](https://twitter.com/1333285748859088896/status/2103891127564894325)
  - Complaint, OpenAI Codex, X search: OpenAI Codex, Codex CLI, Codex app, 2026-09-06: “@thsottiaux fix the remote control on the codex app, it takes 30-50 seconds just to load all my conversations!!” [source](https://twitter.com/10048612/status/2096718173466927217)
  - Complaint, Zed, @zeddotdev, 2026-08-31: “@zeddotdev is this a new update? coz mine takes like half an eternity to boot up” [source](https://twitter.com/1534532597312901120/status/2094330452727435443)
  - Complaint, Google Antigravity, r/cursor, 2026-09-06: “after trying the hype of cursor and antigravity, i've found myself back in plain vscode. it starts up much faster and just works.” [source](https://www.reddit.com/r/cursor/comments/1w84npz/alternative_to_cursor/p833rjg/)

- **Cheap plans buy volume, not speed** ([Latency, throughput and fast mode](https://feedbackbench.com/criteria/rel.response_speed.md)). Low-cost or bundled model access often comes with throughput users call too slow to work with.
  Posts complain about bundled models capped near 60 tokens per second and hosted APIs taking minutes on tasks a direct provider finishes in seconds. Some users accept the wait as the price of cheap access. Others route urgent work to a paid model. Requests run both ways, with users asking for premium ultra-fast tiers and for slower, cheaper modes to stretch usage.
  Evidence:
  - Complaint, Pi, r/PiCodingAgent, 2026-09-13: “it’s good to have this free but actually deepseek api is also way too cheaper. your api is doing a task in 10 min that a direct deepseek api is doing in 30 seconds.” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wegibk/surprisingly_still_seeing_real_demand_for_v4/p9gfxyb/)
  - Complaint, Cline, @cline, 2026-09-01: “@cline the speed of the deepseek v4 flash model of clinepass is only about 60t/s. although it's cheap, this is too disappointing, trading speed for quantity. additionally, it's not an official model but is priced according to the official model's time-based pricing. very disappointed.” [source](https://twitter.com/1839325856231297024/status/2094919315976020335)
  - Complaint, Pi, r/PiCodingAgent, 2026-09-25: “i like it too. it’s slow though (i am on a measly plus plan) but can afford to wait. i have a claude pointing at ds 4.1 flash when i need some speed and don’t mind spending a few cents.” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wpcis6/opus_55_vs_gpt6_sol_luna_in_piagent_results_on_my/pbza1lj/)
  - Praise, Cursor, r/cursor, 2026-09-27: “i think you’re exactly right. grok 4.7 isn’t as “good” as opus 5.5 but on the other hand it’s a lot faster and doesn’t seem to overthink even simple instructions. i find it very useful for a large number of coding tasks.” [source](https://www.reddit.com/r/cursor/comments/1wrktgc/am_i_the_only_one_who_thinks_grok_47_is_actually/pch445d/)

- **Same model, different harness speed** ([Latency, throughput and fast mode](https://feedbackbench.com/criteria/rel.response_speed.md)). Users notice that one model feels quick in one client and slow in another, so the harness gets blamed too.
  A subscription model reportedly runs faster in its native CLI than through a third-party agent. A strong model gets dragged down by a laggy, buggy harness. A desktop app feels fast while the CLI from the same vendor feels slow. One user cut a review workflow sharply just by switching harness. Speed is a stack property, and users judge the whole stack.
  Evidence:
  - Praise, OpenAI Codex, X search: OpenAI Codex, Codex CLI, Codex app, 2026-09-27: “is it me or chatgpt sub is slower on opencode? workes much faster on codex cli/desktop and pi.” [source](https://twitter.com/195303291/status/2104097784505008346)
  - Complaint, Cline, @cline, 2026-09-26: “@cline not sure if in the desktop app is fast, but in cli feels slow” [source](https://twitter.com/1318567296974073866/status/2103736845507129631)
  - Complaint, OpenCode, r/opencode, 2026-09-17: “i have to say, the output is good but it's painfully slow!!!! it's trash through opencode though, error galore! on openrouter it works smoothly but still slowwww!!” [source](https://www.reddit.com/r/opencode/comments/1wijqfv/union_alpha_is_not_a_model_at_all_it_is_a_2_tier/padtmt4/)
  - Praise, Pi, @pidotdev, 2026-09-17: “in a personal ci code-review experiment, migrating the agents to @pidotdev cut the closest like-for-like review from 593 seconds to 120 seconds. that is one sample, not a universal benchmark. the interesting result is where the time went.” [source](https://twitter.com/1473970545486176256/status/2100571921242792034)

### Who stands out

- **Zed (stronger)**. Zed's speed praise is about the editor itself, which users call faster and lighter than mainstream editors.
  Posts praise raw editor responsiveness and memory efficiency. One user jokes the only flaw is that it is not slow enough to allow distraction. The complaints that do exist target boot time and a feeling of bloat, not model latency. Zed wins here on client snappiness rather than on any model advantage.
  Evidence:
  - Praise, Zed, @zeddotdev, 2026-09-25: “@theo use @zeddotdev ? it is faster than vs code and not intrusive at all.” [source](https://twitter.com/15903760/status/2103287859876712589)
  - Praise, Zed, @zeddotdev, 2026-09-25: “@zeddotdev the only thing to hate about zed is that it is isn’t slow enough. i miss the waiting time to get distracted” [source](https://twitter.com/377812931/status/2103316200688345574)
  - Praise, Zed, @zeddotdev, 2026-09-16: “@strzibnyj i just now installed zed, very memory efficient &amp; very fast @zeddotdev” [source](https://twitter.com/1888252538836955136/status/2100218449238827397)
  - Complaint, Zed, @zeddotdev, 2026-08-31: “@zeddotdev is this a new update? coz mine takes like half an eternity to boot up” [source](https://twitter.com/1534532597312901120/status/2094330452727435443)

- **Claude Code (stronger)**. Claude Code users credit its newest top model with finishing faster, which also makes it cheaper per task.
  Users report the flagship model completing hard tasks in roughly half the time of alternatives, so total cost drops despite higher prices. Parallel sessions and subagents get praised for aggregate throughput. The exception is another model family in the lineup, which users call excruciatingly slow and say they would swap away from.
  Evidence:
  - Praise, Claude Code, r/ClaudeCode, 2026-09-22: “i've been using it for the last few hours. it seems excellent, fast, cheaper on usage and i don't think it's said load-bearing once!” [source](https://www.reddit.com/r/ClaudeCode/comments/1wne9dx/introducing_claude_opus_55/pbgqvdp/)
  - Praise, Claude Code, r/ClaudeCode, 2026-09-23: “i did a similar test today and got similar results. for a difficult task in my workflow, with all models on high effort, opus 5 and fable 5.1 took substantially longer than opus 5.5. weighting token use by what it would have cost with api pricing, the ratios were comparable to yours (opus 5 slightly more expensive than fable 5.1, which was more than double the cost of 5.5). all converged on the same answer. what's even crazier is i did a second test with a medium difficulty task comparing opus 5.5 and sonnet 5. once again opus 5.5 was cheaper because it finished in half the time. so right now opus 5.5 seems like a real one-size-fits-all model until the newer version of the other models come out. i hope they don't nerf it.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wnq2de/opus_55_is_pure/pbhbc8i/)
  - Praise, Claude Code, r/ClaudeCode, 2026-09-07: “i find that coded burns through way more tokens. even if they're more tokens/sec it's still way slower because it uses a bunch more. also i notice that claude is much faster with multiple sessions. i still can't do multiple sessions on same screen with codex/chatgpt windows, in constantly running 4-5 all on screen with claude. if you're getting 55 tok/s on fable with 1 session, 5 is getting 200 tok/s. plus with a ton of subagents i burn through 1m tokens quick... way over 2000 tok/s how many tokens do you get astra vs fable per week?” [source](https://www.reddit.com/r/ClaudeCode/comments/1w9nlqr/can_claude_max_20x_still_compete_with_astra_with/p8ets9i/)
  - Complaint, Claude Code, @ClaudeDevs, 2026-09-06: “good lord. fable 5.1 is excruciatingly slow. i might swap back to fable 5 as a main. pls @anthropicai @claudedevs do something <strict_link>” [source](https://twitter.com/1518127723335790592/status/2096396255169761465)

- **OpenAI Codex (weaker)**. OpenAI Codex draws the largest pile of latency complaints, centred on slow starts, stalling runs and sluggish app loading.
  Users describe the CLI running slowly while repeatedly auto-compacting, and the app taking long to load conversations. It leads every speed request on this page, including faster responses, faster startup and faster history loading. Praise exists for fast tiers and a cleaned-up app, and one user traced their own slowness to a misconfigured proxy.
  Evidence:
  - Complaint, OpenAI Codex, X search: OpenAI Codex, Codex CLI, Codex app, 2026-09-14: “@thsottiaux in the last 1-2 days, codex cli has been working very poorly - it takes a long time to run, constantly performs autocompact, and repeatedly finds and fixes its own errors” [source](https://twitter.com/1838953404648882176/status/2099469586831647174)
  - Complaint, OpenAI Codex, X search: OpenAI Codex, Codex CLI, Codex app, 2026-09-06: “@thsottiaux fix the remote control on the codex app, it takes 30-50 seconds just to load all my conversations!!” [source](https://twitter.com/10048612/status/2096718173466927217)
  - Praise, OpenAI Codex, r/codex, 2026-09-05: “the speed is excellent and the verbosity is right down. loving it so far.” [source](https://www.reddit.com/r/codex/comments/1w7l3yj/gpt6_astra_is_now_out_to_all_plus_business_pro/p7vxrqr/)
  - Praise, OpenAI Codex, X search: OpenAI Codex, Codex CLI, Codex app, 2026-09-13: “@bu7emba this was self imposed on my part actually; i was running a proxy for the api calls on background priority instead of interactive priority. codex app is fast for me now.” [source](https://twitter.com/162208766/status/2099139564380225652)

- **Devin (mixed)**. Devin's new model earns raves for response time, but users say the surrounding harness lags and drags it down.
  Posts praise near-instant replies and a model that outpaces its base. The same users call the harness laggy, buggy and slow, especially on Windows, and question why its token speed ranks low. Devin users also lead requests for more server capacity, suggesting the speed is real but not yet dependable.
  Evidence:
  - Complaint, Devin, @cognition, 2026-09-11: “i have to tell you that swe-2 model looks great. but having to use devin coding harness is almost making me to abandon it. it's laggy, kinda buggy, slow and lacks some important features that i use in a daily basis. @cognition needs to improve asap...” [source](https://twitter.com/64041638/status/2098391118211665969)
  - Praise, Devin, @cognition, 2026-09-16: “@dabit3 @darega0mae @cognition it works! thank you nader. unreal response time!” [source](https://twitter.com/1588954793124548615/status/2100310161730535762)
  - Praise, Devin, @cognition, 2026-09-11: “learnt about @cognition for a long time and just tried their subscription because of their latest swe-2 model. boy o boy... this thing is really good and fast, it was post trained on kimi k3 but i never had such speed on kimi k3 ever! and there is currently no usage limit on models like swe-2 and glm 5.2 on devin. devin's pro $20 plan seems like the best plan money can buy on the market right now” [source](https://twitter.com/197714768/status/2098381449724403775)
  - Complaint, Devin, @cognition, 2026-09-12: “@getaskclaw @cognition why is the speed ranked last? can't they achieve 100tps?” [source](https://twitter.com/2073035346707968000/status/2098572491828666850)

- **Pi (stronger)**. Pi users praise the harness as lean and quick, while its hosted free model draws slowness complaints.
  One user reports a like-for-like review job dropping from minutes to about two minutes after moving agents to Pi. Others call its speed brutal compared with rival agents. The weak spot is Pi's own hosted model access, which users say runs far slower than calling the provider directly.
  Evidence:
  - Praise, Pi, @pidotdev, 2026-09-17: “in a personal ci code-review experiment, migrating the agents to @pidotdev cut the closest like-for-like review from 593 seconds to 120 seconds. that is one sample, not a universal benchmark. the interesting result is where the time went.” [source](https://twitter.com/1473970545486176256/status/2100571921242792034)
  - Praise, Pi, @pidotdev, 2026-09-11: “la verdad es que llevo varias semanas con @pidotdev y la velocidad que tiene es brutal vs a claude, cursor, codex, opencode...” [source](https://twitter.com/300503103/status/2098355293243511006)
  - Complaint, Pi, r/PiCodingAgent, 2026-09-13: “it is slow as hell and also making your other model offering slow.” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wegibk/surprisingly_still_seeing_real_demand_for_v4/p9hnqo6/)
  - Complaint, Pi, r/PiCodingAgent, 2026-09-25: “i like it too. it’s slow though (i am on a measly plus plan) but can afford to wait. i have a claude pointing at ds 4.1 flash when i need some speed and don’t mind spending a few cents.” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wpcis6/opus_55_vs_gpt6_sol_luna_in_piagent_results_on_my/pbza1lj/)

### Fine print

- Many posts credit or blame a specific model rather than the agent, so harness and model speed are hard to separate.
- Several agents have too few posts here to rank. Their quotes illustrate patterns but carry little weight.
- Evidence clusters in one month, so short outages and launch spikes may be overrepresented.

## Top requests

What users ask to add or change, most asked first. 173 author-weeks ask for something. Requests do not change the Feedback Score. Rule: A separate pass by Claude Sonnet 5 reads every counted post and extracts what the author asks the agent or its vendor to add or change, with the criteria it maps to and a short normalised wording; it does not touch the labels or the Feedback Score. Claude Opus 5.5 groups the wordings within each criterion (the first criterion the request maps to) into themes; code counts them. A theme counts distinct author-weeks that ask for it, per agent; across agents, one author-week per agent. Themes asked in fewer than 2 author-weeks, and requests that share no theme, are not shown. Examples: up to 3 posts per theme from different authors, without slurs, preferring posts of 60 to 450 characters, most recent first.

| Rank | Request | Author-weeks | Posts | Agents (author-weeks) |
|---|---|---|---|---|
| 1 | Faster model response speed | 76 | 76 | OpenAI Codex 20, Google Antigravity 19, Claude Code 12, OpenCode 11, Cursor 6, Devin 3, Cline 2, Pi 2, Kiro 1 |
| 2 | Faster application and CLI startup | 12 | 12 | OpenAI Codex 5, Cursor 4, Google Antigravity 2, Pi 1 |
| 3 | Add or expand fast mode availability | 8 | 8 | OpenAI Codex 3, Cursor 2, Devin 2, Factory 1 |
| 4 | Faster speed on lower or current plans | 6 | 6 | Amp 1, Google Antigravity 1, Cline 1, OpenAI Codex 1, Devin 1, Pi 1 |
| 5 | More server capacity for speed | 6 | 6 | Devin 3, OpenAI Codex 1, Cursor 1, OpenCode 1 |
| 6 | Premium ultra-fast mode at higher cost | 6 | 6 | OpenAI Codex 2, Devin 2, Google Antigravity 1, Cursor 1 |
| 7 | Slower cheaper mode to save usage | 6 | 6 | OpenAI Codex 4, GitHub Copilot 1, Cursor 1 |
| 8 | Faster completion of simple tasks | 5 | 5 | OpenAI Codex 3, Google Antigravity 1, Claude Code 1 |
| 9 | Faster desktop app matching CLI speed | 5 | 5 | OpenAI Codex 2, Claude Code 1, Cline 1, Cursor 1 |
| 10 | Higher token throughput and output speed | 5 | 5 | Claude Code 2, Google Antigravity 1, OpenAI Codex 1, OpenCode 1 |
| 11 | Restore previous response speed | 5 | 5 | OpenAI Codex 2, Google Antigravity 1, Cursor 1, OpenCode 1 |
| 12 | Faster loading of chats and history | 4 | 4 | OpenAI Codex 3, Cursor 1 |

### 1. Faster model response speed

- Claude Code, 2026-09-27, r/ClaudeCode (Reddit): “so here is also where i am stuck. i am limited with these, the faster i want to work, either ai needs to speed up with its results instead of letting me wait for so long. but even after it instantly returns the correct way, it will be limited by your own capacity. so the other thing is to fully trust the ai to make decisions, and here is the difficult part. till what extend can you do this? without losing track of how it works under the hood.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wradi9/why_normal_discussions_has_so_many_haters/pcg0zse/)
- Google Antigravity, 2026-09-26, r/google_antigravity (Reddit): “that's a lot of text, so i just wanna ask first. is gemini flash 3.8's speed fixed? it's so slow it should be changed to gemini slow 3.8.” [source](https://www.reddit.com/r/google_antigravity/comments/1wqbyzk/antigravity_cli_release_v127_v1211/pc333k5/)
- Cline, 2026-09-26, @cline (X): “@cline this model very super slow, tell the developer of this model 🗿, fix that speed” [source](https://twitter.com/1883868794000695296/status/2103703396515483978)

### 2. Faster application and CLI startup

- Cursor, 2026-09-26, @cursor_ai (X): “@cursor_ai can you guys fix fucking fix your cli, the startup time is absolutely horrendous just let me use your sub in grok build if its trouble” [source](https://twitter.com/1333285748859088896/status/2103891127564894325)
- Google Antigravity, 2026-09-24, r/LocalLLaMA (Reddit): “unfortunately this. if anyone thinks antigravity was bad, gemini cli is in even worse state (it takes forever just to load even though it is just a cli, no update for new models, can't even use their subscription). it seems like this end of ai support in google has been struggling (other than the model itself, which is magnificent).” [source](https://www.reddit.com/r/LocalLLaMA/comments/1wof9kk/introducing_support_for_local_ai_models_in_the/pbocdmf/)
- Cursor, 2026-09-17, @cursor_ai (X): “hey, @cursor_ai i like your product, but could you please stop making it worse? it now takes &gt; 30 seconds to start up and it never starts in the ide window i last used, but in the agents view (without a simple way to open the repo in the ide).” [source](https://twitter.com/2717214655/status/2100458360545935804)

### 3. Add or expand fast mode availability

- Devin, 2026-09-16, @cognition (X): “@cognition i am really loving swe-2.0 as my go to for quick ops tasks and answering questions about my codebase. really really fun model for this slice in the agentic dev stack. would be cool to have a fast mode like the swe-2.0 model :)” [source](https://twitter.com/4893651116/status/2100366132217823299)
- OpenAI Codex, 2026-09-12, r/codex (Reddit): “let me use it during peak hours at least. but the only thing is it must save more than 50% tokens if its 50% speed. because at that point it makes zero difference.” [source](https://www.reddit.com/r/codex/comments/1wdsjpe/please_give_us_a_slow_mode/p99ew8t/)
- OpenAI Codex, 2026-09-11, r/codex (Reddit): “i'm hoping they have now optimized it so well that they will replace the codex 5.3 spark on cerebras for pro users. sol at 700 tps, fucking give it to me.” [source](https://www.reddit.com/r/codex/comments/1wcpz5j/gpt6sol_staged_in_openai_api/p93ggfi/)

### 4. Faster speed on lower or current plans

- Pi, 2026-09-25, r/PiCodingAgent (Reddit): “i like it too. it’s slow though (i am on a measly plus plan) but can afford to wait. i have a claude pointing at ds 4.1 flash when i need some speed and don’t mind spending a few cents.” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wpcis6/opus_55_vs_gpt6_sol_luna_in_piagent_results_on_my/pbza1lj/)
- Amp, 2026-09-22, @AmpCode (X): “@thorstenball brother i started orb'n today and first of all thank you to the entire team of @ampcode and especially @sqs (i'm waiting for his next live stream, probably you should too) /feedback, the orb's terminal on mw plan seems too slow though” [source](https://twitter.com/1940807256377118724/status/2102466139293126912)
- Google Antigravity, 2026-09-17, @antigravity (X): “@googledeepmind @geminiapp @antigravity please, please , please fix this speed issue. barely getting any throughput and my ultra plan is burning down at ~2x the speed it usually does.” [source](https://twitter.com/178086973/status/2100655461972472243)

### 5. More server capacity for speed

- OpenAI Codex, 2026-09-24, r/codex (Reddit): “we need fast infra. nvda dgx ain't it. 5090 is stupid expensive” [source](https://www.reddit.com/r/codex/comments/1wpd1gc/moarrrrr_higher_tier_pro_plans_are_forthcoming/pbuus2s/)
- Devin, 2026-09-18, @cognition (X): “@fei2411 @cognition @cognition please expand the capacity of your devin servers to boost the speed of swe-2. it's currently too slow.” [source](https://twitter.com/1159835302275346433/status/2100823113516941327)
- Devin, 2026-09-18, @cognition (X): “the server capacity of devin still seems insufficient, swe-2 is particularly slow now, and it has already become a bit less intelligent 🫠 quickly increase the computing power @cognition 🥺” [source](https://twitter.com/2055515632146472960/status/2100800965997928941)

### 6. Premium ultra-fast mode at higher cost

- OpenAI Codex, 2026-09-22, r/codex (Reddit): “it's really nice, but it's missing an "ultra max ultrafast" mode” [source](https://www.reddit.com/r/codex/comments/1wnm7f6/created_custom_animation_bar_for_codex_on_public/pbg4hcx/)
- Cursor, 2026-09-22, @cursor_ai (X): “it would be awesome to have a hyper-speed toggle or turbo option for power users who want maximum speed @grok @cursor_ai <strict_link>” [source](https://twitter.com/1959973200881737728/status/2102416779373126126)
- Devin, 2026-09-15, @cognition (X): “ngl swe-1.7-lightning is so addictive @cognition @cerebras it's amazing to be able to burn tokens at the speed of thought (or faster) i can't wait to see what swe-2 is able to do on lightning mode... i'd gladly pay orders of magnitude more in order to have even faster compute” [source](https://twitter.com/1771302147348660224/status/2099945957962154249)

### 7. Slower cheaper mode to save usage

- Cursor, 2026-09-23, r/cursor (Reddit): “yeah a slow/overnight lane for the same model would be useful for batch refactors and evals. right now fast is mostly a priority/queue upsell, not a smarter model, so the missing downsell feels intentional. closest diy is off-peak runs, smaller contexts, or pinning a cheaper explicit model instead of auto. worth asking support/feature voting; a lot of people want the same thing.” [source](https://www.reddit.com/r/cursor/comments/1wo1gt6/why_can_i_pay_more_for_fast_but_not_pay_less_for/pbjk9ws/)
- OpenAI Codex, 2026-09-13, r/codex (Reddit): “lol this is not really a codex issue, my ci is long ass. just wanted to slow down codex if possible” [source](https://www.reddit.com/r/codex/comments/1tjfxcf/anyone_else_ask_here_about_current_codex_issues/p9k09gx/)
- OpenAI Codex, 2026-09-09, r/codex (Reddit): “can we get ultraslow? astra is fast as hell, i would be fine with slower responses if it cut down on token use.” [source](https://www.reddit.com/r/codex/comments/1wbn954/i_got_ultrafast_on_codex/p8rj6lg/)

### 8. Faster completion of simple tasks

- OpenAI Codex, 2026-09-23, r/codex (Reddit): “also way slower, takes like 3 hours to do a single basic task” [source](https://www.reddit.com/r/codex/comments/1wnocd6/how_does_gpt6_sol_fare/pbgz6ad/)
- Google Antigravity, 2026-09-18, @antigravity (X): “@antigravity 3.8 flash on high is unusable. it takes 10 minutes to change 1 line of copy. also, can we please work on improving the app. you still can't reorder messages in a queue.” [source](https://twitter.com/52840428/status/2100869553437671458)
- OpenAI Codex, 2026-09-07, r/codex (Reddit): “is anyone having speed issues. gtp 5.6 medium. most basic prompts now taking 30 mins, friday same work was taking 4-5 mins. is ther an issue are are they trying to get us to use astra?” [source](https://www.reddit.com/r/codex/comments/1w3i0mm/codex_usage_and_operation_discussion_last_updated/p8bqcp6/)

### 9. Faster desktop app matching CLI speed

- Cline, 2026-09-21, @cline (X): “@cline please fix ssh session bugs. and desktop output is so slow compared to cli.” [source](https://twitter.com/1927003633196969984/status/2102176503857664095)
- Cursor, 2026-09-11, @cursor_ai (X): “@cursor_ai @bot what about creating project to speed up you desktop app as this complete laggy nightmare on m3 macbook? also your harness is very very slow with even fast llms outside cursor like gemini fast.” [source](https://twitter.com/1089070555/status/2098276461056499753)
- Claude Code, 2026-09-08, @ClaudeDevs (X): “@claudedevs claude moves in slowmotion these day's.... work on interface and optimisation...” [source](https://twitter.com/1430186063570558994/status/2097370310605484122)

### 10. Higher token throughput and output speed

- Claude Code, 2026-09-23, @ClaudeDevs (X): “@claudedevs cool, could you make token output faster as well?” [source](https://twitter.com/2093080824820244481/status/2102881622949585367)
- OpenAI Codex, 2026-09-22, r/codex (Reddit): “honestly 5.6 luna max is already enough for everything i need, i don't care about anything other than cost-benefit on real use and tps(tokens per second) would be great to improve too” [source](https://www.reddit.com/r/codex/comments/1wnica9/gpt_6_sol_6_and_luna_19_in_aai_index/pbfg3bl/)
- OpenCode, 2026-09-21, r/google_antigravity (Reddit): “i hit claude limit after 4mins of opus usage on pro plan. i have already cancelled gemini pro sub. switching to opencode. google strangles usage limits at their desire without informing users. ds4.1f has been a huge step up over 3.8 flash. i wish opencode could serve it at 300tps.” [source](https://www.reddit.com/r/google_antigravity/comments/1wm7g5r/weekly_quotas_known_issues_support_september_21/pb7a4a2/)

### 11. Restore previous response speed

- Google Antigravity, 2026-09-22, r/google_antigravity (Reddit): “yeah it is working slow since the last week, but the token usage remains the same for me.” [source](https://www.reddit.com/r/google_antigravity/comments/1wm9ouj/is_agy_slow_again_for_anyone_else_or_is_it_just_me/pbd7hlb/)
- OpenAI Codex, 2026-09-12, r/codex (Reddit): “its so slow for me as well now, astra on medium could get tasks done in 20 minutes it cant do in over an hour now” [source](https://www.reddit.com/r/codex/comments/1we551s/this_nerf_is_getting_out_of_hand_look_at_this/p9bay6w/)
- OpenAI Codex, 2026-09-12, r/codex (Reddit): “i feel like they are trying to channel the traffic to astra instead of other models to get people more hooked on astra. yes astra is pretty good but it's very slow. luna was very fast like it almost instantly answers even on non fast mode but it's currently at the speed of astra. that's what i don't understand.” [source](https://www.reddit.com/r/codex/comments/1we6ece/is_the_overall_slowness_from_lack_of_compute/p9baflg/)

### 12. Faster loading of chats and history

- Cursor, 2026-09-11, @cursor_ai (X): “@cursor_ai @bot would be great if it actually worked. loading repositories while creating a project takes forever.” [source](https://twitter.com/2184642108/status/2098552550568173614)
- OpenAI Codex, 2026-09-06, X search: OpenAI Codex, Codex CLI, Codex app (X): “@thsottiaux fix the remote control on the codex app, it takes 30-50 seconds just to load all my conversations!!” [source](https://twitter.com/10048612/status/2096718173466927217)
- OpenAI Codex, 2026-09-02, X search: OpenAI Codex, Codex CLI, Codex app (X): “hmm. why does typing @ in codex app feels laggy and how do i make it not try and load the world every time i do? lul” [source](https://twitter.com/28952775/status/2094977020862300194)

## Every agent

| Agent | Overall rank | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|---|
| [Zed](https://feedbackbench.com/agents/zed.md) | =8 | Better than peers | 0.603 | 0.576–0.629 | 59 | 51 | 8 |
| [Claude Code](https://feedbackbench.com/agents/claude-code.md) | 1 | Better than peers | 0.576 | 0.545–0.606 | 348 | 165 | 183 |
| [Pi](https://feedbackbench.com/agents/pi.md) | 7 | Better than peers | 0.542 | 0.514–0.568 | 54 | 31 | 23 |
| [Cline](https://feedbackbench.com/agents/cline.md) | =8 | Typical | 0.518 | 0.491–0.544 | 59 | 30 | 29 |
| [OpenCode](https://feedbackbench.com/agents/opencode.md) | 3 | Typical | 0.492 | 0.462–0.519 | 532 | 200 | 332 |
| [Google Antigravity](https://feedbackbench.com/agents/antigravity.md) | =5 | Typical | 0.482 | 0.453–0.510 | 525 | 181 | 344 |
| [Cursor](https://feedbackbench.com/agents/cursor.md) | 4 | Typical | 0.468 | 0.436–0.500 | 211 | 68 | 143 |
| [Devin](https://feedbackbench.com/agents/devin.md) | =5 | Worse than peers | 0.468 | 0.438–0.497 | 70 | 21 | 49 |
| [OpenAI Codex](https://feedbackbench.com/agents/codex.md) | 2 | Worse than peers | 0.452 | 0.429–0.475 | 982 | 295 | 687 |
| [GitHub Copilot](https://feedbackbench.com/agents/copilot.md) | =8 | Too few posts | – | – | 22 | 10 | 12 |
| [Amp](https://feedbackbench.com/agents/amp.md) | =11 | Too few posts | – | – | 13 | 5 | 8 |
| [Factory](https://feedbackbench.com/agents/factory.md) | =11 | Too few posts | – | – | 11 | 7 | 4 |
| [Kiro](https://feedbackbench.com/agents/kiro.md) | 13 | Too few posts | – | – | 8 | 2 | 6 |
| [Warp](https://feedbackbench.com/agents/warp.md) | 15 | Too few posts | – | – | 7 | 5 | 2 |
| [Grok Build](https://feedbackbench.com/agents/grok-build.md) | 16 | Too few posts | – | – | 3 | 1 | 2 |
| [Conductor](https://feedbackbench.com/agents/conductor.md) | 14 | Too few posts | – | – | 2 | 1 | 1 |
| [Augment Code](https://feedbackbench.com/agents/augment.md) | 17 | Too few posts | – | – | 0 | 0 | 0 |

## Posts

Receipts rule: The 5 most recent praise and complaint posts per area (first 700 characters) and 3 per criterion (first 450 characters).

### Zed

- Praise, 2026-09-27, r/ZedEditor (Reddit): “as long as it’s fast, responsive and no memory bloat, i’ll keep using it - biggest reason why i moved away from vs code. does zed have some kind of task manager or something? would love to see the effect of every extension i install haha.” [source](https://www.reddit.com/r/ZedEditor/comments/1wris6u/pls_dont_turn_into_vs_code/pcekb3x/)
- Praise, 2026-09-27, @zeddotdev (X): “@0xprajwal_ i use @zeddotdev, btw. it’s fast, lightweight, consumes less memory, and uses the gpu for rendering.” [source](https://twitter.com/1542127649463898113/status/2104217257153110457)
- Praise, 2026-09-26, @zeddotdev (X): “@zeddotdev yk what zed? i moved from vscode to zed quite a while ago and it's fucking amazing it starts faster looks better lets me disable ai features bullshit completely honestly, thanks for making my life a bit more better” [source](https://twitter.com/1188417649203539968/status/2103694133131202673)
- Complaint, 2026-09-27, r/ZedEditor (Reddit): “you can disable them but still zed is slower than gram” [source](https://www.reddit.com/r/ZedEditor/comments/1wris6u/pls_dont_turn_into_vs_code/pce43ch/)
- Complaint, 2026-09-25, r/ZedEditor (Reddit): “yeah, it's great but not so performant; also, i face some random issues with wsl.” [source](https://www.reddit.com/r/ZedEditor/comments/1wpr20d/is_there_any_zed_terminal/pbxz2np/)
- Complaint, 2026-09-21, @zeddotdev (X): “@pukno_ai @zeddotdev too slow” [source](https://twitter.com/140255012/status/2102027043412340744)

### Claude Code

- Praise, 2026-09-27, r/ClaudeCode (Reddit): “i keep running out of session with opus55. much faster.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wqnh3q/me_after_opus_55_release/pca1d74/)
- Praise, 2026-09-27, r/ClaudeCode (Reddit): “5.5 is really good but it does have the same flaws as all the other agents. i think one of the biggest appeals (any why most love it) is the huge reduction in its draw on subs (5x just became more like 30x if you set it on medium effort) and it speaks plain language for the most part, oh and its faster and less verbose, all while being a solid reasoning machine. its fast, cheap, and works. that makes for happy coders.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wqjuv2/am_i_the_only_one_who_doesnt_like_opus_55/pca70cd/)
- Praise, 2026-09-27, r/ClaudeCode (Reddit): “it is truly vastly different from opus 5. the cost and response speed are also completely different. that helps me stay better focused.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wqzt2x/opus_55_experience_of_an_engineer_at_big_tech/pcb4klf/)
- Complaint, 2026-09-27, r/ClaudeCode (Reddit): “um. around the same i think? opus 5.5 is like around 30 tokens/ a sec for me. but it warries quite a bit. 30 is properly on the lower end” [source](https://www.reddit.com/r/ClaudeCode/comments/1wriidg/opus_55_vs_sol_in_terms_of_speed/pccwa0e/)
- Complaint, 2026-09-27, r/ClaudeCode (Reddit): “opus 5.5 feels painfully slow, but it does a good job. if i want quick iteration i use astra, because opus is so slow that it's like pulling teeth for smaller tasks.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wriidg/opus_55_vs_sol_in_terms_of_speed/pcdouht/)
- Complaint, 2026-09-27, r/ClaudeCode (Reddit): “but its still slow? because you make most decisions?” [source](https://www.reddit.com/r/ClaudeCode/comments/1wradi9/why_normal_discussions_has_so_many_haters/pcft0al/)

### Pi

- Praise, 2026-09-27, r/PiCodingAgent (Reddit): “yeah. on all my tests accuracy was quite shit, and jev was consistently outperformed by glm 5.3 flash, for anything requiring decision making. but hey, it's fast 🤷♂️” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wnl7da/pi_can_now_use_jev_and_more/pce99cd/)
- Praise, 2026-09-27, r/PiCodingAgent (Reddit): “pretty good, these days my pi is stock plus one extension showing context usage in detail and one custom footer override to apply custom styling. i do maintain one patch for the llama.cpp provider to enable per model and session thinking levels. performance is fast but i use a dual radeon r9700 setup. i pretty much work offline and with no delays.” [source](https://www.reddit.com/r/PiCodingAgent/comments/1vjurqz/considering_claude_code_pi_worth_it/pch2ftv/)
- Praise, 2026-09-26, r/PiCodingAgent (Reddit): “nice work. i tried using it. it is fast. i tried using few extensions but they don’t seem to work with pig. mcp-adapter, pi-web-access, pi-rtk-optimizer” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wqe2v1/meet_pig_the_pi_coding_harness_that_is_yours_but/pc3o4sz/)
- Complaint, 2026-09-26, @pidotdev (X): “@pigcodingagent @pidotdev pi could be faster indeed” [source](https://twitter.com/207233683/status/2103895513695236527)
- Complaint, 2026-09-26, @pidotdev (X): “@pigcodingagent @pidotdev holy ! ty for this, i've been wishing pi would be available in rust or something other than js, the startup time can be annoying sometime” [source](https://twitter.com/1318274063199064064/status/2103905472277557442)
- Complaint, 2026-09-25, r/PiCodingAgent (Reddit): “i like it too. it’s slow though (i am on a measly plus plan) but can afford to wait. i have a claude pointing at ds 4.1 flash when i need some speed and don’t mind spending a few cents.” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wpcis6/opus_55_vs_gpt6_sol_luna_in_piagent_results_on_my/pbza1lj/)

### Cline

- Praise, 2026-09-27, @cline (X): “@cline love the speed in the new desktop app” [source](https://twitter.com/1640928038954336258/status/2104336348568342539)
- Praise, 2026-09-26, @cline (X): “@cline really good! after integrating cline for free, it crushes the next.js tasks of kimi k3—pixel canary action is really fast.” [source](https://twitter.com/1744672948135321600/status/2103663198188749016)
- Praise, 2026-09-26, @cline (X): “@cline that speed boost in the new cline desktop app is so satisfying to use” [source](https://twitter.com/1889631970667405317/status/2103747460858302559)
- Complaint, 2026-09-27, @cline (X): “i know stealth models seem to be the in thing right now, but i'm not sure they're even worth messing around with sometimes. trying to use pixel canary on @cline, and it's just so slow. i mean, 24 hours now, no closer to the task, and it keeps stopping and starting. it's horrible!” [source](https://twitter.com/25673607/status/2104177440129991012)
- Complaint, 2026-09-26, @cline (X): “@cline it's too slow fr” [source](https://twitter.com/1949510520748249088/status/2103662097926393956)
- Complaint, 2026-09-26, @cline (X): “@cline it's also quite slow at the auto/highest reasoning 🤔🤔” [source](https://twitter.com/1803325156955480064/status/2103667261743759603)

### OpenCode

- Praise, 2026-09-27, r/opencodeCLI (Reddit): “it's only getting faster..” [source](https://www.reddit.com/r/opencodeCLI/comments/1wrddp1/minimax_is_actively_testing_m31flashpreview/pccp56a/)
- Praise, 2026-09-27, r/opencode (Reddit): “yea, fast is the thing i love most tbh” [source](https://www.reddit.com/r/opencode/comments/1wroplu/xiaomi_mimo_26_flash_vs_glm_53_flash/pcfapui/)
- Praise, 2026-09-27, r/opencode (Reddit): “i felt the speed performance definitely. the other issue i have with it is high rate of compactions in relation to codex. but definitely a step in the right direction.” [source](https://www.reddit.com/r/opencode/comments/1wrj9e6/deepseekopencode_is_great/pcfx9ox/)
- Complaint, 2026-09-27, r/opencode (Reddit): “how's the speed? 5.3 flash on go is like a turtle. can't stand it.” [source](https://www.reddit.com/r/opencode/comments/1wr8w3t/cheepseek_phase_2_full_precision/pcbnf96/)
- Complaint, 2026-09-27, r/opencode (Reddit): “not sure what you mean. opencode go hosts this model through official enterprise gateways, they serve the lossless base checkpoint bit-for-bit. what you are upset about is the difference in speed between deepseek api and opencode api. the likely cause for this is opencode's architecture (re-routing) and the context re-reading vs. deepseek's native kv caching. so no, you do not get quantized tokens. only how the tokens arrive to you differs.” [source](https://www.reddit.com/r/opencode/comments/1wrg81n/go_subscription_is_slower_deepseek_flash_41/pccdoi8/)
- Complaint, 2026-09-27, r/opencode (Reddit): “its getting slower too in my experience” [source](https://www.reddit.com/r/opencode/comments/1wr8w3t/cheepseek_phase_2_full_precision/pce5q8s/)

### Google Antigravity

- Praise, 2026-09-27, r/google_antigravity (Reddit): “subjective choice here. 1. codex cant pick a side its always beating around the bushes and it just steals your project information like zcode did before it was caught. 2. yes cc is good but its not as fast as gemini 3.8 flash which is a quick workhorse model. cc is for longhorizon no coding projects but i still need to design my architectures and coding projects so i prefer using gemini. also claude pretends to be the superman of modern llms whic” [source](https://www.reddit.com/r/google_antigravity/comments/1wrbsby/antigravitys_gemini_models_are_worst_in_among_all/pce8fep/)
- Praise, 2026-09-27, r/google_antigravity (Reddit): “who told you i dont have codex or claude max plan, i have them also, but antigravity, and only antigravity is the one that complements them, also i am using deepthinking in gemini chat, and it is so good, not what you think, i can get real time feedback with antigravity, solve problems together, spawn tens of agents and get things done in minutes, while with codex or claude, i need to put them on task before bed and wake up to see the results, th” [source](https://www.reddit.com/r/google_antigravity/comments/1wrn9q5/codex_has_ultra_thinking_level_claude_code_has/pcencos/)
- Praise, 2026-09-27, r/google_antigravity (Reddit): “i have claude max and i still use antigravity for coding. i use claude for claude design, antigravity for everything else. i have used claude code and i still prefer antigravity. the result for me are the same and i prefer the output speed of gemini, it allows me to iterate fast.” [source](https://www.reddit.com/r/google_antigravity/comments/1whvudv/people_who_are_complaining_about_gemini/pch3gm0/)
- Complaint, 2026-09-27, @antigravity (X): “@antigravity i haven't used a plan since 2025... google is as fast as a merchant ship from 1920” [source](https://twitter.com/2482947307/status/2104200639890780481)
- Complaint, 2026-09-26, r/google_antigravity (Reddit): “that's a lot of text, so i just wanna ask first. is gemini flash 3.8's speed fixed? it's so slow it should be changed to gemini slow 3.8.” [source](https://www.reddit.com/r/google_antigravity/comments/1wqbyzk/antigravity_cli_release_v127_v1211/pc333k5/)
- Complaint, 2026-09-26, r/google_antigravity (Reddit): “why is 3.8 terribly slow after the first few days of launching?” [source](https://www.reddit.com/r/google_antigravity/comments/1wqb0ii/dedicated_planning_mode_in_antigravity_is_here/pc3rbru/)

### Cursor

- Praise, 2026-09-27, r/cursor (Reddit): “you are right about part of it but wrong about fast being slower than normal . that could never happen. fast will have priority on resources. normal comes seconds . in rare cases normal could be same speed as fast . the speed of normal varies up and down. best case its same as fast . but fast almost always same speed for a certain window of time .” [source](https://www.reddit.com/r/cursor/comments/1wp2ped/i_compared_cursor_composer_25_normal_vs_fast/pccmxhn/)
- Praise, 2026-09-27, r/cursor (Reddit): “sorry but this sounds like an operator problem. my system was running slow at one point so i prompted it to optimize my system. haven’t had a problem since, that was about 5 months ago.” [source](https://www.reddit.com/r/cursor/comments/1wqppuh/grok_is_shutting_down_apps_now/pcf5mxo/)
- Praise, 2026-09-27, r/cursor (Reddit): “for my web development usage i find cursor's models like composer to be much faster for simple tasks. i also like the inbuilt browser and ide interface, despite how it seems like they are trying to make it an afterthought in the app. opus 5.5 is next level but i've only found myself reaching for that for tougher tasks.” [source](https://www.reddit.com/r/cursor/comments/1wrx24u/1_year_of_cursor_switched_to_claude_best_decision/pcgnlh4/)
- Complaint, 2026-09-27, r/cursor (Reddit): “a few minutes/instant. switched to claude yesterday, fully operational on 3 very different projects.” [source](https://www.reddit.com/r/cursor/comments/1wq3m71/im_out/pc9w7tt/)
- Complaint, 2026-09-27, r/cursor (Reddit): “i don't feel similarly. much slower, way less accurate” [source](https://www.reddit.com/r/cursor/comments/1wrktgc/am_i_the_only_one_who_thinks_grok_47_is_actually/pcdnkab/)
- Complaint, 2026-09-26, @cursor_ai (X): “@cursor_ai can you guys fix fucking fix your cli, the startup time is absolutely horrendous just let me use your sub in grok build if its trouble” [source](https://twitter.com/1333285748859088896/status/2103891127564894325)

### Devin

- Praise, 2026-09-23, @DevinAI (X): “@guybedo @devinai @zeddotdev it's been pretty snappy for me” [source](https://twitter.com/896906084014845952/status/2102564705009275354)
- Praise, 2026-09-22, r/windsurf (Reddit): “it's completely free now, faster after few minutes of planning and runs in cloud seemlessly. i am running it almost 24x7 till it's free. my 20 dollar investment is working out now!” [source](https://www.reddit.com/r/windsurf/comments/1wd432p/swe2_first_experiences/pbaufyd/)
- Praise, 2026-09-22, @cognition (X): “@cognition it can really be said to be full of sincerity. i ran geekbench 7 on claude code on the web, grok bot, and devin web (ubuntu/macos/windows) respectively, using my own main device m2 max as a reference. devin web can completely match my own machine in zed compilation tests, and it can run 4 instances in parallel! even more astonishing is that the macos runner actually has a gpu! <strict_link>” [source](https://twitter.com/1649366440808681474/status/2102307587639144533)
- Complaint, 2026-09-27, @DevinAI (X): “@markfenner @devinai why devin take so much time while building?” [source](https://twitter.com/2278324309/status/2104309916702048679)
- Complaint, 2026-09-26, @cognition (X): “@dabit3 i must say swe-2 is still slow but its also really good, thank you for this @cognition” [source](https://twitter.com/1973083865607708673/status/2103901813158355257)
- Complaint, 2026-09-24, r/windsurf (Reddit): “curious, how well does swe2 works for you? it's very slow, for me. it takes minutes to do some tasks, which would take seconds for other models.” [source](https://www.reddit.com/r/windsurf/comments/1wn47sb/devin_free_daily_quota_is_now_gone_and_only_small/pbrlqwl/)

### OpenAI Codex

- Praise, 2026-09-27, r/codex (Reddit): “yes you will. it's very fast and easy for daybreak blue.” [source](https://www.reddit.com/r/codex/comments/1wqxiro/daybreak_issue/pca2ul8/)
- Praise, 2026-09-27, r/codex (Reddit): “it's a decent daily driver. sometimes not waiting 20min for the task to complete is the only thing between you and your task being done. 3.8 does fine for execution and medium complexity. it's cheap and very fast.” [source](https://www.reddit.com/r/codex/comments/1wqtd1g/openai_gpu_are_really_cooling_down_theo_just/pcbd6sl/)
- Praise, 2026-09-27, r/codex (Reddit): “claude models are also slower in general because their harness lack the web sockets connection that makes codex models so much faster as well as it generating far more reasoning tokens. op please update us once claude is done cooking so we have a baseline to compare both models usage in terms of actual work done.” [source](https://www.reddit.com/r/codex/comments/1wre9dq/codex_usage_vs_claude/pcbwouh/)
- Complaint, 2026-09-27, r/codex (Reddit): “are you joking man? codex is already so slow. it already is in slow mode ffs. we were asking for slow mode before when it was actually fast.” [source](https://www.reddit.com/r/codex/comments/1wr1olr/new_idea_codex_slow_mode/pc9tdu1/)
- Complaint, 2026-09-27, r/codex (Reddit): “is slow af, took 11min for a task that gemini3.8 could have done in a minute. its some simple as task” [source](https://www.reddit.com/r/codex/comments/1wravvu/sol_6_is_at_capacity/pcb5bkd/)
- Complaint, 2026-09-27, r/codex (Reddit): “i am getting significantly more done tbh. i used to spend an entire 5h period doing 1 task. now it's 2-3 tasks per 5h period. my only complaint is it seems much slower than before in terms of similar tickets getting done. so i used to be able to do 1 ticket in the 5h limit - and it would take 1h to burn it. now i can do 3 in the same overall capacity but those 3 will take the entire 5h and maybe more. but still only burn ~5% weekly each (which ma” [source](https://www.reddit.com/r/codex/comments/1wrb9p9/luna_6_isnt_as_bad_as_the_whiners_say/pcbc0x1/)

### GitHub Copilot

- Praise, 2026-09-24, r/GithubCopilot (Reddit): “haiku is still fast. for quick inline edits it is still the fastest model. that said, luna is close. maybe 6.0 luna will be even better? not sure, not enabled yet -.-” [source](https://www.reddit.com/r/GithubCopilot/comments/1wovttn/have_you_disabled_haiku/pbq7ox1/)
- Praise, 2026-09-24, r/GithubCopilot (Reddit): “speed is irrelevant to me, i always have 5-6 sessions opened at any point in time. if anything if they are slow it give some more room to breathe” [source](https://www.reddit.com/r/GithubCopilot/comments/1wovttn/have_you_disabled_haiku/pbqi9zy/)
- Praise, 2026-09-24, r/GithubCopilot (Reddit): “haiku is 10x the price of luna and way worse. put luna on low thinking if you just need an line edit, it will be equally fast. on openrouter, their recorded throughput is roughly the same.” [source](https://www.reddit.com/r/GithubCopilot/comments/1wovttn/have_you_disabled_haiku/pbqzq48/)
- Complaint, 2026-09-27, r/OpenAI (Reddit): “open your excel workbook, goto the data tab on your toolbar ribbon, click sort, pick the column you want to sort by, click ok. you're done. with a tiny bit of practice (like once or twice) you can do that faster before most folks can even issue a prompt to copilot to do it for you. and from previous experience i can tell you with 100% confidence, that even the most fumble fingered excel user can do all that before copilot will deign to provide a” [source](https://www.reddit.com/r/OpenAI/comments/1wr8r6g/gpt6luna_comes_out_on_top_in_puppy_kill_bench/pcgywmc/)
- Complaint, 2026-09-24, r/GithubCopilot (Reddit): “no my experience even with xhigh luna is slower than sonnet 5” [source](https://www.reddit.com/r/GithubCopilot/comments/1wnioqo/openais_gpt6_sol_and_gpt6_luna_now_available/pbukmkp/)
- Complaint, 2026-09-21, r/GithubCopilot (Reddit): “terrible and [slow.it](<strict_link>) cannot code or fix errors. it needs removed.” [source](https://www.reddit.com/r/GithubCopilot/comments/1uwkc1m/we_want_your_feedback_how_is_maicode1flash_in/pb8k385/)

### Amp

- Praise, 2026-09-23, @AmpCode (X): “@sqs @purefunctor @ampcode i’ve never opened amp as fast as i just did” [source](https://twitter.com/1443430538/status/2102751907387498733)
- Praise, 2026-09-23, @AmpCode (X): “@sqs @ampcode wow!! so quick! i'll give it a shot! you're amazing @sqs” [source](https://twitter.com/268615009/status/2102838747721277472)
- Praise, 2026-09-16, @AmpCode (X): “@iannuttall @ampcode 100 pages indexed already, that's fast” [source](https://twitter.com/1942068494989983744/status/2100350639935279381)
- Complaint, 2026-09-24, @AmpCode (X): “@sqs @ianlandsman @ampcode i will try tailscale as well, curious how that fits together. but it would be great if portals would be fast and usable. what causes the slowness here? it's one of the main things that's keeping me from moving everything to amp.” [source](https://twitter.com/1330980620617601026/status/2103028237299290550)
- Complaint, 2026-09-24, @AmpCode (X): “@sqs @ampcode some of the speed difference might be slowness in amp from routing through my chatgpt subscription” [source](https://twitter.com/132882990/status/2103057206040261029)
- Complaint, 2026-09-22, @AmpCode (X): “@thorstenball brother i started orb'n today and first of all thank you to the entire team of @ampcode and especially @sqs (i'm waiting for his next live stream, probably you should too) /feedback, the orb's terminal on mw plan seems too slow though” [source](https://twitter.com/1940807256377118724/status/2102466139293126912)

### Factory

- Praise, 2026-09-27, @FactoryAI (X): “the speed of deepseek v4.1 flash in @droid is mind blowing. makes me rethink the ollama subscription - when @factoryai has it all. <strict_link>” [source](https://twitter.com/1590702228234391552/status/2104263092364615842)
- Praise, 2026-09-27, @droid (X): “the speed of deepseek v4.1 flash in @droid is mind blowing. makes me rethink the ollama subscription - when @factory has it all. <strict_link>” [source](https://twitter.com/1590702228234391552/status/2104258455695741401)
- Praise, 2026-09-24, @FactoryAI (X): “@trevorbmurkp @harrystuck77 @factoryai @droid @badlogicgames @ampcode @cursor_ai @opencode @anthropicai @claudeai @cognition @xai @grok @build we’ve done a decent bit of perf improvements in the past few weeks with more incoming!” [source](https://twitter.com/1721143727043887104/status/2102916465167159772)
- Complaint, 2026-09-22, G2 (G2): “q: what problems is the product solving and how is that benefiting you? a: factory ai help me reduce manual coding work and save development time. it help with repetitive tasks, debugging and building features faster. i can focus more on important work instead of doing everything manually. it make my daily workflow more easy and productive. q: what do you like best about the product? a: factory ai is helpful for automating development work. it sa” [source](https://www.g2.com/products/factory-ai/reviews/factory-ai-review-13384158)
- Complaint, 2026-09-20, @FactoryAI (X): “unofficial finding of the day: gpt-6 astra on @factoryai is far more token-efficient than codex. love to @thsottiaux and the codex team — but routing astra through this harness makes usage noticeably slower. micro (read affordable) write-up coming. <strict_link>” [source](https://twitter.com/2093430933026148352/status/2101586928176873560)
- Complaint, 2026-09-08, @FactoryAI (X): “@droid @factoryai factory is full of features. makes it slow.” [source](https://twitter.com/1547527916401082368/status/2097332872663412781)

### Kiro

- Praise, 2026-09-11, r/kiroIDE (Reddit): “for what i have been using kiro-cli it returns responses much faster than claude code but i think that the way it perfoms and the customization layer is under what claude code can offer” [source](https://www.reddit.com/r/kiroIDE/comments/1wchsii/frontier_models_on_other_providers_vs_the_same/p9325ex/)
- Praise, 2026-09-05, r/kiroIDE (Reddit): “just downgrade to v 0.12 -doesnt have agent focus but feeels snappier” [source](https://www.reddit.com/r/kiroIDE/comments/1w5nh3m/bug_report/p7x19jn/)
- Complaint, 2026-09-27, r/kiroIDE (Reddit): “kiro-cli timed out a few times yesterday. i lost two heavy work sessions. all i found is a ticket so it appears it does not exist (<strict_link>) how do people check downtime? does kiro do rate limiting if you are doing heavy work? it felt slow.” [source](https://www.reddit.com/r/kiroIDE/comments/1wrsp5y/kiro_status_page/)
- Complaint, 2026-09-24, r/kiroIDE (Reddit): “if we can use gpt-5.6, that's still better. we can only use sonnet4.6 here. it's a completely delayed service and is unusable.” [source](https://www.reddit.com/r/kiroIDE/comments/1wp3xji/why_kiro_like_this/pbuv1e6/)
- Complaint, 2026-09-23, r/kiroIDE (Reddit): “man kiro is becoming less and less usable everyday. no fast mode, no latest models, no mobile use (in beta or opt in?)” [source](https://www.reddit.com/r/kiroIDE/comments/1wnep1f/give_us_opus_55_atleast/pbh4l3t/)

### Warp

- Praise, 2026-09-19, @warpdotdev (X): “@samueljmcd only because @warpdotdev is so snappy and i can use berkeley mono (font) from @usgraphics in it 🫶 <strict_link>” [source](https://twitter.com/1598005566701551617/status/2101350790611042323)
- Praise, 2026-09-11, @warpdotdev (X): “@warpdotdev fastest cli to add gets the daily driver spot” [source](https://twitter.com/1945115184072105984/status/2098448929515786421)
- Praise, 2026-09-08, @warpdotdev (X): “wow @warpdotdev with glm 5.3 is super cost efficient and fast. no brainer 💆♂️” [source](https://twitter.com/84324573/status/2097445164159483988)
- Complaint, 2026-09-13, @warpdotdev (X): “@warpdotdev here is what i am talking about: like bro... what are you "checking..." , "installing..." just let me connect. me: `ssh named_config_host&gt;` warp: <strict_link>” [source](https://twitter.com/1309409339824840704/status/2099090914744434972)
- Complaint, 2026-09-09, @warpdotdev (X): “i miss when @warpdotdev was incredible was excited about the opensourcing and the concept of 0z and everything but its diabolically bad, slow and laggy when it used to be truely blazingly fast might have to fork and rip out all the bs or just drop it” [source](https://twitter.com/2970558232/status/2097834207552618664)

### Grok Build

- Praise, 2026-09-05, r/codex (Reddit): “is anyone using astra/sol and grok 4.6 high via grok build? i haven’t received astra yet, but i do use sol. and as of late, i’ve almost entirely shifted to grok build (grok 4.6 high) - mostly because it’s way faster than gpt and very good at completing tasks end-to-end for my flutter project. when there is access to both llms, speed does get the veto from me, and i mostly use codex/gpt only as my general ai driver, and not for my flutter project” [source](https://www.reddit.com/r/codex/comments/1w7on0j/astra_is_absolutely_incredible/p7ys22g/)
- Complaint, 2026-09-22, r/cursor (Reddit): “i didn't find it very expensive when used in grok build, but it's very slow.” [source](https://www.reddit.com/r/cursor/comments/1wmswj9/grok_47_is_about_25_times_as_expensive_as_46/pbecymc/)
- Complaint, 2026-09-11, r/ClaudeCode (Reddit): “i use fable as my orchestrator, sonnet recon and headless grok build. will shift to sonnet/opus build once the heavy grok discount runs out. it is slower, a lot slower but token efficient and all the models add a layer of checking the others work. less baby sitting and cleaner code.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wd9eio/fable_as_orchestrator_and_opussonnet_as_executers/p9574gb/)

### Conductor

- Praise, 2026-09-15, @conductor_build (X): “my stack now is @conductor_build + qwen models for building apps. fast and free!” [source](https://twitter.com/1382193359641481217/status/2099810273439715525)
- Complaint, 2026-09-10, @conductor_build (X): “@kaden_hyatt @conductor_build when i'm stuck waiting i zone out 😮💨 and that passivity thing you mentioned hits close” [source](https://twitter.com/318021830/status/2098048462772142487)
