# Reliability and speed (`reliability`)

Feedback Bench, coding agents, built 2026-10-01, window 2026-08-31 to 2026-09-27. Web page: https://feedbackbench.com/#/criterion/reliability

Area of 4 criteria. Does it stay up and respond quickly?

Criteria: [Outages, server errors and capacity or rate errors](https://feedbackbench.com/criteria/rel.service_errors.md), [Latency, throughput and fast mode](https://feedbackbench.com/criteria/rel.response_speed.md), [Client crashes, freezes and failed tool execution](https://feedbackbench.com/criteria/rel.client_failures.md), [Updates break working setups](https://feedbackbench.com/criteria/rel.update_breakage.md)

Rated author-weeks, all agents: 7447. Complaint share: 81%.

## The brief

Written by Claude Opus 5.5 from 110 labelled posts and the numbers on this page. Interpretation, not measurement: every quote is verbatim and links to its post.

**Speed wins fans, but outages and bloated apps drive the complaints.**

TL;DR:

- Complaints dominate everywhere. Server errors and capacity failures stall sessions on Claude Code, OpenAI Codex and Cursor.
- OpenAI Codex draws the heaviest fire for laggy desktop apps, stream disconnects and at-capacity errors.
- Zed and Pi earn praise for instant launches. Fast Gemini Flash keeps Google Antigravity mixed.

In plain terms: Expect to lose time. Sessions drop mid-task with server errors, desktop apps eat memory and lag after updates, and some models take minutes to answer. The tools users rave about are the lightweight ones that launch instantly.

### How it breaks

- **Server errors take sessions down** ([Outages, server errors and capacity or rate errors](https://feedbackbench.com/criteria/rel.service_errors.md)). Server errors arrive in waves and kill whole sessions, sometimes on models the status page said were fine.
  Claude Code users report 529s and 500s spreading to mid-tier models while status notes blamed only high-end ones. Codex users describe stream disconnects that never recover. Forking the chat does not help, and only a fresh project restores work for a few hours. Cursor users see failures even on auto model routing and blame the server side. Codex posts dominate the requests for fewer model-at-capacity errors.
  Evidence:
  - Complaint, Claude Code, r/ClaudeCode, 2026-09-22: “not just the high end models like they said in status, sonnet is down too. got 500 internal errors” [source](https://www.reddit.com/r/ClaudeCode/comments/1wmv6m9/is_claude_down_it_says_claude_is_at_capacity/pba5r41/)
  - Complaint, Claude Code, r/ClaudeCode, 2026-09-22: “yup, some of my sessions just back online, then back to 529. time to do some dishes and clean up my desk, lol.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wmv77z/just_heard_my_cpu_fan_go_to_idle_which_is_weird/pba6tj5/)
  - Complaint, OpenAI Codex, r/codex, 2026-09-13: “i checked the megathread, but nothing seems to match my situation. this is probably the 4th or 5th time this has happened where i'll be just moving right along, everything it great, then it will start getting this : `stream disconnected before completion: error sending request for url` `(<strict_link>)` i see tons of posts about it with various vpns and proxys on github, but i'm not using any of those. everything is fine until a certain point in my project, then it just never is able to recover. forking the chat doesn't help. i have to create an entirely new project, then it will just happen again after 4-5 hours. very frustrating. <strict_link>” [source](https://www.reddit.com/r/codex/comments/1wfb01l/constant_stream_disconnected_before_completion/)
  - Complaint, Cursor, r/cursor, 2026-09-03: “i think the issue is on their server's end. even auto "model" isn't working properly.” [source](https://www.reddit.com/r/cursor/comments/1w67gs9/rate_limiting_by_model_provider/p7kp50t/)

- **Desktop apps leak memory and freeze** ([Client crashes, freezes and failed tool execution](https://feedbackbench.com/criteria/rel.client_failures.md)). Desktop clients balloon in memory and lag, and a full restart is often the only reliable fix users find.
  One Cursor user reports the app idling at most of a 16GB machine. Codex requests to fix laggy UI outnumber those for every other agent combined. Amp heats up older laptops. A machine restart is the common workaround. The contrast is telling. A Claude Code user says the terminal never froze across days-long sessions, and blames desktop app trouble on the lack of a clean restart.
  Evidence:
  - Complaint, Cursor, r/cursor, 2026-09-06: “my cursor app is at idle right now at total cosumed mem right now is 14.4gb out of 16gb.” [source](https://www.reddit.com/r/cursor/comments/1w8558g/cursor_eating_910_gb_of_ram_and_crashing_my_16gb/p8683jy/)
  - Complaint, Amp, @AmpCode, 2026-09-01: “@ampcode amp, you're heating up my laptop! :) but this is on an old laptop i bought in 2017, wonder if that has anything to do with it. <strict_link>” [source](https://twitter.com/12267342/status/2094726363777524022)
  - Praise, Claude Code, @ClaudeDevs, 2026-09-04: “@jackfriks @claudedevs terminal never froze on me even with sessions running for days. desktop app doing remote control for 2 weeks straight sounds like it just never got a clean restart” [source](https://twitter.com/2045518207159525376/status/2095983959528112183)
  - Praise, OpenAI Codex, r/codex, 2026-09-17: “had exactly the same. restarted my machine and instructed the agent to use low verbosity and now it seems back to normal” [source](https://www.reddit.com/r/codex/comments/1wimxpz/banked_reset_on_20x_pro_seems_to_restore_only_a/pac23g5/)

- **Slow models burn minutes per turn** ([Latency, throughput and fast mode](https://feedbackbench.com/criteria/rel.response_speed.md)). Faster responses are the top request in this area, and slow default models are the main reason.
  Devin's SWE-2 is repeatedly called slow, with tasks taking minutes that other models finish in seconds. Free-tier users ask for a paid path that performs. A Claude Code user argues that text streams quickly while the thinking lags, so claims of insane speed feel misleading. Speed also drifts week to week. Codex users notice good weeks and bad ones.
  Evidence:
  - Complaint, Devin, r/windsurf, 2026-09-24: “curious, how well does swe2 works for you? it's very slow, for me. it takes minutes to do some tasks, which would take seconds for other models.” [source](https://www.reddit.com/r/windsurf/comments/1wn47sb/devin_free_daily_quota_is_now_gone_and_only_small/pbrlqwl/)
  - Complaint, Devin, @cognition, 2026-09-12: “swe-2 seems ridiculously slow today. if the fact of being free is the problem, allow a payed version to perform well @cognition” [source](https://twitter.com/64041638/status/2098777825088086023)
  - Complaint, Claude Code, r/ClaudeCode, 2026-09-23: “i honestly don't get the hype. it seems slow for me, really slow. output is faster, as in the writing text to output, but the actually thinking\\processing is slow. which would make the usage window seem to last longer, but it's just trickery. i call bs on these posts claiming "crazy fast", "insane speeds" etc... the claudish gibberish has improved significantly though - so that's a big win. safeguards seem stricter, my sessions never get flagged, but i've had a few just today.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wo7gge/claude_opus_55_how_big_of_an_upgrade_is_it_really/pbkq8de/)
  - Praise, OpenAI Codex, r/codex, 2026-09-16: “yeah, the speed of models is much better this week. if i had to guess, we are getting a reset on tomorrow or friday morning.” [source](https://www.reddit.com/r/codex/comments/1whvec6/any_hope_for_reset/pa5lwio/)

- **Updates break working setups** ([Updates break working setups](https://feedbackbench.com/criteria/rel.update_breakage.md)). Updates regularly break setups that worked yesterday, and rollback is often the only fix.
  A Cline user who updates eagerly reports a release that broke scheduled jobs, leaving rollback or waiting as the only options. Another finds the Windows app broken, including the build from GitHub. Some Cline models halt and demand npm updates again and again. A Copilot user says an upgrade made a large session crash on every real task. Google Antigravity and OpenAI Codex lead requests for a reliable updater.
  Evidence:
  - Complaint, Cline, r/CLine, 2026-09-25: “man, i always update very quickly and fix any issue, however this update messed up all my cron jobs, i cannot fix it, the only option is to go back to 9.4 or wait to 9.7. according to github there are several tickets in progress with that issue. today i got the pain.” [source](https://www.reddit.com/r/CLine/comments/1wq8o1q/openclaw_202696_broke_all_my_cron_jobs/)
  - Complaint, GitHub Copilot, @GitHubCopilot, 2026-09-01: “@ryan_hecht @burkeholland @githubcopilot it happened after i upgraded, i have this one huge session locally that would crash every time i would try to do real work, i exported the session file and made a new session - side note with byok you need to specify— model instead of env var, 2 regressions” [source](https://twitter.com/55700354/status/2094603340449542597)
  - Complaint, Cline, @cline, 2026-09-26: “@cline your all models like gemini etc are stopping automatically after sometime of running and saying update again &amp; again via npm . please fix this @cline” [source](https://twitter.com/1311298190482702338/status/2103761055680012302)
  - Complaint, Cline, @cline, 2026-09-18: “@cline app is broken on windows, installed the lates version from github, that too is also broken <strict_link>” [source](https://twitter.com/1673435210/status/2101048258940329988)

- **Tool calls fail and stall agents** ([Client crashes, freezes and failed tool execution](https://feedbackbench.com/criteria/rel.client_failures.md), [Outages, server errors and capacity or rate errors](https://feedbackbench.com/criteria/rel.service_errors.md)). When tool execution breaks, the agent stops doing work entirely, and posts show this happening during backend changes.
  OpenCode users report every tool call failing at once, plus stuck sessions during what looked like an unannounced backend rollout. Google Antigravity's browser subagent fails to start on Windows because a driver download returns 404. A Copilot user finds assisted-permissions mode routes file reads through extra shell calls and runs slower than local mode. OpenCode also leads asks to stop spurious 429 rate-limit errors.
  Evidence:
  - Complaint, OpenCode, @opencode, 2026-09-16: “@opencode it's not working right now. all tool calls are failing :'(” [source](https://twitter.com/1603858236/status/2100254987754389596)
  - Complaint, Google Antigravity, r/google_antigravity, 2026-09-25: “hey everyone, i've encountered a blocker with the browser subagent in the latest version of antigravity on windows 10/11 (x64). whenever the agent triggers \`open\_browser\_url\` to inspect localhost or do visual validation, the browser environment fails to initialize completely due to a 404 on microsoft's playwright cdn. \### error trace: \`\`\`text failed to create browser context: failed to run playwright manager: failed to install playwright: could not install driver: could not install driver: error: got non 200 status code: 404 (404 not found) from [<strict_link>” [source](https://www.reddit.com/r/google_antigravity/comments/1wpzlyp/antigravity_browser_use_problem/)
  - Complaint, GitHub Copilot, r/GithubCopilot, 2026-09-10: “i saw the same thing. used astra for a few tasks to see what the hype is about. consumed about $12 on the 7 tasks with assisted permissions. i thought something was off and ran the same tasks on the local mode (wiped any memory etc) and it was down to about $8. i also noticed the time went down by roughly 40% in local, and for some reason, in assisted permissions it tends to do a lot of tool calls for tssks where it needs reading through multiple files, like calling powershell or node to read files, while in local mode it just says read filename.js i think they shipped it a bit early and for now it does not pose an advantage - half of the assisted tool calls were unable to be assessed.” [source](https://www.reddit.com/r/GithubCopilot/comments/1wazz5g/using_assisted_permissions_seem_to_increase_token/p8wf1a2/)
  - Complaint, OpenCode, r/opencode, 2026-09-19: “it seems that opencode is updating ther backend system atm. opencode zen is gone, no free models anymore, re-authentications. stucking seasons are also one of the current issues. maybe v2 is rolling out unannounced. or something else is going on.” [source](https://www.reddit.com/r/opencode/comments/1wkv4du/opencode_stopping_then_working_again_after/patqb2t/)

### Who stands out

- **OpenAI Codex (weaker)**. The most-discussed agent here also collects the most reliability grief, from app lag after updates to capacity errors.
  Codex tops nearly every request list in this area, including capacity errors, laggy UI, memory leaks, mid-task freezes and desktop stability. Users describe an update that dragged the whole machine down and sessions that disconnect after hours and never recover. Speed praise exists. Some users notice unusually fast inference in certain weeks, which makes the bad weeks more visible.
  Evidence:
  - Complaint, OpenAI Codex, X search: OpenAI Codex, Codex CLI, Codex app, 2026-09-01: “i just updated my @openai codex app and now it’s causing my computer to lag. cmon guys, what did you break in the app” [source](https://twitter.com/1306080915270103041/status/2094866409264239042)
  - Complaint, OpenAI Codex, r/codex, 2026-09-13: “i checked the megathread, but nothing seems to match my situation. this is probably the 4th or 5th time this has happened where i'll be just moving right along, everything it great, then it will start getting this : `stream disconnected before completion: error sending request for url` `(<strict_link>)` i see tons of posts about it with various vpns and proxys on github, but i'm not using any of those. everything is fine until a certain point in my project, then it just never is able to recover. forking the chat doesn't help. i have to create an entirely new project, then it will just happen again after 4-5 hours. very frustrating. <strict_link>” [source](https://www.reddit.com/r/codex/comments/1wfb01l/constant_stream_disconnected_before_completion/)
  - Complaint, OpenAI Codex, X search: OpenAI Codex, Codex CLI, Codex app, 2026-09-03: “this is odd, @openai codex is down too. some ppl are talking about claude being down as well (works for me now). three major labs being down at the same time is quite an interesting coincidence. <strict_link> <strict_link>” [source](https://twitter.com/1480598876176457728/status/2095526927830319538)
  - Praise, OpenAI Codex, r/codex, 2026-09-02: “i've been seeing super fast inference of models in codex, haven't seen this speed ever. i'm not using the ultrafast mode, so is it default like something as tibo said in berman's video” [source](https://www.reddit.com/r/codex/comments/1w5236b/ultrafast_the_new_default/)

- **Cursor (weaker)**. Fast mode earns real praise, but recurring model errors and idle memory bloat drag Cursor below its peers.
  Users report errors that recur all the time with specific models, failures even on auto routing, and memory consumption that nearly fills a machine at idle. Fast mode with top models draws enthusiastic posts. Some users also feel the composer model was slowed down, which erodes the speed advantage they came for.
  Evidence:
  - Complaint, Cursor, @cursor_ai, 2026-09-22: “@elonmusk whats wrong with @grok 4.7 in @cursor_ai ? this error happens all the time :( <strict_link>” [source](https://twitter.com/1329905429024022530/status/2102341930424451290)
  - Complaint, Cursor, r/cursor, 2026-09-06: “my cursor app is at idle right now at total cosumed mem right now is 14.4gb out of 16gb.” [source](https://www.reddit.com/r/cursor/comments/1w8558g/cursor_eating_910_gb_of_ram_and_crashing_my_16gb/p8683jy/)
  - Complaint, Cursor, r/cursor, 2026-09-03: “i think the issue is on their server's end. even auto "model" isn't working properly.” [source](https://www.reddit.com/r/cursor/comments/1w67gs9/rate_limiting_by_model_provider/p7kp50t/)
  - Praise, Cursor, @cursor_ai, 2026-09-23: “opus 5.5 max fast mode with multitask in @cursor_ai go brrrrrrrrrrrrrrrrrr. <strict_link>” [source](https://twitter.com/1624831101129703424/status/2102783009120567438)

- **Zed (stronger)**. Zed's praise rests on raw responsiveness. Users describe launches so fast the wait disappears.
  Launch speed and low memory use anchor the praise. Users moving from VS Code call it far faster and more responsive, even with AI extensions. Complaints are narrow and local, such as poor behaviour on a VDI, a CLI conflict that triggers crash logs and keybinding bugs. Outages barely feature.
  Evidence:
  - Praise, Zed, @zeddotdev, 2026-08-31: “@zeddotdev that half-bounce hit me. zed comes up so fast it feels like the dock barely noticed. i used to wait through the bounce, then wait again. this is the first editor in years that made launching feel like flipping a light. keep that obsession. it shows.” [source](https://twitter.com/1614212720806486017/status/2094236952019317109)
  - Praise, Zed, @zeddotdev, 2026-09-12: “switched to @zeddotdev! i just fcking love how fast it is - no bloated ui - no unnecessary buttons - no unnecessary memory usage there's no going back now baby 🚀 <strict_link>” [source](https://twitter.com/1856340918250467328/status/2098667784301601049)
  - Praise, Zed, r/ZedEditor, 2026-09-16: “how can a text editor be reliable or unreliable? we just write stuff and press ctrl+s. only reason to use an ide is for the intellisense and autocomplete, and zed is 100x faster than vscode for that. also you can turn off ai slop on zed but not on vscode” [source](https://www.reddit.com/r/ZedEditor/comments/1whv207/i_love_zed_but/pa9ex5w/)
  - Complaint, Zed, @zeddotdev, 2026-09-18: “@christianlempa @zeddotdev zed is good but dont ever use on a vdi” [source](https://twitter.com/1954558913346449408/status/2100917572350677476)

- **Pi (stronger)**. Pi wins on instant start and lean streaming, and runs well even on aging hardware.
  Users praise a client with almost no wait. It starts instantly, input is instant and output streams without decorative animation. One user reports it matching Claude Code CLI speed on a 13-year-old laptop. The friction sits elsewhere. Long agent runs leave users idle, and some third-party extensions fail to work.
  Evidence:
  - Praise, Pi, @pidotdev, 2026-09-08: “"pi starts instantly. input is instant. output streams, done. there’s no animation layer trying to make the wait feel nicer, because there is barely a wait to begin with." @pidotdev - an article written for you. <strict_link>” [source](https://twitter.com/1351335202757423104/status/2097345545564041718)
  - Praise, Pi, r/PiCodingAgent, 2026-09-03: “re bloat with 3 points of interest 1- it don’t merely add modules, it rewrote systems to include that mess from above. 2 from my perspective it’s running as fast as claude code cli 3 it’s on a 13yr old i5 laptop with 16gb of ram. so if it’s bloated on this hunk of junk machine, i can’t tell. oh and it rewrote its core to maintain persistence” [source](https://www.reddit.com/r/PiCodingAgent/comments/1w1wj1l/ive_stopped_using_claude_code/p7ns4nv/)
  - Praise, Pi, r/PiCodingAgent, 2026-09-26: “nice work. i tried using it. it is fast. i tried using few extensions but they don’t seem to work with pig. mcp-adapter, pi-web-access, pi-rtk-optimizer” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wqe2v1/meet_pig_the_pi_coding_harness_that_is_yours_but/pc3o4sz/)
  - Complaint, Pi, r/PiCodingAgent, 2026-09-10: “whenever pi was chewing on something for a minute, i'd grab my phone and come back 20 minutes later to a response that had just been sitting there. so i wrote pi-play: /play opens a game as an overlay right on top of the conversation, and it listens for agent\_settled. the second pi finishes or is waiting on you, the game stops immediately. no alt-tab, no phone. 5 games so far: snake, tetris, 2048, minesweeper, sudoku (6 is on the way, i am implementing chrome dino game rn) to install: `pi install npm:@terminalika/pi-play`” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wcx1w3/piplay_i_made_a_pi_plugin_to_play_simple_games/)

- **Google Antigravity (mixed)**. Gemini Flash speed draws superlatives, while other users wait most of a minute for a greeting.
  Users call Gemini Flash in Antigravity several times faster than rivals, with throughput they describe as redefining fast. Others report a simple hello taking most of a minute, minutes of waiting just to see a quota message, and multi-hour edits on small files. Antigravity users ask for faster responses nearly as often as Codex users, and lead requests for a reliable updater.
  Evidence:
  - Complaint, Google Antigravity, @antigravity, 2026-09-14: “@cleverlypaul @kyuuurius @geminiapp @antigravity a simple hello taked 47seconds” [source](https://twitter.com/954049869521580034/status/2099547763608371469)
  - Praise, Google Antigravity, r/codex, 2026-09-04: “test gemini 3.8 flash on google antigravity, your definition of "fast" will change after testing this thing. gpt 5.6 luna goes at 120 tokens per second and is faster than astra, gemini 3.8 flash goes at 1200 tokens per second... if you want an even greater impression of speed, choose gemini 3.7 flash or 3.6 flash which use much fewer tokens. it's free and very impressive.” [source](https://www.reddit.com/r/codex/comments/1w7gy48/astra_is_fast/p7vbxs3/)
  - Complaint, Google Antigravity, r/google_antigravity, 2026-09-23: “how slow for you guys? here it takes like 5 minutes to just answer i ran out of quota.” [source](https://www.reddit.com/r/google_antigravity/comments/1wmxyjt/yes_you_found_the_post_again_38_is_so_slow_now/pbhniav/)
  - Complaint, Google Antigravity, @antigravity, 2026-09-15: “@kyuuurius @geminiapp @antigravity took almost 2 hours to edit and redesign a 2 pages pdf...” [source](https://twitter.com/1787288502310146048/status/2099674264232251414)

### Fine print

- Posts cover about one month, so a single outage burst can swing an agent's reading in this area.
- Speed praise often credits the underlying model rather than the agent itself.
- Factory, Kiro, Warp, Conductor, Grok Build and Augment Code have too few posts to judge here.

## Top requests

What users ask to add or change, most asked first. 904 author-weeks ask for something. Requests do not change the Feedback Score. Rule: A separate pass by Claude Sonnet 5 reads every counted post and extracts what the author asks the agent or its vendor to add or change, with the criteria it maps to and a short normalised wording; it does not touch the labels or the Feedback Score. Claude Opus 5.5 groups the wordings within each criterion (the first criterion the request maps to) into themes; code counts them. A theme counts distinct author-weeks that ask for it, per agent; across agents, one author-week per agent. Themes asked in fewer than 2 author-weeks, and requests that share no theme, are not shown. Examples: up to 3 posts per theme from different authors, without slurs, preferring posts of 60 to 450 characters, most recent first.

| Rank | Request | Criterion | Author-weeks | Posts | Agents (author-weeks) |
|---|---|---|---|---|---|
| 1 | Faster model response speed | [Latency, throughput and fast mode](https://feedbackbench.com/criteria/rel.response_speed.md) | 76 | 76 | OpenAI Codex 20, Google Antigravity 19, Claude Code 12, OpenCode 11, Cursor 6, Devin 3, Cline 2, Pi 2, Kiro 1 |
| 2 | Stabilize and fix the desktop app | [Client crashes, freezes and failed tool execution](https://feedbackbench.com/criteria/rel.client_failures.md) | 29 | 30 | OpenAI Codex 10, OpenCode 6, Cline 4, Google Antigravity 2, Claude Code 2, Cursor 2, Devin 2, Factory 1 |
| 3 | Fix memory leaks and reduce RAM usage | [Client crashes, freezes and failed tool execution](https://feedbackbench.com/criteria/rel.client_failures.md) | 28 | 31 | OpenAI Codex 12, Google Antigravity 3, Claude Code 3, Cursor 3, Devin 3, OpenCode 2, Conductor 1, Zed 1 |
| 4 | Fix ongoing service outages | [Outages, server errors and capacity or rate errors](https://feedbackbench.com/criteria/rel.service_errors.md) | 28 | 29 | OpenAI Codex 9, Claude Code 6, Cursor 6, Google Antigravity 4, OpenCode 2, Pi 1 |
| 5 | Fix app hanging and freezing mid-task | [Client crashes, freezes and failed tool execution](https://feedbackbench.com/criteria/rel.client_failures.md) | 26 | 32 | OpenAI Codex 11, Cursor 6, Claude Code 2, Devin 2, Pi 2, Google Antigravity 1, OpenCode 1, Warp 1 |
| 6 | Fix laggy UI and slow app performance | [Client crashes, freezes and failed tool execution](https://feedbackbench.com/criteria/rel.client_failures.md) | 24 | 25 | OpenAI Codex 17, Cursor 4, Google Antigravity 1, Claude Code 1, Devin 1 |
| 7 | Fewer model-at-capacity errors | [Outages, server errors and capacity or rate errors](https://feedbackbench.com/criteria/rel.service_errors.md) | 22 | 24 | OpenAI Codex 19, Google Antigravity 2, Devin 1 |
| 8 | Stop spurious 429 rate limit errors | [Outages, server errors and capacity or rate errors](https://feedbackbench.com/criteria/rel.service_errors.md) | 19 | 20 | OpenCode 11, OpenAI Codex 5, Google Antigravity 1, Devin 1, Kiro 1 |
| 9 | Fix frequent app and CLI crashes | [Client crashes, freezes and failed tool execution](https://feedbackbench.com/criteria/rel.client_failures.md) | 18 | 18 | OpenAI Codex 6, Google Antigravity 3, Cline 3, Cursor 3, OpenCode 2, Devin 1 |
| 10 | Reliable update and installer process | [Updates break working setups](https://feedbackbench.com/criteria/rel.update_breakage.md) | 18 | 18 | Google Antigravity 6, OpenAI Codex 6, Cursor 4, Claude Code 2 |
| 11 | Lower CPU, GPU and battery usage | [Client crashes, freezes and failed tool execution](https://feedbackbench.com/criteria/rel.client_failures.md) | 17 | 17 | OpenAI Codex 8, Google Antigravity 2, Amp 1, Claude Code 1, Cursor 1, Factory 1, OpenCode 1, Warp 1, Zed 1 |
| 12 | Fix failed and malformed tool call execution | [Client crashes, freezes and failed tool execution](https://feedbackbench.com/criteria/rel.client_failures.md) | 16 | 16 | OpenAI Codex 6, OpenCode 5, Amp 1, Google Antigravity 1, Claude Code 1, Cursor 1, Warp 1 |

### 1. Faster model response speed

- Claude Code, 2026-09-27, r/ClaudeCode (Reddit): “so here is also where i am stuck. i am limited with these, the faster i want to work, either ai needs to speed up with its results instead of letting me wait for so long. but even after it instantly returns the correct way, it will be limited by your own capacity. so the other thing is to fully trust the ai to make decisions, and here is the difficult part. till what extend can you do this? without losing track of how it works under the hood.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wradi9/why_normal_discussions_has_so_many_haters/pcg0zse/)
- Google Antigravity, 2026-09-26, r/google_antigravity (Reddit): “that's a lot of text, so i just wanna ask first. is gemini flash 3.8's speed fixed? it's so slow it should be changed to gemini slow 3.8.” [source](https://www.reddit.com/r/google_antigravity/comments/1wqbyzk/antigravity_cli_release_v127_v1211/pc333k5/)
- Cline, 2026-09-26, @cline (X): “@cline this model very super slow, tell the developer of this model 🗿, fix that speed” [source](https://twitter.com/1883868794000695296/status/2103703396515483978)

### 2. Stabilize and fix the desktop app

- OpenAI Codex, 2026-09-27, r/ClaudeCode (Reddit): “claude desktop is crap. i wish they took care of their desktop app like openai with codex” [source](https://www.reddit.com/r/ClaudeCode/comments/1wrhtyd/new_to_claude_why_should_we_use_claude_code_over/pcckr78/)
- OpenAI Codex, 2026-09-26, r/codex (Reddit): “same problem after updating today! i'm so frustrated i thought i was the only one. escalated to openai support but i really hope the team notices this soon because now i can't work. well i can use cli but i want my desktop app man... version <phone_number>.0, win 11 25h2 <strict_link>” [source](https://www.reddit.com/r/codex/comments/1wqfiiw/chatgpt_desktop_app_gets_stuck_loading_forever/pc4d5tr/)
- OpenAI Codex, 2026-09-21, X search: OpenAI Codex, Codex CLI, Codex app (X): “@tokengremlin can't wait for this one! maybe bel help them to fix codex app ;)” [source](https://twitter.com/1674056173790679042/status/2102064065614844220)

### 3. Fix memory leaks and reduce RAM usage

- Cursor, 2026-09-26, r/cursor (Reddit): “cursor is highly unoptimised. team needs to now work on optimisation too. it just keeps on filling ram and ssd on its own and then ooms or crashes.” [source](https://www.reddit.com/r/cursor/comments/1wqppuh/grok_is_shutting_down_apps_now/pc5yr3x/)
- Google Antigravity, 2026-09-24, @antigravity (X): “@probiex007 @antigravity @probiex007 @antigravity yeah bro, for 28 gb it better write the whole codebase itself 😅 skipping till they patch this” [source](https://twitter.com/1728403486373789696/status/2103221298763837909)
- Google Antigravity, 2026-09-24, @antigravity (X): “@antigravity pls work on memory management. ide is consuming alot of memory idk why 🙁 <strict_link>” [source](https://twitter.com/1195014217771864064/status/2103213000517939705)

### 4. Fix ongoing service outages

- OpenCode, 2026-09-26, @opencode (X): “@opencode yo intern, this shiii over 24hr now. bls, kindly resolve this. <strict_link>” [source](https://twitter.com/2051959960230088704/status/2103938207142252635)
- OpenAI Codex, 2026-09-25, r/codex (Reddit): “they need to install ssds on their servers, it is insane how long it took them to restart” [source](https://www.reddit.com/r/codex/comments/1wqbnpj/we_are_back/pc2qpla/)
- OpenAI Codex, 2026-09-25, r/codex (Reddit): “i need you nowwww restore now your aws and azure servers openai :dddd” [source](https://www.reddit.com/r/codex/comments/1wqadua/codex_down/pc2gp50/)

### 5. Fix app hanging and freezing mid-task

- OpenAI Codex, 2026-09-27, X search: OpenAI Codex, Codex CLI, Codex app (X): “@openai codex keeps hanging on mac 27.2 beta during context compaction based on my observation. i have already restarted frozen codex 2 dozen times today.” [source](https://twitter.com/304497770/status/2104052196916768981)
- Claude Code, 2026-09-25, @ClaudeDevs (X): “@claudedevs still lags and freezes my machine. cli is better” [source](https://twitter.com/44595437/status/2103303119799210260)
- Cursor, 2026-09-24, @cursor_ai (X): “i'm working in @cursor_ai sometimes running 6 threads at a time. i have to restart the app atleast 10 times a day because the threads are in some infinite loading state also it logs me out when i quit the app thankfully the tasks restart where they left out, but i wish i didn't have to restart the app 10x a day” [source](https://twitter.com/993923490746073088/status/2103104373324709893)

### 6. Fix laggy UI and slow app performance

- OpenAI Codex, 2026-09-27, X search: OpenAI Codex, Codex CLI, Codex app (X): “@thsottiaux @jorilallo the codex app is so laggy now please fix” [source](https://twitter.com/2026063233543462913/status/2104057177396727860)
- OpenAI Codex, 2026-09-27, X search: OpenAI Codex, Codex CLI, Codex app (X): “my codex app on linux has been trying to load chats for at least 15 mins when before it was just a few seconds, still hasn't even loaded smaller chats where there aren't lots of messages. @thsottiaux” [source](https://twitter.com/1674075739895877632/status/2104023351060566107)
- OpenAI Codex, 2026-09-26, r/codex (Reddit): “and with the latest update with the vertical left sidebar, on my m3 imac with 24 gb memory just panning the mouse over the profile menu at the bottom left it is laggy as you mouse over each menu item...” [source](https://www.reddit.com/r/codex/comments/1wqmgm3/what_an_absolute_chonker/pc5o053/)

### 7. Fewer model-at-capacity errors

- Devin, 2026-09-25, r/windsurf (Reddit): “client error: protocol error (unimplemented): we are currently experiencing capacity issues with this serving model. please switch to a different model or try again later. (trace id: <structured_id>) unimplemented? and "with this serving model"? surely it should be "with serving this model". my idea - make the error messages better and fix the "protocol error" bug. more capacity would be nice too of course 😄” [source](https://www.reddit.com/r/windsurf/comments/1wpw8do/error_messages_are_not_their_best_strength/)
- Google Antigravity, 2026-09-24, r/google_antigravity (Reddit): “just experiencing the issue today and it's infuriating at this point, high token usage when it works for simple prompts and most of the time it just says "our servers are experiencing high traffic right now, please try again in a minute." any update on this issue would be great” [source](https://www.reddit.com/r/google_antigravity/comments/1wh8h5u/psa_gemini_38_flash_slowerrors/pbsbuq4/)
- OpenAI Codex, 2026-09-18, r/codex (Reddit): “i am pretty close to refunding my 20x if they don't stop giving selected model is at capacity, ffs i spinned up $20 claude to do the work and haven't hit the 5 hour limit yet” [source](https://www.reddit.com/r/codex/comments/1wj6tl0/in_case_youre_wondering_this_is_what_100_plan/pak3ir7/)

### 8. Stop spurious 429 rate limit errors

- OpenAI Codex, 2026-09-26, r/codex (Reddit): “i have the same error 401 and i couldn't work, and the chatbot said that there were too many requests, because of codex, i literally couldn't work. so i went with claudecode. did anyone else from latam experience the same on september 23-24, 2026?” [source](https://www.reddit.com/r/codex/comments/1wqnejx/openai_codex_offline_again/pc5ojjn/)
- OpenCode, 2026-09-21, r/opencode (Reddit): “command code does but only $20 and the connection is extremely flaky 429 errors all the time” [source](https://www.reddit.com/r/opencode/comments/1wmhlzq/why_doesnt_opencode_go_offer_max_on_muse_13_given/)
- OpenCode, 2026-09-18, @opencode (X): “wth is this @opencode error from provider (console go): upstream request failed: [rate_limit_exceeded] rate limit exceeded. please retry after a brief wait. i only used 2% of my 5-hour limit but get rate limits??” [source](https://twitter.com/1250026646058516481/status/2101001690119876845)

### 9. Fix frequent app and CLI crashes

- Cursor, 2026-09-23, @cursor_ai (X): “@cursor_ai chill on the rollouts, the files tab has been crashing cursor full-screen for a month <strict_link>” [source](https://twitter.com/2464774904/status/2102872782103302216)
- OpenCode, 2026-09-21, r/opencode (Reddit): “i was excited trying out mimo but this is a disaster for me. every time it's using glob or grep, it crashes or freezes and when it does i have a huge bump in context, from 100k to 300k just like that, without explanation. using opencode tui, haven't tested on other harness yet. what's your experience so far?” [source](https://www.reddit.com/r/opencode/comments/1wmsl2p/having_a_bad_experience_with_mimo_26_flash/)
- Cline, 2026-09-19, @cline (X): “@cline cline 3.0.62 cli is dead on arrival on apple silicon. the bundled darwin-arm64 binary has a broken signature, so the kernel sigkills it at launch. both `cline --version` and `cline --help` print nothing and exit 137. fresh `npm install -g cline` on macos 27, arm64, node 26.7.0.” [source](https://twitter.com/1937709229198172160/status/2101359440402518298)

### 10. Reliable update and installer process

- Google Antigravity, 2026-09-26, @antigravity (X): “@antigravity i’m using linux, and i constantly have to manually download and reinstall antigravity just to update it. the “check for updates” button simply doesn’t work. it shouldn't be too difficult to fix. i don't understand why this is being ignored.” [source](https://twitter.com/2031742561392701440/status/2103723644216287683)
- Google Antigravity, 2026-09-26, @antigravity (X): “@rodydavis @antigravity i run antigravity ide, installing the extension google.google-antigravity in antigravity ide doesn't sound like a good idea. antigravity ide's last update was 2.5.5 on august 13, 2026 did the update system get broked when it was renamed from antigravity to antigravity ide? <strict_link>” [source](https://twitter.com/17038251/status/2103671410325639448)
- Cursor, 2026-09-25, @cursor_ai (X): “@cursor_ai please fix the broken update, getting too annoying <strict_link>” [source](https://twitter.com/28009458/status/2103333764927762902)

### 11. Lower CPU, GPU and battery usage

- Zed, 2026-09-25, @zeddotdev (X): “@zeddotdev zed on an old linux laptop is great, but damn the gpu usage shoots up and the fans go full throttle like a jet taking off” [source](https://twitter.com/58782925/status/2103314417576473007)
- Google Antigravity, 2026-09-24, @antigravity (X): “@androidstudio @antigravity this is great, next i wish as didn’t use 700% of my cpu while building or reopening the project . devin / cursor do not have this issue” [source](https://twitter.com/380573269/status/2103194074811629948)
- OpenAI Codex, 2026-09-23, r/codex (Reddit): “i am running debian 13 and vs codium. it does not happens randomly. i am killing the process everytime it reaches 19% cpu. even idle. updated today and same” [source](https://www.reddit.com/r/codex/comments/1suvg6s/has_anyone_noticed_codex_in_vs_code_using_high/pbgy7oh/)

### 12. Fix failed and malformed tool call execution

- OpenAI Codex, 2026-09-25, r/codex (Reddit): “or they could use all that funding to build a decent harness so the user isn't left burning tokens for no reason... somehow claude has the brains to make sure their harness works correctly before a major release. i'm sure astra is great, but codex is pure garbage. damage is done. go look at tibo's x comments lol” [source](https://www.reddit.com/r/codex/comments/1w7x57n/before_blaming_gpt6_astra_read_its_prompting_guide/pbzds5e/)
- OpenCode, 2026-09-22, r/opencode (Reddit): “agreed, i have a bad experience on using the free mimo 2.6 flash. it does a lot of failed tool calling, and sometimes it just kept looping. looks like a harness issue, but other models work fine” [source](https://www.reddit.com/r/opencode/comments/1wmsl2p/having_a_bad_experience_with_mimo_26_flash/pbb0h87/)
- OpenCode, 2026-09-22, @opencode (X): “@opencode fix your tool calls... the idea behind your service is amazing, but harness often gets crazy with the tool calls... mimo just crashed my pc by doing 800+ searches on my pc in a row... then after the restart i've asked it not to do that and it did exactly the same thing... :d” [source](https://twitter.com/996858699812626432/status/2102271295819858277)

## Every agent

| Agent | Overall rank | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|---|
| [Zed](https://feedbackbench.com/agents/zed.md) | =8 | Better than peers | 0.666 | 0.627–0.702 | 112 | 61 | 51 |
| [Pi](https://feedbackbench.com/agents/pi.md) | 7 | Better than peers | 0.592 | 0.549–0.631 | 119 | 43 | 76 |
| [Cline](https://feedbackbench.com/agents/cline.md) | =8 | Better than peers | 0.573 | 0.530–0.612 | 134 | 41 | 93 |
| [Claude Code](https://feedbackbench.com/agents/claude-code.md) | 1 | Better than peers | 0.550 | 0.521–0.576 | 946 | 219 | 727 |
| [GitHub Copilot](https://feedbackbench.com/agents/copilot.md) | =8 | Typical | 0.539 | 0.498–0.578 | 68 | 20 | 48 |
| [Google Antigravity](https://feedbackbench.com/agents/antigravity.md) | =5 | Better than peers | 0.537 | 0.505–0.567 | 906 | 199 | 707 |
| [OpenCode](https://feedbackbench.com/agents/opencode.md) | 3 | Better than peers | 0.533 | 0.506–0.559 | 1251 | 267 | 984 |
| [Devin](https://feedbackbench.com/agents/devin.md) | =5 | Typical | 0.533 | 0.487–0.574 | 149 | 35 | 114 |
| [Amp](https://feedbackbench.com/agents/amp.md) | =11 | Typical | 0.490 | 0.452–0.526 | 74 | 12 | 62 |
| [Cursor](https://feedbackbench.com/agents/cursor.md) | 4 | Worse than peers | 0.446 | 0.409–0.480 | 690 | 102 | 588 |
| [OpenAI Codex](https://feedbackbench.com/agents/codex.md) | 2 | Worse than peers | 0.407 | 0.387–0.426 | 2918 | 398 | 2520 |
| [Factory](https://feedbackbench.com/agents/factory.md) | =11 | Too few posts | – | – | 23 | 7 | 16 |
| [Kiro](https://feedbackbench.com/agents/kiro.md) | 13 | Too few posts | – | – | 23 | 3 | 20 |
| [Warp](https://feedbackbench.com/agents/warp.md) | 15 | Too few posts | – | – | 17 | 5 | 12 |
| [Conductor](https://feedbackbench.com/agents/conductor.md) | 14 | Too few posts | – | – | 14 | 1 | 13 |
| [Grok Build](https://feedbackbench.com/agents/grok-build.md) | 16 | Too few posts | – | – | 3 | 1 | 2 |
| [Augment Code](https://feedbackbench.com/agents/augment.md) | 17 | Too few posts | – | – | 0 | 0 | 0 |

## Posts

Receipts rule: The 5 most recent praise and complaint posts per area (first 700 characters) and 3 per criterion (first 450 characters).

### Zed

- Praise, 2026-09-27, r/ZedEditor (Reddit): “as long as it’s fast, responsive and no memory bloat, i’ll keep using it - biggest reason why i moved away from vs code. does zed have some kind of task manager or something? would love to see the effect of every extension i install haha.” [source](https://www.reddit.com/r/ZedEditor/comments/1wris6u/pls_dont_turn_into_vs_code/pcekb3x/)
- Praise, 2026-09-27, @zeddotdev (X): “@0xprajwal_ i use @zeddotdev, btw. it’s fast, lightweight, consumes less memory, and uses the gpu for rendering.” [source](https://twitter.com/1542127649463898113/status/2104217257153110457)
- Praise, 2026-09-26, @zeddotdev (X): “@zeddotdev yk what zed? i moved from vscode to zed quite a while ago and it's fucking amazing it starts faster looks better lets me disable ai features bullshit completely honestly, thanks for making my life a bit more better” [source](https://twitter.com/1188417649203539968/status/2103694133131202673)
- Praise, 2026-09-26, @zeddotdev (X): “@theo i use @zeddotdev for that, it's fast” [source](https://twitter.com/445781460/status/2103714899180269714)
- Praise, 2026-09-26, @zeddotdev (X): “i tried @zeddotdev again today — clean, speedy, and refreshing. ides now feel like cars: zed = sports car — fast, sleek, fun to drive vs code = suv — versatile, customizable, but heavier jetbrains = luxury sedan — packed with features, less nimble” [source](https://twitter.com/2036660207452028928/status/2103787211770712336)
- Complaint, 2026-09-27, r/ZedEditor (Reddit): “you can disable them but still zed is slower than gram” [source](https://www.reddit.com/r/ZedEditor/comments/1wris6u/pls_dont_turn_into_vs_code/pce43ch/)
- Complaint, 2026-09-27, @zeddotdev (X): “opened @zeddotdev after a while what is this? only happens wiht opencode acp <strict_link>” [source](https://twitter.com/851365565201514498/status/2104122682895921447)
- Complaint, 2026-09-26, r/ZedEditor (Reddit): “tried, but not reliable for everyday use. got some weird issues, saying it can't find neovim, but it works fine in other terminals.” [source](https://www.reddit.com/r/ZedEditor/comments/1wpr20d/is_there_any_zed_terminal/pc539bm/)
- Complaint, 2026-09-26, r/ZedEditor (Reddit): “there is always a node process running when zed is open. i would like to set it up to use bun. i don't even have node installed.” [source](https://www.reddit.com/r/ZedEditor/comments/1wosnp7/zed_still_silently_downloads_binaries_after_two/pc96lwi/)
- Complaint, 2026-09-25, r/ZedEditor (Reddit): “yeah, it's great but not so performant; also, i face some random issues with wsl.” [source](https://www.reddit.com/r/ZedEditor/comments/1wpr20d/is_there_any_zed_terminal/pbxz2np/)

### Pi

- Praise, 2026-09-27, r/PiCodingAgent (Reddit): “yeah. on all my tests accuracy was quite shit, and jev was consistently outperformed by glm 5.3 flash, for anything requiring decision making. but hey, it's fast 🤷♂️” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wnl7da/pi_can_now_use_jev_and_more/pce99cd/)
- Praise, 2026-09-27, r/PiCodingAgent (Reddit): “pretty good, these days my pi is stock plus one extension showing context usage in detail and one custom footer override to apply custom styling. i do maintain one patch for the llama.cpp provider to enable per model and session thinking levels. performance is fast but i use a dual radeon r9700 setup. i pretty much work offline and with no delays.” [source](https://www.reddit.com/r/PiCodingAgent/comments/1vjurqz/considering_claude_code_pi_worth_it/pch2ftv/)
- Praise, 2026-09-26, r/PiCodingAgent (Reddit): “nice work. i tried using it. it is fast. i tried using few extensions but they don’t seem to work with pig. mcp-adapter, pi-web-access, pi-rtk-optimizer” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wqe2v1/meet_pig_the_pi_coding_harness_that_is_yours_but/pc3o4sz/)
- Praise, 2026-09-26, r/PiCodingAgent (Reddit): “that's cool, i'm using pi over bun rn for better performance, i think imma try it” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wqe2v1/meet_pig_the_pi_coding_harness_that_is_yours_but/pc3z3hf/)
- Praise, 2026-09-26, r/PiCodingAgent (Reddit): “damnn it's really fast, lemme take a look for a few days” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wqe2v1/meet_pig_the_pi_coding_harness_that_is_yours_but/pc5h7iz/)
- Complaint, 2026-09-26, r/PiCodingAgent (Reddit): “oof thank you so much definitely was a regression.... hot fix coming. messed up a merge conflict resolve on launch 🤡” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wqe2v1/meet_pig_the_pi_coding_harness_that_is_yours_but/pc3qaf0/)
- Complaint, 2026-09-26, r/PiCodingAgent (Reddit): “every time i turn on my mac and see so many node processes, it’s really terrible. i think i‘ll try.” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wqe2v1/meet_pig_the_pi_coding_harness_that_is_yours_but/pc4svol/)
- Complaint, 2026-09-26, r/PiCodingAgent (Reddit): “would be nice if it worked. reality is that it doesn't tried it with stock settings coloring is bugged out, stock \`read\` tool is failing with jsonschema validation so it's just yet another sloppily coded port from one language to another with no real effort put into the maintenance apart from tons of tokens from the company leeched” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wqe2v1/meet_pig_the_pi_coding_harness_that_is_yours_but/pc9bfwt/)
- Complaint, 2026-09-26, r/PiCodingAgent (Reddit): “i'm not the claiming that i've been working on something for a year where the very foundational tool that harness needs to be able to work fails every query but at least it's faster to start, amirite and no, i'm not gonna contribute to your hardfork if you don't bother with testing the very basics” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wqe2v1/meet_pig_the_pi_coding_harness_that_is_yours_but/pc9duyq/)
- Complaint, 2026-09-26, r/PiCodingAgent (Reddit): “[<strict_link> i mean you really don't deserve anyone helping you, but alas. i have a few tokens to burn [<strict_link> you don't have that handling anywhere, and you remove the existing test suite just so that your code succeeds. bravo. keep it up, good luck any file that is attempted to be read by an llm in full gets a failure and it's very consistent on every turn with astra.” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wqe2v1/meet_pig_the_pi_coding_harness_that_is_yours_but/pc9gdot/)

### Cline

- Praise, 2026-09-27, @cline (X): “@cline love the speed in the new desktop app” [source](https://twitter.com/1640928038954336258/status/2104336348568342539)
- Praise, 2026-09-26, @cline (X): “@cline really good! after integrating cline for free, it crushes the next.js tasks of kimi k3—pixel canary action is really fast.” [source](https://twitter.com/1744672948135321600/status/2103663198188749016)
- Praise, 2026-09-26, @cline (X): “@cline that speed boost in the new cline desktop app is so satisfying to use” [source](https://twitter.com/1889631970667405317/status/2103747460858302559)
- Praise, 2026-09-26, @cline (X): “space bunny alpha is fast and unlimited on @cline cloud; i've been running it for long sessions, even when asleep, for a little side project. btw, thanks for the early cline cloud invite. sorry, i may have overused the free limit 💀👀 <strict_link>” [source](https://twitter.com/834280176313835520/status/2103763278954738038)
- Praise, 2026-09-26, @cline (X): “@cline love how blazing fast it feels in the desktop app amazing lineup” [source](https://twitter.com/2010658787611619328/status/2103900512659857735)
- Complaint, 2026-09-27, r/CLine (Reddit): “windows 11, both computers. and it happens with every model, free and clinepass models. it appears randomly.” [source](https://www.reddit.com/r/CLine/comments/1wrvji9/why_does_cline_agent_keep_showing_this_stupid/pcgewmm/)
- Complaint, 2026-09-27, r/CLine (Reddit): “its been 3 days since i subscribed to clinepass and downloaded the cline agent. but every few hours i keep getting this error "the run failed: hub connection closed (code=1006, reason=connection ended)". its been 3 days, yall couldnt find a solution for this? <strict_link>” [source](https://www.reddit.com/r/CLine/comments/1wrvji9/why_does_cline_agent_keep_showing_this_stupid/)
- Complaint, 2026-09-27, @cline (X): “@flowvsgravity @cline i really don't think it's glm flash. pixel canary is really slow, throws errors every single time, and is so unstable that you can barely even use it right now.” [source](https://twitter.com/2083446628174663680/status/2104116725952188714)
- Complaint, 2026-09-27, @cline (X): “i know stealth models seem to be the in thing right now, but i'm not sure they're even worth messing around with sometimes. trying to use pixel canary on @cline, and it's just so slow. i mean, 24 hours now, no closer to the task, and it keeps stopping and starting. it's horrible!” [source](https://twitter.com/25673607/status/2104177440129991012)
- Complaint, 2026-09-27, r/LocalLLaMA (Reddit): “qwen3.8-flash-next 177b nvfp4(119gib): ssd streaming at 9-10 tok/s on one 16 gb rtx 5060 ti + 32 gb ram we built an inference engine for moe models that don't fit in vram + ram. most of the model stays on the ssd, and experts are read as tokens need them. this started as a proof of concept, and poc worked, we are getting 9-10 tok/s decode on qwen3.8-flash-next nvfp4 (9.06 on the benchmark turn, 10.4 on the best turn). this is just the start. with better ssd streaming, we expect v2 to reach ~14-15 tok/s decode. **model:** qwen3.8-flash-next, 176.9b params, nvfp4 gguf (119 gib): <strict_link> **machine:** rtx 5060 ti 16 gb, ryzen 7 9700x, 32 gb ddr5, gen5 nvme ssd 1tb, windows 11 **where the 1” [source](https://www.reddit.com/r/LocalLLaMA/comments/1wrxap8/qwen38flashnext_177b_nvfp4119gib_ssd_streaming_at/)

### Claude Code

- Praise, 2026-09-27, r/ClaudeCode (Reddit): “i keep running out of session with opus55. much faster.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wqnh3q/me_after_opus_55_release/pca1d74/)
- Praise, 2026-09-27, r/ClaudeCode (Reddit): “5.5 is really good but it does have the same flaws as all the other agents. i think one of the biggest appeals (any why most love it) is the huge reduction in its draw on subs (5x just became more like 30x if you set it on medium effort) and it speaks plain language for the most part, oh and its faster and less verbose, all while being a solid reasoning machine. its fast, cheap, and works. that makes for happy coders.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wqjuv2/am_i_the_only_one_who_doesnt_like_opus_55/pca70cd/)
- Praise, 2026-09-27, r/ClaudeCode (Reddit): “it is truly vastly different from opus 5. the cost and response speed are also completely different. that helps me stay better focused.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wqzt2x/opus_55_experience_of_an_engineer_at_big_tech/pcb4klf/)
- Praise, 2026-09-27, r/ClaudeCode (Reddit): “it depends on the area. i’ve completely replaced it with opus 5.5 in my agent system at work and in web development. but opus 5.5 definitely has its quirks. you first have to adapt your system to it depending on how complex it is and what dependencies are involved. but it saves me a lot of money and above all a lot of time. opus 5.5 is very fast. i use it in xhigh” [source](https://www.reddit.com/r/ClaudeCode/comments/1wrcyz1/so_if_fable_51_was_a_stopgap_for_opus_55_then/pcbjxvg/)
- Praise, 2026-09-27, r/ClaudeCode (Reddit): “not only is it the best model on intellectual level, it's somehow fast and inexpensive” [source](https://www.reddit.com/r/ClaudeCode/comments/1wre78w/opus_55_is_how_its_meant_to_be/pcbu9k1/)
- Complaint, 2026-09-27, r/ClaudeCode (Reddit): “omg stop! i have it controlling a 6 axis robot arm and it says that all day after a crash” [source](https://www.reddit.com/r/ClaudeCode/comments/1wquoxy/fable_51_live_vehicle_diagnostics/pcanovv/)
- Complaint, 2026-09-27, r/ClaudeCode (Reddit): “the "latest" caveat hit me just now; gave the new go phrase, but then followed it up with something else before controller finished launching the authed workflow and it had to stop and say "now that's the last typed message and it doesn't carry the right authorization" 🫠 tested the workflow provenance var - 0 and false definitely don't work. an empty-string theoretically would work, but even when my node and anthropic's bun setup were both registering the empty-string... the haiki probe i sent to test it still recorded both turns. hopefully anthropic gets around to fixing it. workaround is fine just annoying. i'll add my findings to the issue, send a note to anthropic about it.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wr8cw1/claude_code_workflow_harness_change/pcbtb90/)
- Complaint, 2026-09-27, r/ClaudeCode (Reddit): “yes but if you use it for extensively watch for buildup of processes tying up your pc/ laptop and overheating it! i walked in and heard my fan seriously working overtime and cpu at 125%. apparently it kept spinning up more and more helper processes” [source](https://www.reddit.com/r/ClaudeCode/comments/1wls9ul/would_you_actually_use_claude_code_from_your/pcc2733/)
- Complaint, 2026-09-27, r/ClaudeCode (Reddit): “that is helpful, thanks. i have noticed that claude sessions - and claude accounts are now talking to each other and coordinating much more now without being prompted to do so, and this is taking away some of what i needed to do myself. i had a big system crash the other day - had three agents each running multiple sessions, and all my windows went blank and a bunch of my apps dropped out including all of the agent windows. then they started popping back into life. i asked the agents what had happened, and they said they had all been hammering my 10cores at 300% (not sure how that is possible) for 10hours straight and they had finally fallen over. interestingly they said that they have since” [source](https://www.reddit.com/r/ClaudeCode/comments/1wgutzu/orca_for_orchestration_anyone_using_it_here_with/pccif6t/)
- Complaint, 2026-09-27, r/ClaudeCode (Reddit): “um. around the same i think? opus 5.5 is like around 30 tokens/ a sec for me. but it warries quite a bit. 30 is properly on the lower end” [source](https://www.reddit.com/r/ClaudeCode/comments/1wriidg/opus_55_vs_sol_in_terms_of_speed/pccwa0e/)

### GitHub Copilot

- Praise, 2026-09-26, r/GithubCopilot (Reddit): “most likely it launched subagents on a different, more expensive model (luna 5.6 often called gemini on my system). newer vscode/luna fixed that problem. in the meantime, you can disable subagents.” [source](https://www.reddit.com/r/GithubCopilot/comments/1wq39rv/did_they_10x_the_cost_of_luna_56_since_the/pc5o1hu/)
- Praise, 2026-09-26, r/PiCodingAgent (Reddit): “also incase you're interested in the background and how the slop progresses and hardens into slightly less slop. sorry for the spam, last response i swear :) \`\`\` it's the fix-tool-arg-coercion lane's own reproduction of the bug, the failing test it writes first, not an old warning we ignored. the other read failures in the logs are deliberate error-path tests (for example path: \[\]). rca-ca from the logs: \- our own pig sessions: about 4,600 sessions, with \~27,000 model turns, every one through github copilot (gpt and claude models), plus scripted test models. none hit this bug. copilot's models leave unused optional fields out of tool calls entirely. they never send "offset": null” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wqe2v1/meet_pig_the_pi_coding_harness_that_is_yours_but/pc9j62b/)
- Praise, 2026-09-24, r/GithubCopilot (Reddit): “haiku is still fast. for quick inline edits it is still the fastest model. that said, luna is close. maybe 6.0 luna will be even better? not sure, not enabled yet -.-” [source](https://www.reddit.com/r/GithubCopilot/comments/1wovttn/have_you_disabled_haiku/pbq7ox1/)
- Praise, 2026-09-24, r/GithubCopilot (Reddit): “speed is irrelevant to me, i always have 5-6 sessions opened at any point in time. if anything if they are slow it give some more room to breathe” [source](https://www.reddit.com/r/GithubCopilot/comments/1wovttn/have_you_disabled_haiku/pbqi9zy/)
- Praise, 2026-09-24, r/GithubCopilot (Reddit): “haiku is 10x the price of luna and way worse. put luna on low thinking if you just need an line edit, it will be equally fast. on openrouter, their recorded throughput is roughly the same.” [source](https://www.reddit.com/r/GithubCopilot/comments/1wovttn/have_you_disabled_haiku/pbqzq48/)
- Complaint, 2026-09-27, r/GithubCopilot (Reddit): “why would you wanna use copilot harness? it keeps making mistakes, ignore instructions and repeatedly fails read/write ops. another person in this sub posted benchmarks that put it lower than other harnesses like pi/ds/codex” [source](https://www.reddit.com/r/GithubCopilot/comments/1wr5io2/best_cheaper_alternative/pca0pix/)
- Complaint, 2026-09-27, r/OpenAI (Reddit): “open your excel workbook, goto the data tab on your toolbar ribbon, click sort, pick the column you want to sort by, click ok. you're done. with a tiny bit of practice (like once or twice) you can do that faster before most folks can even issue a prompt to copilot to do it for you. and from previous experience i can tell you with 100% confidence, that even the most fumble fingered excel user can do all that before copilot will deign to provide an answer. use ai to elevate your game, not replace your brain.” [source](https://www.reddit.com/r/OpenAI/comments/1wr8r6g/gpt6luna_comes_out_on_top_in_puppy_kill_bench/pcgywmc/)
- Complaint, 2026-09-26, r/GithubCopilot (Reddit): “i’m really new, so bear with me being perhaps a noob. i’ve got github copilot + and visual studio on a macbook that seems to get stuck on “run in terminal” i used it before for a few months so not expert but not totally new and this is first time it’s getting that problem it has run and done similar projects to what i’m doing now before i think but maybe i changed something i see some activity definitely when i run top bash command i’ve tried various coding agents like claude, gpt, and more -same issue and i’ve tried asking them for various solutions” [source](https://www.reddit.com/r/GithubCopilot/comments/1wr56qh/stuck_on_run_in_terminal/)
- Complaint, 2026-09-25, r/GithubCopilot (Reddit): “it's crazy buggy. even the integration with vscode is worse.” [source](https://www.reddit.com/r/GithubCopilot/comments/1wpctlc/vs_code_chat_users_local_or_copilot_harness/pbvujk7/)
- Complaint, 2026-09-25, r/GithubCopilot (Reddit): “has been happening all day to me running claude in the vscode agent window.... and i was so happy to just have found the agent window functionality :( any issue created yet in github one can follow?” [source](https://www.reddit.com/r/GithubCopilot/comments/1usonz3/vscode_agents_window_seems_to_get_stuck/pbyusn8/)

### Google Antigravity

- Praise, 2026-09-27, r/google_antigravity (Reddit): “subjective choice here. 1. codex cant pick a side its always beating around the bushes and it just steals your project information like zcode did before it was caught. 2. yes cc is good but its not as fast as gemini 3.8 flash which is a quick workhorse model. cc is for longhorizon no coding projects but i still need to design my architectures and coding projects so i prefer using gemini. also claude pretends to be the superman of modern llms which is a tendency i am not really fond of.” [source](https://www.reddit.com/r/google_antigravity/comments/1wrbsby/antigravitys_gemini_models_are_worst_in_among_all/pce8fep/)
- Praise, 2026-09-27, r/google_antigravity (Reddit): “who told you i dont have codex or claude max plan, i have them also, but antigravity, and only antigravity is the one that complements them, also i am using deepthinking in gemini chat, and it is so good, not what you think, i can get real time feedback with antigravity, solve problems together, spawn tens of agents and get things done in minutes, while with codex or claude, i need to put them on task before bed and wake up to see the results, that is if there are results, they can take days, also you can get wrong things, if i have to choose only one of my subscriptions, i will go for antigravity without hesitation.” [source](https://www.reddit.com/r/google_antigravity/comments/1wrn9q5/codex_has_ultra_thinking_level_claude_code_has/pcencos/)
- Praise, 2026-09-27, r/google_antigravity (Reddit): “i have claude max and i still use antigravity for coding. i use claude for claude design, antigravity for everything else. i have used claude code and i still prefer antigravity. the result for me are the same and i prefer the output speed of gemini, it allows me to iterate fast.” [source](https://www.reddit.com/r/google_antigravity/comments/1whvudv/people_who_are_complaining_about_gemini/pch3gm0/)
- Praise, 2026-09-27, @antigravity (X): “@ananyairl actually people might be using @antigravity for the crazy usage limits and personally i also like it when you orchestrate it with astra or sol or opus by planning with intelligent models, and execution with gemini flash 3.8 high. its fast and almost reliable.” [source](https://twitter.com/1677941694/status/2104222803167985875)
- Praise, 2026-09-27, @antigravity (X): “people shit on @antigravity a lot - but its very usable free student plan gives a ton of usage its one of the fastest models and it's great when you need smt simple just done. its also pretty good for browser use <strict_link>” [source](https://twitter.com/2031861244118945792/status/2104295241935143042)
- Complaint, 2026-09-27, r/google_antigravity (Reddit): “well i've met a lot of desktop app users who come to my discord server and start screaming that antigravity is shit. then i made them use cli and they've always been happy since. i think the main problem people hate with the gui is the insane token burn and the gui talking up all the resources of your computer. that's why i switched in the first place to get rid of the insane token burn. plus the cli usually only has bugs and doesn't normally experience errors a lot atleast for me” [source](https://www.reddit.com/r/google_antigravity/comments/1wqmd4n/why_is_the_antigravitycli_so_underrated/pcamrq4/)
- Complaint, 2026-09-27, r/google_antigravity (Reddit): “it's a bug, even it's affected half of my accounts” [source](https://www.reddit.com/r/google_antigravity/comments/1wqyjle/does_anybody_else_experience_this_agy_error/pccih98/)
- Complaint, 2026-09-27, r/google_antigravity (Reddit): “i see. i was on linux, not sure if that's the factor. i was on 1.2.11, updated to 1.2.12 still hitting it with \`agy --dangerously-skip-permissions remote-control start\` put \`toolpermission\` and \`permissions\` in both .gemini/config/config.json and .gemini/antigravity-cli/settings.json (not even sure why they have two directory and two different files, maybe gemini 3.8 flash hallucinate when it was troubleshooting it). thanks for discussing anyways. i am still curious if you are using \`agy remote-control start\` (headless mode, which starts a long-live daemon, this is what's broken for me) or \`agy --remote-control\` (this one is fine).” [source](https://www.reddit.com/r/google_antigravity/comments/1wqmd4n/why_is_the_antigravitycli_so_underrated/pccxjbm/)
- Complaint, 2026-09-27, r/google_antigravity (Reddit): “i started a new conversation, but it's still encountering an error.” [source](https://www.reddit.com/r/google_antigravity/comments/1wrd20m/help_orz_i_have_done_everything_i_could/pcd219g/)
- Complaint, 2026-09-27, r/google_antigravity (Reddit): “was wondering same thing yesterday when i switched from agy ide to vs code (because of instability of agy ide and google not updating this product anymore). but i guess we'll have to do with the default (copilot) autocomplete for now.” [source](https://www.reddit.com/r/google_antigravity/comments/1wpw81p/will_tab_autocomplete_be_released_to_the/pcdwa7h/)

### OpenCode

- Praise, 2026-09-27, r/opencodeCLI (Reddit): “very good, if you host it on a server or leave your computer running you can use it with it’s webui or with the ios/android app without losing access to any features from opencode. it’s helped me immensely and i’d strongly recommend at least trying it. also frequent updates and not buggy on the whole ,” [source](https://www.reddit.com/r/opencodeCLI/comments/1wrgmfe/whats_the_deal_with_openchamber/pccaety/)
- Praise, 2026-09-27, r/opencodeCLI (Reddit): “it's only getting faster..” [source](https://www.reddit.com/r/opencodeCLI/comments/1wrddp1/minimax_is_actively_testing_m31flashpreview/pccp56a/)
- Praise, 2026-09-27, r/opencode (Reddit): “yea, fast is the thing i love most tbh” [source](https://www.reddit.com/r/opencode/comments/1wroplu/xiaomi_mimo_26_flash_vs_glm_53_flash/pcfapui/)
- Praise, 2026-09-27, r/opencode (Reddit): “i felt the speed performance definitely. the other issue i have with it is high rate of compactions in relation to codex. but definitely a step in the right direction.” [source](https://www.reddit.com/r/opencode/comments/1wrj9e6/deepseekopencode_is_great/pcfx9ox/)
- Praise, 2026-09-27, @opencode (X): “this is the stealth model space bunny that was on @opencode and @openrouter the model is insanely fast, but it needs clear and specific instructions, otherwise it's too lazy. i tested it here <strict_link> <strict_link>” [source](https://twitter.com/2010272031603068928/status/2104131788578975846)
- Complaint, 2026-09-27, r/opencode (Reddit): “yes, same experience here. i tried to migrate to v2 this friday but all my plugins failed. so, by now is a no-go for me as my workflow depends so much on these plugins.” [source](https://www.reddit.com/r/opencode/comments/1wqppth/opencode_v2/pcag8ri/)
- Complaint, 2026-09-27, r/opencode (Reddit): “was looking for some comments on plugins: i switched to v2 and didn't notice that all the plugins failed, the harness i've built wasn't loading, throwing away tokens instead of saving them... once i notice, took me one or two rounds of claude to adjust everything and now all is fine again. just dropping this to saving you from the bitter drink i had ;-)” [source](https://www.reddit.com/r/opencode/comments/1wqppth/opencode_v2/pcb1yj9/)
- Complaint, 2026-09-27, r/opencode (Reddit): “how's the speed? 5.3 flash on go is like a turtle. can't stand it.” [source](https://www.reddit.com/r/opencode/comments/1wr8w3t/cheepseek_phase_2_full_precision/pcbnf96/)
- Complaint, 2026-09-27, r/opencode (Reddit): “idk about quantization but i'm getting a whole lot more api errors personally” [source](https://www.reddit.com/r/opencode/comments/1wr8w3t/cheepseek_phase_2_full_precision/pcc3yw0/)
- Complaint, 2026-09-27, r/opencode (Reddit): “not sure what you mean. opencode go hosts this model through official enterprise gateways, they serve the lossless base checkpoint bit-for-bit. what you are upset about is the difference in speed between deepseek api and opencode api. the likely cause for this is opencode's architecture (re-routing) and the context re-reading vs. deepseek's native kv caching. so no, you do not get quantized tokens. only how the tokens arrive to you differs.” [source](https://www.reddit.com/r/opencode/comments/1wrg81n/go_subscription_is_slower_deepseek_flash_41/pccdoi8/)

### Devin

- Praise, 2026-09-27, @DevinAI (X): “@hraness @chatgpt @devinai me too and i'm finally getting satisfied after 6 mos of tinkering. especially around reliability and mutli host orchestration. in fact there's so much to orchestrate not just agents.” [source](https://twitter.com/1887172409125658624/status/2104342859843273166)
- Praise, 2026-09-26, @cognition (X): “@agentmasterkey @devinai @cognition @openai @thsottiaux best harness is the one that keeps working when codex is down. failover is the feature” [source](https://twitter.com/2009223361969442816/status/2103644784011407861)
- Praise, 2026-09-23, @DevinAI (X): “@guybedo @devinai @zeddotdev it's been pretty snappy for me” [source](https://twitter.com/896906084014845952/status/2102564705009275354)
- Praise, 2026-09-22, r/windsurf (Reddit): “it's completely free now, faster after few minutes of planning and runs in cloud seemlessly. i am running it almost 24x7 till it's free. my 20 dollar investment is working out now!” [source](https://www.reddit.com/r/windsurf/comments/1wd432p/swe2_first_experiences/pbaufyd/)
- Praise, 2026-09-22, @cognition (X): “@cognition it can really be said to be full of sincerity. i ran geekbench 7 on claude code on the web, grok bot, and devin web (ubuntu/macos/windows) respectively, using my own main device m2 max as a reference. devin web can completely match my own machine in zed compilation tests, and it can run 4 instances in parallel! even more astonishing is that the macos runner actually has a gpu! <strict_link>” [source](https://twitter.com/1649366440808681474/status/2102307587639144533)
- Complaint, 2026-09-27, r/CognitionLabs (Reddit): “looks like there is a bug with the devin tab. im using a m4 macbook air 24gb. the problem started with golden gate update. i asked claude opus 5.5 and after debugging it found this: you're right, it isn't normal. i found the code path responsible, and it's a performance bug inside devin's built-in extension, not something in your setup. **where the 2.2s goes** (newest profile, `exthost-b13005.cpuprofile`, 2240 ms): * 95% of the time is inside the extension's `get usersettings` → `resolveunspecifiedsettings` code. * the callers are the autocomplete features: `getcompletion`, `getquickactions`, `_getsupercompleteitems` and `getrecentclipboardentry`. **why that's slow:** in the extension's `dis” [source](https://www.reddit.com/r/CognitionLabs/comments/1w23p1r/devin_is_simply_too_slow/pcfa29z/)
- Complaint, 2026-09-27, @DevinAI (X): “@markfenner @devinai why devin take so much time while building?” [source](https://twitter.com/2278324309/status/2104309916702048679)
- Complaint, 2026-09-26, @cognition (X): “@dabit3 i must say swe-2 is still slow but its also really good, thank you for this @cognition” [source](https://twitter.com/1973083865607708673/status/2103901813158355257)
- Complaint, 2026-09-25, r/windsurf (Reddit): “client error: protocol error (unimplemented): we are currently experiencing capacity issues with this serving model. please switch to a different model or try again later. (trace id: <structured_id>) unimplemented? and "with this serving model"? surely it should be "with serving this model". my idea - make the error messages better and fix the "protocol error" bug. more capacity would be nice too of course 😄” [source](https://www.reddit.com/r/windsurf/comments/1wpw8do/error_messages_are_not_their_best_strength/)
- Complaint, 2026-09-25, @cognition (X): “@cognition 's server is on fire 🔥🔥 🔥 🔥” [source](https://twitter.com/2047309688077950976/status/2103305136470958095)

### Amp

- Praise, 2026-09-26, @AmpCode (X): “@ampcode team is shipping. i get like 5 relaunch to update per day. and the experience keeps getting better and better. i know orbs and runners are kinda competing products, but would like to have feature parity between them with browser use and everything else.” [source](https://twitter.com/1993188300719636485/status/2103833810978902137)
- Praise, 2026-09-23, @AmpCode (X): “@sqs @purefunctor @ampcode i’ve never opened amp as fast as i just did” [source](https://twitter.com/1443430538/status/2102751907387498733)
- Praise, 2026-09-23, @AmpCode (X): “@sqs @ampcode wow!! so quick! i'll give it a shot! you're amazing @sqs” [source](https://twitter.com/268615009/status/2102838747721277472)
- Praise, 2026-09-18, @AmpCode (X): “nice little popup with auto retry on @ampcode <strict_link>” [source](https://twitter.com/85549810/status/2100910929047162890)
- Praise, 2026-09-18, @AmpCode (X): “@rockorager @ampcode so quick, i left my desk for less than five minutes and it was back up 👍” [source](https://twitter.com/33135576/status/2101085551021654105)
- Complaint, 2026-09-26, @AmpCode (X): “@benvargas @fastchicken @ampcode i set that up too, when i however have too many concurrent claude requests going on it starts giving me errors, do you have issues with that?” [source](https://twitter.com/2416367251/status/2103773774084550768)
- Complaint, 2026-09-25, @AmpCode (X): “@yjsoon @ampcode not just you. we just updated our status page at <strict_link>. seems chatgpt is having some broader issues.” [source](https://twitter.com/1369860113423609866/status/2103621110655004860)
- Complaint, 2026-09-24, @AmpCode (X): “@sqs @ianlandsman @ampcode i will try tailscale as well, curious how that fits together. but it would be great if portals would be fast and usable. what causes the slowness here? it's one of the main things that's keeping me from moving everything to amp.” [source](https://twitter.com/1330980620617601026/status/2103028237299290550)
- Complaint, 2026-09-24, @AmpCode (X): “@sqs @ampcode some of the speed difference might be slowness in amp from routing through my chatgpt subscription” [source](https://twitter.com/132882990/status/2103057206040261029)
- Complaint, 2026-09-23, @AmpCode (X): “@sqs @shavsycle @ampcode since here, can you also check iterm2 nested scroll issue? there's 7mo old post on reddit about this. for me the bug is repeatable with starting new puck thread.” [source](https://twitter.com/1089515309122367489/status/2102847742557196497)

### Cursor

- Praise, 2026-09-27, r/cursor (Reddit): “you are right about part of it but wrong about fast being slower than normal . that could never happen. fast will have priority on resources. normal comes seconds . in rare cases normal could be same speed as fast . the speed of normal varies up and down. best case its same as fast . but fast almost always same speed for a certain window of time .” [source](https://www.reddit.com/r/cursor/comments/1wp2ped/i_compared_cursor_composer_25_normal_vs_fast/pccmxhn/)
- Praise, 2026-09-27, r/cursor (Reddit): “i’m running an m1 air w 16gb and it runs like a breeze. maybe it’s the actual workload you have it running?” [source](https://www.reddit.com/r/cursor/comments/1wmj0pw/does_cursor_make_anyone_elses_computer_extremely/pcdw03l/)
- Praise, 2026-09-27, r/cursor (Reddit): “sorry but this sounds like an operator problem. my system was running slow at one point so i prompted it to optimize my system. haven’t had a problem since, that was about 5 months ago.” [source](https://www.reddit.com/r/cursor/comments/1wqppuh/grok_is_shutting_down_apps_now/pcf5mxo/)
- Praise, 2026-09-27, r/cursor (Reddit): “for my web development usage i find cursor's models like composer to be much faster for simple tasks. i also like the inbuilt browser and ide interface, despite how it seems like they are trying to make it an afterthought in the app. opus 5.5 is next level but i've only found myself reaching for that for tougher tasks.” [source](https://www.reddit.com/r/cursor/comments/1wrx24u/1_year_of_cursor_switched_to_claude_best_decision/pcgnlh4/)
- Praise, 2026-09-27, r/cursor (Reddit): “i think you’re exactly right. grok 4.7 isn’t as “good” as opus 5.5 but on the other hand it’s a lot faster and doesn’t seem to overthink even simple instructions. i find it very useful for a large number of coding tasks.” [source](https://www.reddit.com/r/cursor/comments/1wrktgc/am_i_the_only_one_who_thinks_grok_47_is_actually/pch445d/)
- Complaint, 2026-09-27, r/cursor (Reddit): “a few minutes/instant. switched to claude yesterday, fully operational on 3 very different projects.” [source](https://www.reddit.com/r/cursor/comments/1wq3m71/im_out/pc9w7tt/)
- Complaint, 2026-09-27, r/cursor (Reddit): “cursor hanging on taking longer than expected after the shell already finished is the same false busy lie as waiting for subagent. i kill that agent pane first and reopen the folder so the host actually resets. full app restart helps less than clearing the hung session. if the spinner comes back on the next prompt the host is sick not the model.” [source](https://www.reddit.com/r/cursor/comments/1wquvoq/taking_longer_than_expected/pccfywa/)
- Complaint, 2026-09-27, r/cursor (Reddit): “i don't feel similarly. much slower, way less accurate” [source](https://www.reddit.com/r/cursor/comments/1wrktgc/am_i_the_only_one_who_thinks_grok_47_is_actually/pcdnkab/)
- Complaint, 2026-09-27, @cursor_ai (X): “@grok @bot @cursor_ai been there, done that. nothing in the computer to approve. no tasks running. nothing. we're beyond all that. the bot can now message other bots but cant start a chat. can only reply.” [source](https://twitter.com/1420831171282362374/status/2104346415216689525)
- Complaint, 2026-09-26, r/cursor (Reddit): “cursor is highly unoptimised. team needs to now work on optimisation too. it just keeps on filling ram and ssd on its own and then ooms or crashes.” [source](https://www.reddit.com/r/cursor/comments/1wqppuh/grok_is_shutting_down_apps_now/pc5yr3x/)

### OpenAI Codex

- Praise, 2026-09-27, r/codex (Reddit): “yes you will. it's very fast and easy for daybreak blue.” [source](https://www.reddit.com/r/codex/comments/1wqxiro/daybreak_issue/pca2ul8/)
- Praise, 2026-09-27, r/codex (Reddit): “it's a decent daily driver. sometimes not waiting 20min for the task to complete is the only thing between you and your task being done. 3.8 does fine for execution and medium complexity. it's cheap and very fast.” [source](https://www.reddit.com/r/codex/comments/1wqtd1g/openai_gpu_are_really_cooling_down_theo_just/pcbd6sl/)
- Praise, 2026-09-27, r/codex (Reddit): “claude models are also slower in general because their harness lack the web sockets connection that makes codex models so much faster as well as it generating far more reasoning tokens. op please update us once claude is done cooking so we have a baseline to compare both models usage in terms of actual work done.” [source](https://www.reddit.com/r/codex/comments/1wre9dq/codex_usage_vs_claude/pcbwouh/)
- Praise, 2026-09-27, r/codex (Reddit): “guess i'm lucky, always on latest version, works fine. the only glitch i noticed is sometimes it quits chat and goes on main page, but that may be computer use clicked somewhere” [source](https://www.reddit.com/r/codex/comments/1wra4ev/codex_on_windows_do_you_work_with_wsl_agent/pcbwueb/)
- Praise, 2026-09-27, r/codex (Reddit): “the token efficiency is mainly giving it a leg up in terms of speed, honestly. it’s a very quick model.” [source](https://www.reddit.com/r/codex/comments/1wrcc9j/astra_minor_astra_61_and_devday_we_see_50/pcc54yx/)
- Complaint, 2026-09-27, r/codex (Reddit): “are you joking man? codex is already so slow. it already is in slow mode ffs. we were asking for slow mode before when it was actually fast.” [source](https://www.reddit.com/r/codex/comments/1wr1olr/new_idea_codex_slow_mode/pc9tdu1/)
- Complaint, 2026-09-27, r/codex (Reddit): “ye great idea, great use of my time. i'll just recode the fucking app on a whim because an update randomly removed an intentional login flow. or they could just not make their product consistently worse?” [source](https://www.reddit.com/r/codex/comments/1wr57y7/new_ui_in_codex_is/pc9y80i/)
- Complaint, 2026-09-27, r/codex (Reddit): “highly recommend sticking with that...fun times for the last 48 hours smfh matter of fact, last week or more. haven't been able to send 2 prompts (1, re-log, 1, re-log, etc) now can't even open the app..what an absolute joke <strict_link>” [source](https://www.reddit.com/r/codex/comments/1wr69nb/see_you_soon_guys_probably/pca1i4f/)
- Complaint, 2026-09-27, r/codex (Reddit): “i updated cli today and it's screwed. i can't paste into it. every action it does pops up with blank console windows by the dozen. so if i enter something and want to cancel i can't because of all the popping windows. worst update ever” [source](https://www.reddit.com/r/codex/comments/1wr69nb/see_you_soon_guys_probably/pca6ic6/)
- Complaint, 2026-09-27, r/codex (Reddit): “not me. this is the worst it's ever been. i've never had it go down like it is right now.” [source](https://www.reddit.com/r/codex/comments/1wr71h7/has_anyones_astra_become_more_generous_on_their/pcaa7fk/)

### Factory

- Praise, 2026-09-27, @FactoryAI (X): “the speed of deepseek v4.1 flash in @droid is mind blowing. makes me rethink the ollama subscription - when @factoryai has it all. <strict_link>” [source](https://twitter.com/1590702228234391552/status/2104263092364615842)
- Praise, 2026-09-27, @droid (X): “the speed of deepseek v4.1 flash in @droid is mind blowing. makes me rethink the ollama subscription - when @factory has it all. <strict_link>” [source](https://twitter.com/1590702228234391552/status/2104258455695741401)
- Praise, 2026-09-24, @FactoryAI (X): “@trevorbmurkp @harrystuck77 @factoryai @droid @badlogicgames @ampcode @cursor_ai @opencode @anthropicai @claudeai @cognition @xai @grok @build we’ve done a decent bit of perf improvements in the past few weeks with more incoming!” [source](https://twitter.com/1721143727043887104/status/2102916465167159772)
- Praise, 2026-09-24, @FactoryAI (X): “i used the droid for several hours today. i feel refreshed. it feels faster and more stable than codex and cc. it's better than opencoder and pi. using the open-source models included in them, like core, is also not expensive. praise! 🥳 @factoryai <strict_link>” [source](https://twitter.com/378283252/status/2102974707830276385)
- Praise, 2026-09-15, @droid (X): “@droid @zai_org this morning, i gave a task to sol5.6 high with codex and after 4.5 hours, i had to stop but it was going nowhere and burning my weekly credit. then i used droid + glm5.3 load on my mac studio m3 ultra: it completed very well the task in about 1 hour” [source](https://twitter.com/1954882023769944064/status/2099849372238323983)
- Complaint, 2026-09-27, @FactoryAI (X): “@droid @factoryai it was feature locked until last update.” [source](https://twitter.com/70830663/status/2104209256119775343)
- Complaint, 2026-09-27, @FactoryAI (X): “@droid @factoryai receiving 403s on all requests, fix your system” [source](https://twitter.com/1453931380983820293/status/2104294136631181394)
- Complaint, 2026-09-25, @FactoryAI (X): “@factoryai hi i am getting 400 bad request on gpt 6 luna, solution and opus 5.5. i am on the progress plan. any pointers to check this” [source](https://twitter.com/4499700680/status/2103317535668220377)
- Complaint, 2026-09-23, @FactoryAI (X): “@harrystuck77 @factoryai @droid @badlogicgames @ampcode @cursor_ai @opencode @anthropicai @claudeai @cognition @xai @grok @build can you explain your reasoning? what tasks do you usually give to it using missions? i stopped using droid when it's using too much memory on my servers and it's not worth it.” [source](https://twitter.com/2064081082308599808/status/2102629361249915213)
- Complaint, 2026-09-22, G2 (G2): “q: what problems is the product solving and how is that benefiting you? a: factory ai help me reduce manual coding work and save development time. it help with repetitive tasks, debugging and building features faster. i can focus more on important work instead of doing everything manually. it make my daily workflow more easy and productive. q: what do you like best about the product? a: factory ai is helpful for automating development work. it save my time, reduce manual tasks and help me complete coding work faster. the workflow is easy and useful for daily development. q: what do you dislike about the product? a: sometimes factory ai does not understand my request correctly and i need to g” [source](https://www.g2.com/products/factory-ai/reviews/factory-ai-review-13384158)

### Kiro

- Praise, 2026-09-11, r/kiroIDE (Reddit): “for what i have been using kiro-cli it returns responses much faster than claude code but i think that the way it perfoms and the customization layer is under what claude code can offer” [source](https://www.reddit.com/r/kiroIDE/comments/1wchsii/frontier_models_on_other_providers_vs_the_same/p9325ex/)
- Praise, 2026-09-05, r/kiroIDE (Reddit): “just downgrade to v 0.12 -doesnt have agent focus but feeels snappier” [source](https://www.reddit.com/r/kiroIDE/comments/1w5nh3m/bug_report/p7x19jn/)
- Praise, 2026-09-03, @kirodotdev (X): “everyone is complaining that codex, cursor and claude are down but you can still fully use @kirodotdev” [source](https://twitter.com/1967964218956648448/status/2095580458452910505)
- Complaint, 2026-09-27, r/kiroIDE (Reddit): “kiro-cli timed out a few times yesterday. i lost two heavy work sessions. all i found is a ticket so it appears it does not exist (<strict_link>) how do people check downtime? does kiro do rate limiting if you are doing heavy work? it felt slow.” [source](https://www.reddit.com/r/kiroIDE/comments/1wrsp5y/kiro_status_page/)
- Complaint, 2026-09-25, r/kiroIDE (Reddit): “still now working today.” [source](https://www.reddit.com/r/kiroIDE/comments/1wp1sbg/kiro_subscription_payment/pbxiy8i/)
- Complaint, 2026-09-24, r/kiroIDE (Reddit): “fyi - still not working.” [source](https://www.reddit.com/r/kiroIDE/comments/1wp1sbg/kiro_subscription_payment/pbus63m/)
- Complaint, 2026-09-24, r/kiroIDE (Reddit): “if we can use gpt-5.6, that's still better. we can only use sonnet4.6 here. it's a completely delayed service and is unusable.” [source](https://www.reddit.com/r/kiroIDE/comments/1wp3xji/why_kiro_like_this/pbuv1e6/)
- Complaint, 2026-09-24, r/kiroIDE (Reddit): “cool! gonna try tomorrow. i prefer the vs lsp integration, debugger and interface. kiro only have the free debugger right now and somehow electron-like programs perform much worse then vs in my machine.” [source](https://www.reddit.com/r/kiroIDE/comments/1wp53oc/i_added_native_kiro_support_to_visual_studio_2026/pbvamrb/)

### Warp

- Praise, 2026-09-19, @warpdotdev (X): “@samueljmcd only because @warpdotdev is so snappy and i can use berkeley mono (font) from @usgraphics in it 🫶 <strict_link>” [source](https://twitter.com/1598005566701551617/status/2101350790611042323)
- Praise, 2026-09-11, @warpdotdev (X): “@warpdotdev fastest cli to add gets the daily driver spot” [source](https://twitter.com/1945115184072105984/status/2098448929515786421)
- Praise, 2026-09-08, @warpdotdev (X): “wow @warpdotdev with glm 5.3 is super cost efficient and fast. no brainer 💆♂️” [source](https://twitter.com/84324573/status/2097445164159483988)
- Praise, 2026-09-03, @warpdotdev (X): “@devanshuxi @zeddotdev @warpdotdev @mitchellh yeah it is super fast than iterm !” [source](https://twitter.com/1698865227864113152/status/2095454917637103899)
- Praise, 2026-09-03, @warpdotdev (X): “i’m not sure but i find @warpdotdev terminal much smoother and faster for local system work compared to hermes. i asked hermes to look into this open-design repo and configure it for my deepseek harness agent. it took much longer even though i provided exa ai and firecrawl api for web search. it still took over two minutes and gave me a detailed but cluttered result. in contrast, warp terminal did it in 30 seconds with a much smoother concise reply and suggestion. by the way, i’m using glm 5.3 flash via @deepinfra.” [source](https://twitter.com/859077042129600516/status/2095579049729151354)
- Complaint, 2026-09-19, @warpdotdev (X): “@warpdotdev windows app has unexpected bugs which i have lost count. now i am unable to drag it on the machine to move the window. but warp customer support doesn’t care. how can a company go so pathetic after open sourcing their codebase 🤦🏽♂️” [source](https://twitter.com/1624472822503583744/status/2101339256786677876)
- Complaint, 2026-09-15, @warpdotdev (X): “@vikvang1 @warpdotdev try windows. every app and huge games work fine on this gaming laptop except warp.” [source](https://twitter.com/1624472822503583744/status/2099848972106138049)
- Complaint, 2026-09-14, @warpdotdev (X): “@michael_kove @catalinmpit @warpdotdev moved to ghostty as well and pi as harness. my old intel mac is happy again” [source](https://twitter.com/25074228/status/2099386334116733106)
- Complaint, 2026-09-14, @warpdotdev (X): “warp has fallen so bad @warpdotdev terminal is massively bloated now. new tab opens with a visual lag. and when you provide feedback, some unapologetic dudes reply with no intention to fix. it was once my go to terminal but i am just waiting for my subscription to be over.” [source](https://twitter.com/1624472822503583744/status/2099404156977201621)
- Complaint, 2026-09-13, @warpdotdev (X): “@warpdotdev here is what i am talking about: like bro... what are you "checking..." , "installing..." just let me connect. me: `ssh named_config_host&gt;` warp: <strict_link>” [source](https://twitter.com/1309409339824840704/status/2099090914744434972)

### Conductor

- Praise, 2026-09-15, @conductor_build (X): “my stack now is @conductor_build + qwen models for building apps. fast and free!” [source](https://twitter.com/1382193359641481217/status/2099810273439715525)
- Complaint, 2026-09-21, @conductor_build (X): “@charlieholtz i have no idea whether its an issue with @cursor_ai or @conductor_build but cursor agents sometimes they just stop responding in conductor cloud. it has happened to me thrice in the past 3 days. please figure it out and fix it? i am happy to provide any details for you to debug if needed.” [source](https://twitter.com/1900337293564817408/status/2101825344626208846)
- Complaint, 2026-09-15, r/conductorbuild (Reddit): “agreed. about 50% of my new tabs just spin for a bit, then give me the the ... option to fork into a new tab. it's getting very very very tedious to deal with.” [source](https://www.reddit.com/r/conductorbuild/comments/1wdm8hc/conductor_regressions_are_making_it_harder_to/p9y4n50/)
- Complaint, 2026-09-15, @conductor_build (X): “@conductor_build i'm experiencing some weird behavior, where my window gets black several times when i try to click on it. usually happens when i close conductor, or it crashes, and i open it again. keep getting black several times before stabilizing.” [source](https://twitter.com/2068260048199925760/status/2099818314922897499)
- Complaint, 2026-09-13, r/conductorbuild (Reddit): “yes i would agree. probably one of the most frustrating issues… it always seems 2 steps forward, 1 step back. conductor is my favorite tool but it seems full of constant paper cuts like this.” [source](https://www.reddit.com/r/conductorbuild/comments/1wdm8hc/conductor_regressions_are_making_it_harder_to/p9kmth7/)
- Complaint, 2026-09-13, r/conductorbuild (Reddit): “failed to retry workspace: failed to fetch from origin: fatal: unable to read current working directory: operation not permitted. conductor version 0.85.0 (3d25e11ff9) not sure what is causing this. i had months of work. deeply disappointing.” [source](https://www.reddit.com/r/conductorbuild/comments/1wfkq2n/workspace_initialization_failed/)

### Grok Build

- Praise, 2026-09-05, r/codex (Reddit): “is anyone using astra/sol and grok 4.6 high via grok build? i haven’t received astra yet, but i do use sol. and as of late, i’ve almost entirely shifted to grok build (grok 4.6 high) - mostly because it’s way faster than gpt and very good at completing tasks end-to-end for my flutter project. when there is access to both llms, speed does get the veto from me, and i mostly use codex/gpt only as my general ai driver, and not for my flutter project anymore. how is it for others?” [source](https://www.reddit.com/r/codex/comments/1w7on0j/astra_is_absolutely_incredible/p7ys22g/)
- Complaint, 2026-09-22, r/cursor (Reddit): “i didn't find it very expensive when used in grok build, but it's very slow.” [source](https://www.reddit.com/r/cursor/comments/1wmswj9/grok_47_is_about_25_times_as_expensive_as_46/pbecymc/)
- Complaint, 2026-09-11, r/ClaudeCode (Reddit): “i use fable as my orchestrator, sonnet recon and headless grok build. will shift to sonnet/opus build once the heavy grok discount runs out. it is slower, a lot slower but token efficient and all the models add a layer of checking the others work. less baby sitting and cleaner code.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wd9eio/fable_as_orchestrator_and_opussonnet_as_executers/p9574gb/)
