# Client crashes, freezes and failed tool execution (`rel.client_failures`)

Feedback Bench, coding agents, built 2026-10-01, window 2026-08-31 to 2026-09-27. Web page: https://feedbackbench.com/#/criterion/rel.client_failures

Area: [Reliability and speed](https://feedbackbench.com/criteria/reliability.md)

**Definition.** The local app, CLI or extension crashes, freezes, leaks memory or uses too much CPU, or fails to execute the model's tool calls and shell commands.

**Boundary.** Not this: see [Outages, server errors and capacity or rate errors](https://feedbackbench.com/criteria/rel.service_errors.md) for backend outages and server errors. Not this: see [Updates break working setups](https://feedbackbench.com/criteria/rel.update_breakage.md) when a specific update caused it.

Rated author-weeks, all agents: 2203. Complaint share: 93%.

## The brief

Written by Claude Opus 5.5 from 101 labelled posts and the numbers on this page. Interpretation, not measurement: every quote is verbatim and links to its post.

**Lean terminal clients hold up; heavy desktop apps freeze, bloat and stall.**

TL;DR:

- Codex and Antigravity draw the most crash, freeze and lag complaints, both worse than peers.
- Zed, Pi and Claude Code stand out for light footprints and fewer client failures.
- Recurring pain: hung UIs, runaway RAM, failed tool calls, sessions that silently stop.

In plain terms: Expect the app itself to be the weak link. Users kill frozen processes, restart to clear zombie agents, and watch RAM climb. Many retreat to the CLI, which posts describe as lighter and steadier.

### How it breaks

- **Desktop apps freeze and lock up** ([Client crashes, freezes and failed tool execution](https://feedbackbench.com/criteria/rel.client_failures.md)). The most common failure is a GUI that stops responding mid-task, forcing users to kill the process or close the terminal.
  Posts describe loading spinners that never clear, input that stops registering while the cursor still blinks, and UIs lagging several words behind typing. Devin CLI users report freezes on long file writes that only a fresh process fixes. A Claude Code user says the desktop app lags the whole machine and the CLI is better. Freezing mid-task is one of the most-requested fixes, led by Codex and Cursor users.
  Evidence:
  - Complaint, Claude Code, @ClaudeDevs, 2026-09-25: “@claudedevs still lags and freezes my machine. cli is better” [source](https://twitter.com/44595437/status/2103303119799210260)
  - Complaint, Devin, r/windsurf, 2026-09-03: “when writing longer files, such as markdown plans or html for documentation, devin cli has been constantly freezing up on me. i tried hitting ctrl+c, ctrl+z, ctrl+l, ctrl+shift+l, and nothing happens. the cursor in the chat is still blinking, and i can still scroll, but typing does nothing, refreshing, stopping -- nothing seems to work besides closing the terminal and starting a new process. by the way, i am in windows 11/windows terminal/wsl2/debian trixie/devin 3000.5.20 (2d902011). is anyone else having this problem? some of my teammates are occasionally having this problem, but not nearly as much as i am.” [source](https://www.reddit.com/r/windsurf/comments/1w6eoh8/devin_cli_freezes_but_continues_in_background/)
  - Complaint, Devin, @cognition, 2026-09-12: “i really like what @cognition is building with @devindesktop, but the client experience is really having some state, process management, and ux issues: -sessions disappearing and losing history (!!!) -the ui locking up to 3 *word* input lag -silent devin -p failures (1/3)” [source](https://twitter.com/928932212/status/2098854625323463157)
  - Complaint, OpenAI Codex, r/codex, 2026-09-27: “windows desktop app. unfortunately as a windows user i feel we are often neglected. super buggy - and it's been since the latest release that the bug is high impact (stuck at loading spinner unless we kill codex.exe). even then, codex works but not chatgpt. sometimes i use codex cli (with wsl) for windows, but i mostly use the desktop app. i hope they pay attention to users that are on other os other than macos.” [source](https://www.reddit.com/r/codex/comments/1wra4ev/codex_on_windows_do_you_work_with_wsl_agent/pcc70fw/)

- **Memory and CPU run away** ([Client crashes, freezes and failed tool execution](https://feedbackbench.com/criteria/rel.client_failures.md)). Clients leak memory as sessions grow and leave helper processes alive, turning a coding agent into the heaviest thing on the machine.
  A Cursor user finds stopped sub-agents relaunching on every start and holding gigabytes of RAM. A Devin Ask user watches the tab climb past 8 GB around fifty prompts and suggests virtualizing chat history. Zed users report idle CPU above VS Code. Leaked-process cleanup and RAM reduction both appear on the request list across many vendors.
  Evidence:
  - Complaint, Cursor, @cursor_ai, 2026-09-16: “@cursor_ai having an issue with zombie sub-agents restarting every time i launch cursor. currently eating ~5gb of ram running 11 sub-agents i stopped yesterday. any fix?” [source](https://twitter.com/1481656122645696518/status/2100354201171698135)
  - Complaint, Devin, r/CognitionLabs, 2026-09-15: “i’ve been running into a pretty rough performance issue with devin ask on longer conversations. i’m on windows 10 using operagx, with an older phenom ii x6 1100t be and 12 gb of ddr3 ram, so my machine is definitely not new, but the devin tab starts using 8 gb+ of ram once the conversation gets to around 50 prompts, and input lag becomes noticeable. the ram usage keeps climbing as the context grows. it feels like the chat canvas may be keeping too much of the conversation rendered in the dom at once. a quality-of-life improvement that would help a lot would be virtualizing the chat history, keeping only the most recent messages visible/rendered, maybe around 20 prompt/response pairs, and loading older content back only when scrolling up. other ai chat uis seem to handle long threads this way, and it would make devin ask much more usable on lower-end or older machines.” [source](https://www.reddit.com/r/CognitionLabs/comments/1wgoq0x/perf_issue/)
  - Complaint, Devin, @cognition, 2026-09-23: “i love devin but wtf is this memory consumption? @cognition @dabit3 <strict_link>” [source](https://twitter.com/1333116608609398785/status/2102571264183242968)
  - Complaint, Zed, r/ZedEditor, 2026-09-23: “same issue on sequoia. high cpu usage without any extensions or lsp. cpu usage consistently remains more than vscode with occasional spikes.” [source](https://www.reddit.com/r/ZedEditor/comments/1sbyj79/zed_high_memory_and_cpu_usage_on_mac/pbi6zvh/)

- **Tool calls fail to execute** ([Client crashes, freezes and failed tool execution](https://feedbackbench.com/criteria/rel.client_failures.md)). The model asks for an action and the harness cannot run it, burning credits or retries for nothing.
  Warp users say tool-call failures consume credits and deliver nothing. Antigravity users count retries before a command executes correctly. Copilot users say the harness repeatedly fails read and write operations, and that hung tools waiting on input need manual rules to avoid. Pi users report the stock edit and read tools failing schema validation, with one user building a fix after mining months of edit failures.
  Evidence:
  - Complaint, Warp, @warpdotdev, 2026-09-02: “@warpdotdev when can you guys fix the tool calling failure issues? it burn my credit but deliver nothing.” [source](https://twitter.com/268717001/status/2095001516151132453)
  - Complaint, Google Antigravity, r/google_antigravity, 2026-09-07: “it's meh compared to fable / astra but definitely a massive leg up compared to 3.5 and 3.1 of its own family. i'm comparing "tool usage" how many times does it need to retry to get a command correctly executed” [source](https://www.reddit.com/r/google_antigravity/comments/1w9vwgv/tbh_flash_38_is_really_good_when_it_comes_in/p8e9t6y/)
  - Complaint, GitHub Copilot, r/GithubCopilot, 2026-09-27: “why would you wanna use copilot harness? it keeps making mistakes, ignore instructions and repeatedly fails read/write ops. another person in this sub posted benchmarks that put it lower than other harnesses like pi/ds/codex” [source](https://www.reddit.com/r/GithubCopilot/comments/1wr5io2/best_cheaper_alternative/pca0pix/)
  - Complaint, Pi, r/PiCodingAgent, 2026-09-13: “not hashline edit (imho doesn't work well), but i've had this problem with pi's edit tool for a long time, basically mined every single type of edit failure from 6+ months of usage and then found which ones can be fixed programmatically, the rest i enhanced with better errors the agent can act on. shameless self plug <strict_link>” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wfavpg/which_hash_line_extension_solves_constant_the_old/p9lynql/)

- **Sessions silently stop mid-run** ([Client crashes, freezes and failed tool execution](https://feedbackbench.com/criteria/rel.client_failures.md)). Long runs halt without a clear error, and the only recovery is nudging the agent or starting a new session.
  Cline users report sessions dying around the half-hour to hour mark on every model, sometimes recoverable by switching models. OpenCode users describe the agent waiting forever after running a command until they type something. Conductor and Cursor users see agents that just stop responding or drop queued messages. Unattended work is where this hurts most, and users say it blocks automated workflows.
  Evidence:
  - Complaint, Cline, @cline, 2026-09-26: “@cline this happens a lot for no reason. sometimes changing the model works and able to continue, other times just have to create a new session entirely. usually, happens at 37m or 1hr mark. happens on all (free and paid) models. <strict_link>” [source](https://twitter.com/220288969/status/2103715062653636714)
  - Complaint, OpenCode, r/opencode, 2026-09-08: “when i run some sessions, opencode sometimes just pauses, until i type something it it, and it acts like nothing happened. it's like it's waiting for a response or something and never gets it and just waits forever. it especially happens when it runs a command it sends. sometimes it seems to compact immediately after i wake it up. so maybe that has something to do with it? anyone else having this issue? i'm using llama.cpp, qwen 3.8 27b, 128k context. latest opencode. thanks in advance!” [source](https://www.reddit.com/r/opencode/comments/1wa9yj0/opencode_pausing/)
  - Complaint, Conductor, @conductor_build, 2026-09-21: “@charlieholtz i have no idea whether its an issue with @cursor_ai or @conductor_build but cursor agents sometimes they just stop responding in conductor cloud. it has happened to me thrice in the past 3 days. please figure it out and fix it? i am happy to provide any details for you to debug if needed.” [source](https://twitter.com/1900337293564817408/status/2101825344626208846)
  - Complaint, Cursor, @cursor_ai, 2026-09-22: “@cursor_ai message queueing is broken for about a week now, even after the latest update. the message is posted to the timeline after the agents turn has finished, but then just stops.” [source](https://twitter.com/9172042/status/2102328449134116906)

- **The CLI outlasts the app** ([Client crashes, freezes and failed tool execution](https://feedbackbench.com/criteria/rel.client_failures.md)). Across vendors, users praise the terminal client and blame the desktop app, using the CLI as their stability fallback.
  A Factory user calls the CLI amazing and the app buggy. Codex users say the CLI uses less CPU and memory and lets them run multiple instances freely. An OpenCode user on Linux avoids GUI harnesses entirely for resource reasons. The pattern is consistent enough that it reads as a product signal: the desktop wrappers are where client failures concentrate.
  Evidence:
  - Complaint, Factory, @droid, 2026-09-21: “@ain3sh @benvargas @droid hey a fellow droid user here, the cli is amazing, the app not so much, there were some bugs here and there. only thing with the cli is the rendering can be improved, it's a bit janky on missions, maybe you guys can move to a different framework from ink?” [source](https://twitter.com/1333161864/status/2101883259831632365)
  - Praise, OpenAI Codex, X search: OpenAI Codex, Codex CLI, Codex app, 2026-09-20: “the codex cli actually consumes less cpu and memory than pi, which is somewhat surprising. <strict_link>” [source](https://twitter.com/1732683816685514752/status/2101508430326591969)
  - Praise, OpenAI Codex, X search: OpenAI Codex, Codex CLI, Codex app, 2026-09-07: “personally, i think the opposite... i'm overwhelmingly a cli person. i tried using the desktop for the sol and hackathon topics & implementation, but in the end, i open the terminal for commands and file operations, which crowds the screen. if i use codex cli in the terminal, i can use multiple instances and arrange them freely, and it consumes less memory. also, working with wsl on windows is cumbersome... (the image is gnome) i haven't encountered any benefits other than the cuteness of pets...” [source](https://twitter.com/1743042073962672129/status/2096954964002377745)
  - Praise, OpenCode, r/opencode, 2026-09-20: “maybe , i dont have a mac machine , i have linux and windows , there is no problem in linux using opencode ans other harness on cli or tui and i will not reccomend (spelling mistake) because they f*cking consumes the system resources so heavily , i will prefer the cli or tui still you can install the tui of opencode in windows refer this doc : <strict_link>” [source](https://www.reddit.com/r/opencode/comments/1wlkq6c/windows_app_has_no_plan_mode/pazicsc/)

### Who stands out

- **OpenAI Codex (weaker)**. Codex dominates the complaint volume here, with the desktop app crashing, hanging on launch and lagging while the CLI earns quiet praise.
  Windows users report the app stuck at a loading spinner until they kill the process, and say non-macOS users feel neglected. Others see random crash-and-restart loops and broken projects across machines. Codex users file most requests for fixing laggy UI, 17 of 24. A Cursor user praises Cursor for not hogging the computer like Codex. On the other side, CLI users describe smooth, low-memory runs.
  Evidence:
  - Complaint, OpenAI Codex, X search: OpenAI Codex, Codex CLI, Codex app, 2026-09-23: “codex keeps randomly crashing and restarting on its own, hope they fix this soon @openai #codex <strict_link>” [source](https://twitter.com/1247015624/status/2102672150406525224)
  - Complaint, OpenAI Codex, r/codex, 2026-09-27: “windows desktop app. unfortunately as a windows user i feel we are often neglected. super buggy - and it's been since the latest release that the bug is high impact (stuck at loading spinner unless we kill codex.exe). even then, codex works but not chatgpt. sometimes i use codex cli (with wsl) for windows, but i mostly use the desktop app. i hope they pay attention to users that are on other os other than macos.” [source](https://www.reddit.com/r/codex/comments/1wra4ev/codex_on_windows_do_you_work_with_wsl_agent/pcc70fw/)
  - Complaint, OpenAI Codex, r/codex, 2026-09-25: “i have exactly the same issue on 2 computers now. i got codex to write a work around, but it means i am unable to use projects. if anyone has this issue and knows a proper fix i am desperate.” [source](https://www.reddit.com/r/codex/comments/1wlkggh/codex_stops_working_after_the_first_message_in/pbyknow/)
  - Praise, Cursor, @cursor_ai, 2026-09-06: “been heavily using @cursor_ai desktop app last few days just to try it. gotta say i really like it. the simplicity of it is great, it doesnt have the issues of hogging computer like codex. obviously codex is my fav still cause of the computer use ability. but cursor is def second and if it had equal computer use id prolly use it more” [source](https://twitter.com/539363367/status/2096624059479994547)

- **Google Antigravity (weaker)**. Antigravity users report a harness that breaks at the plumbing level, from missing tool registrations to git index corruption.
  Subagents fail to start because a tool is not found in the registry. Users share aliases to repair corrupted git indexes and question whether the remote feature wears out their SSD. One user left the IDE for Zed after slowness and crashes. Praise exists, with some calling the harness stable and smooth, but complaints vastly outnumber it.
  Evidence:
  - Complaint, Google Antigravity, r/google_antigravity, 2026-09-01: “same, seems like it's just plain broken. >the subagent code review investigator (deepinvestigator) with id \[omitted\] encountered an error and has either stopped or failed to start execution: failed to construct executor: failed to resolve components: unknown component: tool "skill\_search" not found in registry” [source](https://www.reddit.com/r/google_antigravity/comments/1w4d47l/boost_mode_usage/p77zjpm/)
  - Complaint, Google Antigravity, r/google_antigravity, 2026-09-08: “the index corruption is real. `git config --global alias.fix-index '!rm -f .git/index && git reset'` then you can say `git fix-index` at least.” [source](https://www.reddit.com/r/google_antigravity/comments/1wau0ea/how_to_make_file_changes_auto_accept_all/p8l0cos/)
  - Complaint, Google Antigravity, @antigravity, 2026-09-24: “huh, @antigravity 's remote is silently destroying your ssd? <strict_link>” [source](https://twitter.com/1443206574835699715/status/2103251993494179853)
  - Praise, Zed, r/google_antigravity, 2026-09-09: “the ide is slow and crashes. i switched to zed so i can use the models is a stable and snappy ide.” [source](https://www.reddit.com/r/google_antigravity/comments/1waolu4/why_does_antigravity_have_so_few_users/p8r9iig/)

- **Zed (stronger)**. Zed draws users specifically because it stays light and fast, though some report CPU spikes and frozen git panels.
  Posts praise a small memory footprint, GPU rendering and relief for modest laptops, and some arrive fleeing crashier IDEs. The complaints are real but narrower: high idle CPU on macOS, a git panel stuck pushing until restart, and machine-specific freezes on one Linux laptop that the same user's desktop does not show.
  Evidence:
  - Praise, Zed, @zeddotdev, 2026-09-13: “@arpit_bhayani i moved to @zeddotdev from vscode it's really great and runs under 200 mb.” [source](https://twitter.com/1302187537629413376/status/2099140245241594089)
  - Praise, Zed, @zeddotdev, 2026-09-27: “@0xprajwal_ i use @zeddotdev, btw. it’s fast, lightweight, consumes less memory, and uses the gpu for rendering.” [source](https://twitter.com/1542127649463898113/status/2104217257153110457)
  - Praise, Zed, @zeddotdev, 2026-09-25: “@zeddotdev my 16 gb ram development laptop is saved because of you!” [source](https://twitter.com/1115272367154950151/status/2103318390798774306)
  - Complaint, Zed, r/ZedEditor, 2026-09-12: “i have a cachyos laptop, which regularly has a frozen lap or git, or sometimes i just can't open files anymore. however on my cachyos desktop it works fine.” [source](https://www.reddit.com/r/ZedEditor/comments/1we3tul/instability_and_ui_lag/p9bepmk/)

- **Pi (stronger)**. Pi users praise its speed and low memory, and treat breakage as something the agent can fix itself.
  A user who swapped OpenCode for Pi reports it runs much faster and lighter. Another says six months without flicker problems, blaming one extension when it happened. Weak spots are tool-level: subagents failing on macOS while working on Linux, and edit-tool failures users have had to patch themselves.
  Evidence:
  - Praise, Pi, @pidotdev, 2026-09-04: “replaced opencode harness in 20x with @pidotdev and omg it's so much faster and consumes less memory.” [source](https://twitter.com/994170492901773312/status/2095929704993931715)
  - Praise, Pi, r/PiCodingAgent, 2026-09-14: “in six months i've never had this with pi. the only time i had flicker problems it was an extension i was running.” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wfvx2f/any_idea_how_to_resolve_for_this/p9q22uh/)
  - Complaint, Pi, r/PiCodingAgent, 2026-09-24: “just found an issue. everything was great on my linux desktop, but when i moved to my mac (darwin) it just doesn't work at all. a new pane opens, containing just this: echo __pi_herdsman_ready_5b4c042a-589a-48dd-a0c3-1b1e92f215e9__  echo __pi_herdsman_ready_5b4c042a-589a-48dd-a0c3-1b1e92f215e9__ __pi_herdsman_ready_5b4c042a-589a-48dd-a0c3-1b1e92f215e9__ and the subagent simply fails immediately. that pane also stays open. sorry for bad formatting, i'm on phone, not home. i have no idea why this is happening, but maybe it helps you herdr v0.9.1, pi 0.87.1, ghostty” [source](https://www.reddit.com/r/PiCodingAgent/comments/1woif1b/i_built_pi_herdsman_for_async_subagents_with/pbsl1ks/)
  - Complaint, Pi, r/PiCodingAgent, 2026-09-13: “not hashline edit (imho doesn't work well), but i've had this problem with pi's edit tool for a long time, basically mined every single type of edit failure from 6+ months of usage and then found which ones can be fixed programmatically, the rest i enhanced with better errors the agent can act on. shameless self plug <strict_link>” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wfavpg/which_hash_line_extension_solves_constant_the_old/p9lynql/)

- **Claude Code (stronger)**. Claude Code rates better than peers, with users citing modest resource use, though the desktop app still lags and stuck daemons need manual kills.
  Users report zero issues across Linux, macOS and Windows and expect quiet fans because the client stays light. Complaints center on the desktop app: slow side chats, machine-wide lag, and a stuck state one user escaped only by killing a supervisor process from a lock file.
  Evidence:
  - Praise, Claude Code, r/ClaudeCode, 2026-09-04: “yeah bro you have got something else going on the app works great for me zero issues” [source](https://www.reddit.com/r/ClaudeCode/comments/1w76z28/why_cant_anthropic_have_claude_make_their_desktop/p7usuhp/)
  - Praise, Claude Code, r/ClaudeCode, 2026-09-26: “i haven’t had any issues with it at all. i also keep pwsh on linux and maxos and it works on all three major platforms as a result 🤷♂️” [source](https://www.reddit.com/r/ClaudeCode/comments/1wqwi4h/why_claude_code_uses_python_for_everything_and/pc9n3w8/)
  - Complaint, Claude Code, r/ClaudeCode, 2026-09-05: “this is such a bullshit feature, its a bug, if im right, ctrl-x will delete it, the only one way to get unstuck out of it is by find and kill supervisor pid is in \~/.claude/daemon.lock, it quite complicated so ask your ai to do it for you, worked for me” [source](https://www.reddit.com/r/ClaudeCode/comments/1vwtzpf/how_to_completely_exit_a_background_agent/p7vygdv/)
  - Complaint, Claude Code, @ClaudeDevs, 2026-09-23: “@claudedevs @_catwu desktop app /btw (side chat) conversation is still slow, sometimes it shows error. side chats are very important, it helps keeping conversations unpolluted. please make it 3x faster. 🥲” [source](https://twitter.com/490212419/status/2102899328000172040)

### Fine print

- Posts here skew heavily toward complaints for every agent, since users rarely post when a client simply works.
- Some failures involve local model servers or user hardware, which posts cannot always separate from the client itself.
- Several posts tie problems to recent releases; those overlap with update breakage and may be counted there too.

## Top requests

What users ask to add or change, most asked first. 458 author-weeks ask for something. Requests do not change the Feedback Score. Rule: A separate pass by Claude Sonnet 5 reads every counted post and extracts what the author asks the agent or its vendor to add or change, with the criteria it maps to and a short normalised wording; it does not touch the labels or the Feedback Score. Claude Opus 5.5 groups the wordings within each criterion (the first criterion the request maps to) into themes; code counts them. A theme counts distinct author-weeks that ask for it, per agent; across agents, one author-week per agent. Themes asked in fewer than 2 author-weeks, and requests that share no theme, are not shown. Examples: up to 3 posts per theme from different authors, without slurs, preferring posts of 60 to 450 characters, most recent first.

| Rank | Request | Author-weeks | Posts | Agents (author-weeks) |
|---|---|---|---|---|
| 1 | Stabilize and fix the desktop app | 29 | 30 | OpenAI Codex 10, OpenCode 6, Cline 4, Google Antigravity 2, Claude Code 2, Cursor 2, Devin 2, Factory 1 |
| 2 | Fix memory leaks and reduce RAM usage | 28 | 31 | OpenAI Codex 12, Google Antigravity 3, Claude Code 3, Cursor 3, Devin 3, OpenCode 2, Conductor 1, Zed 1 |
| 3 | Fix app hanging and freezing mid-task | 26 | 32 | OpenAI Codex 11, Cursor 6, Claude Code 2, Devin 2, Pi 2, Google Antigravity 1, OpenCode 1, Warp 1 |
| 4 | Fix laggy UI and slow app performance | 24 | 25 | OpenAI Codex 17, Cursor 4, Google Antigravity 1, Claude Code 1, Devin 1 |
| 5 | Fix frequent app and CLI crashes | 18 | 18 | OpenAI Codex 6, Google Antigravity 3, Cline 3, Cursor 3, OpenCode 2, Devin 1 |
| 6 | Lower CPU, GPU and battery usage | 17 | 17 | OpenAI Codex 8, Google Antigravity 2, Amp 1, Claude Code 1, Cursor 1, Factory 1, OpenCode 1, Warp 1, Zed 1 |
| 7 | Fix failed and malformed tool call execution | 16 | 16 | OpenAI Codex 6, OpenCode 5, Amp 1, Google Antigravity 1, Claude Code 1, Cursor 1, Warp 1 |
| 8 | Fix sessions stopping or timing out mid-task | 16 | 16 | OpenAI Codex 4, Cursor 3, Cline 2, OpenCode 2, Google Antigravity 1, Claude Code 1, Conductor 1, Devin 1, Kiro 1 |
| 9 | Fix app failing to launch or blank window | 15 | 15 | OpenAI Codex 12, Amp 1, Google Antigravity 1, Claude Code 1 |
| 10 | Clean up leaked helper processes after sessions | 13 | 14 | OpenAI Codex 7, Claude Code 2, Cursor 2, OpenCode 2 |
| 11 | Fix connection drops and websocket disconnects | 12 | 14 | OpenAI Codex 4, Claude Code 2, Cline 2, OpenCode 2, Google Antigravity 1, Devin 1 |
| 12 | Fix CLI bugs and stability | 12 | 12 | OpenAI Codex 5, Factory 3, Google Antigravity 1, Claude Code 1, Cline 1, Cursor 1 |

### 1. Stabilize and fix the desktop app

- OpenAI Codex, 2026-09-27, r/ClaudeCode (Reddit): “claude desktop is crap. i wish they took care of their desktop app like openai with codex” [source](https://www.reddit.com/r/ClaudeCode/comments/1wrhtyd/new_to_claude_why_should_we_use_claude_code_over/pcckr78/)
- OpenAI Codex, 2026-09-26, r/codex (Reddit): “same problem after updating today! i'm so frustrated i thought i was the only one. escalated to openai support but i really hope the team notices this soon because now i can't work. well i can use cli but i want my desktop app man... version <phone_number>.0, win 11 25h2 <strict_link>” [source](https://www.reddit.com/r/codex/comments/1wqfiiw/chatgpt_desktop_app_gets_stuck_loading_forever/pc4d5tr/)
- OpenAI Codex, 2026-09-21, X search: OpenAI Codex, Codex CLI, Codex app (X): “@tokengremlin can't wait for this one! maybe bel help them to fix codex app ;)” [source](https://twitter.com/1674056173790679042/status/2102064065614844220)

### 2. Fix memory leaks and reduce RAM usage

- Cursor, 2026-09-26, r/cursor (Reddit): “cursor is highly unoptimised. team needs to now work on optimisation too. it just keeps on filling ram and ssd on its own and then ooms or crashes.” [source](https://www.reddit.com/r/cursor/comments/1wqppuh/grok_is_shutting_down_apps_now/pc5yr3x/)
- Google Antigravity, 2026-09-24, @antigravity (X): “@probiex007 @antigravity @probiex007 @antigravity yeah bro, for 28 gb it better write the whole codebase itself 😅 skipping till they patch this” [source](https://twitter.com/1728403486373789696/status/2103221298763837909)
- Google Antigravity, 2026-09-24, @antigravity (X): “@antigravity pls work on memory management. ide is consuming alot of memory idk why 🙁 <strict_link>” [source](https://twitter.com/1195014217771864064/status/2103213000517939705)

### 3. Fix app hanging and freezing mid-task

- OpenAI Codex, 2026-09-27, X search: OpenAI Codex, Codex CLI, Codex app (X): “@openai codex keeps hanging on mac 27.2 beta during context compaction based on my observation. i have already restarted frozen codex 2 dozen times today.” [source](https://twitter.com/304497770/status/2104052196916768981)
- Claude Code, 2026-09-25, @ClaudeDevs (X): “@claudedevs still lags and freezes my machine. cli is better” [source](https://twitter.com/44595437/status/2103303119799210260)
- Cursor, 2026-09-24, @cursor_ai (X): “i'm working in @cursor_ai sometimes running 6 threads at a time. i have to restart the app atleast 10 times a day because the threads are in some infinite loading state also it logs me out when i quit the app thankfully the tasks restart where they left out, but i wish i didn't have to restart the app 10x a day” [source](https://twitter.com/993923490746073088/status/2103104373324709893)

### 4. Fix laggy UI and slow app performance

- OpenAI Codex, 2026-09-27, X search: OpenAI Codex, Codex CLI, Codex app (X): “@thsottiaux @jorilallo the codex app is so laggy now please fix” [source](https://twitter.com/2026063233543462913/status/2104057177396727860)
- OpenAI Codex, 2026-09-27, X search: OpenAI Codex, Codex CLI, Codex app (X): “my codex app on linux has been trying to load chats for at least 15 mins when before it was just a few seconds, still hasn't even loaded smaller chats where there aren't lots of messages. @thsottiaux” [source](https://twitter.com/1674075739895877632/status/2104023351060566107)
- OpenAI Codex, 2026-09-26, r/codex (Reddit): “and with the latest update with the vertical left sidebar, on my m3 imac with 24 gb memory just panning the mouse over the profile menu at the bottom left it is laggy as you mouse over each menu item...” [source](https://www.reddit.com/r/codex/comments/1wqmgm3/what_an_absolute_chonker/pc5o053/)

### 5. Fix frequent app and CLI crashes

- Cursor, 2026-09-23, @cursor_ai (X): “@cursor_ai chill on the rollouts, the files tab has been crashing cursor full-screen for a month <strict_link>” [source](https://twitter.com/2464774904/status/2102872782103302216)
- OpenCode, 2026-09-21, r/opencode (Reddit): “i was excited trying out mimo but this is a disaster for me. every time it's using glob or grep, it crashes or freezes and when it does i have a huge bump in context, from 100k to 300k just like that, without explanation. using opencode tui, haven't tested on other harness yet. what's your experience so far?” [source](https://www.reddit.com/r/opencode/comments/1wmsl2p/having_a_bad_experience_with_mimo_26_flash/)
- Cline, 2026-09-19, @cline (X): “@cline cline 3.0.62 cli is dead on arrival on apple silicon. the bundled darwin-arm64 binary has a broken signature, so the kernel sigkills it at launch. both `cline --version` and `cline --help` print nothing and exit 137. fresh `npm install -g cline` on macos 27, arm64, node 26.7.0.” [source](https://twitter.com/1937709229198172160/status/2101359440402518298)

### 6. Lower CPU, GPU and battery usage

- Zed, 2026-09-25, @zeddotdev (X): “@zeddotdev zed on an old linux laptop is great, but damn the gpu usage shoots up and the fans go full throttle like a jet taking off” [source](https://twitter.com/58782925/status/2103314417576473007)
- Google Antigravity, 2026-09-24, @antigravity (X): “@androidstudio @antigravity this is great, next i wish as didn’t use 700% of my cpu while building or reopening the project . devin / cursor do not have this issue” [source](https://twitter.com/380573269/status/2103194074811629948)
- OpenAI Codex, 2026-09-23, r/codex (Reddit): “i am running debian 13 and vs codium. it does not happens randomly. i am killing the process everytime it reaches 19% cpu. even idle. updated today and same” [source](https://www.reddit.com/r/codex/comments/1suvg6s/has_anyone_noticed_codex_in_vs_code_using_high/pbgy7oh/)

### 7. Fix failed and malformed tool call execution

- OpenAI Codex, 2026-09-25, r/codex (Reddit): “or they could use all that funding to build a decent harness so the user isn't left burning tokens for no reason... somehow claude has the brains to make sure their harness works correctly before a major release. i'm sure astra is great, but codex is pure garbage. damage is done. go look at tibo's x comments lol” [source](https://www.reddit.com/r/codex/comments/1w7x57n/before_blaming_gpt6_astra_read_its_prompting_guide/pbzds5e/)
- OpenCode, 2026-09-22, r/opencode (Reddit): “agreed, i have a bad experience on using the free mimo 2.6 flash. it does a lot of failed tool calling, and sometimes it just kept looping. looks like a harness issue, but other models work fine” [source](https://www.reddit.com/r/opencode/comments/1wmsl2p/having_a_bad_experience_with_mimo_26_flash/pbb0h87/)
- OpenCode, 2026-09-22, @opencode (X): “@opencode fix your tool calls... the idea behind your service is amazing, but harness often gets crazy with the tool calls... mimo just crashed my pc by doing 800+ searches on my pc in a row... then after the restart i've asked it not to do that and it did exactly the same thing... :d” [source](https://twitter.com/996858699812626432/status/2102271295819858277)

### 8. Fix sessions stopping or timing out mid-task

- Kiro, 2026-09-27, r/kiroIDE (Reddit): “kiro-cli timed out a few times yesterday. i lost two heavy work sessions. all i found is a ticket so it appears it does not exist (<strict_link>) how do people check downtime? does kiro do rate limiting if you are doing heavy work? it felt slow.” [source](https://www.reddit.com/r/kiroIDE/comments/1wrsp5y/kiro_status_page/)
- OpenCode, 2026-09-27, @opencode (X): “@opencode @opencode fix your terminal, long running terminal tasks block the chat , they should not block the chat, sending a message terminates the terminal” [source](https://twitter.com/1272168992120045568/status/2104314185320653174)
- OpenCode, 2026-09-26, @opencode (X): “@opencode don't know where to go for official support. with opencode go (but i guess with zen as well?) a streaming response times out after 10 minutes. this makes it impossible to use glm 5.3 flash because it just keeps reasoning for over 10 minutes.” [source](https://twitter.com/3351906549/status/2103783728271040607)

### 9. Fix app failing to launch or blank window

- OpenAI Codex, 2026-09-27, r/codex (Reddit): “windows desktop app. unfortunately as a windows user i feel we are often neglected. super buggy - and it's been since the latest release that the bug is high impact (stuck at loading spinner unless we kill codex.exe). even then, codex works but not chatgpt. sometimes i use codex cli (with wsl) for windows, but i mostly use the desktop app. i hope they pay attention to users that are on other os other than macos.” [source](https://www.reddit.com/r/codex/comments/1wra4ev/codex_on_windows_do_you_work_with_wsl_agent/pcc70fw/)
- OpenAI Codex, 2026-09-27, X search: OpenAI Codex, Codex CLI, Codex app (X): “i am not able to open @openai codex, it is just showing me this loading and then error have tried reinstalling also still same can anyone help here? <strict_link>” [source](https://twitter.com/1153665686524137478/status/2104153434954064227)
- OpenAI Codex, 2026-09-27, X search: OpenAI Codex, Codex CLI, Codex app (X): “the codex desktop application has stopped launching. it seems that i need to fix the codex app with codex cli...” [source](https://twitter.com/2074467637095186432/status/2104003860893495489)

### 10. Clean up leaked helper processes after sessions

- OpenCode, 2026-09-25, r/opencode (Reddit): “bun keeps uploading data at 100% of my bandwidth. i closed all sessions, turned off all plugins and mcps. yet no luck. even after closing opencode, bun keeps running and uploading until i kill it in task manager!” [source](https://www.reddit.com/r/opencode/comments/1wpvj70/bun_is_eating_my_bandwidth_even_after_i_close/)
- OpenCode, 2026-09-20, @opencode (X): “@geoffreyhuntley @opencode @herdrdev opencode is great when it doesn't grind my machine to a halt from spawning all those headless processes.” [source](https://twitter.com/1118175955/status/2101534508290130216)
- OpenAI Codex, 2026-09-16, X search: OpenAI Codex, Codex CLI, Codex app (X): “bug report for the codex team: the chatgpt macos app's codex app-server leaks a skycomputeruseclient + node_repl process pair every ~5 min and never reaps them. after 4h i had 381 pairs, 14.6 gb ram, ~1,850 processes on a 16 gb m4. xcode builds hung until i killed them. @thsottiaux @openaidevs” [source](https://twitter.com/32409020/status/2100218556885721318)

### 11. Fix connection drops and websocket disconnects

- Cline, 2026-09-26, @cline (X): “cline desktop windows new version.37 , fails to connect error code 1006 @cline <strict_link>” [source](https://twitter.com/49906954/status/2103683897502618075)
- Google Antigravity, 2026-09-26, @antigravity (X): “@antigravity hey, we don't know what happened, but your models no, it's slow and easy to disconnect.please resolve this problem.” [source](https://twitter.com/1501685529389445123/status/2103656571603497051)
- Cline, 2026-09-25, r/CLine (Reddit): “the desktop app is too buggy to be used - it’s keeps on disconnecting and the error message is cryptic saying tauri invoked failed for get desktop backend endpoint : desktop backend endpoint not ready. i don’t think it’s usable if i keep losing the session and have debug the ide instead of coding” [source](https://www.reddit.com/r/CLine/comments/1wposc5/buggy_desktop_apps/)

### 12. Fix CLI bugs and stability

- Google Antigravity, 2026-09-23, @antigravity (X): “@rodydavis @antigravity bruh, fix your cli, ide and your models first.” [source](https://twitter.com/1991081729956945920/status/2102738975748481322)
- Factory, 2026-09-19, @droid (X): “add chatgpt sign-in with official plugin for byok, make cli faster and less buggy. communication. especially in github issues, bug reports and new feature requests. i before used droid more but currently using claude code, opencode and pi. if claude didn’t ban, i could just use opencode and pi depending on case” [source](https://twitter.com/1031311555/status/2101343501250163065)
- Factory, 2026-09-19, @droid (X): “@droid i had been using droid a bit and idk why, the tui is so buggy, idk if you guys fix it” [source](https://twitter.com/1482670613051625474/status/2101334639453577399)

## Every agent

| Agent | Overall rank | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|---|
| [Zed](https://feedbackbench.com/agents/zed.md) | =8 | Better than peers | 0.599 | 0.535–0.659 | 50 | 11 | 39 |
| [Pi](https://feedbackbench.com/agents/pi.md) | 7 | Better than peers | 0.580 | 0.519–0.633 | 46 | 10 | 36 |
| [Claude Code](https://feedbackbench.com/agents/claude-code.md) | 1 | Better than peers | 0.563 | 0.508–0.611 | 269 | 30 | 239 |
| [OpenCode](https://feedbackbench.com/agents/opencode.md) | 3 | Typical | 0.558 | 0.496–0.611 | 327 | 32 | 295 |
| [GitHub Copilot](https://feedbackbench.com/agents/copilot.md) | =8 | Typical | 0.507 | 0.465–0.551 | 30 | 3 | 27 |
| [Devin](https://feedbackbench.com/agents/devin.md) | =5 | Typical | 0.497 | 0.446–0.557 | 50 | 3 | 47 |
| [Cline](https://feedbackbench.com/agents/cline.md) | =8 | Typical | 0.491 | 0.434–0.546 | 57 | 3 | 54 |
| [Cursor](https://feedbackbench.com/agents/cursor.md) | 4 | Typical | 0.481 | 0.414–0.544 | 200 | 12 | 188 |
| [Google Antigravity](https://feedbackbench.com/agents/antigravity.md) | =5 | Worse than peers | 0.434 | 0.368–0.498 | 228 | 10 | 218 |
| [OpenAI Codex](https://feedbackbench.com/agents/codex.md) | 2 | Worse than peers | 0.410 | 0.358–0.458 | 883 | 41 | 842 |
| [Amp](https://feedbackbench.com/agents/amp.md) | =11 | Too few posts | – | – | 24 | 1 | 23 |
| [Kiro](https://feedbackbench.com/agents/kiro.md) | 13 | Too few posts | – | – | 12 | 0 | 12 |
| [Factory](https://feedbackbench.com/agents/factory.md) | =11 | Too few posts | – | – | 9 | 0 | 9 |
| [Conductor](https://feedbackbench.com/agents/conductor.md) | 14 | Too few posts | – | – | 9 | 0 | 9 |
| [Warp](https://feedbackbench.com/agents/warp.md) | 15 | Too few posts | – | – | 9 | 0 | 9 |
| [Grok Build](https://feedbackbench.com/agents/grok-build.md) | 16 | Too few posts | – | – | 0 | 0 | 0 |
| [Augment Code](https://feedbackbench.com/agents/augment.md) | 17 | Too few posts | – | – | 0 | 0 | 0 |

## Posts

Receipts rule: The 5 most recent praise and complaint posts per area (first 700 characters) and 3 per criterion (first 450 characters).

### Zed

- Praise, 2026-09-27, @zeddotdev (X): “@0xprajwal_ i use @zeddotdev, btw. it’s fast, lightweight, consumes less memory, and uses the gpu for rendering.” [source](https://twitter.com/1542127649463898113/status/2104217257153110457)
- Praise, 2026-09-25, @zeddotdev (X): “@zeddotdev my 16 gb ram development laptop is saved because of you!” [source](https://twitter.com/1115272367154950151/status/2103318390798774306)
- Praise, 2026-09-18, r/ZedEditor (Reddit): “any base to this claim? it has been rock solid for me, not to mention excellent performance” [source](https://www.reddit.com/r/ZedEditor/comments/1wjp5nq/popular_zed_fork_where_the_community_is_more/pakkq3d/)
- Complaint, 2026-09-27, @zeddotdev (X): “opened @zeddotdev after a while what is this? only happens wiht opencode acp <strict_link>” [source](https://twitter.com/851365565201514498/status/2104122682895921447)
- Complaint, 2026-09-26, r/ZedEditor (Reddit): “tried, but not reliable for everyday use. got some weird issues, saying it can't find neovim, but it works fine in other terminals.” [source](https://www.reddit.com/r/ZedEditor/comments/1wpr20d/is_there_any_zed_terminal/pc539bm/)
- Complaint, 2026-09-26, r/ZedEditor (Reddit): “there is always a node process running when zed is open. i would like to set it up to use bun. i don't even have node installed.” [source](https://www.reddit.com/r/ZedEditor/comments/1wosnp7/zed_still_silently_downloads_binaries_after_two/pc96lwi/)

### Pi

- Praise, 2026-09-26, r/PiCodingAgent (Reddit): “haha time will tell on my intentions, sorry for being salty. negativity only breeds so shouldn't contribute to it. also btw will be fixed in 0.2.1. pig now runs pi's argument normalization (optional nulls, typebox conversion, json schema coercion) before validation, and all of pi's validation tests are ported. thanks again, for the pointers.” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wqe2v1/meet_pig_the_pi_coding_harness_that_is_yours_but/pc9hklu/)
- Praise, 2026-09-22, r/opencode (Reddit): “i consistently use luna for coding tasks. in fact, i rarely use luna max these days. the thinking level i use most often for luna is "high." luna high is actually sufficient to handle the vast majority of modification requests. max simply takes longer without necessarily yielding proportionally better results; in fact, excessive reasoning can sometimes lead to unnecessary extra work (such as "smartly" adding "safety" protocols you didn't ask” [source](https://www.reddit.com/r/opencode/comments/1wm61m3/the_opencode_go_path_is_not_the_right_one_there/pbafplf/)
- Praise, 2026-09-20, r/LocalLLaMA (Reddit): “yeah that was a pita at first. but easily something you can fix with the underltying pi-agent config that dsh uses. set timeout to 5 minutes and 1000 retries, so the harness keeps retrying until the gpu is back up.” [source](https://www.reddit.com/r/LocalLLaMA/comments/1wloora/the_bear_can_dance_qwen_38_27b_on_one_3090_for_3/pb1k0q2/)
- Complaint, 2026-09-26, r/PiCodingAgent (Reddit): “every time i turn on my mac and see so many node processes, it’s really terrible. i think i‘ll try.” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wqe2v1/meet_pig_the_pi_coding_harness_that_is_yours_but/pc4svol/)
- Complaint, 2026-09-26, r/PiCodingAgent (Reddit): “would be nice if it worked. reality is that it doesn't tried it with stock settings coloring is bugged out, stock \`read\` tool is failing with jsonschema validation so it's just yet another sloppily coded port from one language to another with no real effort put into the maintenance apart from tons of tokens from the company leeched” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wqe2v1/meet_pig_the_pi_coding_harness_that_is_yours_but/pc9bfwt/)
- Complaint, 2026-09-26, r/PiCodingAgent (Reddit): “i'm not the claiming that i've been working on something for a year where the very foundational tool that harness needs to be able to work fails every query but at least it's faster to start, amirite and no, i'm not gonna contribute to your hardfork if you don't bother with testing the very basics” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wqe2v1/meet_pig_the_pi_coding_harness_that_is_yours_but/pc9duyq/)

### Claude Code

- Praise, 2026-09-27, r/ClaudeCode (Reddit): “it used to be faster and less buggy, but i very recently switched to the desktop app completely as i don't see any remaining speed or performance issues, and the experience is far better for me.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wrhtyd/new_to_claude_why_should_we_use_claude_code_over/pccjlcy/)
- Praise, 2026-09-27, r/kiroIDE (Reddit): “i work for amazon and is somewhat “strongly recommended” to use kiro. i still use claude code at work and at home. so much better … (auto classifier, transcript details, integrated tooling, open source tooling, cli features, general stability, model fallback, sub agents control, etc …)” [source](https://www.reddit.com/r/kiroIDE/comments/1wqueq0/claude_pro_vs_kiro_pro_subscription_which_is/pccqcjf/)
- Praise, 2026-09-26, r/ClaudeCode (Reddit): “yup. also i prefer claude harness because my underpowered laptop can't handle electron apps well” [source](https://www.reddit.com/r/ClaudeCode/comments/1wqvwry/just_resubscribed_to_max_after_a_long_time_and/pc7y15k/)
- Complaint, 2026-09-27, r/ClaudeCode (Reddit): “omg stop! i have it controlling a 6 axis robot arm and it says that all day after a crash” [source](https://www.reddit.com/r/ClaudeCode/comments/1wquoxy/fable_51_live_vehicle_diagnostics/pcanovv/)
- Complaint, 2026-09-27, r/ClaudeCode (Reddit): “the "latest" caveat hit me just now; gave the new go phrase, but then followed it up with something else before controller finished launching the authed workflow and it had to stop and say "now that's the last typed message and it doesn't carry the right authorization" 🫠 tested the workflow provenance var - 0 and false definitely don't work. an empty-string theoretically would work, but even when my node and anthropic's bun setup were both regist” [source](https://www.reddit.com/r/ClaudeCode/comments/1wr8cw1/claude_code_workflow_harness_change/pcbtb90/)
- Complaint, 2026-09-27, r/ClaudeCode (Reddit): “yes but if you use it for extensively watch for buildup of processes tying up your pc/ laptop and overheating it! i walked in and heard my fan seriously working overtime and cpu at 125%. apparently it kept spinning up more and more helper processes” [source](https://www.reddit.com/r/ClaudeCode/comments/1wls9ul/would_you_actually_use_claude_code_from_your/pcc2733/)

### OpenCode

- Praise, 2026-09-26, r/opencode (Reddit): “i had been running them in parallel (different laptops, projects) and i’ve switched entirely to v2 now. in my experience, v2 has been more stable, and with far fewer annoyance level bugs, than v1 for at least a couple weeks now.” [source](https://www.reddit.com/r/opencode/comments/1wqppth/opencode_v2/pc6ub9p/)
- Praise, 2026-09-25, r/opencodeCLI (Reddit): “i am using opencode inside paseo. never had much of an issue. it wraps the cli.” [source](https://www.reddit.com/r/opencodeCLI/comments/1wpta1k/opencode_free_models_in_your_own_harness_again/pbyim7n/)
- Praise, 2026-09-22, r/opencode (Reddit): “i consistently use luna for coding tasks. in fact, i rarely use luna max these days. the thinking level i use most often for luna is "high." luna high is actually sufficient to handle the vast majority of modification requests. max simply takes longer without necessarily yielding proportionally better results; in fact, excessive reasoning can sometimes lead to unnecessary extra work (such as "smartly" adding "safety" protocols you didn't ask” [source](https://www.reddit.com/r/opencode/comments/1wm61m3/the_opencode_go_path_is_not_the_right_one_there/pbafplf/)
- Complaint, 2026-09-27, r/opencodeCLI (Reddit): “forget vscode + extension, if you want real performance go for wu which is rust (a fork of zed without ai modules) and terminal running opencode in a panel, vscode and extensions give you an overhead of electron and hundreds of megabytes or more than 1 gigabyte of memory versus the 180 or 200 mb of memory that wu consumes [<strict_link> <strict_link>” [source](https://www.reddit.com/r/opencodeCLI/comments/1wrgmfe/whats_the_deal_with_openchamber/pcd1qha/)
- Complaint, 2026-09-27, r/opencode (Reddit): “for the past 6 months, i have 2 ways of running coding agents: either on host, or in docker for auto mode. after not being able to find a solution for a lightweight sandbox that is a middle ground between the two so i can run opencode on host in auto mode, i decided to just create one myself with "bwrap", the same sandbox solution used by claude code and flatpak. here is my repo: <strict_link> note that i could never get "npm i -g bachsofttrick/” [source](https://www.reddit.com/r/opencode/comments/1wrp6xx/i_created_a_sandbox_hook_for_opencode_with_bwrap/)
- Complaint, 2026-09-27, r/codex (Reddit): “i have gpt and opencode go, ppl went crazy today, opencode go is lagging entire day, probably ppl just migrated and testing? (gpt has some real server problems), ehh idea to go for claude again... i feel like we all were there before” [source](https://www.reddit.com/r/codex/comments/1wrw0sg/is_this_proof_were_getting_a_powerful_new_model/pcgo1b1/)

### GitHub Copilot

- Praise, 2026-09-26, r/PiCodingAgent (Reddit): “also incase you're interested in the background and how the slop progresses and hardens into slightly less slop. sorry for the spam, last response i swear :) \`\`\` it's the fix-tool-arg-coercion lane's own reproduction of the bug, the failing test it writes first, not an old warning we ignored. the other read failures in the logs are deliberate error-path tests (for example path: \[\]). rca-ca from the logs: \- our own pig sessions: about 4,” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wqe2v1/meet_pig_the_pi_coding_harness_that_is_yours_but/pc9j62b/)
- Praise, 2026-09-16, r/GithubCopilot (Reddit): “because of this issue, i stopped using \`#\` and switched to \`@\` as all other agents use at-sign anyway.. works fine, and my muscle memory is interchangeable between copilot, codex and claude. but srsly, i use all three of them daily, and copilot still offers best ux in general: diffs, files attachment, tools selection & usage etc. like in claude, i still have no idea how to set 'auto' mode in one of my workspaces. it just keeps sliding back to” [source](https://www.reddit.com/r/GithubCopilot/comments/1wcc1c4/i_m_losing_12hrs_of_life_expectancy_everytime_i/pa4qguf/)
- Praise, 2026-09-02, r/opencode (Reddit): “u/trovebloxian i don't fully understand the reason yet. however, according to my initial findings, dns resolver timeouts start occurring on the router around 11 am (around the time i noticed something strange on the network.), increasing from 2-4 per hour to 25-30 per hour. from 1 pm onwards (the times when services started reporting "unhealthy"), dns queries from my machine (the one running opencode) reach up to 600 per minute. naturally, dns re” [source](https://www.reddit.com/r/opencode/comments/1w5j4vf/goodbye_opencode/p7foht1/)
- Complaint, 2026-09-27, r/GithubCopilot (Reddit): “why would you wanna use copilot harness? it keeps making mistakes, ignore instructions and repeatedly fails read/write ops. another person in this sub posted benchmarks that put it lower than other harnesses like pi/ds/codex” [source](https://www.reddit.com/r/GithubCopilot/comments/1wr5io2/best_cheaper_alternative/pca0pix/)
- Complaint, 2026-09-26, r/GithubCopilot (Reddit): “i’m really new, so bear with me being perhaps a noob. i’ve got github copilot + and visual studio on a macbook that seems to get stuck on “run in terminal” i used it before for a few months so not expert but not totally new and this is first time it’s getting that problem it has run and done similar projects to what i’m doing now before i think but maybe i changed something i see some activity definitely when i run top bash command i’ve tried” [source](https://www.reddit.com/r/GithubCopilot/comments/1wr56qh/stuck_on_run_in_terminal/)
- Complaint, 2026-09-25, r/GithubCopilot (Reddit): “it's crazy buggy. even the integration with vscode is worse.” [source](https://www.reddit.com/r/GithubCopilot/comments/1wpctlc/vs_code_chat_users_local_or_copilot_harness/pbvujk7/)

### Devin

- Praise, 2026-09-16, @DevinAI (X): “@dennisonbertram @devinai @zeddotdev been ok for me! just zed and a mishmash of homegrown cli tools, eg <strict_link>” [source](https://twitter.com/896906084014845952/status/2100305572943835614)
- Praise, 2026-09-11, @cognition (X): “@topitopongsalac @cognition its fixed, during the start it was an issue” [source](https://twitter.com/1957706668034387968/status/2098427101573726667)
- Praise, 2026-09-08, r/windsurf (Reddit): “just thought i'd post this as an example of recovery; i'm using glm. devin errored, i didn't check the details, but it looked worth flagging so i stopped what glm was doing, which i always do if things are going not where i think they should. i flagged up the issue and it handled it well, which is what it generally seems to do. other ide's probably do too, but i appreciate how things don't fall apart. \--- **there was a json error when you were w” [source](https://www.reddit.com/r/windsurf/comments/1wasbyy/devin_recovery_example/)
- Complaint, 2026-09-27, r/CognitionLabs (Reddit): “looks like there is a bug with the devin tab. im using a m4 macbook air 24gb. the problem started with golden gate update. i asked claude opus 5.5 and after debugging it found this: you're right, it isn't normal. i found the code path responsible, and it's a performance bug inside devin's built-in extension, not something in your setup. **where the 2.2s goes** (newest profile, `exthost-b13005.cpuprofile`, 2240 ms): * 95% of the time is inside the” [source](https://www.reddit.com/r/CognitionLabs/comments/1w23p1r/devin_is_simply_too_slow/pcfa29z/)
- Complaint, 2026-09-23, @cognition (X): “i love devin but wtf is this memory consumption? @cognition @dabit3 <strict_link>” [source](https://twitter.com/1333116608609398785/status/2102571264183242968)
- Complaint, 2026-09-23, @cognition (X): “bug report for @cognition devin: subagents chew through ram. i tried to process a batch of files in parallel, and the session dies almost immediately, every time. swe-2 diagnosed oom and pretty much pinned it on the subagents. the same work is fine single-threaded. the machine only has 16gb, sure. still feels like something that can be fixed😃😃😃” [source](https://twitter.com/2065820231470399488/status/2102799952967557239)

### Cline

- Praise, 2026-09-23, r/opencodeCLI (Reddit): “anyone else is getting opencode mimo2.6 to just fail and stop the session after few responses? if i switch to cline api mimo2.6, it keeps going properly until the task is delivered.” [source](https://www.reddit.com/r/opencodeCLI/comments/1wmozts/mimov26pro_debuts_as_the_top_open_weights_model/pbjealn/)
- Praise, 2026-09-22, @cline (X): “you must try @cline desktop app. models like deepseek v4.1 flash, kimi k3, muse spark 1.3 and glm 5.3 flash can be used for free i am using it since few days and it is pretty fast and efficient also it doesn't eats a lot of ram <strict_link>” [source](https://twitter.com/1758147983584161792/status/2102370281792905468)
- Praise, 2026-09-17, @cline (X): “@adityawaslost @cline working perfectly good <strict_link>” [source](https://twitter.com/1485557841880829953/status/2100484677220110748)
- Complaint, 2026-09-27, r/CLine (Reddit): “windows 11, both computers. and it happens with every model, free and clinepass models. it appears randomly.” [source](https://www.reddit.com/r/CLine/comments/1wrvji9/why_does_cline_agent_keep_showing_this_stupid/pcgewmm/)
- Complaint, 2026-09-27, @cline (X): “@flowvsgravity @cline i really don't think it's glm flash. pixel canary is really slow, throws errors every single time, and is so unstable that you can barely even use it right now.” [source](https://twitter.com/2083446628174663680/status/2104116725952188714)
- Complaint, 2026-09-27, r/LocalLLaMA (Reddit): “qwen3.8-flash-next 177b nvfp4(119gib): ssd streaming at 9-10 tok/s on one 16 gb rtx 5060 ti + 32 gb ram we built an inference engine for moe models that don't fit in vram + ram. most of the model stays on the ssd, and experts are read as tokens need them. this started as a proof of concept, and poc worked, we are getting 9-10 tok/s decode on qwen3.8-flash-next nvfp4 (9.06 on the benchmark turn, 10.4 on the best turn). this is just the start. with” [source](https://www.reddit.com/r/LocalLLaMA/comments/1wrxap8/qwen38flashnext_177b_nvfp4119gib_ssd_streaming_at/)

### Cursor

- Praise, 2026-09-27, r/cursor (Reddit): “i’m running an m1 air w 16gb and it runs like a breeze. maybe it’s the actual workload you have it running?” [source](https://www.reddit.com/r/cursor/comments/1wmj0pw/does_cursor_make_anyone_elses_computer_extremely/pcdw03l/)
- Praise, 2026-09-25, r/GoogleAntigravityIDE (Reddit): “i got a cursor ultra from a reseller for essentially nothing compared to what i used to pay . it works flawless. since 2 months no issues yet. now i can't stop thinking about their business model. if resellers can sell it this cheap and still profit, how are they acquiring these accounts at scale?” [source](https://www.reddit.com/r/GoogleAntigravityIDE/comments/1wq9cd3/cursor_ultra_for_18_instead_of_200_seriously_how/)
- Praise, 2026-09-24, @cursor_ai (X): “@the_codewala pretty sire some of my @cursor_ai conversations set to auto land there. and so far it didn’t fail.” [source](https://twitter.com/223391610/status/2103218907930939401)
- Complaint, 2026-09-27, r/cursor (Reddit): “cursor hanging on taking longer than expected after the shell already finished is the same false busy lie as waiting for subagent. i kill that agent pane first and reopen the folder so the host actually resets. full app restart helps less than clearing the hung session. if the spinner comes back on the next prompt the host is sick not the model.” [source](https://www.reddit.com/r/cursor/comments/1wquvoq/taking_longer_than_expected/pccfywa/)
- Complaint, 2026-09-27, @cursor_ai (X): “@grok @bot @cursor_ai been there, done that. nothing in the computer to approve. no tasks running. nothing. we're beyond all that. the bot can now message other bots but cant start a chat. can only reply.” [source](https://twitter.com/1420831171282362374/status/2104346415216689525)
- Complaint, 2026-09-26, r/cursor (Reddit): “cursor is highly unoptimised. team needs to now work on optimisation too. it just keeps on filling ram and ssd on its own and then ooms or crashes.” [source](https://www.reddit.com/r/cursor/comments/1wqppuh/grok_is_shutting_down_apps_now/pc5yr3x/)

### Google Antigravity

- Praise, 2026-09-24, r/LocalLLaMA (Reddit): “> antigravity ide doesn't support itself well too many errors since the release of antigravity 2 i've never had an issue with it. i started using it when they released 3.8 flash and it's been one of the most consistent clients i've used. in terms of performance the only issue i have with 3.8 flash is if i change the context of what it's working on too many times. it will lose the thread and go dumb. i cleaned that up by building a bunch of ded” [source](https://www.reddit.com/r/LocalLLaMA/comments/1wof9kk/introducing_support_for_local_ai_models_in_the/pbomwq8/)
- Praise, 2026-09-21, r/google_antigravity (Reddit): “that's not high usage data, it's like default” [source](https://www.reddit.com/r/google_antigravity/comments/1wm8vt2/high_data_usage_issue/pb4wrcm/)
- Praise, 2026-09-17, r/google_antigravity (Reddit): “3.8 flash (high) confirmed working normal on my machine.” [source](https://www.reddit.com/r/google_antigravity/comments/1wiptxc/is_it_slow_again_or_am_i_just_being_paranoid/pac8sjh/)
- Complaint, 2026-09-27, r/google_antigravity (Reddit): “well i've met a lot of desktop app users who come to my discord server and start screaming that antigravity is shit. then i made them use cli and they've always been happy since. i think the main problem people hate with the gui is the insane token burn and the gui talking up all the resources of your computer. that's why i switched in the first place to get rid of the insane token burn. plus the cli usually only has bugs and doesn't normally exp” [source](https://www.reddit.com/r/google_antigravity/comments/1wqmd4n/why_is_the_antigravitycli_so_underrated/pcamrq4/)
- Complaint, 2026-09-27, r/google_antigravity (Reddit): “i see. i was on linux, not sure if that's the factor. i was on 1.2.11, updated to 1.2.12 still hitting it with \`agy --dangerously-skip-permissions remote-control start\` put \`toolpermission\` and \`permissions\` in both .gemini/config/config.json and .gemini/antigravity-cli/settings.json (not even sure why they have two directory and two different files, maybe gemini 3.8 flash hallucinate when it was troubleshooting it). thanks for discussing” [source](https://www.reddit.com/r/google_antigravity/comments/1wqmd4n/why_is_the_antigravitycli_so_underrated/pccxjbm/)
- Complaint, 2026-09-27, r/google_antigravity (Reddit): “was wondering same thing yesterday when i switched from agy ide to vs code (because of instability of agy ide and google not updating this product anymore). but i guess we'll have to do with the default (copilot) autocomplete for now.” [source](https://www.reddit.com/r/google_antigravity/comments/1wpw81p/will_tab_autocomplete_be_released_to_the/pcdwa7h/)

### OpenAI Codex

- Praise, 2026-09-27, r/codex (Reddit): “that worked after a restart, thank you!” [source](https://www.reddit.com/r/codex/comments/1wq86dk/codex_not_working_today/pcdfqvb/)
- Praise, 2026-09-27, X search: OpenAI Codex, Codex CLI, Codex app (X): “i opened the codex cli without the screen, and it worked well! i should have done it this way.” [source](https://twitter.com/227971543/status/2104175340176466310)
- Praise, 2026-09-24, r/codex (Reddit): “runs just fine for me lol” [source](https://www.reddit.com/r/codex/comments/1woxxj2/this_needs_more_attention/pbrkks9/)
- Complaint, 2026-09-27, r/codex (Reddit): “highly recommend sticking with that...fun times for the last 48 hours smfh matter of fact, last week or more. haven't been able to send 2 prompts (1, re-log, 1, re-log, etc) now can't even open the app..what an absolute joke <strict_link>” [source](https://www.reddit.com/r/codex/comments/1wr69nb/see_you_soon_guys_probably/pca1i4f/)
- Complaint, 2026-09-27, r/codex (Reddit): “desktop windows works really bad, constantly gets stuck, sometimes refuses to send the messages, and now it simply doesnt load. when it decides not to work, i use cli. so half the time cli half the time desktop” [source](https://www.reddit.com/r/codex/comments/1wra4ev/codex_on_windows_do_you_work_with_wsl_agent/pcay1g6/)
- Complaint, 2026-09-27, r/codex (Reddit): “the daemon is a real nightmare. moved the rest of my projects over to wsl finally just because i was so annoyed. i also think arrow hotkeys are annoying. the worse is alt+up. like what the hell were they thinking?” [source](https://www.reddit.com/r/codex/comments/1wr9e2j/for_agents_is_a_bad_feature_and_codex_cli_is/pcayvay/)

### Amp

- Praise, 2026-09-02, @AmpCode (X): “@sqs @ampcode updated and restarted amp, restarted 1password, and now it works like a charm 💯” [source](https://twitter.com/1729720291/status/2095096517602271408)
- Complaint, 2026-09-23, @AmpCode (X): “@sqs @shavsycle @ampcode since here, can you also check iterm2 nested scroll issue? there's 7mo old post on reddit about this. for me the bug is repeatable with starting new puck thread.” [source](https://twitter.com/1089515309122367489/status/2102847742557196497)
- Complaint, 2026-09-20, @AmpCode (X): “@sqs @ampcode the ask question tool works great on the web ui, but is buggy on the tui. i selected the answer, it didn't reach to the model, but then i opened the web, and then chose the same option there, it worked... have a look pls :)” [source](https://twitter.com/1957706668034387968/status/2101695639709167693)
- Complaint, 2026-09-17, @AmpCode (X): “@sqs @nothingrutvik @davidmansaray @ampcode i dont know if i am doing smt wrong here but it's giving me some errors... mind hoping into dm's real quick? tried using puck but didnt got a sucess this time” [source](https://twitter.com/2931128860/status/2100563947384422528)

### Kiro

- Complaint, 2026-09-24, r/kiroIDE (Reddit): “fyi - still not working.” [source](https://www.reddit.com/r/kiroIDE/comments/1wp1sbg/kiro_subscription_payment/pbus63m/)
- Complaint, 2026-09-24, r/kiroIDE (Reddit): “cool! gonna try tomorrow. i prefer the vs lsp integration, debugger and interface. kiro only have the free debugger right now and somehow electron-like programs perform much worse then vs in my machine.” [source](https://www.reddit.com/r/kiroIDE/comments/1wp53oc/i_added_native_kiro_support_to_visual_studio_2026/pbvamrb/)
- Complaint, 2026-09-23, r/kiroIDE (Reddit): “yes, use cli rather than ide as ide is very memory hungry” [source](https://www.reddit.com/r/kiroIDE/comments/1wo0mu2/kiro_vs_claude_code_any_difference_worth_knowing/pbk51y1/)

### Factory

- Complaint, 2026-09-23, @FactoryAI (X): “@harrystuck77 @factoryai @droid @badlogicgames @ampcode @cursor_ai @opencode @anthropicai @claudeai @cognition @xai @grok @build can you explain your reasoning? what tasks do you usually give to it using missions? i stopped using droid when it's using too much memory on my servers and it's not worth it.” [source](https://twitter.com/2064081082308599808/status/2102629361249915213)
- Complaint, 2026-09-21, @droid (X): “@ain3sh @benvargas @droid hey a fellow droid user here, the cli is amazing, the app not so much, there were some bugs here and there. only thing with the cli is the rendering can be improved, it's a bit janky on missions, maybe you guys can move to a different framework from ink?” [source](https://twitter.com/1333161864/status/2101883259831632365)
- Complaint, 2026-09-19, @droid (X): “add chatgpt sign-in with official plugin for byok, make cli faster and less buggy. communication. especially in github issues, bug reports and new feature requests. i before used droid more but currently using claude code, opencode and pi. if claude didn’t ban, i could just use opencode and pi depending on case” [source](https://twitter.com/1031311555/status/2101343501250163065)

### Conductor

- Complaint, 2026-09-21, @conductor_build (X): “@charlieholtz i have no idea whether its an issue with @cursor_ai or @conductor_build but cursor agents sometimes they just stop responding in conductor cloud. it has happened to me thrice in the past 3 days. please figure it out and fix it? i am happy to provide any details for you to debug if needed.” [source](https://twitter.com/1900337293564817408/status/2101825344626208846)
- Complaint, 2026-09-15, r/conductorbuild (Reddit): “agreed. about 50% of my new tabs just spin for a bit, then give me the the ... option to fork into a new tab. it's getting very very very tedious to deal with.” [source](https://www.reddit.com/r/conductorbuild/comments/1wdm8hc/conductor_regressions_are_making_it_harder_to/p9y4n50/)
- Complaint, 2026-09-15, @conductor_build (X): “@conductor_build i'm experiencing some weird behavior, where my window gets black several times when i try to click on it. usually happens when i close conductor, or it crashes, and i open it again. keep getting black several times before stabilizing.” [source](https://twitter.com/2068260048199925760/status/2099818314922897499)

### Warp

- Complaint, 2026-09-19, @warpdotdev (X): “@warpdotdev windows app has unexpected bugs which i have lost count. now i am unable to drag it on the machine to move the window. but warp customer support doesn’t care. how can a company go so pathetic after open sourcing their codebase 🤦🏽♂️” [source](https://twitter.com/1624472822503583744/status/2101339256786677876)
- Complaint, 2026-09-15, @warpdotdev (X): “@vikvang1 @warpdotdev try windows. every app and huge games work fine on this gaming laptop except warp.” [source](https://twitter.com/1624472822503583744/status/2099848972106138049)
- Complaint, 2026-09-14, @warpdotdev (X): “@michael_kove @catalinmpit @warpdotdev moved to ghostty as well and pi as harness. my old intel mac is happy again” [source](https://twitter.com/25074228/status/2099386334116733106)
