# Reviewing and approving the agent's changes (`verify.change_review_ui`)

Feedback Bench, coding agents, built 2026-10-01, window 2026-08-31 to 2026-09-27. Web page: https://feedbackbench.com/#/criterion/verify.change_review_ui

Area: [Checking and finishing](https://feedbackbench.com/criteria/checking.md)

**Definition.** Diff views, per-file approval, edit-acceptance prompts, and how manageable the amount of change is for a human reviewer.

**Boundary.** Not this: see [IDE and editor integration](https://feedbackbench.com/criteria/setup.ide_integration.md) for general editor features. Not this: see [Length and clarity of replies, summaries and comments](https://feedbackbench.com/criteria/work.response_verbosity.md) for prose length.

Rated author-weeks, all agents: 323. Complaint share: 62%.

## The brief

Written by Claude Opus 5.5 from 67 labelled posts and the numbers on this page. Interpretation, not measurement: every quote is verbatim and links to its post.

**Agents write faster than humans can read, and review tools lag**

TL;DR:

- The core complaint is volume. Agents produce more change than one reviewer can absorb.
- Diff panels that compare the wrong base or drift out of sync break trust.
- OpenAI Codex is the only agent rated worse than peers. Its terminal review draws the most heat.

In plain terms: Users finish an agent run facing a pile of edits. The diff panel may compare the wrong branch, and there is no clean way to comment back. The people who cope shrink patches on purpose.

### How it breaks

- **Change volume outruns human review** ([Reviewing and approving the agent's changes](https://feedbackbench.com/criteria/verify.change_review_ui.md)). The worst failure is not a broken button. Agents generate more diff than a person can read, so review quietly becomes skimming or skipping.
  Users describe waking up to days of review after one long run. Some admit they no longer read teammates' large agent PRs and test behaviour instead. Others flag a subtler risk. A change can pass tests and still encode a decision nobody approved. The fix users report is procedural, not product. They ask for smaller patches, scope the work before it runs, and add gates that reject oversized diffs without a risk note.
  Evidence:
  - Complaint, Cursor, @cursor_ai, 2026-09-12: “kicked off a @cursor_ai /goal last night and woke up this morning overwhelmed with the amount of changes it made. it's going to take days to review.” [source](https://twitter.com/16858549/status/2098835728063226150)
  - Complaint, Claude Code, r/ChatGPTCoding, 2026-09-01: “for me, the biggest problem has become review. codex and claude code can produce changes so quickly that i sometimes become the bottleneck. a change can be technically reasonable and pass all the tests, but still not be what i actually intended. what’s helped me is making the process more explicit: keep the change small, review the scope before implementation, let the agent work, verify the result, then either revise it or accept it. the part that makes me most uncomfortable is when an agent makes a perfectly reasonable decision that i never actually approved. so ai has made implementation much faster for me, but it has also made “what exactly are we agreeing to change?” much more important.” [source](https://www.reddit.com/r/ChatGPTCoding/comments/1w045is/ai_coding_has_made_me_dramatically_faster_but_im/p76558d/)
  - Praise, Cursor, r/cursor, 2026-09-24: “this matches what i see too. generation is cheap compared to reading a noisy diff. i started asking for smaller patches on purpose and review got way faster. the bottleneck is attention, not tokens.” [source](https://www.reddit.com/r/cursor/comments/1woujtl/i_timed_agent_diff_review_for_5_days_generation/pbq7sca/)
  - Praise, Cursor, r/cursor, 2026-09-04: “we added a review gate that rejects diffs over 200 lines or more than 8 files unless the agent includes a risk note; it stopped “tests pass” from being the whole review. do you also block agent-generated tests from counting as evidence?” [source](https://www.reddit.com/r/cursor/comments/1w76958/agent_mode_quietly_made_me_a_worse_code_reviewer/p7tuuny/)

- **Diffs that compare the wrong thing** ([Reviewing and approving the agent's changes](https://feedbackbench.com/criteria/verify.change_review_ui.md)). A diff view is only useful if it is correct, and users report panels showing the wrong base, missing changes, or wildly inflated counts.
  Claude Code desktop users say the diff tab compares against main instead of the current branch, or misses changes made on it. An Amp user reports the changes tab on a large repo jumping to tens of thousands of phantom changes. Google Antigravity users point to diff bugs in the extension after the agent finishes. Each case pushes people back to raw git, which defeats the point of a built-in reviewer.
  Evidence:
  - Complaint, Claude Code, @ClaudeDevs, 2026-09-03: “@claudedevs when will the diff feature in claude desktop be fixed so that it correctly shows the changes in the current branch instead of comparing against the main branch?” [source](https://twitter.com/3036525114/status/2095514513823137904)
  - Complaint, Amp, @AmpCode, 2026-09-25: “@sqs @ampcode the changes tab on very large repos gets weirdly out of sync showing like 80k+ changes or something. sometimes running git pull or other commands fix it, other times get worse. i think it needs some way to refresh on the ui. could also be comparing wrong commit” [source](https://twitter.com/1857935142670450688/status/2103610446720942394)
  - Complaint, Google Antigravity, r/google_antigravity, 2026-09-05: “yeah there are bugs in extension related to diff after the agent as done its work hopefully they will solve this in the next update” [source](https://www.reddit.com/r/google_antigravity/comments/1w7vqfk/ag_extension_causing_git_errors/p7ymaj8/)
  - Praise, Cursor, r/ClaudeCode, 2026-08-31: “ok its been weeks now, i have already refreshed to latest claude code desktop version. i make changes to a branch and i cannot see in the diff tab all the changes it has made? with codex and cursor i can clearly see my diffs. i know i can use git but i want to see my file changes in the right side panel the ux is not intuitive surely i am not the only one? why is cc not prioritising this is a most have for any production code.” [source](https://www.reddit.com/r/ClaudeCode/comments/1w33h7n/why_is_claude_code_not_fixing_the_code_changes/)

- **Terminal-only review feels like a downgrade** ([Reviewing and approving the agent's changes](https://feedbackbench.com/criteria/verify.change_review_ui.md)). When the only diff lives in a terminal or a tool log, users leave the agent and open an editor just to read the changes.
  The single most requested fix here is a full panel showing every changed file, asked for in 17 author-weeks. Cline users say changes hide behind line-number clicks. Codex CLI users ask why verification needs VS Code or Zed. Even fans of terminal agents say a terminal is a rough place to read diffs. A related ask is LSP navigation inside review, since a bare git diff gives no way to jump to definitions.
  Evidence:
  - Complaint, Cline, @cline, 2026-09-20: “@cline right now, i can only see the code changes by clicking the line numbers, which shows additions/deletions in the top-right. a proper code editor with a clear diff view would make the agent workflow much better.” [source](https://twitter.com/1874710292577292288/status/2101765049316421672)
  - Praise, OpenAI Codex, X search: OpenAI Codex, Codex CLI, Codex app, 2026-09-25: “@theo switched back yesterday. opus 5.5 needs way less correcting and the usage feels generous. what i miss is the codex app for reviewing changes. a terminal is still a rough place to read diffs.” [source](https://twitter.com/1773624715547922432/status/2103391407570624592)
  - Complaint, OpenAI Codex, X search: OpenAI Codex, Codex CLI, Codex app, 2026-09-10: “why does it require the use of a gui application like vs code or zed for code verification even though i'm using codex cli?” [source](https://twitter.com/15722742/status/2097870755564503381)
  - Complaint, OpenCode, @opencode, 2026-09-25: “@badlogicgames i think this is something which is missing from all ai tools like @opencode desktop and @ampcode i want to review/read the code with lsp and code navigation. all of them just shows git diff only” [source](https://twitter.com/1158785224299335680/status/2103409717959705080)

- **Approval gates missing or unreliable** ([Reviewing and approving the agent's changes](https://feedbackbench.com/criteria/verify.change_review_ui.md)). Per-file and per-hunk accept or reject is what careful reviewers want, and they notice fast when it disappears or misbehaves.
  Users praise line-level accept and reject because unapproved hunks can go back to the agent for rework. Complaints cluster where that control vanished. Devin users miss inline proposed changes in notebooks. GitHub Copilot users report a keep-edits prompt that resurfaces later, where either choice can overwrite newer work or remove everything. A gate that cannot be trusted is worse than none, because it invites the wrong click.
  Evidence:
  - Complaint, GitHub Copilot, r/GithubCopilot, 2026-09-24: “local. not having the ability to approve the changed files is unacceptable. if i need to do something across multiple repos, i use the github copilot app” [source](https://www.reddit.com/r/GithubCopilot/comments/1wpctlc/vs_code_chat_users_local_or_copilot_harness/pbvangl/)
  - Complaint, GitHub Copilot, r/GithubCopilot, 2026-09-02: “vs code chat is full of frustrating bugs. chats randomly disappear, and deleted chats sometimes come back. the edit count is often wrong, and some chats just appear and disappear while vs code is open. the worst part is the **keep edits** issue. you can click “keep,” continue editing the same files multiple times, then reopen the chat later and see the “keep edits” prompt again. at that point, it’s impossible to know what to do: clicking “keep” can overwrite your newer changes, while clicking “no” can remove everything. and with all these serious bugs, what are the developers focused on? adding chat backgrounds. total trash!” [source](https://www.reddit.com/r/GithubCopilot/comments/1tfkawl/vs_code_silently_loses_all_your_copilot_chat/p7f1bpn/)
  - Complaint, Devin, r/windsurf, 2026-09-10: “is there any chance we can get this feature in devin? cascade handle it wonderfully, but now devin does not. the notebook\_edit tool was the feature that showed proposed changes in line in a jupyter notebook. this was incredibly useful when running large notebooks. now devin can only directly modify a cell, without showing me the proposed changes. makes tracking the changes so much harder now.” [source](https://www.reddit.com/r/windsurf/comments/1wc55gv/notebook_edit_too_in_devin_local/)
  - Praise, Google Antigravity, r/google_antigravity, 2026-09-02: “ide actually gives better results and can be guided around better, and being able to accept and reject individual lines is important if you plan on actually reading and maintaining your code. plus, i think you can leave some changes unaccepted and ask if to fix it refactor what you haven't yet approved. ide just needs remote control functionality and pause/resume functionality to be perfect. oh, and it would be nice if changing the model in one window didn't change it in all of them. some of us juggle more than one workspace at a time.” [source](https://www.reddit.com/r/google_antigravity/comments/1w52ypl/am_i_the_only_one_who_prefers_the_ide_over/p7c6o0b/)

- **Reviews that talk back to the agent** ([Reviewing and approving the agent's changes](https://feedbackbench.com/criteria/verify.change_review_ui.md)). The most praised pattern turns the diff into a conversation, with comments pinned to exact lines and sent straight back to the agent.
  Users like pointing at a precise line in the real git change and firing it back without switching to an editor. Linking each change to the thread that produced it makes agent work reviewable instead of opaque. Amp's built-in diff commenter and Zed's diff feedback both get credit. Where commenting is absent, users call it the missing feature of the moment and ask for batched inline comments.
  Evidence:
  - Complaint, Zed, @zeddotdev, 2026-09-25: “@zeddotdev you **really** need a review feature though (with commenting) in this day &amp; age…” [source](https://twitter.com/1394007551348461570/status/2103372080511066375)
  - Praise, Claude Code, r/ClaudeCode, 2026-09-10: “yep, that exact-line handoff is the part i didn’t want to compromise on. i want the real git change beside the agent, point at exactly what i care about, fire it back, keep moving. no editor pilgrimage required 😄 and good shout on the plugin catalog, i’ll take a look.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wcn50x/i_got_tired_of_staring_at_claude_code_in_the/p8zdnk9/)
  - Praise, Amp, @AmpCode, 2026-09-03: “@jtaby @sethmills21 we use @ampcode, which has solutions to both: * their cloud agents can rpc to a local mac to build and report back (we do this for our ios app) * they have a built-in diff viewer/commenter and multiplayer for other team members other cloud agents might have this solved too?” [source](https://twitter.com/1183203638/status/2095596826544259359)
  - Praise, Zed, @zeddotdev, 2026-09-16: “@zeddotdev linking every code change back to the conversation is the key detail. it makes agent work reviewable instead of magical, and multiplayer could speed up the review loop. curious to see how this feels on larger repos.” [source](https://twitter.com/60312892/status/2100296365515448392)

### Who stands out

- **OpenAI Codex (weaker)**. The only agent rated worse than peers here, with complaints centred on reading and approving changes from the CLI.
  Users call its diff hard to work with and say the CLI lacks an approval flow, easy revert, or file browser. Several route review through VS Code instead. A recurring theme is that fast output plus thin review tooling nudges people toward accepting code they have not read. The bright spot is in-editor and desktop use, where users say it shows every change and lets them approve one by one or all at once.
  Evidence:
  - Praise, GitHub Copilot, r/GithubCopilot, 2026-09-10: “i use copilot in work and codex at home. i prefer copilot in vscode. i like copying symbols into the chat for the ais context and codex diff is beyond terrible. but i could be using it wrong so take that with a pinch of salt. to keep the credits i use the cheaper model, it's not too much different from the higher models” [source](https://www.reddit.com/r/GithubCopilot/comments/1wcj4ce/i_cancelled_my_github_copilot_subscription_are/p914kn1/)
  - Complaint, Cursor, r/cursor, 2026-09-13: “i am in the exact same position. but i just tried codex and it’s so not intuitive to work with. no file browser, no approval process, no way to easily revert, no way to discuss, plan and execute on md plans and tec specs. the idea to use codex inside cursor is not a bad one. is this flawless and smooth? plus a $20 subscription from cursor gives us a second model to bounce back ideas. grok medium probably is perfect for this. if we use codex api inside cursor do their models appear to be selected from the model list? or how does it work? did you find any video or article about this?” [source](https://www.reddit.com/r/cursor/comments/1wfkoba/codex_vs_cursor/p9njlhp/)
  - Complaint, OpenAI Codex, r/codex, 2026-09-13: “fast mode is a trap because unless you're making slop or having the ai do something like fuzzy data matching, you should be reviewing the output and its a lot harder to do that if its churning away at high speed.” [source](https://www.reddit.com/r/codex/comments/1wfilwz/i_was_running_astra_on_high_and_extra_high_used/p9mod8i/)
  - Praise, OpenAI Codex, r/codex, 2026-08-31: “if you want to manually review all the code it generates it's easiest to use codex from within your editor, e.g. vscode. it will show you all the changes it made and you can approve them one by one or all at once. there's no need to generate a diff and then apply it as two steps.” [source](https://www.reddit.com/r/codex/comments/1w3duyc/is_there_a_local_way_to_apply_a_codexproposed_diff/p6zmdo4/)

- **Zed (mixed)**. Users love that the diff sits beside the agent thread as a review guide, then hit a wall when they need branch-level comparison.
  Praise targets the shape. Changes link back to the conversation, and review happens in something smarter than a terminal. One user made it their main AI coding app for the review process alone. Complaints are about depth. Zed users lead the requests for diffing against a branch or commit, built-in PR review, and commenting. Others want the whole file behind the diff toggle, and some find the delta view confusing.
  Evidence:
  - Praise, Zed, @zeddotdev, 2026-09-11: “@zeddotdev agent turning the diff into a review guide beside the thread is the right shape.” [source](https://twitter.com/1569824026444599297/status/2098491713518002438)
  - Complaint, Zed, @zeddotdev, 2026-09-25: “@zeddotdev wish these pls · compare all changed files against a commit/branch/revision in one multi-file diff view · open a file’s history and diff two versions, or compare an old version with my working tree · make these commands so we can bind our own shortcuts” [source](https://twitter.com/2543890370/status/2103369091142803785)
  - Complaint, Zed, @zeddotdev, 2026-09-10: “@zeddotdev in the agentic era people need solid git panel and good review capabilities. unfortunately zed lacks it, saying as a zed user” [source](https://twitter.com/2060822033387139076/status/2097970958854136069)
  - Complaint, Zed, @zeddotdev, 2026-09-25: “@zeddotdev behind the toggle...allow us to view the whole file..not just the changed code” [source](https://twitter.com/1364105804996087809/status/2103379339643584547)

- **Cursor (mixed)**. Often named as the place where diffs are easy to see, yet its autonomous modes reproduce the same review flood users complain about elsewhere.
  Users switching from other agents cite Cursor as the reference for a clear side-panel diff, and some praise smooth PR review on the go. The in-loop editing style keeps reviewers close to the code. The friction shows up in the hands-off features. Users describe projects that hide plans, PRs, and diffs, and long goals that leave days of review. A stacked-branch diff is a specific ask.
  Evidence:
  - Praise, Cursor, @cursor_ai, 2026-09-02: “in addition to the super @grok heavy benefits, @bot can deploy subagents on @cursor_ai, and then allows me to review/merge pr’s quite easily! with the app running on the vm and @tailscale setup, i can test things on the go! <strict_link> <strict_link>” [source](https://twitter.com/34196099/status/2095216084572164588)
  - Praise, Cursor, r/ClaudeCode, 2026-08-31: “ok its been weeks now, i have already refreshed to latest claude code desktop version. i make changes to a branch and i cannot see in the diff tab all the changes it has made? with codex and cursor i can clearly see my diffs. i know i can use git but i want to see my file changes in the right side panel the ux is not intuitive surely i am not the only one? why is cc not prioritising this is a most have for any production code.” [source](https://www.reddit.com/r/ClaudeCode/comments/1w33h7n/why_is_claude_code_not_fixing_the_code_changes/)
  - Complaint, Cursor, @cursor_ai, 2026-09-18: “gave @cursor_ai projects an honest shot with a multi pr undertaking for 3 days. here's the review: great idea, crappy execution. abstracts away very important details, very hard to see the code, no clarity on associated artefacts (plans, prs, diffs), no worktree management, no subagent statuses. got me to ship slop i eventually had to clean up manually. it's really addicting to keep sending messages and not worry too much - the end result looks good enough and the minimal communication from orchestrator makes it sound smarter than it is. someone should build a well thought out version of it.” [source](https://twitter.com/1248167246771261440/status/2100826154559234359)
  - Complaint, Cursor, @cursor_ai, 2026-09-03: “feature request for @cursor_ai: in the changes view, let me see the diff between the current branch and its parent in a graphite stack. right now you can see the diff between two commits but you can't see the diff between 2 branches. this would be very helpful! @poteto @spacexai” [source](https://twitter.com/1455554960616214530/status/2095311289530650678)

- **Claude Code (mixed)**. The loudest voice on this page, with real enthusiasm for line-level handoff and a stubborn complaint about its desktop diff tab.
  Users welcome keeping the patch visible while the agent keeps working, saying scope drift becomes obvious before a final scroll hunt. Exact-line feedback sent back to the agent wins praise. Against that, users say the desktop diff tab compares the wrong base or misses branch changes. Shell edits arrive for approval without syntax highlighting. Handoff reports covering plan, tests, and changed files are a frequent request.
  Evidence:
  - Praise, Claude Code, @ClaudeDevs, 2026-09-11: “@claudedevs my second screen would stay on the diff. keeping the patch visible while claude works makes scope drift obvious before the final review turns into a giant scroll hunt.” [source](https://twitter.com/828440649535873025/status/2098303419589280064)
  - Praise, Claude Code, @ClaudeDevs, 2026-09-11: “@claudedevs this fixes the actual friction point — reviewing a diff used to mean staring at the same window claude needed to keep working in.” [source](https://twitter.com/1748396781405106176/status/2098310525926944975)
  - Complaint, Claude Code, r/ClaudeCode, 2026-09-07: “bash/shell scripts are one more indirection from the user's point of view. you have to parse the quotes, escapes and shell code around the actual code if meant to find or edit. even with syntax highlighting, which is missing when you are to approve those, but present after you've approved, you have more displayed than necessary, a distraction from the code.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wa2ezw/anthropic_push_to_default_auto_mode_coincide_with/p8fhle2/)
  - Complaint, Claude Code, r/ClaudeCode, 2026-09-15: “i don't/can't. i do like to check the data and structures in the db as that is the most important thing imo. other than that i mostly rely on testing thoroughly by doing it myself and spinning up agents to test via playwright. there is no way i can review my teammates 64 file change pr, too much velocity is expected from us.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wgrtt6/do_yall_still_read_lines_of_code/pa1p9qm/)

### Fine print

- Most agents here have too few posts to rate, so their absence from the leaders says little.
- Many posts compare several agents, so a post tagged to one agent sometimes describes another's diff tool.

## Top requests

What users ask to add or change, most asked first. 117 author-weeks ask for something. Requests do not change the Feedback Score. Rule: A separate pass by Claude Sonnet 5 reads every counted post and extracts what the author asks the agent or its vendor to add or change, with the criteria it maps to and a short normalised wording; it does not touch the labels or the Feedback Score. Claude Opus 5.5 groups the wordings within each criterion (the first criterion the request maps to) into themes; code counts them. A theme counts distinct author-weeks that ask for it, per agent; across agents, one author-week per agent. Themes asked in fewer than 2 author-weeks, and requests that share no theme, are not shown. Examples: up to 3 posts per theme from different authors, without slurs, preferring posts of 60 to 450 characters, most recent first.

| Rank | Request | Author-weeks | Posts | Agents (author-weeks) |
|---|---|---|---|---|
| 1 | Full diff panel of all changed files | 17 | 17 | OpenAI Codex 4, Cursor 3, Pi 3, Claude Code 2, Google Antigravity 1, Cline 1, Devin 1, OpenCode 1, Zed 1 |
| 2 | Reliable, accurate, in-sync diff display | 12 | 14 | Google Antigravity 4, Claude Code 3, OpenAI Codex 2, Amp 1, Cline 1, Cursor 1 |
| 3 | Batched inline review comments sent to agent | 9 | 10 | Zed 4, Claude Code 3, Pi 2 |
| 4 | Diff against branch, commit, or stack parent | 9 | 9 | Zed 6, Cursor 2, Claude Code 1 |
| 5 | Handoff report of plan, tests, and changed files | 9 | 9 | Claude Code 5, GitHub Copilot 1, Cursor 1, Devin 1, Zed 1 |
| 6 | Per-file accept/reject of agent changes | 8 | 10 | Google Antigravity 3, Claude Code 2, GitHub Copilot 1, Cursor 1, OpenCode 1 |
| 7 | Preview and approve diffs before applying | 7 | 7 | Devin 2, Google Antigravity 1, Claude Code 1, OpenAI Codex 1, Cursor 1, Zed 1 |
| 8 | Built-in pull request review tools | 6 | 6 | Zed 4, GitHub Copilot 1, Cursor 1 |
| 9 | Code review with LSP navigation and full file | 4 | 4 | Amp 1, OpenAI Codex 1, OpenCode 1, Zed 1 |
| 10 | Change size metrics for reviewers | 3 | 3 | Claude Code 2, Devin 1 |
| 11 | Per-session execution logs and diffs | 3 | 3 | Claude Code 1, OpenAI Codex 1, Cursor 1 |
| 12 | Diff history across checkpoints and versions | 2 | 2 | OpenAI Codex 1, Zed 1 |

### 1. Full diff panel of all changed files

- Google Antigravity, 2026-09-26, r/google_antigravity (Reddit): “same. hard to track changes in vs code with the ag extension and alarmingly, the only good ide ag ide is now no longer showing changed files either. i must track it via git changes. well...” [source](https://www.reddit.com/r/google_antigravity/comments/1wqlqcg/bug_generated_file_changes_disappear_after/pc6vbqr/)
- Zed, 2026-09-25, @zeddotdev (X): “@zeddotdev good...please get the diff viewer something like vscode...i wanna migrate to zed” [source](https://twitter.com/1364105804996087809/status/2103379110026440950)
- Pi, 2026-09-22, @pidotdev (X): “@pidotdev the number of extension with "diff"/"review" in title or description. that's a clear signal to improve diff.” [source](https://twitter.com/618819434/status/2102497902077603913)

### 2. Reliable, accurate, in-sync diff display

- Amp, 2026-09-25, @AmpCode (X): “@sqs @ampcode the changes tab on very large repos gets weirdly out of sync showing like 80k+ changes or something. sometimes running git pull or other commands fix it, other times get worse. i think it needs some way to refresh on the ui. could also be comparing wrong commit” [source](https://twitter.com/1857935142670450688/status/2103610446720942394)
- Google Antigravity, 2026-09-23, r/google_antigravity (Reddit): “there has been quite some time but no change log on ide extension and its fundamental issues still not been fixed. when i go to the old chats, the changes that the last message has done are re-shown and re-applied and it fucks up my code as all my changes got removed because of this. also, i have to accept the changes after every turn or i cannot run the code itself as it shows both old and new code in the file itself duplicated” [source](https://www.reddit.com/r/google_antigravity/comments/1wnqoya/antigravity_2_release_v2160/pbifc4n/)
- Claude Code, 2026-09-18, @ClaudeDevs (X): “@bcherny @claudedevs @lydiahallie if you look through these 3 screenshots, this is what i mean, should have been more clear. there are more uncommitted changes, but they don't automatically refresh in the diff, you have to physically click refresh for them to show. would feel much better if i didn't have to. 😀 <strict_link>” [source](https://twitter.com/732783174166536192/status/2100754062433976749)

### 3. Batched inline review comments sent to agent

- Zed, 2026-09-25, @zeddotdev (X): “@zeddotdev you **really** need a review feature though (with commenting) in this day &amp; age…” [source](https://twitter.com/1394007551348461570/status/2103372080511066375)
- Claude Code, 2026-09-24, @ClaudeDevs (X): “claude's annotation game is shit. they should learn from codex. @claudedevs” [source](https://twitter.com/1437350362822836235/status/2103127339656003727)
- Zed, 2026-09-19, r/ZedEditor (Reddit): “okay, this looks really neat. any improved workflows to improve the agent thread/session to pr process? would also love to be able to batch add comments to a diff and have the agent work on them.” [source](https://www.reddit.com/r/ZedEditor/comments/1v74260/flint_a_terminalagentfocused_fork_of_zed/paq8my4/)

### 4. Diff against branch, commit, or stack parent

- Zed, 2026-09-25, @zeddotdev (X): “@zeddotdev wish these pls · compare all changed files against a commit/branch/revision in one multi-file diff view · open a file’s history and diff two versions, or compare an old version with my working tree · make these commands so we can bind our own shortcuts” [source](https://twitter.com/2543890370/status/2103369091142803785)
- Zed, 2026-09-25, @zeddotdev (X): “@zeddotdev can you make it easier to review worktrees/branches ? somehow there's no file picker/file browser when clicking view branch diff (worktree/branch a -&gt; main). everything is in a single clunky "changed since main" tab :/” [source](https://twitter.com/24510559/status/2103369080975876486)
- Zed, 2026-09-24, @zeddotdev (X): “@zeddotdev i need to be able to see changes from my feature branch to develop branch. as github pr diff view shows it. is it too hard to build?” [source](https://twitter.com/2060822033387139076/status/2103076253519478903)

### 5. Handoff report of plan, tests, and changed files

- Claude Code, 2026-09-26, r/ClaudeCode (Reddit): “exactly. i want to know exactly what lines it changed, where it changed, how it tested and is that the right test. i could have the 4.6 explain me it's choice and decisions simply. opus 5....not so much. this is why i felt unproductive or slow. i had to slam my head against the table and ask it 100 times.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wqzt2x/opus_55_experience_of_an_engineer_at_big_tech/pc94atn/)
- Zed, 2026-09-25, @zeddotdev (X): “@shadowfetch @zeddotdev a useful companion is a review mode that shows the task contract, changed files, and verification status beside the diff. less prompt chrome is great, but the trust signal is an explicit gate before merge.” [source](https://twitter.com/2099871292480421888/status/2103334568006947155)
- Claude Code, 2026-09-21, r/ClaudeCode (Reddit): “git diff in intellij. claude report of what files it plans to change before coding, match with what actually changed. the hardest one is when it does a task but in an unexpected way. i had claude design some screen layouts and put them in figma. it was looking pretty good. until i noticed they were images and not figma elements. claude had built a rasteriser and rendered the images beforehand in python before uploading them to figma lol” [source](https://www.reddit.com/r/ClaudeCode/comments/1wmh1ix/how_do_you_guys_know_if_claude_code_did_anything/pb7c9rg/)

### 6. Per-file accept/reject of agent changes

- Google Antigravity, 2026-09-25, r/google_antigravity (Reddit): “the extension retains the commands from agy 2.x, such as /boost. however, the review workflow in vs code is frustrating: change acceptance is strictly all-or-nothing, and files are not saved beforehand, which triggers compilation errors. they still need to refine the extension significantly before sunsetting their standalone ide.” [source](https://www.reddit.com/r/google_antigravity/comments/1wmxq14/antigravity_product_release_time/pbwcqhz/)
- GitHub Copilot, 2026-09-24, r/GithubCopilot (Reddit): “local. not having the ability to approve the changed files is unacceptable. if i need to do something across multiple repos, i use the github copilot app” [source](https://www.reddit.com/r/GithubCopilot/comments/1wpctlc/vs_code_chat_users_local_or_copilot_harness/pbvangl/)
- OpenCode, 2026-09-22, r/opencodeCLI (Reddit): “wondering you found any solotion for this? opencode just make me a blind vibe coder and i still prefer ghcp in vscode to see changes and then accept or reject them” [source](https://www.reddit.com/r/opencodeCLI/comments/1t9zdv7/how_can_i_view_diffs_and_acceptreject_changes/pbe4buz/)

### 7. Preview and approve diffs before applying

- Zed, 2026-09-25, @zeddotdev (X): “@zeddotdev can't use zed. i don't want to keep changing the repo to view where ai made the changes. hope zed add this soon. <strict_link>” [source](https://twitter.com/707623871847997440/status/2103436229702517011)
- Google Antigravity, 2026-09-22, r/google_antigravity (Reddit): “i did, and i feel like it's asking more questions for commands than the ide. i was used not to use them at all. maybe i didn't configure it the same way. it also doesn't give the inline diffs in the editor, just changes them automatically. i think i'll test some more when 3.8 doesn't burn all my tokens.” [source](https://www.reddit.com/r/google_antigravity/comments/1wn2xsg/token_usage_between_ide_and_extensions/pbcbh9s/)
- Devin, 2026-09-19, @cognition (X): “@brandon_galang @cognition @devinai the sidebar progress is doing more work than the harness. if the agent shows what it is about to run before it runs it, you review instead of debugging. most harnesses only show you what already broke.” [source](https://twitter.com/913700556253753345/status/2101304314082083282)

### 8. Built-in pull request review tools

- Zed, 2026-09-25, @zeddotdev (X): “@zeddotdev you have pull request / code review tools? #lazyweb” [source](https://twitter.com/1126321/status/2103297889003033061)
- Zed, 2026-09-17, @zeddotdev (X): “@zeddotdev when pr reviews inside zed? only thing keeping me from swithing to zed” [source](https://twitter.com/948542359557541888/status/2100698564175257736)
- Cursor, 2026-09-14, @cursor_ai (X): “@simonlind @cursor_ai @openai would an in-app pr view make it feel more polished?” [source](https://twitter.com/334714036/status/2099452971084067263)

### 9. Code review with LSP navigation and full file

- OpenCode, 2026-09-25, @opencode (X): “@badlogicgames i think this is something which is missing from all ai tools like @opencode desktop and @ampcode i want to review/read the code with lsp and code navigation. all of them just shows git diff only” [source](https://twitter.com/1158785224299335680/status/2103409717959705080)
- Zed, 2026-09-25, @zeddotdev (X): “@zeddotdev behind the toggle...allow us to view the whole file..not just the changed code” [source](https://twitter.com/1364105804996087809/status/2103379339643584547)
- OpenAI Codex, 2026-09-15, X search: OpenAI Codex, Codex CLI, Codex app (X): “@arpit_bhayani px0 sounds great but theres a reason i'm sticking to vscode, even codex app has a built-in diff viewer that has similar problems, i want to view large diffs without glitches, see types on symbol hover, go-to definition, see git blame, pr comments, edit diff” [source](https://twitter.com/1551268048694579200/status/2099982444275618088)

### 10. Change size metrics for reviewers

- Devin, 2026-09-22, @cognition (X): “@cognition over-scoping is the useful failure mode here. teams need the split by task type plus the extra files, edits, tests, and tokens the agent introduced, because a near-pass that expands the change surface can cost more to review than a clean miss.” [source](https://twitter.com/1029077130850660352/status/2102292711717933446)
- Claude Code, 2026-09-24, r/ClaudeCode (Reddit): “is there a way to not include docs in this diff counter” [source](https://www.reddit.com/r/ClaudeCode/comments/1wov62z/got_mogged_by_claude_opus/pbr78uk/)
- Claude Code, 2026-09-15, r/ClaudeCode (Reddit): “it's a novel idea, and the numbers show you can build. but as a maintainer, i wouldn't use it... **my bottleneck is time, not tokens.** writing the fix isn't the hard part for me. reviewing is. your ledger sends me more code from a stranger's agent to review, and i pay credit for it. if the pr is wrong, i pay twice... **trust...** maintainers are already drowning in plausible-but-wrong ai prs. with this, i'd have to read a stranger's pr like it c” [source](https://www.reddit.com/r/ClaudeCode/comments/1whc3pj/claude_code_as_the_lead_wrote_1150_commits_of_my/pa1c5ix/)

### 11. Per-session execution logs and diffs

- Claude Code, 2026-09-24, @ClaudeDevs (X): “@claudedevs 로컬 실행이 가능해질수록 프로젝트별 파일·셸·네트워크 권한을 기본 거부하고, 스레드별 diff와 실행 로그를 남겨야 병렬 작업이 편해져도 사고 범위를 통제할 수 있습니다.” [source](https://twitter.com/1619634971848892418/status/2102983921353052556)
- Cursor, 2026-09-15, @cursor_ai (X): “@cursor_ai project grouping can reduce context pollution, but it is not equivalent to being auditable. after task switching or failed retries, can the scope of changes, permissions, and costs be replayed with one click?” [source](https://twitter.com/2088076772931731456/status/2099698868824813844)
- OpenAI Codex, 2026-09-02, X search: OpenAI Codex, Codex CLI, Codex app (X): “@hude_icp by adding @codex at the beginning, it calls the codex cli, and the main usage is to consider discussions and check functions with claude in chat. the integration of features into mulmoclaude is yet to come. specifically, it is proposed to codex: codex execution logs codex file editing git diff approval/discard review” [source](https://twitter.com/96413942/status/2095124947907776589)

### 12. Diff history across checkpoints and versions

- Zed, 2026-09-25, @zeddotdev (X): “@zeddotdev wish these pls · compare all changed files against a commit/branch/revision in one multi-file diff view · open a file’s history and diff two versions, or compare an old version with my working tree · make these commands so we can bind our own shortcuts” [source](https://twitter.com/2543890370/status/2103369091142803785)
- OpenAI Codex, 2026-08-31, X search: OpenAI Codex, Codex CLI, Codex app (X): “@thsottiaux in the codex app i really need to see the diffs from the previous checkpoints, at least the snapshot... in a goal mode, it becomes impossible to review without git diffs.” [source](https://twitter.com/587759076/status/2094300039975719148)

## Every agent

| Agent | Overall rank | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|---|
| [Cursor](https://feedbackbench.com/agents/cursor.md) | 4 | Typical | 0.526 | 0.498–0.552 | 52 | 26 | 26 |
| [Zed](https://feedbackbench.com/agents/zed.md) | =8 | Typical | 0.517 | 0.494–0.542 | 37 | 20 | 17 |
| [Claude Code](https://feedbackbench.com/agents/claude-code.md) | 1 | Typical | 0.498 | 0.473–0.524 | 107 | 38 | 69 |
| [OpenAI Codex](https://feedbackbench.com/agents/codex.md) | 2 | Worse than peers | 0.474 | 0.452–0.497 | 45 | 11 | 34 |
| [Google Antigravity](https://feedbackbench.com/agents/antigravity.md) | =5 | Too few posts | – | – | 25 | 6 | 19 |
| [Devin](https://feedbackbench.com/agents/devin.md) | =5 | Too few posts | – | – | 21 | 3 | 18 |
| [OpenCode](https://feedbackbench.com/agents/opencode.md) | 3 | Too few posts | – | – | 12 | 7 | 5 |
| [GitHub Copilot](https://feedbackbench.com/agents/copilot.md) | =8 | Too few posts | – | – | 10 | 6 | 4 |
| [Pi](https://feedbackbench.com/agents/pi.md) | 7 | Too few posts | – | – | 4 | 1 | 3 |
| [Amp](https://feedbackbench.com/agents/amp.md) | =11 | Too few posts | – | – | 4 | 2 | 2 |
| [Cline](https://feedbackbench.com/agents/cline.md) | =8 | Too few posts | – | – | 2 | 0 | 2 |
| [Conductor](https://feedbackbench.com/agents/conductor.md) | 14 | Too few posts | – | – | 2 | 1 | 1 |
| [Factory](https://feedbackbench.com/agents/factory.md) | =11 | Too few posts | – | – | 1 | 0 | 1 |
| [Warp](https://feedbackbench.com/agents/warp.md) | 15 | Too few posts | – | – | 1 | 1 | 0 |
| [Kiro](https://feedbackbench.com/agents/kiro.md) | 13 | Too few posts | – | – | 0 | 0 | 0 |
| [Grok Build](https://feedbackbench.com/agents/grok-build.md) | 16 | Too few posts | – | – | 0 | 0 | 0 |
| [Augment Code](https://feedbackbench.com/agents/augment.md) | 17 | Too few posts | – | – | 0 | 0 | 0 |

## Posts

Receipts rule: The 5 most recent praise and complaint posts per area (first 700 characters) and 3 per criterion (first 450 characters).

### Cursor

- Praise, 2026-09-26, r/cursor (Reddit): “i'd switch. if the ui is what's slowing you down, that costs you more than the quota ever will, and cursor's multi-chat and diff review are genuinely nicer. just don't cancel codex yet: run cursor on your real repo for a few heavy days and watch the usage meter. if you're burning it on tiny edits, that's the workflow leaking, not the plan.” [source](https://www.reddit.com/r/cursor/comments/1wptv9n/thinking_of_switching_from_codex_to_cursor/pc4r5px/)
- Praise, 2026-09-25, r/cursor (Reddit): “hello there. so regarding your question, from experience, i have been subscribed with cursor for around two years, a yearly subscription. the usage limit actually is the best you will ever get. i will share with you a photo from my usage, so you can see that if you use the composer 2.5, you get around 2 billion tokens from my current workload, which is as a full-time developer working on multiple projects. it's more than enough. even now i'm tryi” [source](https://www.reddit.com/r/cursor/comments/1wptv9n/thinking_of_switching_from_codex_to_cursor/pbywqpt/)
- Praise, 2026-09-24, r/cursor (Reddit): “this matches what i see too. generation is cheap compared to reading a noisy diff. i started asking for smaller patches on purpose and review got way faster. the bottleneck is attention, not tokens.” [source](https://www.reddit.com/r/cursor/comments/1woujtl/i_timed_agent_diff_review_for_5_days_generation/pbq7sca/)
- Complaint, 2026-09-27, @cursor_ai (X): “@cjbell_ @cursor_ai agent branch commits hidden till pr is frustrating, i've hit that markdown plan viewer shuffle too” [source](https://twitter.com/184674873/status/2104338238085509151)
- Complaint, 2026-09-24, @cursor_ai (X): “@cursor_ai the monitoring plan is the easy part. the hard part is deciding which regression is worth rolling back vs shipping a hotfix while users already feel it.” [source](https://twitter.com/8611542/status/2103111364206383520)
- Complaint, 2026-09-23, @cursor_ai (X): “@kamellperry_ @cursor_ai cursor ate a weekend. monday was just reading the diff” [source](https://twitter.com/2029192070183960579/status/2102690653423448403)

### Zed

- Praise, 2026-09-27, @zeddotdev (X): “is this the pr replacement we’ve been waiting for in the agentic age? @zeddotdev team has already turned off pull requests on delta’s own repository. they’re building, reviewing and merging changes inside shared agent conversations instead. delta entered public beta on september 16. you spend an hour with an agent investigating a problem, ruling out approaches and working through the fix. then you open a pr and try to explain all that to someone” [source](https://twitter.com/36634050/status/2104270370635452902)
- Praise, 2026-09-27, @zeddotdev (X): “is this the pr replacement we’ve been waiting for in the agentic age? @zeddotdev team has already turned off pull requests on delta’s own repository. they’re building, reviewing and merging changes inside shared agent conversations instead. delta entered public beta on september 16. you spend an hour with an agent investigating a problem, ruling out approaches and working through the fix. then you open a pr and try to explain all that to someone” [source](https://twitter.com/36634050/status/2104272571055349948)
- Praise, 2026-09-25, r/ZedEditor (Reddit): “i work by myself most of the time and i really like the review process (probably you could get something similar with a skill), but i also get better usage there than on the codex app with my codex sub (probably context or cache), so it became my first ai coding app this week” [source](https://www.reddit.com/r/ZedEditor/comments/1wq03mv/has_anyone_tried_delta/pc17ujd/)
- Complaint, 2026-09-26, r/ZedEditor (Reddit): “i just tried it out for the first time and i don't really get it... the ui is not really intuitive and i have threads... subthreads... and so on. also need to pay extra for it and can't use my claude code subscription (yes thats anthropic who is blocking that) then there is the change panel who does show nothing.. beside the agent is already changing the code... edit: it did now show changes after a while... but the stranges thing is i don't see” [source](https://www.reddit.com/r/ZedEditor/comments/1wq03mv/has_anyone_tried_delta/pc3xwob/)
- Complaint, 2026-09-25, @zeddotdev (X): “@shadowfetch @zeddotdev a useful companion is a review mode that shows the task contract, changed files, and verification status beside the diff. less prompt chrome is great, but the trust signal is an explicit gate before merge.” [source](https://twitter.com/2099871292480421888/status/2103334568006947155)
- Complaint, 2026-09-25, @zeddotdev (X): “@zeddotdev fix the search pls. i stopped using bcz of search and diff viewer” [source](https://twitter.com/2065733203663659008/status/2103345319836815865)

### Claude Code

- Praise, 2026-09-26, r/ClaudeCode (Reddit): “i have it make human review tasks of what it did so i can review and give feedbakck. it litterally makes me checklists if what i need to do with step by strp instructions. doesn't mean i don't do other stuff to but it's actually quite important i do the things it tells me to do to check its work” [source](https://www.reddit.com/r/ClaudeCode/comments/1wqfne5/whats_the_strategy_to_understand_what_your_app_is/pc42adm/)
- Praise, 2026-09-26, r/ClaudeCode (Reddit): “yes, claude using its edit mode displays a diff. that's exactly why op prefers it over claude using python” [source](https://www.reddit.com/r/ClaudeCode/comments/1wqwi4h/why_claude_code_uses_python_for_everything_and/pc7xv5y/)
- Praise, 2026-09-26, @ClaudeDevs (X): “@claudedevs a visible review state turns “submitted” into an acceptance check, not a black box” [source](https://twitter.com/2078041339099578369/status/2103813247300165882)
- Complaint, 2026-09-26, r/ClaudeCode (Reddit): “i gave up to review code already as ai generated too much code already for me. instead i rely on testing and document. i review document which much easier to read, and let the ai keep sync well between code and document, which much more efficient than review code directly” [source](https://www.reddit.com/r/ClaudeCode/comments/1wpovtv/the_slow_collapse_of_code_reviews_how_do_you_deal/pc3st6g/)
- Complaint, 2026-09-26, r/ClaudeCode (Reddit): “this is why its better to build it step by step or if you do try to one shot something do it then start polishing it until you get familiar with it. waste of time to manually review code” [source](https://www.reddit.com/r/ClaudeCode/comments/1wqfne5/whats_the_strategy_to_understand_what_your_app_is/pc3tzm9/)
- Complaint, 2026-09-26, r/ClaudeCode (Reddit): “not reviewing code is insane to me.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wowilt/i_still_dont_understand_this_agentic_workflow/pc43gdb/)

### OpenAI Codex

- Praise, 2026-09-25, X search: OpenAI Codex, Codex CLI, Codex app (X): “@theo switched back yesterday. opus 5.5 needs way less correcting and the usage feels generous. what i miss is the codex app for reviewing changes. a terminal is still a rough place to read diffs.” [source](https://twitter.com/1773624715547922432/status/2103391407570624592)
- Praise, 2026-09-25, X search: OpenAI Codex, Codex CLI, Codex app (X): “using ai to modify code in the terminal, i crashed two tasks last night and wasted several hundred tokens. fortunately, today i tried the latest refactored codex cli, which has fixed both major issues. the two most annoying things about ai coding are: first, not being able to see the specific diff code changes, relying on luck when pressing enter; second, when there are too many background tasks, the token bill skyrockets. the latest version of c” [source](https://twitter.com/1805553619703578625/status/2103624715927794114)
- Praise, 2026-09-16, r/codex (Reddit): “what? you can inspect the files, open terminals etc inside the desktop app. the file change view is also much better.” [source](https://www.reddit.com/r/codex/comments/1wi7jaa/what_is_the_current_state_of_codex_cli_vs_desktop/pa8ccyw/)
- Complaint, 2026-09-26, r/codex (Reddit): “i wanted to try codex since claude code sometimes makes weird choices so i wanted to do a code review with codex. my current workflow is linux andvscode and claude code extensionnand its exactly what i want. i tried downloading codex on linux and i am confused. added current project that i work on that has 20+ files changed but not commited. codex bugged out in endless loading spinner with /review command and i cannot see the changes. also i hav” [source](https://www.reddit.com/r/codex/comments/1wqo1s6/confused_how_to_use_codex/)
- Complaint, 2026-09-25, X search: OpenAI Codex, Codex CLI, Codex app (X): “@moiiikaaa yep, full access ai model with codex cli reviews were not in a user friendly diff in codex environment for that i use vs code. also to test some new models i use open router extention, previous used codex extention in vs code but verbose wasn't much effective so switch to cli.” [source](https://twitter.com/1378862873003225090/status/2103332407315403183)
- Complaint, 2026-09-25, X search: OpenAI Codex, Codex CLI, Codex app (X): “@theo i've been running both for a month and honestly claude won my terminal while codex kept the ide. the thing i miss most from the codex app is the review ux, claude's diff review still feels like reading a receipt. but claude's planning mode is the part i can't give up anymore” [source](https://twitter.com/1085377722237546504/status/2103462497294463337)

### Google Antigravity

- Praise, 2026-09-22, r/google_antigravity (Reddit): “i still use the ide version. because i feel that the cli uses more token and because i prefer make little change by myself in the code instead of burning token for minor task. and its also easily to review de code .” [source](https://www.reddit.com/r/google_antigravity/comments/1wmi7qb/why_is_cli_being_used_by_most/pbdbhnf/)
- Praise, 2026-09-18, @antigravity (X): “@google @antigravity @googleaistudio harness updates are the boring part that actually matters. still leaving a human on the last pass.” [source](https://twitter.com/2095150665442164736/status/2100769146317222395)
- Praise, 2026-09-10, r/google_antigravity (Reddit): “i dont even look at code.. maybe sometimes and the 2.0 lets you see difs.” [source](https://www.reddit.com/r/google_antigravity/comments/1wcachv/can_someone_explain_how_you_actually_work_with/p8wdbkj/)
- Complaint, 2026-09-27, @antigravity (X): “@antigravity i believe that it is important to get approval before execution rather than just making a plan. it would be better if we could also check the changes again when the plan changes after approval.” [source](https://twitter.com/2978197789/status/2104080470883614974)
- Complaint, 2026-09-26, r/google_antigravity (Reddit): “it harder to review code, and code generated by gemini is dangerous if not reviewed, at least for the 3.7 and 3.8 flash” [source](https://www.reddit.com/r/google_antigravity/comments/1wqmd4n/why_is_the_antigravitycli_so_underrated/pc5b0hh/)
- Complaint, 2026-09-26, r/google_antigravity (Reddit): “same. hard to track changes in vs code with the ag extension and alarmingly, the only good ide ag ide is now no longer showing changed files either. i must track it via git changes. well...” [source](https://www.reddit.com/r/google_antigravity/comments/1wqlqcg/bug_generated_file_changes_disappear_after/pc6vbqr/)

### Devin

- Praise, 2026-09-15, @cognition (X): “@cognition testflight link plus screen recording is the useful part. you can inspect what the agent built before trusting the handoff” [source](https://twitter.com/2032890486571372544/status/2099897386696860149)
- Praise, 2026-09-15, @cognition (X): “@cognition the screen recording is the part that changes how this gets used. a non-technical stakeholder can't read a diff but can absolutely tell you the button is in the wrong place, and that shortens the review loop more than the building does.” [source](https://twitter.com/2066905512692633600/status/2099954647679041574)
- Praise, 2026-09-15, r/opencode (Reddit): “i‘m a dev for >10 years. i use llm‘s nowadays. i used windsurf/devin ide and whenever the cascade llm touched a file, that file was opened in the ide, each hunk was highlighted (red = deletions, green = insertions), i could manually accept or decline them. now i use positron (vscode fork). whenever the llm makes a change, the file is not automatically opened. but i can click in the „source control“ panel on the file, i see the hunks highlighte” [source](https://www.reddit.com/r/opencode/comments/1wfxfyj/is_there_a_way_to_see_diffs_of_edits_in_the_files/p9xx2gm/)
- Complaint, 2026-09-27, @DevinAI (X): “@ryancarson @devinai @linear @hellountangle zero local dev just relocates the humans to the one place that still matters: review. which makes the reviewer the production line — and the only one holding the loss when the diff reads fine and isn't.” [source](https://twitter.com/2065683882587144192/status/2104001307874885845)
- Complaint, 2026-09-27, @DevinAI (X): “@hraness @chatgpt @devinai @claudeai @cursor_ai @bot @zeddotdev 39 terminals make human review the bottleneck, not code generation” [source](https://twitter.com/1803494630366785536/status/2104144285939404963)
- Complaint, 2026-09-25, G2 (G2): “q: what problems is the product solving and how is that benefiting you? a: for ebiquity, i see the most value for data and engineering teams by reducing repetitive development, debugging and maintenance work. it could help teams move through smaller backlog tasks faster while allowing developers to focus on more complex work. q: what do you like best about the product? a: devin can take a development task from the initial request through coding,” [source](https://www.g2.com/products/devin-ai/reviews/devin-ai-review-13609931)

### OpenCode

- Praise, 2026-09-27, @opencode (X): “@thewritingdev @opencode opencode has a gui app too but hermes is just a general agent, it doesn't have any concept of open pr, diff file view, etc. it's jsut not the right tool for the job. it has other bot related features.” [source](https://twitter.com/412133001/status/2104160113313398978)
- Praise, 2026-09-24, @opencode (X): “opencode 2.0 is out from @opencode, with a built-in shell and a diff view. good basics. cockpit is what i built for past the basics: four open-source instruments that make long-running, agent-driven work something you can watch and steer. <strict_link>” [source](https://twitter.com/1939891306349637632/status/2102942966654407064)
- Praise, 2026-09-23, @opencode (X): “@opencode love you opencode ❤️ the best coding agent! pls do not change the ui of how change code will be shown. this is the easiest, fastest and most comfortable for developers to read them, for example claude code and codex now are terrible for devs to follow. i use opencode everyday” [source](https://twitter.com/2085450705804972033/status/2102791527349039307)
- Complaint, 2026-09-26, @opencode (X): “sometimes i find it hard to navigate between file diffs in @opencode when they’re large and stacked in one long scroll. exploring a persistent file list on the left bar, with one diff at a time on the right panel. thoughts? <strict_link>” [source](https://twitter.com/1075661598960873473/status/2103912455118405672)
- Complaint, 2026-09-25, @opencode (X): “@badlogicgames i think this is something which is missing from all ai tools like @opencode desktop and @ampcode i want to review/read the code with lsp and code navigation. all of them just shows git diff only” [source](https://twitter.com/1158785224299335680/status/2103409717959705080)
- Complaint, 2026-09-25, @opencode (X): “@nivekithans @badlogicgames @opencode @ampcode exactly, none of the agentic envs currently ship proper code exploration for some reason. i don't want to switch between 2 apps just to navigate code. the only reasonable way currently is pi + herdr + nvim in all in one window <strict_link>” [source](https://twitter.com/2076386152953565184/status/2103431958298562734)

### GitHub Copilot

- Praise, 2026-09-10, r/GithubCopilot (Reddit): “i use copilot in work and codex at home. i prefer copilot in vscode. i like copying symbols into the chat for the ais context and codex diff is beyond terrible. but i could be using it wrong so take that with a pinch of salt. to keep the credits i use the cheaper model, it's not too much different from the higher models” [source](https://www.reddit.com/r/GithubCopilot/comments/1wcj4ce/i_cancelled_my_github_copilot_subscription_are/p914kn1/)
- Praise, 2026-09-10, r/ClaudeCode (Reddit): “i can offer nothing but sympathy. i really enjoyed how github copilot worked in vs code. i loved seeing and approving the diffs. claude is better in every way except its vs code integration.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wbyrlh/anyone_else_absolutely_hates_the_vs_code_extension/p8vr0w7/)
- Praise, 2026-09-07, r/windsurf (Reddit): “my use is a lot like yours. i've tried most of the options out there. but phpstorm is objectively the best ide and i'm currently using using copilot for the agent, in its plugin, and kimi3. it gives you a diff, each turn. devin isnt really for guys like us.” [source](https://www.reddit.com/r/windsurf/comments/1w6ycby/is_the_20_devin_pro_tier_worth_it_if_i_dont_need/p8dvejm/)
- Complaint, 2026-09-25, r/GithubCopilot (Reddit): “we have supercov security check in [agents.md](<strict_link>) before commiting. usually takes <10s for full repo scan for deps/mcps we use dependabot on prs but i dont like it. agree that agents need verification and pr stage is too late” [source](https://www.reddit.com/r/GithubCopilot/comments/1wprq7d/best_application_security_tools_for_ai_generated/pc20ewi/)
- Complaint, 2026-09-24, r/GithubCopilot (Reddit): “local. not having the ability to approve the changed files is unacceptable. if i need to do something across multiple repos, i use the github copilot app” [source](https://www.reddit.com/r/GithubCopilot/comments/1wpctlc/vs_code_chat_users_local_or_copilot_harness/pbvangl/)
- Complaint, 2026-09-15, r/GithubCopilot (Reddit): “for most of my recent chats, the history shows that a number of lines are added and removed. when i open such a chat, it does not show a keep button, nor can i individually keep/accept the changes. for older chats (over 4 weeks ago) i don't see these changes listed for the chats. <strict_link> is this a bug or is there anything i can do to have all changes accepted and these number cleared?” [source](https://www.reddit.com/r/GithubCopilot/comments/1wgy49g/why_cant_i_keep_changes/)

### Pi

- Praise, 2026-09-23, @pidotdev (X): “@miguelriosen @pidotdev keeping the workflow as a dsl file means it shows up in code review as a diff, which a drag-and-drop canvas never gives you.” [source](https://twitter.com/2058824892238209024/status/2102900988915093764)
- Complaint, 2026-09-22, @pidotdev (X): “@pidotdev the number of extension with "diff"/"review" in title or description. that's a clear signal to improve diff.” [source](https://twitter.com/618819434/status/2102497902077603913)
- Complaint, 2026-09-09, r/PiCodingAgent (Reddit): “ok looks cool but... what value does it actually bring check a session change visually? not a bad concept, but imho an entire application for a functionallity that pi users barely check (yes, i'm pointing to loop/harness users) won't bring so much help. maybe a smaller case as plugin for vs code or obsidian may make more sense” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wbef1h/i_opened_a_pi_session_as_an_editable_map/p8q2d78/)
- Complaint, 2026-08-31, r/PiCodingAgent (Reddit): “👍 yeah, i just struggle with how slow and tedious it is. and after all that then i gotta survive sending it to someone else for them to then go thru the same process of peeling it apart.” [source](https://www.reddit.com/r/PiCodingAgent/comments/1w1zw2n/surviving_code_review/p6yi8l7/)

### Amp

- Praise, 2026-09-11, @AmpCode (X): “`amp sync` is yet another awesome addition from @ampcode - lets you (temporarily) sync changes from a thread to your machine and deletes them when you exit. great for review in your preferred diff tool.” [source](https://twitter.com/1491081/status/2098282433674092676)
- Praise, 2026-09-03, @AmpCode (X): “@jtaby @sethmills21 we use @ampcode, which has solutions to both: * their cloud agents can rpc to a local mac to build and report back (we do this for our ios app) * they have a built-in diff viewer/commenter and multiplayer for other team members other cloud agents might have this solved too?” [source](https://twitter.com/1183203638/status/2095596826544259359)
- Complaint, 2026-09-25, @AmpCode (X): “@badlogicgames i think this is something which is missing from all ai tools like @opencode desktop and @ampcode i want to review/read the code with lsp and code navigation. all of them just shows git diff only” [source](https://twitter.com/1158785224299335680/status/2103409717959705080)
- Complaint, 2026-09-25, @AmpCode (X): “@sqs @ampcode the changes tab on very large repos gets weirdly out of sync showing like 80k+ changes or something. sometimes running git pull or other commands fix it, other times get worse. i think it needs some way to refresh on the ui. could also be comparing wrong commit” [source](https://twitter.com/1857935142670450688/status/2103610446720942394)
- Complaint, 2026-09-25, @AmpCode (X): “@beyang @sqs @ampcode you're not offering that, but i don't mind beta testing. :) this one bugs me a lot in a certain huge ass monorepo” [source](https://twitter.com/1857935142670450688/status/2103615119880224923)

### Cline

- Complaint, 2026-09-20, @cline (X): “@cline right now, i can only see the code changes by clicking the line numbers, which shows additions/deletions in the top-right. a proper code editor with a clear diff view would make the agent workflow much better.” [source](https://twitter.com/1874710292577292288/status/2101765049316421672)
- Complaint, 2026-09-10, @cline (X): “i've been experimenting with your codebase, and there are some serious issues with cline cli. * your search codebase tool alone isn't enough. introduce glob and grep instead. * your edit tools diff returned is insanely noisy, if a edit is made to the top of a file everything after the edit is also shown in the tool result. * even the search used for edit tools is quite bad, there are no fallback searches like fuzzy; which other morden harnesses h” [source](https://twitter.com/1183625711401066497/status/2097901009209311546)

### Conductor

- Praise, 2026-09-11, r/PiCodingAgent (Reddit): “hey! i haven't tried this myself yet, just looked at the video, and from that it looks really nice and sleek. i've noticed others complained about the ui being "more of the same", but i like it and i think the codex inspiration was a nice call, their ui is great. i'm also a fan of you wanting to keep this project with a high quality bar and focused on pi. this is what i've been looking for, but i hope you can manage to introduce more features tha” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wcp3b7/supernova_a_minimal_opinionated_and_sleek/p94jb88/)
- Complaint, 2026-09-10, r/conductorbuild (Reddit): “for dark mode, the colours in the diff are terrible, i cannot read anything at all (specially the green), can you take a look into this? <strict_link>” [source](https://www.reddit.com/r/conductorbuild/comments/1wciucn/bug_report_terrible_issue_in_diff_colors/)

### Factory

- Complaint, 2026-09-11, G2 (G2): “q: what problems is the product solving and how is that benefiting you? a: a lot of engineering time still gets eaten up by repetitive, multi-file work that isn’t difficult, just time-consuming—small refactors, test fixes, pr cleanup, documentation updates, and straightforward ticket implementation. factory lets me hand those pieces off to droids so i can stay focused on design decisions, tougher bugs, and review. the result is less context switc” [source](https://www.g2.com/products/factory-ai/reviews/factory-ai-review-13441655)

### Warp

- Praise, 2026-09-12, @warpdotdev (X): “@warpdotdev remote-control plus the code review panel in the same shell is what sells it. i want the agent session where i already type, not in a second window.” [source](https://twitter.com/1894226409356496903/status/2098610214651982320)
