# Length and clarity of replies, summaries and comments (`work.response_verbosity`)

Feedback Bench, coding agents, built 2026-10-01, window 2026-08-31 to 2026-09-27. Web page: https://feedbackbench.com/#/criterion/work.response_verbosity

Area: [Doing the work](https://feedbackbench.com/criteria/work.md)

**Definition.** How long and how readable the agent's prose and summaries are, and how many code comments and doc notes it writes.

**Boundary.** Not this: see [Reviewing and approving the agent's changes](https://feedbackbench.com/criteria/verify.change_review_ui.md) for the volume of code changes. Not this: see [Does unrequested work or over-engineers](https://feedbackbench.com/criteria/work.scope_overreach.md) for extra code features.

Rated author-weeks, all agents: 940. Complaint share: 79%.

## The brief

Written by Claude Opus 5.5 from 56 labelled posts and the numbers on this page. Interpretation, not measurement: every quote is verbatim and links to its post.

**Users want answers in plain English, and Claude Code talks in jargon.**

TL;DR:

- Claude Code draws most complaints here: dense, coined jargon and long replies users must decode.
- OpenAI Codex earns praise for calmer, shorter prose, though some summaries turn too terse.
- Custom output styles and newer models help. Users still want a concise mode that holds.

In plain terms: Expect to reread. Users describe reports that string familiar words into unclear sentences, invent terms, and run long enough to cost tokens. Others hit the opposite problem: clipped summaries that list paths but skip what changed.

### How it breaks

- **Coined jargon users cannot parse** ([Length and clarity of replies, summaries and comments](https://feedbackbench.com/criteria/work.response_verbosity.md)). The loudest complaint is not length but legibility: replies built from invented terms that users must decode before they can act.
  Posts describe bug-fix and code explanations that read fine word by word but make no sense as a whole. Users report spending extra turns asking what the agent meant, or giving up. Some write custom output styles with rules against coined terms just to get readable reports. Others switched agents because the code was fine but the explanation was not.
  Evidence:
  - Complaint, Claude Code, r/ClaudeCode, 2026-09-12: “definitely. i feel like i'm having a stroke trying to understand what any current claude model is reporting back to me. i know all the words it's using but together they don't make sense. i also hate that it creates its own jargon but that's a whole other topic.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wdtrsa/im_done_with_opus_5/p99i9st/)
  - Complaint, Claude Code, r/ClaudeCode, 2026-09-12: “the problem with it is the mumbo jumbo. if you ask it to define everything you can trudge through it. i don’t care to spend my time parsing garbage though so i give it a custom output style with a bunch of rules to correct its tendency to output meandering piles of coined terms and jargon.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wdy70c/opus_is_fine_actually/p99vczy/)
  - Complaint, Claude Code, r/ClaudeAI, 2026-09-22: “how do you deal with long boring text the claude code shows when working with sdlc use case ? i have recently found that opus model (which i am forced to use due to the anthropic dumming sonnet down) that it gives me response of any bug fixes , feature work or explanation of code in a language which is very hard to understand in first read. for that either i give up or spend time in asking what do you mean by that ? how does people manage to let cc answer in understandable words when working with sdlc use cases ?” [source](https://www.reddit.com/r/ClaudeAI/comments/1wmy6n2/how_do_you_deal_with_long_boring_text_the_claude/)
  - Praise, OpenAI Codex, r/codex, 2026-09-05: “i had to stop using opus when i couldn't even understand the slop it would explain to me about how something worked. the code itself was fine but so is sol's, except sol actually talks like a normal person without so called claude-isms. astra is great but i haven't ran into anything that sol couldn't handle already, what are people using it for?” [source](https://www.reddit.com/r/codex/comments/1w7wgqg/massive_claude_fan_congrats_openai/p7y62dk/)

- **Walls of text that burn tokens** ([Length and clarity of replies, summaries and comments](https://feedbackbench.com/criteria/work.response_verbosity.md)). Long, chatty replies cost users time and quota, and several only get brevity by asking for it in every prompt.
  Users describe a simple phrase coming back as a newspaper, and output so long they must cap it explicitly, for example by requesting a fixed paragraph count. Others tie the verbosity directly to token spend. Shorter, less verbose responses is the top request on this page, and most of those asks target Claude Code. Some users also call out page-long thinking-out-loud that never answers the question.
  Evidence:
  - Complaint, Google Antigravity, r/google_antigravity, 2026-09-03: “the same problem as always, it continues to be too verbose and therefore consumes too many tokens.” [source](https://www.reddit.com/r/google_antigravity/comments/1w5i29e/review_of_gemini_38_flash_from_a_person_who/p7oacst/)
  - Complaint, OpenCode, r/opencode, 2026-09-11: “extremely and so annoyingly verbose, it prints a newspaper from a simple phrase, and actually not that good for what i used it. for 4x the usuage it's just a fair value not a great value, you get what you pay for that's it” [source](https://www.reddit.com/r/opencode/comments/1wd8rn5/is_deepseek_41_flash_not_as_good_as_glm_53_kimi/p9999rv/)
  - Complaint, Amp, @AmpCode, 2026-09-14: “it would be _very_ cool if the amp thread also suggested like 2-3 follow on prompts that are buttons that i could just click. right now a lot of the issue is each prompt gives a wall of text unless i explicitly prompt "2-3 paragraphs". i suspect more concise prompts with suggested follow-on prompts is a nicer ux” [source](https://twitter.com/721234540/status/2099486641966531069)
  - Complaint, GitHub Copilot, r/GithubCopilot, 2026-09-05: “opus 5 is unusable.. i had to fire it. it was very verbose, obfuscating. its stopping mid research with "if this is true, then ... " instead of reading the backing code, requiring me to do the research for it. it was almost like its purposefully using up wasted tokens. pedantic, panicky, or “adhd-like” output; overcomplicating minor issues; requiring constant clarification; treating small problems as major scope-creeping ones; and a distinctive, hard-to-parse voice full of phrases like “worth noting/naming.” sometime it never answered the question asked but spent a full page of output of thinking out loud. with opus 4.5 i was actually getting work done.” [source](https://www.reddit.com/r/GithubCopilot/comments/1w55bze/why_did_copilot_kill_opus_45_almost_3_months/p807mu5/)

- **Summaries clipped past usefulness** ([Length and clarity of replies, summaries and comments](https://feedbackbench.com/criteria/work.response_verbosity.md)). Cutting length can overshoot into terse, half-sentence reports that list links and paths without saying what finished.
  Some users now complain about the opposite failure. Posts describe a short, punchy style that reads like half a sentence, and end-of-task summaries that dump doc paths instead of explaining what was done and what comes next. Users suspect token saving and say it makes progress hard to track. Brevity without structure trades one comprehension cost for another.
  Evidence:
  - Complaint, Cursor, r/cursor, 2026-09-19: “it's not the technical jargon, it's the weird short punchy talk. i've been an it engineer for 30 years and it had been confusing me lately. its like reading half a sentence.” [source](https://www.reddit.com/r/cursor/comments/1wjy8nh/is_there_any_way_to_make_cursor_grok_46_speak/paop584/)
  - Complaint, OpenAI Codex, r/codex, 2026-09-22: “i use both claude / codex 200 usd plans, sol 6 xhigh, just word vomits links and doc-paths instead of actually explaining what it's doing / what finished. even astra has 'unceremonious' endings to a lot of my repo work. maybe it's saving tokens, but it's seriously frustrating to keep track of what it's done / what's next. who knows maybe because i just need to add additional [agents.md](<strict_link>) lines for it. but if by default it works this way, idk” [source](https://www.reddit.com/r/codex/comments/1wnkoh8/anyone_tried_gpt6_sol_yet/pbfqqh9/)
  - Complaint, OpenAI Codex, r/codex, 2026-09-12: “it's harder to understand what it is saying sometimes than astra” [source](https://www.reddit.com/r/codex/comments/1wdo7ys/whats_everyone_doing_while_we_wait/p9dz98x/)

- **Verbal tics on repeat** ([Length and clarity of replies, summaries and comments](https://feedbackbench.com/criteria/work.response_verbosity.md)). Pet words and stock phrases repeat until users notice them more than the content.
  Users single out recurring words and filler phrases that no person would say that often, from a favourite adjective to a habitual sentence opener. The tics make output feel machine-written and add noise to every report. Requests to stop repetitive filler and to tone down quirky personality show up for both Claude Code and OpenAI Codex.
  Evidence:
  - Complaint, OpenAI Codex, r/codex, 2026-09-15: “"bounded"!! that word gives me nightmares! never in my life have i heard a real person use this word, yet codex vomits it all over me every chance it gets!!” [source](https://www.reddit.com/r/codex/comments/1wgg7zu/2030b_tokens_a_week_to_1b_are_they_for_real/p9yx1fe/)
  - Complaint, OpenCode, r/opencode, 2026-09-17: “same experience for me, the repeated "actually," drive me nuts!” [source](https://www.reddit.com/r/opencode/comments/1wiu75f/free_models/paga0ac/)
  - Complaint, GitHub Copilot, r/GithubCopilot, 2026-09-05: “opus 5 is unusable.. i had to fire it. it was very verbose, obfuscating. its stopping mid research with "if this is true, then ... " instead of reading the backing code, requiring me to do the research for it. it was almost like its purposefully using up wasted tokens. pedantic, panicky, or “adhd-like” output; overcomplicating minor issues; requiring constant clarification; treating small problems as major scope-creeping ones; and a distinctive, hard-to-parse voice full of phrases like “worth noting/naming.” sometime it never answered the question asked but spent a full page of output of thinking out loud. with opus 4.5 i was actually getting work done.” [source](https://www.reddit.com/r/GithubCopilot/comments/1w55bze/why_did_copilot_kill_opus_45_almost_3_months/p807mu5/)

- **Output styles and newer models help** ([Length and clarity of replies, summaries and comments](https://feedbackbench.com/criteria/work.response_verbosity.md)). Users who set concise output styles, or moved to newer models, report replies that lead with the point and use plain words.
  The praise here is mostly about fixes. Users report a concise output style removing the fluff, and output styles adding diagrams and plainer explanations. A newer model release drew posts saying it puts the important points first and returns clean bullet lists. The catch is reliability. Users still ask for a concise mode that consistently holds.
  Evidence:
  - Praise, Claude Code, r/ClaudeCode, 2026-09-02: “i also set concise output style today and it with 5.1 has 0 fluff bullshit, straight to the point, lovely” [source](https://www.reddit.com/r/ClaudeCode/comments/1w4qdbe/the_end_of_claudish/p7bsx6e/)
  - Praise, Claude Code, r/ClaudeCode, 2026-09-17: “how are you getting the same results when you can change the literal results? my opus has been incredibly easy to read for weeks now. i had the same complaints, but now as part of the output styles, it draws more ascii diagrams, and gives better descriptions and context before explaining things in actual top-2,000-common-word english. it’s great.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wiujum/claude_code_is_falling_behind_codex_not_because/padvchj/)
  - Praise, Claude Code, r/ClaudeCode, 2026-09-23: “been using opus 5.5 for over a day now and i have found it to be really good to interact with. it keeps things concise and puts the important points at the top. so far, i am pleased with this model. what does everyone else feel ?” [source](https://www.reddit.com/r/ClaudeCode/comments/1woddlf/initial_thoughts_on_opus_55_it_is_a_considerable/)
  - Praise, Claude Code, r/ClaudeCode, 2026-09-22: “didn’t know opus 5.5 was released, i immediately picked up an item i left open from a few days ago, asked it to summarize what was going on and provide a starting point for the task. it did it beautifully, no word salad, no load-bearing objects, no item that bites. just a clean and simple bullet point list with everything i asked.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wnie5o/opus_55_finally_someone_who_doesnt_yap_as_much/pbfdnc0/)

### Who stands out

- **Claude Code (weaker)**. Claude Code collects most of the complaints, from hard-to-parse jargon to chatty replies, even from users who value its tooling.
  Posts call its prose verbose, cutesy and full of self-invented terms. Several users keep the subscription for the CLI and harness while preferring another agent's writing. Some see an improvement with a newer model and with custom output styles. Requests for shorter responses and clearer plain-language explanations are concentrated here.
  Evidence:
  - Praise, Kiro, r/kiroIDE, 2026-09-10: “i’ve tried a bunch of different clients and models out there, and prefer kiro + one of the claude 4.8 models to most of them, including claude code cli which i find too cutesy and verbose by comparison. honestly this is more of a matter of tastes + skills issue on your part.” [source](https://www.reddit.com/r/kiroIDE/comments/1wchsii/frontier_models_on_other_providers_vs_the_same/p8zeoia/)
  - Complaint, Claude Code, r/codex, 2026-09-13: “i have the complete opposite experience. the only good thing about claude is claude design. claude code with opus is unusuable. i use it sometimes for second opinions and reviews but the amount of false positives makes it output unusuable. also in my experience it is way slower than chatgpt so i really don't understand how you came to your conclusions. last the chatty responses by opus are a major annoyance for me. i prefer the direct and calm responses of chatgpt models.” [source](https://www.reddit.com/r/codex/comments/1wetm28/we_switched_from_claude_code_to_codex_at_work/p9j686v/)
  - Praise, OpenAI Codex, r/ClaudeCode, 2026-09-11: “there’s really something wrong with it. it seems like it acts like an overqualified post doc intern who cares more about proving he’s super intelligent, than actually doing the job he’s asked to do. for instance: talking in a non intelligible way, or being overly rigid in following any kind of process. i have a claude max x20 sub that i struggle to keep within weekly limits, so i decided to take a codex sub for a month to try out astra. and this what made me realize how crazy unintelligible opus can be. fable is a bit better, but still incomparable to astra. the only thing that makes me keep my claude sub is how better is the cli/tooling/harness/etc.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wdtrsa/im_done_with_opus_5/)
  - Praise, OpenAI Codex, r/codex, 2026-09-10: “i like the way codex speaks, claude just vomits words it's currenlty the sloppiest writer, but good implementer. honestly gemini right now is the best human like writer it gives you just the right amount of prose.” [source](https://www.reddit.com/r/codex/comments/1wcmwb1/so_what_are_everyones_general_thoughts_on_astra/p914h53/)

- **OpenAI Codex (stronger)**. OpenAI Codex wins on tone, with users praising direct, calm replies and less rambling, though some find its summaries too terse.
  Users contrast its normal-sounding prose with Claude's, and note verbosity coming down in recent models. The complaints are narrower. Users flag a few overused words and end-of-task reports that list paths instead of explaining the work. Its weakness is clarity at the edges, not volume.
  Evidence:
  - Praise, OpenAI Codex, r/codex, 2026-09-05: “the speed is excellent and the verbosity is right down. loving it so far.” [source](https://www.reddit.com/r/codex/comments/1w7l3yj/gpt6_astra_is_now_out_to_all_plus_business_pro/p7vxrqr/)
  - Praise, OpenAI Codex, r/codex, 2026-09-10: “i like the way codex speaks, claude just vomits words it's currenlty the sloppiest writer, but good implementer. honestly gemini right now is the best human like writer it gives you just the right amount of prose.” [source](https://www.reddit.com/r/codex/comments/1wcmwb1/so_what_are_everyones_general_thoughts_on_astra/p914h53/)
  - Praise, OpenAI Codex, r/codex, 2026-09-22: “i feel that at this time, when the series of tasks i've launched luna 6 are still ongoing, and that the consequence of its work will not be seen or measured in any relevant way at least until a few days, trying to say "it's better/it's worse" is as reliable as diving into chicken guts to predict the weather of tomorrow. only thing i can say is that is yapps less in paragraphs of non sense adressed only at itself.” [source](https://www.reddit.com/r/codex/comments/1wnmhnq/gpt56_luna_fans_how_does_gpt6_luna_feel_so_far/pbg4cx8/)
  - Complaint, OpenAI Codex, r/codex, 2026-09-22: “i use both claude / codex 200 usd plans, sol 6 xhigh, just word vomits links and doc-paths instead of actually explaining what it's doing / what finished. even astra has 'unceremonious' endings to a lot of my repo work. maybe it's saving tokens, but it's seriously frustrating to keep track of what it's done / what's next. who knows maybe because i just need to add additional [agents.md](<strict_link>) lines for it. but if by default it works this way, idk” [source](https://www.reddit.com/r/codex/comments/1wnkoh8/anyone_tried_gpt6_sol_yet/pbfqqh9/)

- **OpenCode (mixed)**. OpenCode's quiet harness gets praise, but verbosity complaints follow whichever model users plug into it.
  Users like that it gets to work without chatter, and some say even a wordy model talks less inside it. The complaints target specific models run through OpenCode, described as the most verbose they have used, with repeated tics. The experience depends on the model as much as the tool.
  Evidence:
  - Praise, OpenCode, @opencode, 2026-09-09: “@0xlars_ @opencode the silence is a feature, i hate models that talk too much” [source](https://twitter.com/1823746967543144449/status/2097520413714821502)
  - Praise, OpenCode, @opencode, 2026-09-03: “spent an hour or so with muse spark 1.3 doing coding with @opencode i like how quiet it is. if the prompt is clear enough, it'll just get to work without asking any questions or being overly verbose about what it is doing not sure yet about code quality - will check that next” [source](https://twitter.com/367701406/status/2095499643220377824)
  - Praise, OpenCode, r/ClaudeCode, 2026-09-01: “while opusplaining is inevitable with opus 5, it is a lot less worse when you use him in opencode. though i would still not use it as the agent i’m actively talking to.” [source](https://www.reddit.com/r/ClaudeCode/comments/1w3rxkj/average_opus_5_response/p75al6j/)
  - Complaint, OpenCode, r/opencodeCLI, 2026-09-12: “it's dumb and incredibly verbose, the most verbose model ever and nothing comes close” [source](https://www.reddit.com/r/opencodeCLI/comments/1wdznnp/deepseek_v41_token_efficiency_is_so_cracked/p9fue8t/)

### Fine print

- Posts skew heavily to Claude Code, so other agents' patterns rest on far fewer users.
- Many posts blame a model rather than the agent, so harness and model effects are hard to separate.
- Most agents here have too few posts to rank with confidence.

## Top requests

What users ask to add or change, most asked first. 123 author-weeks ask for something. Requests do not change the Feedback Score. Rule: A separate pass by Claude Sonnet 5 reads every counted post and extracts what the author asks the agent or its vendor to add or change, with the criteria it maps to and a short normalised wording; it does not touch the labels or the Feedback Score. Claude Opus 5.5 groups the wordings within each criterion (the first criterion the request maps to) into themes; code counts them. A theme counts distinct author-weeks that ask for it, per agent; across agents, one author-week per agent. Themes asked in fewer than 2 author-weeks, and requests that share no theme, are not shown. Examples: up to 3 posts per theme from different authors, without slurs, preferring posts of 60 to 450 characters, most recent first.

| Rank | Request | Author-weeks | Posts | Agents (author-weeks) |
|---|---|---|---|---|
| 1 | Shorter, less verbose responses | 47 | 48 | Claude Code 38, OpenAI Codex 5, Google Antigravity 2, Devin 1, Pi 1 |
| 2 | Clearer plain-language explanations | 22 | 22 | Claude Code 14, OpenAI Codex 5, OpenCode 2, Devin 1 |
| 3 | Fewer and less verbose code comments | 11 | 12 | Claude Code 9, OpenAI Codex 2 |
| 4 | Reliable concise mode setting | 7 | 7 | Claude Code 6, OpenCode 1 |
| 5 | Option to disable progress update messages | 6 | 6 | Claude Code 3, OpenAI Codex 2, OpenCode 1 |
| 6 | Better overall writing quality | 5 | 6 | Claude Code 4, Cursor 1 |
| 7 | Stop repetitive verbal tics and filler words | 4 | 4 | Claude Code 3, OpenAI Codex 1 |
| 8 | Restore previous personality and tone | 2 | 5 | Claude Code 2 |
| 9 | Concise, clear documentation output | 2 | 2 | Claude Code 2 |
| 10 | Concise, well-structured plan output | 2 | 2 | Claude Code 2 |
| 11 | Less quirky personality in responses | 2 | 2 | Claude Code 1, OpenAI Codex 1 |
| 12 | Less verbose reasoning output | 2 | 2 | Claude Code 2 |

### 1. Shorter, less verbose responses

- Claude Code, 2026-09-25, @ClaudeDevs (X): “@claudedevs the usage limits are great currently with opus 5.5 but they would go even further if it wouldn't dump a novel full of claude-speak at me in every answer. you need to get that verbosity under control!” [source](https://twitter.com/1823237976295383041/status/2103604406109266061)
- OpenAI Codex, 2026-09-24, r/codex (Reddit): “i'd pay max tokens for a gpt-6 stfu model. i dont want to read a novel for a simple question. i dont want it to tell me 'youre right' 5000 times. i dont want it to suggest things i didnt explicitly ask. seriously, shut tf up!” [source](https://www.reddit.com/r/codex/comments/1wp2bov/chatgpt_manipulates_you_to_keep_chatting/pbrrfs0/)
- Claude Code, 2026-09-24, @ClaudeDevs (X): “@claudedevs but can it reply with a simple yes or no and avoid dissertation length responses….nope. context would be far lighter if it could manage this one simple thing” [source](https://twitter.com/1959700170771144704/status/2103172816732660155)

### 2. Clearer plain-language explanations

- Claude Code, 2026-09-23, r/ClaudeCode (Reddit): “i am glad i am not the only one frustrated with reading those in big claude response. why can't claude present it better may be under additional considerations heading, and concise in a way that is also easy to catch with a glance. . edit: while writing my thought i realised i could have fixed this behaviour with claude global instructions.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wnt4d3/they_fucking_cooked_yo_opus_55_is_a_massive/pbi4fhv/)
- OpenAI Codex, 2026-09-22, r/codex (Reddit): “help with overengineered scientific reports i am trying to write my phd thesis together with codex. it does great job at analysis and plots, but when it comes to describe a method it is describing it in a very technical and too-detailed manner. did anyone face this and has a quick fix? i would love it if there was a plugin or something similar that i could use to make codex write simpler but still informal and engaging scientific reports.” [source](https://www.reddit.com/r/codex/comments/1wnb773/help_with_overengineered_scientific_reports/)
- Claude Code, 2026-09-21, r/ClaudeCode (Reddit): “@anthropic: please, i’m begging you, make it speak plainly, i just can’t deal with opus 5 right now. ![gif](giphy|l0nwnrl4btdd7jcx2)” [source](https://www.reddit.com/r/ClaudeCode/comments/1wmeayt/anthropic_is_currently_stealth_testing_opus_55/pb6e35n/)

### 3. Fewer and less verbose code comments

- Claude Code, 2026-09-22, r/ClaudeCode (Reddit): “no more code comments claude, they’re hurtful and destructive” [source](https://www.reddit.com/r/ClaudeCode/comments/1wn1uau/in_sept_2026_are_code_comments_useful_or_hurtful/pbbv7yf/)
- Claude Code, 2026-09-16, r/ClaudeCode (Reddit): “my biggest issue is not matter what rules and skills i put in, convincing the model to not write novels of code comments.” [source](https://www.reddit.com/r/ClaudeCode/comments/1whspya/the_only_thing_ive_seen_work_on_big_claude_code/pa546z7/)
- Claude Code, 2026-09-13, r/ClaudeCode (Reddit): “i [just posted](<strict_link>) about the comment thing, too. it's incredibly frustrating; i feel like i can sort of drag it kicking and screaming to the right behavior with the right guardrails, sometimes, but the way it writes comments is just categorically bad and doesn't feel like something that we as users should need to solve.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wfihk3/why_is_it_so_bad_at_basic_coding_skills/p9mo8gn/)

### 4. Reliable concise mode setting

- Claude Code, 2026-09-22, r/ClaudeCode (Reddit): “i hope that this becomes standard. the amount of times i have to tell opus 5 to trim verbosity and be more concise is too damn high” [source](https://www.reddit.com/r/ClaudeCode/comments/1wnfz86/proof_that_opus_55_is_easier_to_talk_todeal_with/pbfkm0g/)
- Claude Code, 2026-09-14, @ClaudeDevs (X): “is it possible @claudedevs @claudeai to have claude only give me executive summary type of responses? i really don't need to sift through a 10 page response to find the answer to my yes/no question and my instructions are not being remembered between prompts” [source](https://twitter.com/482112385/status/2099551097262121272)
- OpenCode, 2026-09-13, @opencode (X): “@jvr0x @miaai_lab @grok @opencode @nousresearch opencode could have inject a prompt to answer concisely/thinkless” [source](https://twitter.com/1997718730214957056/status/2099090414439694778)

### 5. Option to disable progress update messages

- Claude Code, 2026-09-23, @ClaudeDevs (X): “@claudedevs is there a way to disable the prompt injection that tells claude to update the user every so often on what it's currently doing?” [source](https://twitter.com/2088198392635666433/status/2102696593535451348)
- OpenCode, 2026-09-15, r/opencode (Reddit): “too much request and long winded. result is good but goddamn it looks and call every single tools and command” [source](https://www.reddit.com/r/opencode/comments/1wh412m/muse_spark_13_vs_gemini_31_pro/p9zdkt3/)
- Claude Code, 2026-09-12, @ClaudeDevs (X): “@claudedevs whatever update got pushed recently that makes the model give a summary after every tool call is garbage. using claude-in-chrome through the cli is 95% noise and wasted tokens.” [source](https://twitter.com/2064383974340751360/status/2098602524156527068)

### 6. Better overall writing quality

- Claude Code, 2026-09-21, r/ClaudeCode (Reddit): “i don't need it to get better at coding, i just want it to get better at writing text and comments that don't make me want to claw my eyes out. "honest caveat: it's not the goal, it's the seam that exposes the smoking gun"” [source](https://www.reddit.com/r/ClaudeCode/comments/1wmeayt/anthropic_is_currently_stealth_testing_opus_55/pb918nn/)
- Cursor, 2026-09-17, @cursor_ai (X): “@collisionadv @cursor_ai @grok experimenting with this myself. cursor projects is fantastic. wish 4.6 writing was better. we’ll see what we get with 4.7.” [source](https://twitter.com/42959431/status/2100404051204517974)
- Claude Code, 2026-09-04, @ClaudeDevs (X): “@claudedevs we would appreciate it if “mannered prose” was no longer a thing when using claude. bring back human prose! that’s it, that’s the feature request that lands, not a complaint.” [source](https://twitter.com/9673282/status/2095677112493416537)

### 7. Stop repetitive verbal tics and filler words

- Claude Code, 2026-09-23, @ClaudeDevs (X): “@claudedevs too bad claude didn’t decide to stop generating those em-dashes instead of fixing the performance of their display ;-)” [source](https://twitter.com/14351189/status/2102878066875736510)
- OpenAI Codex, 2026-09-15, r/codex (Reddit): “"bounded"!! that word gives me nightmares! never in my life have i heard a real person use this word, yet codex vomits it all over me every chance it gets!!” [source](https://www.reddit.com/r/codex/comments/1wgg7zu/2030b_tokens_a_week_to_1b_are_they_for_real/p9yx1fe/)
- Claude Code, 2026-09-15, @ClaudeDevs (X): “privately: why does every sentence now start with "privately"? privately: what i develop is not a classified government secret. privately, you see how annoying it is. please stop @claudeai @claudedevs, privately 😅” [source](https://twitter.com/229266877/status/2099806482594234786)

### 8. Restore previous personality and tone

- Claude Code, 2026-09-22, r/ClaudeCode (Reddit): “and it makes sense! that was a dark time. please dont do that again.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wne9k9/well_its_official_its_55_and_not_51/pbf2mto/)
- Claude Code, 2026-09-16, @ClaudeDevs (X): “@claudedevs please bring back old buttery voice and personality! this new one is just a friendly fluffy chatgpt level.” [source](https://twitter.com/2050968414269693952/status/2100175002934944207)

### 9. Concise, clear documentation output

- Claude Code, 2026-09-17, r/ClaudeCode (Reddit): “my work currently is porting some legacy apps into modern api. at the end of a session when i ask them to document worthwhile findings, they spit out shitty docs. it's genuinely not readable, like a word dump with a smattering of magic words. so i'm looking for a skill/tooling that can help claude write docs that's \*\*easy to read\*\* and \*\*structured\*\*. like those technical blogs we find online. i'm not sure why they're more pleasing to re” [source](https://www.reddit.com/r/ClaudeCode/comments/1wir537/skillstools_for_documenting/)
- Claude Code, 2026-09-17, r/ClaudeCode (Reddit): “it's designed for safety critical manuals, and thus overly explanatory. it talks to you like you are dumb. it's too slow for me. for example: > before you deploy a new version to the production environment, run the automated tests. make sure that all tests are successful. also make sure that all required database migrations are complete. do not deploy the new version if the tests fail or if a required migration is not complete. > after the deploy” [source](https://www.reddit.com/r/ClaudeCode/comments/1wiujum/claude_code_is_falling_behind_codex_not_because/pag9xcz/)

### 10. Concise, well-structured plan output

- Claude Code, 2026-09-08, r/ClaudeCode (Reddit): “i have to say that this doesn't really work either. it will start generating a massive multipage plan on a simple question.” [source](https://www.reddit.com/r/ClaudeCode/comments/1waoxg5/how_do_you_instruct_claude_to_answer_rather_than/p8ky2mj/)
- Claude Code, 2026-09-18, r/ClaudeCode (Reddit): “i'm messing around with antigravity to scratch that productivity itch.... and seeing it take a prompt, distill down in thought to the action items to take, and then see it actually do those action items..... my god, the crystal clear introspection... my god, the straight, clean, clear-cut verbiage... why can't you do this, claude code? <strict_link> i'm not one to venture to other pastures once i find something i like..... but...... anthropic..” [source](https://www.reddit.com/r/ClaudeCode/comments/1wjex73/having_run_out_of_tokens_for_the_week_like/)

### 11. Less quirky personality in responses

- Claude Code, 2026-09-24, r/ClaudeCode (Reddit): “i dont care about a little nerf in intelligence as long as they don't make it speak claudish ever again 😭” [source](https://www.reddit.com/r/ClaudeCode/comments/1woba1t/opus_55/pboyqc1/)
- OpenAI Codex, 2026-09-04, r/codex (Reddit): “yeah but openai models aren't being so sassy and having annoying personality quirks. if their focus is on coding users then stop trying to give it personality. it makes more sense for chat session with those who wants to date their chatbot.” [source](https://www.reddit.com/r/codex/comments/1w6o4xf/i_like_sam_altman_way_better_than_dario/p7p25kh/)

### 12. Less verbose reasoning output

- Claude Code, 2026-09-22, r/ClaudeCode (Reddit): “it's something they added to bloat your bill, burn your usage credits, and boost revenue. you have to tell claude to stop the fluff, end the thinking process paragraphs, and work like an engineer on a tight schedule. i've literally watched claude argue with itself for 15 minutes over changing a symbol's name, and i'm tired of it.” [source](https://www.reddit.com/r/ClaudeCode/comments/1tfzamb/claude_code_has_been_thinking_too_long_in_other/pbgfwge/)
- Claude Code, 2026-09-21, r/ClaudeCode (Reddit): “> guidance to our shared configs on overengineering less would love to know what this looks like. i agree that just keeping a close eye on the model is non-negotiable at the moment. but even e.g. planning is made painful by the fact that models don't seem to know what to prioritise, and will give you novels of reasoning to sift through for almost any task or change, regardless of its size. i am just complaining atp, but if there was a way to get” [source](https://www.reddit.com/r/ClaudeCode/comments/1wm15qm/how_do_you_keep_complexity_out/pb3gg7s/)

## Every agent

| Agent | Overall rank | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|---|
| [OpenAI Codex](https://feedbackbench.com/agents/codex.md) | 2 | Better than peers | 0.588 | 0.553–0.620 | 163 | 57 | 106 |
| [OpenCode](https://feedbackbench.com/agents/opencode.md) | 3 | Typical | 0.503 | 0.476–0.532 | 33 | 8 | 25 |
| [Claude Code](https://feedbackbench.com/agents/claude-code.md) | 1 | Worse than peers | 0.433 | 0.409–0.455 | 673 | 106 | 567 |
| [Cursor](https://feedbackbench.com/agents/cursor.md) | 4 | Too few posts | – | – | 22 | 3 | 19 |
| [Google Antigravity](https://feedbackbench.com/agents/antigravity.md) | =5 | Too few posts | – | – | 18 | 12 | 6 |
| [Devin](https://feedbackbench.com/agents/devin.md) | =5 | Too few posts | – | – | 12 | 6 | 6 |
| [GitHub Copilot](https://feedbackbench.com/agents/copilot.md) | =8 | Too few posts | – | – | 7 | 1 | 6 |
| [Pi](https://feedbackbench.com/agents/pi.md) | 7 | Too few posts | – | – | 4 | 2 | 2 |
| [Amp](https://feedbackbench.com/agents/amp.md) | =11 | Too few posts | – | – | 4 | 3 | 1 |
| [Cline](https://feedbackbench.com/agents/cline.md) | =8 | Too few posts | – | – | 1 | 1 | 0 |
| [Zed](https://feedbackbench.com/agents/zed.md) | =8 | Too few posts | – | – | 1 | 0 | 1 |
| [Factory](https://feedbackbench.com/agents/factory.md) | =11 | Too few posts | – | – | 1 | 1 | 0 |
| [Kiro](https://feedbackbench.com/agents/kiro.md) | 13 | Too few posts | – | – | 1 | 1 | 0 |
| [Conductor](https://feedbackbench.com/agents/conductor.md) | 14 | Too few posts | – | – | 0 | 0 | 0 |
| [Warp](https://feedbackbench.com/agents/warp.md) | 15 | Too few posts | – | – | 0 | 0 | 0 |
| [Grok Build](https://feedbackbench.com/agents/grok-build.md) | 16 | Too few posts | – | – | 0 | 0 | 0 |
| [Augment Code](https://feedbackbench.com/agents/augment.md) | 17 | Too few posts | – | – | 0 | 0 | 0 |

## Posts

Receipts rule: The 5 most recent praise and complaint posts per area (first 700 characters) and 3 per criterion (first 450 characters).

### OpenAI Codex

- Praise, 2026-09-24, r/codex (Reddit): “100%. from a lead engineer standpoint, 6 sol has very obvious better output, is way cheaper, less verbose, finishes a task faster, etc. the cycle of “any new model feels dumber than the last one” has been going on for the last year and models have only gotten better… vibe coder paranoia is getting too much attention.” [source](https://www.reddit.com/r/codex/comments/1wodcz2/omg_no_wayy/pbq9s7o/)
- Praise, 2026-09-24, r/codex (Reddit): “i don't share that view. lately, i’ve gone the other way, switching from anthropic to openai. if you're just doing "vibe-coding," then anything works, but i find opus really hard to follow. its explanations are verbose and meaningless. i spend more time trying to understand what it's saying than i would writing the code myself. by comparison, openai's models are much clearer in their explanations and dialogue.” [source](https://www.reddit.com/r/codex/comments/1worwfr/sol_6_was_insufferable_glad_to_be_back_to_56/pbqcc17/)
- Praise, 2026-09-24, r/codex (Reddit): “i love claude for prose, but for coding i prefer openai's short and precise. openai can be kind of the boring corporate, but extemely effective bot anthropic is more creative, colourful, interesting” [source](https://www.reddit.com/r/codex/comments/1wp3bzl/you_need_to_try_opus_55/pbsieqe/)
- Complaint, 2026-09-27, r/codex (Reddit): “4.7 absolutley was. verbose, argumentative, lazy. 5 wasn't even worth considering with where 5.6 was” [source](https://www.reddit.com/r/codex/comments/1wr1oir/they_are_aware_and_working_on_it_apparently_just/pcdu3gf/)
- Complaint, 2026-09-27, r/codex (Reddit): “im giving the $20 claude a go and gpt is giving me sass in the handover ive never seen it be so negative and unhelpful in the response lol” [source](https://www.reddit.com/r/codex/comments/1wrw0sg/is_this_proof_were_getting_a_powerful_new_model/pcgiiaq/)
- Complaint, 2026-09-26, r/codex (Reddit): “yes. they became verbose trying to be like chatgpt” [source](https://www.reddit.com/r/codex/comments/1wqbnpj/we_are_back/pc3njv7/)

### OpenCode

- Praise, 2026-09-25, r/opencode (Reddit): “i used it today for the first time for coding a vite/supabase app and it has been fantastic. granted i have not used openai or anthropic for a long while and mostly but do use gemini 3.8 pro and flash. i'd pick space bunny over it. what i like about is a mostly matter of fact default communication style with clear conclusions and next steps.” [source](https://www.reddit.com/r/opencode/comments/1wp1y9t/i_fingerprinted_space_bunny_alpha_for_us_vs_cn/pc2no11/)
- Praise, 2026-09-13, r/codex (Reddit): “switched to some free opencode ones for curiosity yesterday and gone is all the bullshit chit chat and i'm getting must more logical sensibly structured product and it is following my guardrail from the get go. really surprised” [source](https://www.reddit.com/r/codex/comments/1wetm28/we_switched_from_claude_code_to_codex_at_work/p9jjoxl/)
- Praise, 2026-09-09, @opencode (X): “@0xlars_ @opencode the silence is a feature, i hate models that talk too much” [source](https://twitter.com/1823746967543144449/status/2097520413714821502)
- Complaint, 2026-09-27, r/opencode (Reddit): “yea it's choppy style is horrible, you have to tell it to stop replying with status lines.” [source](https://www.reddit.com/r/opencode/comments/1won61w/real_life_performance_of_muse_spark_13/pc9x8c4/)
- Complaint, 2026-09-25, r/opencode (Reddit): “nobody said v2's cache management or dynamic tool injection is bad — that's literally why we forked **v2.0.16** instead of staying on v1, and we kept 100% of the v2 runtime, caching, and tool engine intact. everything listed in the post is verifiable directly in the v2.0.16 source tree: * `packages/core/src/plugin/identity.ts` (`# your model:` block injected on every request) * `packages/core/src/instruction-discovery.ts` (`instructions from: ${f” [source](https://www.reddit.com/r/opencode/comments/1wpe80m/inside_opencode_v2s_source_code_how_it_sabotages/pbxg2o4/)
- Complaint, 2026-09-24, r/opencode (Reddit): “i like it. i'm using it for re/vr tasks and it does a pretty solid job. i'm finding it functionally comparable to luna. maybe slightly underperforming it? however i don't really like its output. it's not wrong, but it reads a little technical/terse. it's weird to say an ai is not outputting enough, but its kind of techno-babbly in short sentences.” [source](https://www.reddit.com/r/opencode/comments/1won61w/real_life_performance_of_muse_spark_13/pbp73io/)

### Claude Code

- Praise, 2026-09-27, r/ClaudeCode (Reddit): “5.5 is really good but it does have the same flaws as all the other agents. i think one of the biggest appeals (any why most love it) is the huge reduction in its draw on subs (5x just became more like 30x if you set it on medium effort) and it speaks plain language for the most part, oh and its faster and less verbose, all while being a solid reasoning machine. its fast, cheap, and works. that makes for happy coders.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wqjuv2/am_i_the_only_one_who_doesnt_like_opus_55/pca70cd/)
- Praise, 2026-09-27, r/ClaudeCode (Reddit): “i find the exact opposite. and if i'm ever confused i just ask what it did in plain language.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wr91sx/i_lose_track_with_opus_55/pcao86v/)
- Praise, 2026-09-27, r/ClaudeCode (Reddit): “i think they penalized the model on output length making it be more concise in its answers therefore reducing token usage and making it yap less than opus 5” [source](https://www.reddit.com/r/ClaudeCode/comments/1wrpq38/any_idea_what_anthropic_figured_out/pcgzcbv/)
- Complaint, 2026-09-27, r/ClaudeCode (Reddit): “coming from a long-time codex user, i feel the opposite. astra and sol do exactly what you just said, but opus sends a report every minute or so.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wr91sx/i_lose_track_with_opus_55/pcav9vk/)
- Complaint, 2026-09-27, r/ClaudeCode (Reddit): “i tried the learning mode but found it rather frustrating - it never quite gave me enough information for me understand the intent.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wqk827/im_doing_agentic_development_all_wrong_im_writing/pcbq052/)
- Complaint, 2026-09-27, r/ClaudeCode (Reddit): “opus 5 also continued forever. every single reply...” [source](https://www.reddit.com/r/ClaudeCode/comments/1wre78w/opus_55_is_how_its_meant_to_be/pccis8l/)

### Cursor

- Praise, 2026-09-26, r/cursor (Reddit): “hi, i do not feel like 4.7 is doing a bettere job. i do feel like it writes more text for the user. which in my case is good since im a scum vibe code indie game dev, so i have no clue about coding and only make small games with it. but i am wondering what you all feel: is 4.7 any better than 4.6? which one do you prefer and why? thanks” [source](https://www.reddit.com/r/cursor/comments/1wqhcse/grok_46_vs_47/)
- Praise, 2026-09-24, r/cursor (Reddit): “actually the chat outputs are very much human readable and on point. grok output is hard to comprehend.” [source](https://www.reddit.com/r/cursor/comments/1wp9gpx/composer_25_fast_is_still_really_good_as_always/pbumjnp/)
- Praise, 2026-09-16, @cursor_ai (X): “@tippmann777 @cursor_ai yes. task in, result out. no walls of text, no lectures. just code.” [source](https://twitter.com/1720665183188922368/status/2100069271174828227)
- Complaint, 2026-09-25, r/cursor (Reddit): “the answers it gives are so much more opaque than 4.6. i can ask about a piece of code, and it will give me a ton of useless hard-to-digest information about the thing, and still not even answer the question i asked. i'm using the same rules as with the 4.6 model, and it's significantly worse in this regard. there's also something about its writing style that requires more mental parsing, like it's using ambiguous words too often.” [source](https://www.reddit.com/r/cursor/comments/1woad6b/grok_47_performance_in_cursor/pc0k0j2/)
- Complaint, 2026-09-24, r/cursor (Reddit): “an analysis benchmarking models capability on frontier intelligence tasks is not the right spot to test cost efficiency. also 4.7 is much more verbose than 4.6. so may have some regression there” [source](https://www.reddit.com/r/cursor/comments/1wp6j44/is_this_supposed_to_be_good_news/pbuqlh4/)
- Complaint, 2026-09-23, r/cursor (Reddit): “absolute word-salad responses nowadays.” [source](https://www.reddit.com/r/cursor/comments/1woad6b/grok_47_performance_in_cursor/pblkqyk/)

### Google Antigravity

- Praise, 2026-09-21, r/google_antigravity (Reddit): “keeps the ais answers more streamliked. i don’t have a lot of external stuff installed but this one i got.” [source](https://www.reddit.com/r/google_antigravity/comments/1wmmzit/top_10_antigravity_skill_repos/pb88n7a/)
- Praise, 2026-09-20, r/google_antigravity (Reddit): “context ? i have 3 subscriptions, claude, copilot and gemini, use all of them regularly, gemini went to number one in speed and accuracy followed by claude. when comes to building uis for example, gemini is far superior, the agy’s summaries with screenshots are absolutely excellent. overall. incredible speed improvement” [source](https://www.reddit.com/r/google_antigravity/comments/1wllxyb/this_thing_became_a_coding_beast/pb0mvsd/)
- Praise, 2026-09-19, r/google_antigravity (Reddit): “i like the extra "koala express" 😛 ahh the writer that is sonnet just has to write something 😅” [source](https://www.reddit.com/r/google_antigravity/comments/1wk8x4f/anyone_else_getting_gemini_4_under_flash_38_model/paq9a9w/)
- Complaint, 2026-09-16, r/google_antigravity (Reddit): “yes, although my suspicion is that the harness is more of a problem than the model itself. the permissions checks alone are maddening - the sandbox is an improvement but it desperately needs an auto review mode like cc and codex have. beyond that, i find it goes in circles a lot, especially if given directives more vague than “look at this file for this thing”. it benchmarks well but in practice feels like using luna in a much worse harness than” [source](https://www.reddit.com/r/google_antigravity/comments/1whvudv/people_who_are_complaining_about_gemini/pa6xu4g/)
- Complaint, 2026-09-15, r/google_antigravity (Reddit): “yea the gemini flash models like to write so much thats why people say it behaves bad for exemple glm 5.3 flash wouldnt” [source](https://www.reddit.com/r/google_antigravity/comments/1wgsxwn/how_to_make_antigravity_slow_down_and_stop/p9wxlqd/)
- Complaint, 2026-09-08, r/google_antigravity (Reddit): “google models are decent but its hard to find the setup that outperforms other models usually it writes too much, too optimistic. if i want to write an algorithm, i like to use a 3 or 5-shot on gemini and then pass to claude / chatgpt write the final version” [source](https://www.reddit.com/r/google_antigravity/comments/1waolu4/why_does_antigravity_have_so_few_users/p8m6pon/)

### Devin

- Praise, 2026-09-27, @DevinAI (X): “@ashtweetsonx @devinai makes e2e tests really easy, can run the whole app on both windows/mac with audio support other than that explanations are more coherent than claude” [source](https://twitter.com/792026036716175360/status/2104030537170174179)
- Praise, 2026-09-26, @cognition (X): “@markknd @cognition i like the quick and clean unraveling of the story from the initial bullet point.” [source](https://twitter.com/2017384758452621312/status/2103907809670971431)
- Praise, 2026-09-22, @cognition (X): “@1kartikkabadi1 @cognition the summary breakdown on that pr looks crazy clean tbh” [source](https://twitter.com/1749093765078523904/status/2102202305219281253)
- Complaint, 2026-09-24, @DevinAI (X): “@devinai less yap more taps <strict_link>” [source](https://twitter.com/1364957012367446026/status/2103254261908074910)
- Complaint, 2026-09-15, @DevinAI (X): “i don’t understand how it differs from claude code’s workflows or “advisor mode”? i’m on 5x max plan and almost exclusive fire a task with fable high and it runs a swarm of opus 5 agents. the flow is exactly the same as fusion by default. dirty work on opus, decisions on fable. usage is brilliant, opus is very good and efficient model for code, it’s just painful to speak with.” [source](https://twitter.com/119424087/status/2099901937432842495)
- Complaint, 2026-09-12, @cognition (X): “i love the overload of ai generated comments on the split for planning and execution🤣 why bother anyway, good work, now we need to see the acceptance of a completed task after human review, i bet it wouldn’t hold and we would need more requests to complete a task and costs will stay the same / similar” [source](https://twitter.com/71076617/status/2098707201313406986)

### GitHub Copilot

- Praise, 2026-09-21, r/GithubCopilot (Reddit): “i’ve been a fan of grok because its verbiage is the right balance. it’s honestly a good all round model.” [source](https://www.reddit.com/r/GithubCopilot/comments/1wmho7l/grok_47_is_now_available_in_github_copilot/pb7yjui/)
- Complaint, 2026-09-25, r/GithubCopilot (Reddit): “i agree. i find the praise for luna baffling. it may be cheap, but most people seem to recommend running it at xhigh or max, where it takes forever to reason and generates bloated output.” [source](https://www.reddit.com/r/GithubCopilot/comments/1wp1ygm/tiers_of_auto_now_available/pbwf62p/)
- Complaint, 2026-09-24, r/GithubCopilot (Reddit): “my only complaint is the excess commenting but i think it is a bit subdued over opus 5.0” [source](https://www.reddit.com/r/GithubCopilot/comments/1wngkdo/opus_55_rollout_and_review/pbs8iin/)
- Complaint, 2026-09-18, r/devops (Reddit): “yes, i'm in networking but we are migrating to a new vendor and my senior just handed me like 15 documents, each 5-15 pages long entirely of copilot slop. its unreadable and i have to re-write it so we can understand” [source](https://www.reddit.com/r/devops/comments/1wjqrl3/rant_has_anyone_experienced_vibe_planning/palo558/)

### Pi

- Praise, 2026-09-14, r/PiCodingAgent (Reddit): “depends on what's important to you as the end result summary. i built a metrics tracking extension that's specific to my workflow and a little cli interface for me or the agent to query it. per task i'm usually happy with a files edited/diff and a two sentence summary of what it does and how it maintains alignment with project intent. i don't worry too much about tool calls and other deterministic activity per task because it's all controlled by” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wgaln5/should_pi_have_a_humanfacing_execution_summary/p9sujes/)
- Praise, 2026-09-02, @pidotdev (X): “kinda loving using @pidotdev i reply "works" it responds "nice" no thesis statement or gigantic 5,000tk responses. just "nice" or "sweet", similar.” [source](https://twitter.com/1916104214398267392/status/2095266005157294364)
- Complaint, 2026-09-22, @pidotdev (X): “@pidotdev remote control native support. better compaction. human-writing like :)” [source](https://twitter.com/89646598/status/2102407468450296138)
- Complaint, 2026-09-20, @pidotdev (X): “@nnennahacks yeh i use it in my @pidotdev setup some of the models just talk too damn much lol” [source](https://twitter.com/1377012764464517130/status/2101695233679380693)

### Amp

- Praise, 2026-09-10, @AmpCode (X): “appreciate @ampcode on the little things - slack message was succinct, everything relevant linked <strict_link>” [source](https://twitter.com/2000470921329709056/status/2098014756703531053)
- Praise, 2026-09-01, @AmpCode (X): “for bug fixes and troubleshooting, @ampcode really nails the kind of coding agent i want out of the box. it's precise, minimal, and non-verbose. it fixes the issue, changes what needs changing, and gets out of the way.” [source](https://twitter.com/347237526/status/2094883834227470647)
- Praise, 2026-08-31, @AmpCode (X): “codex cli now has `codex remote-control`. kinda `--no-tui` in @ampcode but much more verbose.” [source](https://twitter.com/103563938/status/2094376768371036477)
- Complaint, 2026-09-14, @AmpCode (X): “it would be _very_ cool if the amp thread also suggested like 2-3 follow on prompts that are buttons that i could just click. right now a lot of the issue is each prompt gives a wall of text unless i explicitly prompt "2-3 paragraphs". i suspect more concise prompts with suggested follow-on prompts is a nicer ux” [source](https://twitter.com/721234540/status/2099486641966531069)

### Cline

- Praise, 2026-09-12, r/codex (Reddit): “my 2 cents... a while back i switched from claude and codex to cline (vscode) and deliberately chose to use chinese llms. initially deepseek v4 flash / pro (v good) , recently glm 5.3 flash (excellent+ all-rounder) qwen 3.8 max (excellent ++ at frontends /good general coding) and now v4.1 flash (excellent++ agentic all-rounder) etc primarily for the cost aspect. i've gone from spending €4/600 month to €40/60 .. more importantly the code quality /” [source](https://www.reddit.com/r/codex/comments/1wenzkp/codex_vs_zcode/p9fhlc6/)

### Zed

- Complaint, 2026-09-20, r/ZedEditor (Reddit): “i'm not reading like 9 full pages of llm output to understand why this fork exists and the screenshot just looks like themed zed running opencode in a terminal thread.” [source](https://www.reddit.com/r/ZedEditor/comments/1wjdiff/i_am_bulding_dez_a_fork_on_zed/pavxzoy/)

### Factory

- Praise, 2026-09-22, @FactoryAI (X): “@factoryai @anthropicai wait fewer tokens and clearer answers is a nasty combo” [source](https://twitter.com/1572280352093126657/status/2102461541341700114)

### Kiro

- Praise, 2026-09-10, r/kiroIDE (Reddit): “i’ve tried a bunch of different clients and models out there, and prefer kiro + one of the claude 4.8 models to most of them, including claude code cli which i find too cutesy and verbose by comparison. honestly this is more of a matter of tastes + skills issue on your part.” [source](https://www.reddit.com/r/kiroIDE/comments/1wchsii/frontier_models_on_other_providers_vs_the_same/p8zeoia/)
