# Augment Code (Augment)

Feedback Bench, coding agents, built 2026-10-01, window 2026-08-31 to 2026-09-27. Web page: https://feedbackbench.com/#/agent/augment

| Measure | Value |
|---|---|
| Rank | 17 of 17 (rank range 17–17) |
| Feedback Score | 9.7 (95% interval 9.7–9.7) |
| Popularity | 0.019 (share of voice 0.04%) |
| Customer love | 0.495 (95% interval 0.491–0.499) |
| Top quadrant | no |
| Authors | 40 |
| Posts counted | 46 |
| Posts that judge the agent | 32 |
| Criteria better / worse than peers | 0 / 0 of 63 |

## The brief

Written by Claude Opus 5.5 from 12 labelled posts and the numbers on this page. Interpretation, not measurement: every quote is verbatim and links to its post.

**A context engine users love, behind credits they cannot predict.**

TL;DR:

- Users say it pulls the right files from huge codebases without being told where to look.
- Every pricing post is a complaint. Credits burn fast, meters surprise, and plans feel expensive.
- Retrieval quality draws people in. Cost pushes them out, and posts show some already leaving.

### What hurts

- **Credits drain within a single session** ([Single prompt, model or effort level consumes disproportionate quota](https://feedbackbench.com/criteria/limits.burn_rate.md), [How much use a plan's price buys](https://feedbackbench.com/criteria/limits.plan_value.md)). Heavy runs eat plan credits faster than users expect, so a monthly sub can feel like a daily one.
  The sharpest report has a 22-minute session with one model consuming a large slice of a plan's credits. The user asks whether the plan will last a day.
  
  G2 reviewers who otherwise praise the tool say the same thing more calmly. Back-to-back multi-file refactors and all-day multi-agent work force them to watch the credit counter instead of the code.
  Evidence:
  - Complaint, Augment Code, @augmentcode, 2026-09-08: “@augmentcode lol im using kimi 2.7 code and 22mins, it used up 12% of my plans credits. so a $20 sub wont even last 24hours?” [source](https://twitter.com/2002397531226148864/status/2097377115071287410)
  - Praise, Augment Code, G2, 2026-09-23: “q: what problems is the product solving and how is that benefiting you? a: it takes away all the pain of digging through massive codebases manually tracking down cross-file dependencies. i save about 1-2 hours of grunt work every single day just in refactoring and boilerplate. q: what do you like best about the product? a: the context is really amazing. in contrast to a simple ai autocomplete function, it creates an index of all of your multi-repo architecture and understands how your microvaves and dependencies work. the extension works great without any delays for vs code and jetbrains. q: what do you dislike about the product? a: the credit system feels a bit restrictive when you run heavy multi-file refactoring tasks back-to-back. also, support is mostly ticket-based on standard plans, so you won't get instant live help if you run into edge-case bugs.” [source](https://www.g2.com/products/augment-code/reviews/augment-code-review-13582326)
  - Praise, Augment Code, G2, 2026-09-26: “q: what problems is the product solving and how is that benefiting you? a: it completely eliminates the tedious manual effort of navigating unfamiliar codebases, writing repetitive boilerplate, and tracing cross-field dependencies saving me 1-2 hours of grunt work every day. q: what do you like best about the product? a: the context engine is really amazing. in contrast to a simple ai autocomplete function, it creates an index of all of your multi-repo architecture and understands how your microservices and dependencies work. the extension works great without any delays for vs code and jetbrains. q: what do you dislike about the product? a: the usage pricing and token credit limit require active monitoring if you are running multi-agent tasks all day. additionally, support in standard tiers relies on ticket portals rather than instant live assistance.” [source](https://www.g2.com/products/augment-code/reviews/augment-code-review-13618001)
  - Complaint, Augment Code, r/cscareerquestions, 2026-09-18: “have you tried switching to lower models? the higher ones are overkill if you're using it to augment the engineering that you are doing rather than trying to rely on them to do the design/software architecting for you. sonnet 5 is plenty capable when given clear accurate instructions, and it's faster and much more token efficient. i'll use opus occasionally for some harder tasks, but i find both opus and fable tend to over complicate everything, take forever to do what you asked, and burn through tokens for minimal gain.” [source](https://www.reddit.com/r/cscareerquestions/comments/1wjk5vg/anyone_frustrated_with_how_ai_harnesses_are/paklj9u/)

- **Price outgrew the value for some** ([How much use a plan's price buys](https://feedbackbench.com/criteria/limits.plan_value.md), [Price, allowance or plan terms changed](https://feedbackbench.com/criteria/limits.allowance_change.md)). Users who liked the indexing say the tool became too expensive to keep using, and the cheaper terms they remember did not last.
  Posts frame cost as the reason to stop, not quality. One user calls the codebase indexing fast and accurate, then says it got too expensive.
  
  Another recalls a fixed request bundle that was too good to pass up and notes it ended. A G2 reviewer who rates it above competitors still lists expense as the main dislike.
  Evidence:
  - Complaint, Augment Code, r/PiCodingAgent, 2026-09-01: “augment code does it pretty well. it automatically returns the right code very quickly (codebase indexing) but it got too expensive to use.” [source](https://www.reddit.com/r/PiCodingAgent/comments/1w4hefn/extensionstools_for_context_retrieval/p77wk8r/)
  - Praise, Augment Code, r/ClaudeCode, 2026-09-13: “glm 5.3 and glm 5.3 flash are no go for me, i tried their promo weekend tokens and found them underwhelming in more ways than i can tolerate. most importantly, they embedded a lot of confidently incomplete information in designs which failed audits and got blocked during implementation. (i used augmentcode before switching to claudecode, then after a month with claudecode i went back to augmentcode, again this was about a year ago and i found augmentcode context engineering much better at that time and their fixed pricing for 1500 requests were too good to pass which didn't last of course )” [source](https://www.reddit.com/r/ClaudeCode/comments/1wf9uuf/wth_is_going_on_with_claude_usage_limits/p9lym97/)
  - Complaint, Augment Code, G2, 2026-09-10: “q: what problems is the product solving and how is that benefiting you? a: code analyzing, editing, autocompleting, very good for professional developers. good as compared to other software in the market. q: what do you like best about the product? a: it has excellent fetching quality of code database. i could easily integrate it with github. the speed is good as compared to other products in the market. it also edits and autocompletes your code. q: what do you dislike about the product? a: it is an expensive software also i have used it i could feel a little bugs in customer support system.” [source](https://www.g2.com/products/augment-code/reviews/augment-code-review-13433460)

- **Usage meter leaves users guessing** ([Usage meter visibility and accuracy](https://feedbackbench.com/criteria/limits.usage_meter.md), [Pricing and plan terms stated clearly and consistently](https://feedbackbench.com/criteria/billing.pricing_clarity.md)). Users run out of tokens without warning because the pricing structure and remaining balance are hard to read.
  The complaint is not only the price but the surprise. One user ran out of tokens unexpectedly, calls the structure unclear, and says they would not recommend the product.
  
  A G2 reviewer describes the credit limit as something that needs active monitoring. That is the opposite of a meter that does the watching for you.
  Evidence:
  - Complaint, Augment Code, @augmentcode, 2026-09-18: “@augmentcode regardless of whether it merges code or offers optimizations, the subscription for augment is very expensive. i ran out of tokens unexpectedly without even realizing it, and the pricing structure is unclear—i wouldn't recommend using it.” [source](https://twitter.com/955391612632252418/status/2100826713957810177)
  - Praise, Augment Code, G2, 2026-09-26: “q: what problems is the product solving and how is that benefiting you? a: it completely eliminates the tedious manual effort of navigating unfamiliar codebases, writing repetitive boilerplate, and tracing cross-field dependencies saving me 1-2 hours of grunt work every day. q: what do you like best about the product? a: the context engine is really amazing. in contrast to a simple ai autocomplete function, it creates an index of all of your multi-repo architecture and understands how your microservices and dependencies work. the extension works great without any delays for vs code and jetbrains. q: what do you dislike about the product? a: the usage pricing and token credit limit require active monitoring if you are running multi-agent tasks all day. additionally, support in standard tiers relies on ticket portals rather than instant live assistance.” [source](https://www.g2.com/products/augment-code/reviews/augment-code-review-13618001)

- **Ships scope nobody asked for** ([Does unrequested work or over-engineers](https://feedbackbench.com/criteria/work.scope_overreach.md)). Users say cheaper runs do not fix the agent adding invented scope, because nothing in the loop refuses work outside the request.
  The scope complaint is a gate problem, not a cost problem. Users say lowering the price per run leaves the overreach in place.
  
  A related post on model choice says higher-tier models tend to overcomplicate tasks and burn tokens for minimal gain. That ties overreach straight back to the credit drain.
  Evidence:
  - Complaint, Augment Code, @augmentcode, 2026-09-22: “@augmentcode @anthropicai 40% cheaper still ships invented scope if nobody owns refuse. cost is the easy dial. the gate isn't.” [source](https://twitter.com/1610522467264565249/status/2102485835211821392)
  - Complaint, Augment Code, r/cscareerquestions, 2026-09-18: “have you tried switching to lower models? the higher ones are overkill if you're using it to augment the engineering that you are doing rather than trying to rely on them to do the design/software architecting for you. sonnet 5 is plenty capable when given clear accurate instructions, and it's faster and much more token efficient. i'll use opus occasionally for some harder tasks, but i find both opus and fable tend to over complicate everything, take forever to do what you asked, and burn through tokens for minimal gain.” [source](https://www.reddit.com/r/cscareerquestions/comments/1wjk5vg/anyone_frustrated_with_how_ai_harnesses_are/paklj9u/)

### What works

- **Finds the right files unprompted** ([Finding the right files in the codebase](https://feedbackbench.com/criteria/context.codebase_retrieval.md)). The context engine fetches relevant files from massive codebases with no file references in the prompt, and users rank it above Copilot's.
  Users describe tossing a bare prompt at it and getting the correct files back almost instantly. One Reddit user compares it directly with Copilot and says Augment wins on large codebases.
  
  Another user switched to a different agent, then came back because Augment's context engineering was stronger.
  Evidence:
  - Praise, Augment Code, r/GithubCopilot, 2026-09-13: “he's basically saying that augment code has a better context engine than what copilot is offering. copilot's is good too, it's just not as good as augment on those massive codebases. you can literally just toss any prompt at it without specifying a single file or reference and it will fetch the correct files in 1 second” [source](https://www.reddit.com/r/GithubCopilot/comments/1ovwwlk/context_engine_for_github_copilot/p9miyd7/)
  - Praise, Augment Code, r/ClaudeCode, 2026-09-13: “glm 5.3 and glm 5.3 flash are no go for me, i tried their promo weekend tokens and found them underwhelming in more ways than i can tolerate. most importantly, they embedded a lot of confidently incomplete information in designs which failed audits and got blocked during implementation. (i used augmentcode before switching to claudecode, then after a month with claudecode i went back to augmentcode, again this was about a year ago and i found augmentcode context engineering much better at that time and their fixed pricing for 1500 requests were too good to pass which didn't last of course )” [source](https://www.reddit.com/r/ClaudeCode/comments/1wf9uuf/wth_is_going_on_with_claude_usage_limits/p9lym97/)
  - Praise, Augment Code, G2, 2026-09-23: “q: what problems is the product solving and how is that benefiting you? a: it takes away all the pain of digging through massive codebases manually tracking down cross-file dependencies. i save about 1-2 hours of grunt work every single day just in refactoring and boilerplate. q: what do you like best about the product? a: the context is really amazing. in contrast to a simple ai autocomplete function, it creates an index of all of your multi-repo architecture and understands how your microvaves and dependencies work. the extension works great without any delays for vs code and jetbrains. q: what do you dislike about the product? a: the credit system feels a bit restrictive when you run heavy multi-file refactoring tasks back-to-back. also, support is mostly ticket-based on standard plans, so you won't get instant live help if you run into edge-case bugs.” [source](https://www.g2.com/products/augment-code/reviews/augment-code-review-13582326)

- **Rivals' users want it copied** ([Finding the right files in the codebase](https://feedbackbench.com/criteria/context.codebase_retrieval.md), [IDE and editor integration](https://feedbackbench.com/criteria/setup.ide_integration.md)). Users of other agents publicly ask those vendors to build an Augment-style context engine, which makes it the feature others get measured against.
  One post begs a competing vendor to add a context engine like Augment's, or to buy Augment outright.
  
  G2 reviewers explain the pull. The tool indexes multi-repo architecture, maps cross-service dependencies, and saves them hours of grunt work. They also say the VS Code and JetBrains extensions run without lag.
  Evidence:
  - Praise, Augment Code, @augmentcode, 2026-09-08: “@bcherny @addyosmani @anthropicai boris please implement a context engine to your models. better yet take over @augmentcode and become unstoppable. i fucking beg you” [source](https://twitter.com/1795256572102209537/status/2097359764678214092)
  - Praise, Augment Code, G2, 2026-09-23: “q: what problems is the product solving and how is that benefiting you? a: it takes away all the pain of digging through massive codebases manually tracking down cross-file dependencies. i save about 1-2 hours of grunt work every single day just in refactoring and boilerplate. q: what do you like best about the product? a: the context is really amazing. in contrast to a simple ai autocomplete function, it creates an index of all of your multi-repo architecture and understands how your microvaves and dependencies work. the extension works great without any delays for vs code and jetbrains. q: what do you dislike about the product? a: the credit system feels a bit restrictive when you run heavy multi-file refactoring tasks back-to-back. also, support is mostly ticket-based on standard plans, so you won't get instant live help if you run into edge-case bugs.” [source](https://www.g2.com/products/augment-code/reviews/augment-code-review-13582326)
  - Praise, Augment Code, G2, 2026-09-26: “q: what problems is the product solving and how is that benefiting you? a: it completely eliminates the tedious manual effort of navigating unfamiliar codebases, writing repetitive boilerplate, and tracing cross-field dependencies saving me 1-2 hours of grunt work every day. q: what do you like best about the product? a: the context engine is really amazing. in contrast to a simple ai autocomplete function, it creates an index of all of your multi-repo architecture and understands how your microservices and dependencies work. the extension works great without any delays for vs code and jetbrains. q: what do you dislike about the product? a: the usage pricing and token credit limit require active monitoring if you are running multi-agent tasks all day. additionally, support in standard tiers relies on ticket portals rather than instant live assistance.” [source](https://www.g2.com/products/augment-code/reviews/augment-code-review-13618001)

### Fine print

- Fewer than fifty posts in the window. Every criterion is flagged as too few posts, so treat patterns as directional.
- Support complaints appear only inside long G2 reviews, so no card leads on them despite consistent ticket-only grumbles.
- One retrieval post recalls an experience from about a year earlier, not this window.

## Top requests

What users ask to add or change, most asked first. 2 author-weeks ask for something. Requests do not change the Feedback Score. Rule: A separate pass by Claude Sonnet 5 reads every counted post and extracts what the author asks the agent or its vendor to add or change, with the criteria it maps to and a short normalised wording; it does not touch the labels or the Feedback Score. Claude Opus 5.5 groups the wordings within each criterion (the first criterion the request maps to) into themes; code counts them. A theme counts distinct author-weeks that ask for it, per agent; across agents, one author-week per agent. Themes asked in fewer than 2 author-weeks, and requests that share no theme, are not shown. Examples: up to 3 posts per theme from different authors, without slurs, preferring posts of 60 to 450 characters, most recent first.

No request is asked for in enough author-weeks to show.

## Facts

| Fact | Value |
|---|---|
| Version | Cosmos 'Unified Agents Platform' (team agent fleets across the SDLC); Intent desktop workspace (public beta, macOS); Auggie CLI 0.36.0 (2026-08-21); VS Code extension 0.901.1 and IntelliJ plugin v0.491.0 (2026-09-14); Context Engine MCP (GA 2026-02-06). Inline Completions and Next Edit sunset 2026-03-31 on Indie/Standard/Legacy plans |
| Released | IDE extensions VS Code 0.901.1 / IntelliJ v0.491.0: 2026-09-14; Cosmos launch: 2026-06-05; Intent public beta: 2026-02-26; new $20/mo Standard tier first seen on pricing page 2026-09-20 (exact go-live date not confirmed) |
| Price | Standard $20/mo (includes $20 of usage, up to 50 seats, pooled), Business $100/mo (includes $100 of usage, up to 50 seats), Enterprise custom; overage billed at LLM provider list price plus 40% service fee, plus Context Engine and Cosmos compute; top-ups valid 12 months |
| Model | Multi-vendor model picker: Claude (Fable 5.1, Fable 5, Opus 5, Opus 4.6-4.8, Sonnet 5, Sonnet 4.6, Haiku 4.5), Gemini (3.1 Pro, 3.7/3.8 Flash), OpenAI GPT, xAI Grok, Zhipu GLM, Moonshot Kimi, plus Augment's own 'Prism' routing; Intent also runs BYO agents (Claude Code, Codex, OpenCode) |
| Surface | VS Code and JetBrains extensions, Auggie CLI (terminal, also runs as MCP server), Intent desktop app (macOS beta), Cosmos web platform, Context Engine MCP for third-party agents |

## Sources

| Channel | Source | Posts |
|---|---|---|
| X | @augmentcode | 28 |
| Reddit | Posts that name it | 12 |
| G2 | G2 | 6 |

## Better than peers on

None.

## Worse than peers on

None.

## All 63 criteria

Criterion love: 0.5 is the category norm. n: rated author-weeks.

### Paying and limits: Too few posts (customer love 0.487, n 8)

| Criterion | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|
| [How much use a plan's price buys](https://feedbackbench.com/criteria/limits.plan_value.md) | Too few posts | 0.495 | 0.486–0.506 | 5 | 1 | 4 |
| [Single prompt, model or effort level consumes disproportionate quota](https://feedbackbench.com/criteria/limits.burn_rate.md) | Too few posts | 0.494 | 0.488–0.499 | 4 | 0 | 4 |
| [Price, allowance or plan terms changed](https://feedbackbench.com/criteria/limits.allowance_change.md) | Too few posts | 0.497 | 0.494–0.500 | 2 | 0 | 2 |
| [Usage meter visibility and accuracy](https://feedbackbench.com/criteria/limits.usage_meter.md) | Too few posts | 0.497 | 0.493–0.500 | 2 | 0 | 2 |
| [Pricing and plan terms stated clearly and consistently](https://feedbackbench.com/criteria/billing.pricing_clarity.md) | Too few posts | 0.499 | 0.496–0.500 | 1 | 0 | 1 |
| [Short rolling usage window blocks or interrupts work](https://feedbackbench.com/criteria/limits.window_interrupts_work.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Quota reset timing and bonus or banked resets](https://feedbackbench.com/criteria/limits.reset_schedule.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Prompt cache hits, misses and invalidation](https://feedbackbench.com/criteria/limits.prompt_cache.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Pay-as-you-go overage, fallback billing and spend caps](https://feedbackbench.com/criteria/billing.overage_charges.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Free tier and free model availability and limits](https://feedbackbench.com/criteria/billing.free_tier.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Using an existing subscription across tools](https://feedbackbench.com/criteria/billing.subscription_portability.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |

Most recent posts:

- Praise, 2026-09-13, r/ClaudeCode (Reddit): “glm 5.3 and glm 5.3 flash are no go for me, i tried their promo weekend tokens and found them underwhelming in more ways than i can tolerate. most importantly, they embedded a lot of confidently incomplete information in designs which failed audits and got blocked during implementation. (i used augmentcode before switching to claudecode, then after a month with claudecode i went back to augmentcod…” [source](https://www.reddit.com/r/ClaudeCode/comments/1wf9uuf/wth_is_going_on_with_claude_usage_limits/p9lym97/)
- Complaint, 2026-09-26, G2 (G2): “q: what problems is the product solving and how is that benefiting you? a: it completely eliminates the tedious manual effort of navigating unfamiliar codebases, writing repetitive boilerplate, and tracing cross-field dependencies saving me 1-2 hours of grunt work every day. q: what do you like best about the product? a: the context engine is really amazing. in contrast to a simple ai autocomplete…” [source](https://www.g2.com/products/augment-code/reviews/augment-code-review-13618001)
- Complaint, 2026-09-23, G2 (G2): “q: what problems is the product solving and how is that benefiting you? a: it takes away all the pain of digging through massive codebases manually tracking down cross-file dependencies. i save about 1-2 hours of grunt work every single day just in refactoring and boilerplate. q: what do you like best about the product? a: the context is really amazing. in contrast to a simple ai autocomplete func…” [source](https://www.g2.com/products/augment-code/reviews/augment-code-review-13582326)
- Complaint, 2026-09-18, @augmentcode (X): “@augmentcode regardless of whether it merges code or offers optimizations, the subscription for augment is very expensive. i ran out of tokens unexpectedly without even realizing it, and the pricing structure is unclear—i wouldn't recommend using it.” [source](https://twitter.com/955391612632252418/status/2100826713957810177)

### Setting up and connecting: Too few posts (customer love 0.499, n 5)

| Criterion | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|
| [IDE and editor integration](https://feedbackbench.com/criteria/setup.ide_integration.md) | Too few posts | 0.504 | 0.500–0.509 | 3 | 3 | 0 |
| [Onboarding, discoverability and documentation](https://feedbackbench.com/criteria/setup.onboarding_docs.md) | Too few posts | 0.499 | 0.494–0.504 | 2 | 1 | 1 |
| [Install, launch and sign-in](https://feedbackbench.com/criteria/setup.install_signin.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Connecting own API keys, local models and custom endpoints](https://feedbackbench.com/criteria/setup.provider_byok_local.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [MCP servers, plugins, skills and hooks](https://feedbackbench.com/criteria/setup.extensions_mcp.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |

Most recent posts:

- Praise, 2026-09-26, G2 (G2): “q: what problems is the product solving and how is that benefiting you? a: it completely eliminates the tedious manual effort of navigating unfamiliar codebases, writing repetitive boilerplate, and tracing cross-field dependencies saving me 1-2 hours of grunt work every day. q: what do you like best about the product? a: the context engine is really amazing. in contrast to a simple ai autocomplete…” [source](https://www.g2.com/products/augment-code/reviews/augment-code-review-13618001)
- Praise, 2026-09-23, G2 (G2): “q: what problems is the product solving and how is that benefiting you? a: it takes away all the pain of digging through massive codebases manually tracking down cross-file dependencies. i save about 1-2 hours of grunt work every single day just in refactoring and boilerplate. q: what do you like best about the product? a: the context is really amazing. in contrast to a simple ai autocomplete func…” [source](https://www.g2.com/products/augment-code/reviews/augment-code-review-13582326)
- Praise, 2026-09-17, G2 (G2): “q: what problems is the product solving and how is that benefiting you? a: i use it to build, refactor and manage a large codebase. work like that gets tedious when changes touch a lot of files, and having augment code help with it makes those tasks more manageable. it's only been a week or two, so i don't have hard numbers, but it has made refactoring work feel less of a chore. q: what do you lik…” [source](https://www.g2.com/products/augment-code/reviews/augment-code-review-13367494)
- Complaint, 2026-09-01, G2 (G2): “q: what problems is the product solving and how is that benefiting you? a: my workflow involves generating code across multiple platforms, which often means the output needs refinement before it's actually usable. augment code solves that last-mile problem, it takes rough, generated code and polishes it into something cleaner and more reliable. the benefit is that i'm spending less time manually r…” [source](https://www.g2.com/products/augment-code/reviews/augment-code-review-13394203)

### Choosing models: Too few posts (customer love 0.498, n 1)

| Criterion | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|
| [Quality got worse or better over time](https://feedbackbench.com/criteria/models.quality_drift.md) | Too few posts | 0.498 | 0.494–0.500 | 1 | 0 | 1 |
| [Which models are offered on a plan and when](https://feedbackbench.com/criteria/models.catalog_access.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Automatic model routing and fallback](https://feedbackbench.com/criteria/models.routing_auto.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Reasoning effort setting and its defaults](https://feedbackbench.com/criteria/models.effort_control.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |

Most recent posts:

- Complaint, 2026-09-12, @augmentcode (X): “@simplygandan @augmentcode also, sonnet used to feel good enough. now, even opus feels dumb! maybe, augment code was that good or we are spoilt by fable and astra” [source](https://twitter.com/2959524282/status/2098864871924478317)

### Instructing and context: Too few posts (customer love 0.520, n 8)

| Criterion | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|
| [Finding the right files in the codebase](https://feedbackbench.com/criteria/context.codebase_retrieval.md) | Too few posts | 0.517 | 0.507–0.529 | 8 | 8 | 0 |
| [Persistent project rules files are read and obeyed](https://feedbackbench.com/criteria/context.instruction_files.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Direct in-prompt instructions and caps are followed](https://feedbackbench.com/criteria/context.instruction_following.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Asks the user versus guessing](https://feedbackbench.com/criteria/context.clarifying_questions.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Output degrades as the context window fills](https://feedbackbench.com/criteria/context.long_context_decay.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Context compaction keeps what matters, cheaply and quickly](https://feedbackbench.com/criteria/context.compaction.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Memory and state carried across sessions](https://feedbackbench.com/criteria/context.session_memory.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Images, PDFs and file attachments as input](https://feedbackbench.com/criteria/context.attachments.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |

Most recent posts:

- Praise, 2026-09-26, G2 (G2): “q: what problems is the product solving and how is that benefiting you? a: it completely eliminates the tedious manual effort of navigating unfamiliar codebases, writing repetitive boilerplate, and tracing cross-field dependencies saving me 1-2 hours of grunt work every day. q: what do you like best about the product? a: the context engine is really amazing. in contrast to a simple ai autocomplete…” [source](https://www.g2.com/products/augment-code/reviews/augment-code-review-13618001)
- Praise, 2026-09-23, G2 (G2): “q: what problems is the product solving and how is that benefiting you? a: it takes away all the pain of digging through massive codebases manually tracking down cross-file dependencies. i save about 1-2 hours of grunt work every single day just in refactoring and boilerplate. q: what do you like best about the product? a: the context is really amazing. in contrast to a simple ai autocomplete func…” [source](https://www.g2.com/products/augment-code/reviews/augment-code-review-13582326)
- Praise, 2026-09-13, r/ClaudeCode (Reddit): “glm 5.3 and glm 5.3 flash are no go for me, i tried their promo weekend tokens and found them underwhelming in more ways than i can tolerate. most importantly, they embedded a lot of confidently incomplete information in designs which failed audits and got blocked during implementation. (i used augmentcode before switching to claudecode, then after a month with claudecode i went back to augmentcod…” [source](https://www.reddit.com/r/ClaudeCode/comments/1wf9uuf/wth_is_going_on_with_claude_usage_limits/p9lym97/)

### Doing the work: Too few posts (customer love 0.495, n 7)

| Criterion | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|
| [Can do the user's kind of task](https://feedbackbench.com/criteria/work.capability.md) | Too few posts | 0.500 | 0.491–0.509 | 4 | 3 | 1 |
| [Does unrequested work or over-engineers](https://feedbackbench.com/criteria/work.scope_overreach.md) | Too few posts | 0.497 | 0.493–0.500 | 2 | 0 | 2 |
| [Risky or irreversible actions without confirmation](https://feedbackbench.com/criteria/work.destructive_actions.md) | Too few posts | 0.499 | 0.495–0.500 | 1 | 0 | 1 |
| [Frontend and visual UI output](https://feedbackbench.com/criteria/work.frontend_ui.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Diagnosing and fixing reported bugs](https://feedbackbench.com/criteria/work.bug_diagnosis.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Breaks existing code or reintroduces bugs](https://feedbackbench.com/criteria/work.regressions_introduced.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Spins, loops or gets stuck without progress](https://feedbackbench.com/criteria/work.stuck_loops.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Stops mid-task or answers instead of acting](https://feedbackbench.com/criteria/work.premature_stop.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Long unattended runs and goal/loop mode](https://feedbackbench.com/criteria/work.long_running_autonomy.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Subagents, parallel agents and orchestrators](https://feedbackbench.com/criteria/work.multi_agent_orchestration.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Games checks instead of fixing the problem](https://feedbackbench.com/criteria/work.reward_hacking.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Git commits, branches and sync](https://feedbackbench.com/criteria/work.git_workflow.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Computer use and browser control](https://feedbackbench.com/criteria/work.computer_browser_use.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Safety filters block legitimate coding tasks](https://feedbackbench.com/criteria/work.safety_refusals.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Tool approval prompts and autonomy modes](https://feedbackbench.com/criteria/work.permission_prompts.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Plan-before-edit mode](https://feedbackbench.com/criteria/work.plan_mode.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Length and clarity of replies, summaries and comments](https://feedbackbench.com/criteria/work.response_verbosity.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Caves to or argues with the user's judgement](https://feedbackbench.com/criteria/work.sycophancy_pushback.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |

Most recent posts:

- Praise, 2026-09-23, r/ExperiencedDevs (Reddit): “i would definitely start getting comfortable with it, you don't need to let it be an agent and do everything for you. it can be fun to figure out where that boundary is. i work in a small team that owns and maintains several software systems, from vendor based to integration layers and some full stack software with both internal and customer users, so knowing everything about everything is effecti…” [source](https://www.reddit.com/r/ExperiencedDevs/comments/1wo4d8p/job_requiresuses_very_little_ai_sinking_ship_or/pblbka0/)
- Praise, 2026-09-17, G2 (G2): “q: what problems is the product solving and how is that benefiting you? a: i use it to build, refactor and manage a large codebase. work like that gets tedious when changes touch a lot of files, and having augment code help with it makes those tasks more manageable. it's only been a week or two, so i don't have hard numbers, but it has made refactoring work feel less of a chore. q: what do you lik…” [source](https://www.g2.com/products/augment-code/reviews/augment-code-review-13367494)
- Praise, 2026-09-01, G2 (G2): “q: what problems is the product solving and how is that benefiting you? a: my workflow involves generating code across multiple platforms, which often means the output needs refinement before it's actually usable. augment code solves that last-mile problem, it takes rough, generated code and polishes it into something cleaner and more reliable. the benefit is that i'm spending less time manually r…” [source](https://www.g2.com/products/augment-code/reviews/augment-code-review-13394203)
- Complaint, 2026-09-22, @augmentcode (X): “@augmentcode @anthropicai 40% cheaper still ships invented scope if nobody owns refuse. cost is the easy dial. the gate isn't.” [source](https://twitter.com/1610522467264565249/status/2102485835211821392)
- Complaint, 2026-09-19, @augmentcode (X): “the fleet fixing ci and conflicts is the write. a briefing can look complete while a conflict resolution already pushed the wrong change into the branch. humans approve and merge only if that merge is still unforced. stage the fix before the briefing is handed over. the evidence packet is not the gate. the push is.” [source](https://twitter.com/1870072035608584192/status/2101324639938965913)
- Complaint, 2026-09-18, r/cscareerquestions (Reddit): “have you tried switching to lower models? the higher ones are overkill if you're using it to augment the engineering that you are doing rather than trying to rely on them to do the design/software architecting for you. sonnet 5 is plenty capable when given clear accurate instructions, and it's faster and much more token efficient. i'll use opus occasionally for some harder tasks, but i find both o…” [source](https://www.reddit.com/r/cscareerquestions/comments/1wjk5vg/anyone_frustrated_with_how_ai_harnesses_are/paklj9u/)

### Checking and finishing: Too few posts (customer love 0.503, n 3)

| Criterion | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|
| [Agent-performed code review finds real issues](https://feedbackbench.com/criteria/verify.agent_code_review.md) | Too few posts | 0.503 | 0.500–0.508 | 2 | 2 | 0 |
| [Claims work is done or fixed when it is not](https://feedbackbench.com/criteria/verify.false_completion.md) | Too few posts | 0.499 | 0.496–0.500 | 1 | 0 | 1 |
| [Builds, tests or runs its own changes](https://feedbackbench.com/criteria/verify.self_testing.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Reviewing and approving the agent's changes](https://feedbackbench.com/criteria/verify.change_review_ui.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |

Most recent posts:

- Praise, 2026-09-23, r/ExperiencedDevs (Reddit): “i would definitely start getting comfortable with it, you don't need to let it be an agent and do everything for you. it can be fun to figure out where that boundary is. i work in a small team that owns and maintains several software systems, from vendor based to integration layers and some full stack software with both internal and customer users, so knowing everything about everything is effecti…” [source](https://www.reddit.com/r/ExperiencedDevs/comments/1wo4d8p/job_requiresuses_very_little_ai_sinking_ship_or/pblbka0/)
- Praise, 2026-09-10, @augmentcode (X): “line-by-line code review will soon disappear. the future is a risk-gated handoff between humans and agents, where people get pulled in only for the reviews that need judgment. @augmentcode published a schematic of how that works and it's a stellar blueprint for this new architecture. 𝐑𝐢𝐬𝐤 𝐫𝐨𝐮𝐭𝐢𝐧𝐠 𝐛𝐞𝐟𝐨𝐫𝐞 𝐭𝐡𝐞 𝐪𝐮𝐞𝐮𝐞: every pr gets classified first. docs and config auto-approve with a written justific…” [source](https://twitter.com/771267202762670081/status/2098097592001548319)
- Complaint, 2026-09-19, @augmentcode (X): “the fleet fixing ci and conflicts is the write. a briefing can look complete while a conflict resolution already pushed the wrong change into the branch. humans approve and merge only if that merge is still unforced. stage the fix before the briefing is handed over. the evidence packet is not the gate. the push is.” [source](https://twitter.com/1870072035608584192/status/2101324639938965913)

### Interface and sessions: Too few posts (customer love 0.502, n 2)

| Criterion | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|
| [Mobile, remote-control and voice access](https://feedbackbench.com/criteria/surfaces.remote_mobile.md) | Too few posts | 0.503 | 0.500–0.508 | 1 | 1 | 0 |
| [Cloud and remote sandbox execution](https://feedbackbench.com/criteria/surfaces.cloud_sessions.md) | Too few posts | 0.498 | 0.495–0.500 | 1 | 0 | 1 |
| [How the interface shows work, and what the user can configure](https://feedbackbench.com/criteria/ui.display_settings.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Saving, switching, resuming and rewinding sessions](https://feedbackbench.com/criteria/ui.session_history.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Stopping and steering a running agent](https://feedbackbench.com/criteria/ui.interrupt_steer.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |

Most recent posts:

- Praise, 2026-09-02, @augmentcode (X): “@augmentcode easily cloud computing and developing from my phone” [source](https://twitter.com/1159581796423483392/status/2094950663113572397)
- Complaint, 2026-09-17, G2 (G2): “q: what problems is the product solving and how is that benefiting you? a: i use it to build, refactor and manage a large codebase. work like that gets tedious when changes touch a lot of files, and having augment code help with it makes those tasks more manageable. it's only been a week or two, so i don't have hard numbers, but it has made refactoring work feel less of a chore. q: what do you lik…” [source](https://www.g2.com/products/augment-code/reviews/augment-code-review-13367494)

### Reliability and speed: Too few posts (customer love 0.500, n 0)

| Criterion | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|
| [Outages, server errors and capacity or rate errors](https://feedbackbench.com/criteria/rel.service_errors.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Latency, throughput and fast mode](https://feedbackbench.com/criteria/rel.response_speed.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Client crashes, freezes and failed tool execution](https://feedbackbench.com/criteria/rel.client_failures.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Updates break working setups](https://feedbackbench.com/criteria/rel.update_breakage.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |

### Account and support: Too few posts (customer love 0.495, n 3)

| Criterion | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|
| [Support, refunds and issue handling](https://feedbackbench.com/criteria/account.support.md) | Too few posts | 0.495 | 0.490–0.500 | 3 | 0 | 3 |
| [Wrong charges, failed payments and plan provisioning](https://feedbackbench.com/criteria/account.billing_errors.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Account bans and access restrictions](https://feedbackbench.com/criteria/account.bans_restrictions.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Data retention, training use and deployment isolation](https://feedbackbench.com/criteria/account.data_privacy.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |

Most recent posts:

- Complaint, 2026-09-26, G2 (G2): “q: what problems is the product solving and how is that benefiting you? a: it completely eliminates the tedious manual effort of navigating unfamiliar codebases, writing repetitive boilerplate, and tracing cross-field dependencies saving me 1-2 hours of grunt work every day. q: what do you like best about the product? a: the context engine is really amazing. in contrast to a simple ai autocomplete…” [source](https://www.g2.com/products/augment-code/reviews/augment-code-review-13618001)
- Complaint, 2026-09-23, G2 (G2): “q: what problems is the product solving and how is that benefiting you? a: it takes away all the pain of digging through massive codebases manually tracking down cross-file dependencies. i save about 1-2 hours of grunt work every single day just in refactoring and boilerplate. q: what do you like best about the product? a: the context is really amazing. in contrast to a simple ai autocomplete func…” [source](https://www.g2.com/products/augment-code/reviews/augment-code-review-13582326)
- Complaint, 2026-09-10, G2 (G2): “q: what problems is the product solving and how is that benefiting you? a: code analyzing, editing, autocompleting, very good for professional developers. good as compared to other software in the market. q: what do you like best about the product? a: it has excellent fetching quality of code database. i could easily integrate it with github. the speed is good as compared to other products in the…” [source](https://www.g2.com/products/augment-code/reviews/augment-code-review-13433460)
