# Usage meter visibility and accuracy (`limits.usage_meter`)

Feedback Bench, coding agents, built 2026-10-01, window 2026-08-31 to 2026-09-27. Web page: https://feedbackbench.com/#/criterion/limits.usage_meter

Area: [Paying and limits](https://feedbackbench.com/criteria/paying.md)

**Definition.** Whether the product shows how much quota and tokens were used and how much remains, per task and per model, and whether that figure is accurate. Covers meters that climb while idle and hidden token counts.

**Boundary.** Not this: see [Pay-as-you-go overage, fallback billing and spend caps](https://feedbackbench.com/criteria/billing.overage_charges.md) for money actually charged. Not this: see [Single prompt, model or effort level consumes disproportionate quota](https://feedbackbench.com/criteria/limits.burn_rate.md) when the meter is believed and the complaint is the amount consumed.

Rated author-weeks, all agents: 1821. Complaint share: 92%.

## The brief

Written by Claude Opus 5.5 from 75 labelled posts and the numbers on this page. Interpretation, not measurement: every quote is verbatim and links to its post.

**A usage meter that drifts on its own is worse than none.**

TL;DR:

- Codex draws the bulk of complaints: meters that reset wrongly, reverse, or drain after work stops.
- Claude Code earns credit for per-session and sub-agent token views, though weekly meter spikes persist.
- The top request everywhere is a meter that matches real consumption, shown where users work.

In plain terms: Users watch a percentage they cannot verify. It jumps overnight, climbs after they stop, or splits silently across models. When a tool shows real tokens per session, users say they change habits and spend less.

### How it breaks

- **Meters that jump, reverse, or reset** ([Usage meter visibility and accuracy](https://feedbackbench.com/criteria/limits.usage_meter.md)). The most damaging failure is a meter that changes without cause, which pushes users into decisions based on numbers that later prove wrong.
  One Codex user saw the meter near empty, spent a banked reset, then watched the old balance return an hour later with the reset gone. Another saw the meter snap back to its prior level. A third logged token totals around a reset and found the same percentage covering roughly half the tokens. Claude Code users report weekly meters dropping sharply overnight. A Cursor user describes an allowance used up in one day, then reset for no stated reason.
  Evidence:
  - Complaint, OpenAI Codex, r/codex, 2026-09-09: “got double fucked, went from 4% -> 0% and just thought i used it all, so i then spent a banked reset to get it back to 100%. then i get an alert an hour later my usage went back down to 4% and the banked reset was still gone...” [source](https://www.reddit.com/r/codex/comments/1wbsgrf/unexpected_usage_limit_resets/p8sqoqe/)
  - Praise, OpenAI Codex, r/codex, 2026-09-09: “update - mine reverted back to 74% which is where is was before.” [source](https://www.reddit.com/r/codex/comments/1wbr0xt/codex_issue_vaporized_weekly_limit_out_of_nowhere/p8se0su/)
  - Complaint, OpenAI Codex, r/codex, 2026-09-06: “|model|before: 85% → 99%|after: 0% → 14%| |:-|:-|:-| |gpt-6 astra|592 calls / **95.47m** total tokens|381 calls / **52.75m** total tokens| |gpt-5.6 terra|51 / **6.18m**|38 / **4.59m**| |gpt-5.6 sol|9 / **744,761**|42 / **5.14m**| |`codex-auto-review`|86 / **12.48m**|28 / **1.40m**| |**all models**|**738 / 114.88m**|**489 / 63.88m**| i had that max reasoning doublechecked. you see that before the reset 592 astra calls at 95m tokens were consumed for 14$ after the reset 381 calls with 52m tokens are consumed for the same 14% the billing is supposed to be token consumption based, that's what they claim everywhere. input, output and cached input were double or more than double in the pre-reset time.” [source](https://www.reddit.com/r/codex/comments/1w8zbz9/i_analyzed_the_allowance_a_banked_reset_gives_vs/p86ndwa/)
  - Complaint, Claude Code, r/ClaudeCode, 2026-09-16: “i confirm, my account went from 70% (yesterday) to 50% weekly usage on fable with a x5 subscription.” [source](https://www.reddit.com/r/ClaudeCode/comments/1whrr18/limits_are_fixed/pa5pywo/)

- **Quota drains while nothing runs** ([Usage meter visibility and accuracy](https://feedbackbench.com/criteria/limits.usage_meter.md)). Users report meters climbing after they stop or step away, so the meter tracks activity they never started.
  A Codex user says a stopped chat kept working on the backend and consumed the full allowance. A Factory user stepped away and came back to find both the short-window and weekly meters higher. On Google Antigravity, a post says the agent keeps looping once tokens run out and stops sending notifications. In each case the meter is the only signal, and users read it as consumption they did not trigger.
  Evidence:
  - Complaint, Factory, @droid, 2026-09-14: “ayo @droid homies, ur weekly limits raise % while im afk, can we fix this please? i literally went afk 40 min found my 5h limit going up 3 % and weekly 2%? doesn't make any sense.” [source](https://twitter.com/1834314883510226944/status/2099646975146660255)
  - Complaint, OpenAI Codex, r/codex, 2026-09-14: “there’s a bug today i stopped my work chat multiple times and chatgpt acknowledged it had stopped work i came back 3 hours later and 100% of my 20x was gone it stopped work at 98% but keeps working on the backend. i escalated it for human review” [source](https://www.reddit.com/r/codex/comments/1wgg7zu/2030b_tokens_a_week_to_1b_are_they_for_real/p9u87xg/)
  - Complaint, Google Antigravity, @antigravity, 2026-09-23: “i've detected a bug. when the tokens run out, the software stops providing notifications. it just keeps running in a loop. @antigravity” [source](https://twitter.com/1414218035028758533/status/2102657396501795024)

- **Hidden links across models and accounts** ([Usage meter visibility and accuracy](https://feedbackbench.com/criteria/limits.usage_meter.md)). Quota pools connect in ways no meter shows, so one model or account silently drains another.
  A Factory user found that hitting the cap on one model tier also blocked standard models that showed no usage. On Google Antigravity, choosing Opus reportedly spawned Gemini sub-agents that consumed a large share of a separate quota. Another user says draining one account drains a second at the same time. Cline users ask what a subscription actually includes per model. Users want a pre-run estimate or a per-model split.
  Evidence:
  - Complaint, Factory, @droid, 2026-09-07: “@droid possible limits bug: using droid core before standard models seems to link their usage limits once droid core hits its cap, standard models are also blocked despite having no usage thanks” [source](https://twitter.com/2034772443500314624/status/2096921125833867534)
  - Complaint, Google Antigravity, r/google_antigravity, 2026-09-25: “yeah, this is the part that annoys me too. if i explicitly pick opus, i’d expect the usage to come out of my opus allowance, not have it silently spin up a bunch of gemini agents in the background and nuke 40% of my 5-hour gemini quota. i’m guessing that’s exactly what happened here. the opus run itself probably exhausted its allowance, while whatever sub-agents/tools it spawned were billed against gemini separately. would be way better if antigravity showed something like “this run may use x opus + y gemini agents” before you hit go, or at least gave us a toggle to disable cross-model agents. burning that much quota from one prompt with basically zero visibility is pretty rough.” [source](https://www.reddit.com/r/google_antigravity/comments/1wpvi9v/opus_using_a_large_amount_of_my_gemini_allowance/pbypxr7/)
  - Complaint, Google Antigravity, r/google_antigravity, 2026-09-22: “it \*looks\* that way but what happens is that using the bucket in one account will drain the other account's simultaneously.” [source](https://www.reddit.com/r/google_antigravity/comments/1wni05g/we_need_a_usage_reset_now/pbfmp0w/)
  - Complaint, Cline, @cline, 2026-09-21: “@cline any chance we could get a transparent model-by-model usage breakdown for clinepass? how much mimo-v2.6-pro, glm, deepseek, etc. usage do we actually get with the subscription? other similar subscriptions publish this pretty clearly - would be really useful for comparing plans.” [source](https://twitter.com/1770069362608648192/status/2102165413500973469)

- **Percentages with no token math** ([Usage meter visibility and accuracy](https://feedbackbench.com/criteria/limits.usage_meter.md)). Bare percentages without token counts or a published formula leave users unable to tune prompts or predict when they will run out.
  A Google Antigravity user wants token-per-prompt figures to tighten prompts and says the product withholds them. A Claude Code user says there is no transparency on what a token is. Augment Code users report running out of tokens without warning. Cursor users note other-model usage values quietly removed from the dashboard. Showing actual token counts, not just percentages, recurs as a request, led by Codex users.
  Evidence:
  - Complaint, Google Antigravity, r/google_antigravity, 2026-09-08: “no matter what the model is id like to know the tokens to prompt usage so i can tweak my prompt and make them more efficient, and i feel that there is a lot of smoke in the mirrors and antigravity does not give you that information. has anyone figured out a good way to see this?” [source](https://www.reddit.com/r/google_antigravity/comments/1wau6di/best_way_to_see_prompt_vs_tokens_used/)
  - Complaint, Claude Code, @ClaudeDevs, 2026-09-25: “@claudedevs why did it not have this to begin with. it would waste all your credits if the large task burned through the usage… forever. no transparency on what a token is. that’s the problem. #stryker336 it’s whatever they feel like” [source](https://twitter.com/1741904545716785153/status/2103574390302810154)
  - Complaint, Augment Code, @augmentcode, 2026-09-18: “@augmentcode regardless of whether it merges code or offers optimizations, the subscription for augment is very expensive. i ran out of tokens unexpectedly without even realizing it, and the pricing structure is unclear—i wouldn't recommend using it.” [source](https://twitter.com/955391612632252418/status/2100826713957810177)
  - Complaint, Cursor, @cursor_ai, 2026-09-12: “between aug 24th-29th @cursor_ai's "other model usage" values were quietly removed. <strict_link> <strict_link>” [source](https://twitter.com/2073328259681632256/status/2098900087338594647)

- **The meter lives somewhere else** ([Usage meter visibility and accuracy](https://feedbackbench.com/criteria/limits.usage_meter.md)). Many meters sit behind a browser tab or settings page, so users learn they are out only when work stops.
  Devin users have to leave the editor and cross two domains to check usage, and the app and CLI disagree on remaining quota. Zed users ask for balance inside the app instead of a browser. A GitHub Copilot user gets a daily-limit warning but cannot find daily or weekly numbers anywhere. Requests for an always-visible meter in the statusline or chat come from most agents.
  Evidence:
  - Complaint, Devin, @cognition, 2026-09-14: “@cognition @devindesktop please update the usage limits so that we can easily just see it in the app, not in the web. also this button is too close together 😭 great overall bytheway <strict_link>” [source](https://twitter.com/1828265572884467712/status/2099302134277865776)
  - Complaint, Devin, @cognition, 2026-09-07: “@cognition why is the devin app and api so terrible still? cant check usage without leaving vscode and bouncing off 2 different domains ..the app and cli dont seem to communicate and one says im out of quota but the other doesnt. total hot mess” [source](https://twitter.com/488298142/status/2096819957350948888)
  - Complaint, Zed, @zeddotdev, 2026-09-04: “hey @zeddotdev guys, this is probably really far down on the priority list for you guys rn, but please put the usage/balance of the zedvip tokens in the delta app somehow. having to open a browser window to see my balance is a few too many steps. otherwise i’m a huge fan of the app so far 👍👍” [source](https://twitter.com/30111156/status/2095791764284281340)
  - Complaint, GitHub Copilot, r/GithubCopilot, 2026-09-01: “i received a pop-up message saying that i have reached 85% of my daily limit (or something to that effect, it is now gone). last time that happened, doing very little else caused my weekly limit to be hit, and that was it for the rest of the week. if i click on the widget on the lower right of the screen it shows i have used 12% of my monthly usage and that it resets on the 30th. where can i see my daily and weekly numbers? i can't find it in either the widget in vsc, nor in the page on gh.” [source](https://www.reddit.com/r/GithubCopilot/comments/1w4jwhz/vscode_on_the_mac_where_is_the_weekly_and_daily/)

### Who stands out

- **OpenAI Codex (weaker)**. Codex carries the heaviest complaint load here, driven by meters users stop trusting after unexplained resets and backend drain.
  Posts describe a stopped chat consuming the whole allowance, a CLI reporting no quota during a reset, and token logs where the same percentage covers very different volumes. The usage analysis view and in-CLI limit guidance draw real praise; one user used it to see one model consuming far more than another. Codex users ask most for an accurate meter and an explanation of how usage is calculated.
  Evidence:
  - Complaint, OpenAI Codex, r/codex, 2026-09-14: “there’s a bug today i stopped my work chat multiple times and chatgpt acknowledged it had stopped work i came back 3 hours later and 100% of my 20x was gone it stopped work at 98% but keeps working on the backend. i escalated it for human review” [source](https://www.reddit.com/r/codex/comments/1wgg7zu/2030b_tokens_a_week_to_1b_are_they_for_real/p9u87xg/)
  - Complaint, OpenAI Codex, X search: OpenAI Codex, Codex CLI, Codex app, 2026-08-31: “the codex cli told me that the model has no quota when resetting, which startled me.” [source](https://twitter.com/2283852535/status/2094259229435843035)
  - Complaint, OpenAI Codex, r/codex, 2026-09-06: “|model|before: 85% → 99%|after: 0% → 14%| |:-|:-|:-| |gpt-6 astra|592 calls / **95.47m** total tokens|381 calls / **52.75m** total tokens| |gpt-5.6 terra|51 / **6.18m**|38 / **4.59m**| |gpt-5.6 sol|9 / **744,761**|42 / **5.14m**| |`codex-auto-review`|86 / **12.48m**|28 / **1.40m**| |**all models**|**738 / 114.88m**|**489 / 63.88m**| i had that max reasoning doublechecked. you see that before the reset 592 astra calls at 95m tokens were consumed for 14$ after the reset 381 calls with 52m tokens are consumed for the same 14% the billing is supposed to be token consumption based, that's what they claim everywhere. input, output and cached input were double or more than double in the pre-reset time.” [source](https://www.reddit.com/r/codex/comments/1w8zbz9/i_analyzed_the_allowance_a_banked_reset_gives_vs/p86ndwa/)
  - Praise, OpenAI Codex, r/codex, 2026-09-04: “very interesting perspective. with the release of the "usage analysis", i could see the difference between my sol usage and my luna usage. sol is devouring almost 20x my luna consumption, even luna having many more instances than sol. i dont know if this, for itself, justifies using a mid tier model or reasoning, but could at least show me, that a i need to improve maybe the orchestration context window.” [source](https://www.reddit.com/r/codex/comments/1w73b4v/subagents_on_luna_max_or_astra_medium/p7sfjci/)

- **Claude Code (stronger)**. Claude Code wins credit for making token spend visible per session and per sub-agent, which users say changes how they work.
  Users praise a view of sub-agent token consumption, warnings when the prompt cache has expired, and a cost calculator reachable from /usage. One team says seeing where session tokens went led them to shorten session starts rather than cut sessions. Complaints remain about weekly meters spiking, and one reply suggests the cause lies in metering rather than real burn. Per-session usage logs top the requests for this agent.
  Evidence:
  - Praise, Claude Code, @ClaudeDevs, 2026-08-31: “@claudedevs visibility changes behaviour. once we could see where a session's tokens went, the fix was not fewer sessions, it was a shorter start: a one-page read of rules and current state instead of scrolled history. same work, smaller spend, nothing important left implicit.” [source](https://twitter.com/2011830476051529734/status/2094223845662118180)
  - Praise, Claude Code, r/ClaudeAI, 2026-09-27: “no one talks about these claude code features, but they save me so many tokens anthropic released 2 features in the past 2 weeks in their vs code plugin: 1) ui to see sub-agents spawned, their token consumption, and a green indicator if they're still cached. cherry on the cake, it even supports codex agents. 2) notification at the end of a development conversation warning when a cache has expired >*idle 1h 10m. the prompt cache has likely expired, so your next message will re-cache about 342k tokens.* this sounds like basic stuff, but it's so much easier to optimise cache and tokens with some visibility. if sonnet and haiku 5.5 follow the leap of opus 5.5, we may never dread running out of weekly tokens with good agent delegation.” [source](https://www.reddit.com/r/ClaudeAI/comments/1wrxzgc/no_one_talks_about_these_claude_code_features_but/)
  - Praise, Claude Code, @ClaudeDevs, 2026-09-26: “opus 5.5 is 20% cheaper per input/output token than opus 5 and 60% cheaper on cache reads, and @claudedevs built a calculator so you can run your own task costs from /usage. long, cache-heavy #claudecode sessions should feel this most 👀 <strict_link>” [source](https://twitter.com/1190407448/status/2103817516510531985)
  - Complaint, Claude Code, r/ClaudeCode, 2026-09-12: “seen a handful of similar reports today, and the pattern in the replies here (fable/astra tasks spiking both the 5-hour and weekly meters at once) points at the metering side rather than actual token burn. since your support case is already queued, the account-level usage log should let them diff metered units against real requests — would be great if you post their explanation.” [source](https://www.reddit.com/r/ClaudeCode/comments/1weo0w1/max_20x_weekly_usage_suddenly_jumped_from_7_to_50/p9g6dcb/)

- **Cursor (weaker)**. Cursor users report a dashboard that shows phantom usage, misses recent chats, and drops data without notice.
  Posts describe morning usage appearing for hours the user did not work, recent chat usage missing, and other-model values quietly removed. Several users call usage opaque even while they stay on the product. Plan usage meters drew thanks when added, and the usage tab helps some users spot heavy model consumption. Requests center on an accurate meter and cost per completed task.
  Evidence:
  - Complaint, Cursor, r/cursor, 2026-09-24: “please help. i’m using it right now, and i opened the dashboard just to check my usage. it’s showing usage from the morning, but i didn’t use it this morning. is this a bug? also, my recent chat usage isn’t showing up.” [source](https://www.reddit.com/r/cursor/comments/1wowhh7/i_didnt_use_it_in_morningwhy_its_showing_usage/)
  - Complaint, Cursor, @cursor_ai, 2026-09-12: “between aug 24th-29th @cursor_ai's "other model usage" values were quietly removed. <strict_link> <strict_link>” [source](https://twitter.com/2073328259681632256/status/2098900087338594647)
  - Complaint, Cursor, r/cursor, 2026-09-24: “for pm work the ui is half the product so i get sticking with cursor. on external models i mostly use them when i need a cheaper long-context pass, and keep the main agent on whatever is stable that week. usage still feels opaque though, not just you.” [source](https://www.reddit.com/r/cursor/comments/1wouq64/usage_for_external_models/pbq7hxw/)
  - Praise, Cursor, @cursor_ai, 2026-09-25: “thanks @cursor_ai &amp; @spacexai for adding plan usage meters 👏 <strict_link>” [source](https://twitter.com/2073328259681632256/status/2103390000805294309)

- **OpenCode (mixed)**. OpenCode's cost tracker and usage API win praise, but users still cannot easily see remaining tokens or reset times.
  Users like the TUI cost tracker and an API that lets other apps show usage limits, and other tools build live quota views on top of it. Yet posts say remaining tokens and reset times are hard to find, the actual per-model quota is tucked behind a details link, and earlier history goes missing when the time window changes. Per-model usage breakdown is its most requested fix.
  Evidence:
  - Complaint, OpenCode, r/opencode, 2026-09-01: “although i love using open code, i can't figure out how many tokens i have left or when they will reset, and the chat often goes into a loop.” [source](https://www.reddit.com/r/opencode/comments/1w48fqo/meta_introduced_coding_plans_for_muse_spark_12/p767ulx/)
  - Praise, OpenCode, @opencode, 2026-09-22: “i just love @opencode tui, theme is beautiful and cost tracker is super useful! <strict_link>” [source](https://twitter.com/1367864633290199041/status/2102336198098485705)
  - Complaint, OpenCode, r/opencode, 2026-09-03: “yeah it's annoying and imo too misleading. under your usage there's a "show details". it shows the actual quota of the model you've used.” [source](https://www.reddit.com/r/opencode/comments/1w5s8yn/doubt_about_the_limit_go_plan/p7jl73s/)
  - Complaint, OpenCode, r/opencode, 2026-09-22: “ya the time window changes but the data for before today is misisng entirely” [source](https://www.reddit.com/r/opencode/comments/1wn8fz8/new_usage_dashboard_is_unusable/pbdgm63/)

### Fine print

- Users report meter readings they rarely verify; few can compare metered units against actual requests.
- Devin, Pi, GitHub Copilot, Amp and smaller agents have too few posts to rank here.
- Posts e58 and e59 are the same post, tagged once for Pi and once for GitHub Copilot.

## Top requests

What users ask to add or change, most asked first. 580 author-weeks ask for something. Requests do not change the Feedback Score. Rule: A separate pass by Claude Sonnet 5 reads every counted post and extracts what the author asks the agent or its vendor to add or change, with the criteria it maps to and a short normalised wording; it does not touch the labels or the Feedback Score. Claude Opus 5.5 groups the wordings within each criterion (the first criterion the request maps to) into themes; code counts them. A theme counts distinct author-weeks that ask for it, per agent; across agents, one author-week per agent. Themes asked in fewer than 2 author-weeks, and requests that share no theme, are not shown. Examples: up to 3 posts per theme from different authors, without slurs, preferring posts of 60 to 450 characters, most recent first.

| Rank | Request | Author-weeks | Posts | Agents (author-weeks) |
|---|---|---|---|---|
| 1 | Accurate usage meter matching real consumption | 71 | 73 | OpenAI Codex 27, Claude Code 22, Cursor 8, OpenCode 7, Google Antigravity 3, Amp 1, GitHub Copilot 1, Devin 1, Factory 1 |
| 2 | Transparent explanation of usage calculation | 43 | 43 | OpenAI Codex 23, Claude Code 13, Google Antigravity 2, Cursor 2, OpenCode 2, Cline 1 |
| 3 | Clear display of remaining quota and reset time | 41 | 46 | OpenAI Codex 16, Claude Code 6, Cursor 6, OpenCode 5, Google Antigravity 3, Cline 1, Devin 1, Kiro 1, Pi 1, Zed 1 |
| 4 | Always-visible usage meter in statusline or chat | 37 | 38 | OpenAI Codex 14, Claude Code 9, Cursor 5, Google Antigravity 4, OpenCode 3, GitHub Copilot 1, Zed 1 |
| 5 | Show actual token counts, not just percentages | 34 | 34 | OpenAI Codex 17, Claude Code 8, Google Antigravity 5, Devin 3, OpenCode 1 |
| 6 | Per-model usage breakdown | 21 | 22 | OpenCode 9, Cursor 4, OpenAI Codex 3, Cline 2, Google Antigravity 1, Claude Code 1, Factory 1 |
| 7 | Per-task usage breakdown | 19 | 23 | OpenAI Codex 8, Claude Code 6, Cursor 2, Amp 1, Google Antigravity 1, OpenCode 1 |
| 8 | Cost per completed task metrics | 19 | 19 | Cursor 6, Claude Code 5, OpenAI Codex 3, Devin 3, OpenCode 2 |
| 9 | Visible context window usage indicator | 18 | 19 | Google Antigravity 8, Claude Code 6, OpenAI Codex 2, Cursor 1, Factory 1 |
| 10 | Usage shown in dollar cost | 18 | 18 | OpenAI Codex 7, Claude Code 4, Cursor 3, OpenCode 2, Devin 1, Pi 1 |
| 11 | Visibility into what consumes usage quota | 17 | 17 | Claude Code 4, OpenAI Codex 4, Cursor 4, Google Antigravity 2, GitHub Copilot 1, OpenCode 1, Pi 1 |
| 12 | Per-session usage logs and breakdown | 16 | 17 | Claude Code 11, OpenCode 2, Google Antigravity 1, OpenAI Codex 1, Cursor 1 |

### 1. Accurate usage meter matching real consumption

- OpenAI Codex, 2026-09-26, r/codex (Reddit): “hey did anybody else notice their usage drop to zero after yesterday's outage? i had 29% and now i'm sitting at zero. i did one job after the outage yesterday. typically cost me 2 to 5%. so i fully expect to be around 20% and i'm pretty sure tibo said something about a reset so why am i at zero.” [source](https://www.reddit.com/r/codex/comments/1wqqg0k/on_the_reset_situation_written_by_astra/pc65ta0/)
- Google Antigravity, 2026-09-26, @antigravity (X): “@antigravity i don't know what's wrong with the app; the quota always shows as 100% and doesn't reset like it's supposed to, even though updates have been released almost daily. i've already uninstalled and reinstalled it—deleting the folder in the process—but nothing works.” [source](https://twitter.com/1544718019263447041/status/2103665685633118252)
- GitHub Copilot, 2026-09-25, r/GithubCopilot (Reddit): “i believe copilot app is bugged. ui shows 20 credits used, yet i lose hundreads.” [source](https://www.reddit.com/r/GithubCopilot/comments/1wq39rv/did_they_10x_the_cost_of_luna_56_since_the/pc1hxi9/)

### 2. Transparent explanation of usage calculation

- OpenAI Codex, 2026-09-27, r/codex (Reddit): “exactly these black box usage limits need to be regulated, it's insanely scummy behavior.” [source](https://www.reddit.com/r/codex/comments/1wrr1z8/plans_are_no_longer_x5_or_x20_overall_limit/pcg5d27/)
- OpenAI Codex, 2026-09-27, r/codex (Reddit): “good to know! last time i've checked i was under the impression that there is no official info on pro-model limits anywhere. i then asked oa's support ai and it was like "there is a limit that is not shared with codex" and no quantification at all. this limit however is still not tracked anywhere or am i missing something (again)?” [source](https://www.reddit.com/r/codex/comments/1wrk74r/gpt_6_pro_on_chat_is_now_spawning_subagents/pcdu0bm/)
- Claude Code, 2026-09-27, r/ClaudeCode (Reddit): “i understand, but i wish for a quantifiable multiplier, since you measure and monitor usage and tokens” [source](https://www.reddit.com/r/ClaudeCode/comments/1wra7t6/good_while_it_lasts/pcdl511/)

### 3. Clear display of remaining quota and reset time

- OpenAI Codex, 2026-09-27, r/GithubCopilot (Reddit): “i switched to codex and it seems to get me many times farther. they kind of obscure how much usage you’re really getting, but $100/mo with codex gets me waaaaay more than $200/mo with copilot, not to mention that gpt 6 is on the pareto frontier anyway” [source](https://www.reddit.com/r/GithubCopilot/comments/1wr5io2/best_cheaper_alternative/pcf7r1z/)
- OpenAI Codex, 2026-09-27, r/codex (Reddit): “pro usage still has its own limit. you just find out the hard way when you hit it as it is not shown anywhere.” [source](https://www.reddit.com/r/codex/comments/1wrk74r/gpt_6_pro_on_chat_is_now_spawning_subagents/pcdo80p/)
- Pi, 2026-09-22, @pidotdev (X): “@pidotdev remaining quota display widget, plannotator review in vscode + herdr orchestration. <strict_link>” [source](https://twitter.com/947783691320905729/status/2102411360743080258)

### 4. Always-visible usage meter in statusline or chat

- OpenAI Codex, 2026-09-27, X search: OpenAI Codex, Codex CLI, Codex app (X): “the new codex app ui misses a critical feature. a lot of people are anxious about how much of their weekly limit is remaining. this information should be always visible like the battery charge information is always visible on a laptop. @thsottiaux @reach_vb” [source](https://twitter.com/15043081/status/2104080512650469439)
- OpenAI Codex, 2026-09-24, r/codex (Reddit): “makes me unhappy that i need to have another app/window open just to monitor usage :(” [source](https://www.reddit.com/r/codex/comments/1wo00mm/usage_indicator_disappeared_from_vs_code_after/pbr979k/)
- Cursor, 2026-09-23, @cursor_ai (X): “@theo @viticci @t3dotcodes @cursor_ai @jullerino @gabrielelpidio an easier way to see limits would be great! right now its settings &gt; usage. i think a shortcut from chat would be great” [source](https://twitter.com/1278354375304413184/status/2102834606915354934)

### 5. Show actual token counts, not just percentages

- Google Antigravity, 2026-09-23, r/google_antigravity (Reddit): “tbf a 20usd sub isn't that big anymore these days. back when original ag launched last year you had near infinite limits. today, a 20usd sub is maxxed out easily. but yeah if google's ag2 team could implement this counting tokens featue i would like it.” [source](https://www.reddit.com/r/google_antigravity/comments/1woefec/just_found_the_number_of_tokens_used_in/pbmj4oo/)
- Claude Code, 2026-09-22, r/google_antigravity (Reddit): “how to see the exact number of token usage in antigravity like as same as we can see in claude code” [source](https://www.reddit.com/r/google_antigravity/comments/1wn39uk/how_to_see_the_exact_usage_of_number_of_tokens/)
- OpenAI Codex, 2026-09-22, r/codex (Reddit): “considering they control the spout and don't give concrete numbers, you may see no difference period” [source](https://www.reddit.com/r/codex/comments/1wngvak/gpt_6_sol_and_luna_prices_wtf/pbgrevn/)

### 6. Per-model usage breakdown

- OpenCode, 2026-09-27, r/opencode (Reddit): “when a model was selected, it showed the various costs in tokens. do we have something similar on opencode?” [source](https://www.reddit.com/r/opencode/comments/1wrkid7/ua_cosa_che_ghithub_copilot_ha_e_che_su_opencode/)
- OpenCode, 2026-09-22, r/opencode (Reddit): “the new usage dashboard is just missing so much in functionality, in looks, and in data. i'm only able to see data from today onwards, past data seems to be gone i can't see usage by api keys. i had different api keys set up for different devices and now i can't see the usage across them. i can't see usage in the chart by models as well i can't see model wise usage remaining is it really this lacking or am i missing something here?” [source](https://www.reddit.com/r/opencode/comments/1wn8fz8/new_usage_dashboard_is_unusable/)
- OpenCode, 2026-09-22, r/opencode (Reddit): “<strict_link> before, per usage, i could see a breakdown alongside the total model quota. now, they have removed it from the console. this is deliberate.” [source](https://www.reddit.com/r/opencode/comments/1wmz6yn/can_no_longer_see_the_usage_per_model_in_the/)

### 7. Per-task usage breakdown

- Claude Code, 2026-09-26, @ClaudeDevs (X): “@claudedevs putting cost per task in /usage instead of a pricing page is the right call, i think. one layer i'd add for plan users: how much of the 5h window a task eats. a lot of the replies here think in windows, not dollars.” [source](https://twitter.com/1912898758083514368/status/2103864336385007926)
- OpenCode, 2026-09-24, r/codex (Reddit): “some harnesses show the equivalent cost even for subscription plans (such as grok 4.7 on opencode), which makes sense since providers often scale usage quotas based on actual compute cost. anyway, the primary metric for the user should be the percentage of their quota consumed by the submitted task, not "amount of tokens".” [source](https://www.reddit.com/r/codex/comments/1wp3bzl/you_need_to_try_opus_55/pbs6smc/)
- OpenAI Codex, 2026-09-17, r/codex (Reddit): “their codex session probably ran longer than anticipated. they could give us an accurate timeline if they implement my idea to have a task-by-task quota estimation/tracking tool built in.” [source](https://www.reddit.com/r/codex/comments/1wirnly/so_no_resets_this_week_it_seems/padxhw9/)

### 8. Cost per completed task metrics

- OpenAI Codex, 2026-09-27, r/codex (Reddit): “i wonder if "api value" is still the best way to compare these plans as models get more efficient. if a newer model uses fewer tokens or needs fewer retries to finish the same coding task, a lower dollar value could still produce more useful work. i'd love to see something like "tasks completed before hitting the limit" tracked alongside this. that might show whether users are actually getting less value.” [source](https://www.reddit.com/r/codex/comments/1wqw2qq/warning_codex_allowances_have_dropped_about_20/pcb0vn6/)
- OpenAI Codex, 2026-09-27, r/codex (Reddit): “api cost benchmarks are meaningless, nobody works api based. we need quoata usage analysis.” [source](https://www.reddit.com/r/codex/comments/1wnx8i9/i_gave_astra_sol_and_opus_55_the_same_rust_task/pca4vpk/)
- OpenCode, 2026-09-27, @opencode (X): “@opencode $60 permanent cheepseek sounds great, but i want to see the unit price for a successful delivery once - sometimes the cheaper model takes more rounds and can end up being more expensive than the expensive model in one go.” [source](https://twitter.com/2259799350/status/2104306716720644156)

### 9. Visible context window usage indicator

- Google Antigravity, 2026-09-27, @antigravity (X): “@antigravity can you show the context usage? it's hard to know how much content is being used.” [source](https://twitter.com/1925388589241966594/status/2104010402829041896)
- Claude Code, 2026-09-26, r/ClaudeCode (Reddit): “never used claude code: the context widget to follow your context window usage is specific to claude code ? because in chat mode i’d love to have the same metric” [source](https://www.reddit.com/r/ClaudeCode/comments/1wkjnrz/instant_claude_code_compaction_is_my_favorite_use/pc4oz6o/)
- Claude Code, 2026-09-24, r/ClaudeCode (Reddit): “claude code shows the output tokens there. showing the context window is usually common on open source harnesses like pi and amp. i'm not sure why cc is hiding it but it always bugged me not being able to see it directly.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wp9nga/the_agentic_loop_is_outdated/pbuqato/)

### 10. Usage shown in dollar cost

- OpenCode, 2026-09-27, r/opencode (Reddit): “it’s the most frustrating thing i can’t see model costs and how much cash left in my account in the ui of open code.” [source](https://www.reddit.com/r/opencode/comments/1wrs0xl/openchamber_is_so_much_better_than_opencode/pcfrjqr/)
- Claude Code, 2026-09-26, @ClaudeDevs (X): “@claudedevs it would be good if claude could find out the usage rate, or where we stand in terms of expenditure.” [source](https://twitter.com/1293505674178170880/status/2103736755815850151)
- OpenAI Codex, 2026-09-25, r/codex (Reddit): “not op. but they could say, for example this many tokens per month. or this many dollars of api equivalent spend. whatever just something we can track, work off and verify.” [source](https://www.reddit.com/r/codex/comments/1wpniqq/did_they_reduce_gpt_6_sol_usage/pbx79jt/)

### 11. Visibility into what consumes usage quota

- Claude Code, 2026-09-24, @ClaudeDevs (X): “@claudedevs will billing logs show which category triggered the block? that would make unexpected charges much easier to investigate.” [source](https://twitter.com/1491654782091735041/status/2103174419904684170)
- OpenAI Codex, 2026-09-19, r/codex (Reddit): “i'll die on the hill that oai releasing a tool that goes through and points out *what* caused the most token burn for a prompt/conversation would be more useful than half the resets we've gotten. "hey dummy, you don't need to load 50 pages of documentation for every task", "you asked me to make gta6 from scratch with no further details, i had to reason out the plan and you're damn right you're paying for that"” [source](https://www.reddit.com/r/codex/comments/1wkm8yl/reset_culture_is_terrible_for_subscription_users/pasmvj9/)
- Pi, 2026-09-19, @pidotdev (X): “@pidotdev i want which of those four tools still burned the most tokens.” [source](https://twitter.com/1307899154560151552/status/2101245668635369829)

### 12. Per-session usage logs and breakdown

- Claude Code, 2026-09-26, @claude_code (X): “@bcherny smart usage insights page to see usage by conversation, @claude_code session, etc @claudedevs” [source](https://twitter.com/312170411/status/2103723623777538403)
- OpenCode, 2026-09-25, @opencode (X): “@opencode shouldn't logs just be under usage , saves me a click” [source](https://twitter.com/1587769745054396417/status/2103408317846659282)
- Google Antigravity, 2026-09-24, r/google_antigravity (Reddit): “i suspect this is pure hallucination, at least the agy cli does not log token usage in the session transcript json files. with pi agent harness, you can calculate this from the session json files, it keeps track of the usage with fields like this: "usage":{"input":4208,"output":38,"cacheread":0,"cachewrite":0,"reasoning":0,"totaltokens":4246...” [source](https://www.reddit.com/r/google_antigravity/comments/1woefec/just_found_the_number_of_tokens_used_in/pbsah28/)

## Every agent

| Agent | Overall rank | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|---|
| [Claude Code](https://feedbackbench.com/agents/claude-code.md) | 1 | Better than peers | 0.593 | 0.552–0.627 | 434 | 60 | 374 |
| [Google Antigravity](https://feedbackbench.com/agents/antigravity.md) | =5 | Typical | 0.527 | 0.477–0.585 | 57 | 8 | 49 |
| [OpenCode](https://feedbackbench.com/agents/opencode.md) | 3 | Typical | 0.471 | 0.420–0.524 | 124 | 9 | 115 |
| [Cursor](https://feedbackbench.com/agents/cursor.md) | 4 | Worse than peers | 0.449 | 0.400–0.494 | 132 | 8 | 124 |
| [OpenAI Codex](https://feedbackbench.com/agents/codex.md) | 2 | Worse than peers | 0.390 | 0.343–0.433 | 998 | 39 | 959 |
| [Devin](https://feedbackbench.com/agents/devin.md) | =5 | Too few posts | – | – | 27 | 3 | 24 |
| [Pi](https://feedbackbench.com/agents/pi.md) | 7 | Too few posts | – | – | 11 | 6 | 5 |
| [GitHub Copilot](https://feedbackbench.com/agents/copilot.md) | =8 | Too few posts | – | – | 11 | 2 | 9 |
| [Amp](https://feedbackbench.com/agents/amp.md) | =11 | Too few posts | – | – | 9 | 6 | 3 |
| [Cline](https://feedbackbench.com/agents/cline.md) | =8 | Too few posts | – | – | 8 | 1 | 7 |
| [Factory](https://feedbackbench.com/agents/factory.md) | =11 | Too few posts | – | – | 6 | 0 | 6 |
| [Zed](https://feedbackbench.com/agents/zed.md) | =8 | Too few posts | – | – | 2 | 0 | 2 |
| [Augment Code](https://feedbackbench.com/agents/augment.md) | 17 | Too few posts | – | – | 2 | 0 | 2 |
| [Kiro](https://feedbackbench.com/agents/kiro.md) | 13 | Too few posts | – | – | 0 | 0 | 0 |
| [Conductor](https://feedbackbench.com/agents/conductor.md) | 14 | Too few posts | – | – | 0 | 0 | 0 |
| [Warp](https://feedbackbench.com/agents/warp.md) | 15 | Too few posts | – | – | 0 | 0 | 0 |
| [Grok Build](https://feedbackbench.com/agents/grok-build.md) | 16 | Too few posts | – | – | 0 | 0 | 0 |

## Posts

Receipts rule: The 5 most recent praise and complaint posts per area (first 700 characters) and 3 per criterion (first 450 characters).

### Claude Code

- Praise, 2026-09-27, r/ClaudeCode (Reddit): “the readme bit about not capping a rebuilt estimate at 100% caught my eye. a wrong number could make the agent stop halfway through a job, so having it say "unknown, refresh the reading" is useful. showing the snapshot’s age alongside the estimate also helps the user tell whether they’re looking at an old estimate or a fresh account reading” [source](https://www.reddit.com/r/ClaudeCode/comments/1wrlvez/built_a_plugin_for_claude_code_that_lets_claude/pcf4jm7/)
- Praise, 2026-09-27, @ClaudeDevs (X): “@claudedevs nice, this makes the math easier to sanity-check. i usually assume savings get eaten by retries, but a calculator helps separate that from real task cost. would love to see examples where the drop felt biggest.” [source](https://twitter.com/1721322400237940736/status/2104051736596095092)
- Praise, 2026-09-27, @ClaudeDevs (X): “i will not switch default on sticker price. i compare cost per finished task via /usage. if cache reads stay under half of input, i keep sonnet for short edits and use opus 5.5 only for supervised multi-file work. i also freeze the comparison after two full weeks of logged tasks before any default change.” [source](https://twitter.com/1862041092352532480/status/2104194729223106703)
- Complaint, 2026-09-27, r/ClaudeCode (Reddit): “2. llm reasoning summary is pointless. eventually llms will go-away from readable thinking processes (if they haven't already, like the astra model). all you'd be doing is allowing for distillation, with no benefit to the user. 3. nobody knows what usage is based off. tokens? it's an internal magic value. all we know is you get a significantly better deal vs paying for tokens. if they specified the actual token usage exactly, then they'd might a” [source](https://www.reddit.com/r/ClaudeCode/comments/1wr67vx/i_agree_with_dario_regulating_llms/pca00m4/)
- Complaint, 2026-09-27, r/ClaudeCode (Reddit): “ok. my reset is on wed. i used it to 1% and reset it this morning using the free. i was at about 82-83% earlier. than when i was going to check usage, i notice it went down to 1%. also i double checked on website, its the same. i'm max 20x.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wr9xke/i_think_we_just_got_a_reset/pcax64m/)
- Complaint, 2026-09-27, r/ClaudeCode (Reddit): “the fact that you measure in "number of prompts" is very very telling.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wradi9/why_normal_discussions_has_so_many_haters/pcb2lvp/)

### Google Antigravity

- Praise, 2026-09-27, r/google_antigravity (Reddit): “so curiously agy told me to use vertex veo/ ai and imagen - i can ask it obviously ask why it didn’t say remotion but for the layman can you explain the diff? the quality is absolutely phenomenal it linked to my gcloud made skills for both so i can say much like generating ui skill > create an image and it links to image gen or create video it links to veo. it also does multi shots and stitch etc. lets me know the estimated cost. i would assume s” [source](https://www.reddit.com/r/google_antigravity/comments/1wr9gwg/how_can_i_make_a_short_animation_video/pcbhhoj/)
- Praise, 2026-09-26, r/google_antigravity (Reddit): “i think the token graph is nice but id like to see an output comparison as well. i still feel like this could change output quality especially if "caveman" is used by name in the skill etc.” [source](https://www.reddit.com/r/google_antigravity/comments/1wppu90/caveman_multi_agent_efficiency/pc7f3zf/)
- Praise, 2026-09-24, @antigravity (X): “@mannacodeai @jonsouyang @antigravity exactly — once the meter is visible mid-run, “keep going” stops feeling free. soft prompts are polite; a hard stop outside the loop is what actually ends the spend.” [source](https://twitter.com/1437246414845841409/status/2103252551743258946)
- Complaint, 2026-09-27, r/google_antigravity (Reddit): “bro in my hermes it says resource exhausted or quota exhausted even if the quota is full ?? any help” [source](https://www.reddit.com/r/google_antigravity/comments/1wrh0mx/gemini_38_in_alternative_harnesses/pcf336g/)
- Complaint, 2026-09-27, @antigravity (X): “@antigravity can you show the context usage? it's hard to know how much content is being used.” [source](https://twitter.com/1925388589241966594/status/2104010402829041896)
- Complaint, 2026-09-26, @antigravity (X): “@antigravity i don't know what's wrong with the app; the quota always shows as 100% and doesn't reset like it's supposed to, even though updates have been released almost daily. i've already uninstalled and reinstalled it—deleting the folder in the process—but nothing works.” [source](https://twitter.com/1544718019263447041/status/2103665685633118252)

### OpenCode

- Praise, 2026-09-25, r/opencode (Reddit): “for sure. that's why they refuse to build an actually useful dashboard like opencode with detailed usage metrics. sadly, opencode keeps refusing my cards, so i can't even go back even if i wanted to, but my god is command code such a shitty and shady provider.” [source](https://www.reddit.com/r/opencode/comments/1wozjjf/with_deepseek_v41_flash_it_feels_impossible_to/pbw771g/)
- Praise, 2026-09-25, r/opencodeCLI (Reddit): “their not because cline is not at all transparent when it comes to usage transparent scale opencode > command code > cline” [source](https://www.reddit.com/r/opencodeCLI/comments/1wpp83n/deep_comparison_for_heavy_use_command_code_vs/pbxbzqi/)
- Praise, 2026-09-22, @opencode (X): “i just love @opencode tui, theme is beautiful and cost tracker is super useful! <strict_link>” [source](https://twitter.com/1367864633290199041/status/2102336198098485705)
- Complaint, 2026-09-27, r/opencode (Reddit): “it’s the most frustrating thing i can’t see model costs and how much cash left in my account in the ui of open code.” [source](https://www.reddit.com/r/opencode/comments/1wrs0xl/openchamber_is_so_much_better_than_opencode/pcfrjqr/)
- Complaint, 2026-09-27, r/opencode (Reddit): “the desktop app tells me when i've \*hit\* a limit ("go limit reached") — but i couldn't find anywhere that shows how much of my 5h / weekly caps i have \*\*left\*\* before i run into it. so i built a sidebar that shows it, plus the other numbers i kept alt-tabbing to check. \*\*what's in it\*\* \- opencode go + openai caps — 5h / weekly / monthly: used, left, and reset times \- estimated per-model share of each go cap window (so i can see which” [source](https://www.reddit.com/r/opencode/comments/1wrfhdx/i_built_a_telemetry_sidebar_for_the_opencode/)
- Complaint, 2026-09-27, @opencode (X): “i have been with the sub of @openai and go of @opencode with tools like pi, raycast, hermes, and others that work in byok mode. i wanted a quick way to check how much inference i had left, and that's why i made mana. <strict_link>” [source](https://twitter.com/36750304/status/2104256556183286036)

### Cursor

- Praise, 2026-09-25, @cursor_ai (X): “thanks @cursor_ai &amp; @spacexai for adding plan usage meters 👏 <strict_link>” [source](https://twitter.com/2073328259681632256/status/2103390000805294309)
- Praise, 2026-09-15, @cursor_ai (X): “@coscosmico @cursor_ai @claudedevs @openai useful feature for the multi-agent setups most developers end up running. auto-detecting the source tool and only surfacing the aggregated counts when 2+ are active keeps the report clean and relevant.” [source](https://twitter.com/1720665183188922368/status/2099774828299301339)
- Praise, 2026-09-15, @cursor_ai (X): “@kylezantos @cursor_ai @grok @openai @claudeai this is really cool and helpful at the same time. no more clicking the settings just to check the usage.” [source](https://twitter.com/2041919964312170504/status/2099858300984832261)
- Complaint, 2026-09-27, r/cursor (Reddit): “i feel you — that usage discrepancy is brutal and the dashboard does a poor job of explaining what's actually happening. the 'other models' bucket covers any non-cursor model you call (gpt-4, claude, etc.), and those burn through the allowance way faster than the native cursor models. if you have 'auto' model selection on or you're hitting cmd+k / chat with a non-cursor model selected, you'll chew through that quota in minutes. check settings → m” [source](https://www.reddit.com/r/cursor/comments/1wrddz6/why_is_the_other_models_usage_is_being_used_so/pcd5hao/)
- Complaint, 2026-09-27, r/cursor (Reddit): “so here's my issue with grok models up until 3.x\~ they were training their own models on their own data. in house model, in house data. spacex sees what cursor is doing, which is basically using all the enterprise data flow and their retail use towards training their own in-house model and they want in on it too. here's the kicker, they all start with the same base, it's all kimi 2.5 under the hood. spacex agrees to "buy" cursor, provides bigges” [source](https://www.reddit.com/r/cursor/comments/1wrktgc/am_i_the_only_one_who_thinks_grok_47_is_actually/pcdbwh3/)
- Complaint, 2026-09-27, @cursor_ai (X): “hi elon — multiple paid subscriber here: x premium - ~$40/month @x @elonmusk - $4.00/month. cursor pro - ~$20/mo @cursor_ai i love grokbot and i’m using it more and more but i think it is going to cost me a small fortune (total unknown) to get all my finances and household tracking into grokbot control/management. it’s not easy to tell which plan you should pay for, or how much each kind of use counts against your allotted tokens. showing usage” [source](https://twitter.com/562176566/status/2104314460584501652)

### OpenAI Codex

- Praise, 2026-09-27, r/codex (Reddit): “nope but it’s back at 100%” [source](https://www.reddit.com/r/codex/comments/1wqvl66/reset_just_came_in/pcb8tlc/)
- Praise, 2026-09-27, r/codex (Reddit): “usage is actually very transparent. you can check your entire consumption with codex, and you can even see what a quota of 100 points represents in codex credits, rather than relying on token counts, which can vary depending on cache usage. so far, they have never changed the quota itself. it's just that some models, like astra for example, cost much more in codex credits. but gpt-5.6 sol has been very stable at around 60k codex credits per full” [source](https://www.reddit.com/r/codex/comments/1wrazdi/i_cannot_take_this_anymore/pcbbii9/)
- Praise, 2026-09-26, r/codex (Reddit): “same for me. still visible through web gpt interface though” [source](https://www.reddit.com/r/codex/comments/1wr1lqe/did_the_usage_meter_disappear_from_vs_code/pc93bsd/)
- Complaint, 2026-09-27, r/codex (Reddit): “that’s my observation as well. i swear it seems like how fast it drains depends on the resource they have available or how many people are using it. or something else. it’s pretty annoying not having any reliable way to estimate how much work i can do over a given timeframe. the user experience is just so bad. can i use astra for this task? should i switch to a smaller model just in case astra decides to eat 10%? the “range anxiety” they’re crea” [source](https://www.reddit.com/r/codex/comments/1wr4cp6/more_resets_incoming_next_week/pc9ya13/)
- Complaint, 2026-09-27, r/codex (Reddit): “same, one account has 90% the other has 50% it sucks” [source](https://www.reddit.com/r/codex/comments/1wqvl66/reset_just_came_in/pca0rga/)
- Complaint, 2026-09-27, r/codex (Reddit): “api cost benchmarks are meaningless, nobody works api based. we need quoata usage analysis.” [source](https://www.reddit.com/r/codex/comments/1wnx8i9/i_gave_astra_sol_and_opus_55_the_same_rust_task/pca4vpk/)

### Devin

- Praise, 2026-09-19, @cognition (X): “@brandon_galang @cognition @devinai one thing i really like about it is it tells you the size of your thread and how many acu its currently cost so u can change to a new thread. it seems like they dont have compaction on cloud” [source](https://twitter.com/1665450872363708417/status/2101143353551388871)
- Praise, 2026-09-19, @cognition (X): “@ashen_one @cognition codex and k3 both flash empty, then cognition walks in with a prettier cli and fable 5.1 like it paid the rent.the limit meter is still the only honest product roadmap. <strict_link>” [source](https://twitter.com/1371440359608320002/status/2101312910039650418)
- Praise, 2026-09-12, r/windsurf (Reddit): “you really need to learn how to ration. glm 5.2 can do most things and it's compleatly free on devin desktop. you can see your percentage of weekly usage and so you can use a little of the non free models but you need to make sure not to use more than 20% of your weekly usage per day (if you work 5 days). i would mainly or entirely avoid the western models to do this as they are very expensive. when you want the occasional prompt nore powerful th” [source](https://www.reddit.com/r/windsurf/comments/1wdu5dn/usage_limits_on_20_plan/p9bhoeh/)
- Complaint, 2026-09-27, @cognition (X): “couple more gripes with @cognition. please don't have cloud usage only visible on the web app. if we can spin up local and cloud from desktop we should be able to see usage for both from there” [source](https://twitter.com/4108256724/status/2104358548629295526)
- Complaint, 2026-09-25, @cognition (X): “@cognition another common w by devin! now please add subscription usage t_t” [source](https://twitter.com/1289980460244713473/status/2103536340747104278)
- Complaint, 2026-09-21, @cognition (X): “everything else is going great so far @cognition swe-2 is so good at coding it’s crazy - if you could get a decent amount of usage for that on just local for the $20 dollar plan that would be crazy, also on the usage tracker why does it only show the cloud version ?” [source](https://twitter.com/860967132489801728/status/2101944818511290570)

### Pi

- Praise, 2026-09-24, r/PiCodingAgent (Reddit): “thank you! it actually made me realize i have an issue with my custom subagent extension (usage is not seen by the usage tab) which a me problem, so extra kudos for orbit!” [source](https://www.reddit.com/r/PiCodingAgent/comments/1woela9/orbit_pi_v0015_is_out/pbsh0le/)
- Praise, 2026-09-24, r/PiCodingAgent (Reddit): “haha, nice 😂 at least orbit helped you find another bug along the way 😄 glad the usage tab is actually useful! good luck tracking down the subagent issue.” [source](https://www.reddit.com/r/PiCodingAgent/comments/1woela9/orbit_pi_v0015_is_out/pbsnn2p/)
- Praise, 2026-09-18, @pidotdev (X): “habilité el cache miss notice, y un analisis de sesiones en @pidotdev para ver cuánto cuesta cada cosa que le pido. sumé el último incidente que cerré: 36 sesiones, 484m tokens, ~usd 478. ( el 94% fue contexto releído porque no cerraba sesiones. no es que el modelo sea caro, es que lo usé mal. cuando cambien los precios, el que no aprendió queda frito. a seguir aprendiendo.” [source](https://twitter.com/13267532/status/2100961830172545529)
- Complaint, 2026-09-22, @pidotdev (X): “waiting for @pidotdev to add opus 5.5 so i can obliterate my usage meter in one sitting <strict_link>” [source](https://twitter.com/3251128820/status/2102440522849599783)
- Complaint, 2026-09-13, r/codex (Reddit): “i'm seeing unusually fast weekly quota depletion on the $100 pro 5x plan. my remaining allowance went from 100% on september 12 at 11:35 to 24% on september 13 at 06:53, then approximately 13% later that day. times are utc+3. a lengthy browser e2e test explains roughly the final 10 percentage points, but the earlier depletion felt very different from my previous usage. i checked both codex desktop/cli logs and pi agent logs, separating openai usa” [source](https://www.reddit.com/r/codex/comments/1w9w4tj/codex_usage_and_operation_discussion_last_updated/p9k6yjj/)
- Complaint, 2026-09-09, @pidotdev (X): “@pidotdev what got me was one level down: a call that charges you and says nothing about it in the response. the two endpoints that did return a billable count were the ones nobody integrates against. reconciling an invoice against a number that isn't there is guessing.” [source](https://twitter.com/2379961783/status/2097803856512205066)

### GitHub Copilot

- Praise, 2026-09-25, r/ClaudeAI (Reddit): “another bias opinion. copilot works great for many things, like any tool, it has its pro's and cons. if a company is a pure ms shop, copilot makes sense with its integration, sure claude has m365 connector but it is limited in the access it gives and now you need to add custom mcp servers to do more. if you are a pure coder and a good one, sure claude excels there, but for most, copilot paid and choose opus model will do most of what people need.” [source](https://www.reddit.com/r/ClaudeAI/comments/1wq2hsf/claude_in_enterprise/pc0vf1c/)
- Praise, 2026-09-03, r/PiCodingAgent (Reddit): “at work, i use copilot, and the model costs are always visible, but on pi, only the names are shown. so i created model-costs, a small extension that adds the /model-cost command: \- input/output cost per 1 million tokens for each model, directly in the selector \- context window, maximum output, cache read/write speeds, and price tiers \- approximate search, so both /model-cost claude and /model-cost $0.00 work \- current model prices always vis” [source](https://www.reddit.com/r/PiCodingAgent/comments/1w65uyb/i_got_tired_of_not_knowing_what_a_model_costs/)
- Complaint, 2026-09-27, r/GithubCopilot (Reddit): “i switched to codex and it seems to get me many times farther. they kind of obscure how much usage you’re really getting, but $100/mo with codex gets me waaaaay more than $200/mo with copilot, not to mention that gpt 6 is on the pareto frontier anyway” [source](https://www.reddit.com/r/GithubCopilot/comments/1wr5io2/best_cheaper_alternative/pcf7r1z/)
- Complaint, 2026-09-25, r/GithubCopilot (Reddit): “i believe copilot app is bugged. ui shows 20 credits used, yet i lose hundreads.” [source](https://www.reddit.com/r/GithubCopilot/comments/1wq39rv/did_they_10x_the_cost_of_luna_56_since_the/pc1hxi9/)
- Complaint, 2026-09-19, r/GithubCopilot (Reddit): “you get to explain why you used up the group allotment to a team at your company. i got an email saying that we had 75dollars of usage, github enable credits view and i supposedly have 300 dollars of usage a month. 1 month, just for the hell of it used claude for everything and ran my usage to 150, normally use 5.3 codex and barely run over 70-80 ususally” [source](https://www.reddit.com/r/GithubCopilot/comments/1wkdn9s/copilot_is_allowing_27_in_overage_despite_my/pati89v/)

### Amp

- Praise, 2026-09-25, @AmpCode (X): “my pet peeve right now is "ask ai" features on ai providers that don't tell me about my usages. please follow @ampcode on how to do your own puck.” [source](https://twitter.com/913700214183124995/status/2103360003914818025)
- Praise, 2026-09-24, @AmpCode (X): “@brianevanmiller @ampcode i think that's the most nicely distributed orb chart i've ever seen” [source](https://twitter.com/1369860113423609866/status/2103228181197299934)
- Praise, 2026-09-24, @AmpCode (X): “gotta love @ampcode 's usage graph... 😍 <strict_link>” [source](https://twitter.com/2062509956574650368/status/2103237817073565750)
- Complaint, 2026-09-27, @AmpCode (X): “hey @ampcode slight bug with the usage graph thingy. starts breaking if you use 200 threads in a day. there was maybe 30 on saturday the rest is all sunday and literally spilling out onto the saturday box lol <strict_link>” [source](https://twitter.com/2095743155312467968/status/2104268844986671135)
- Complaint, 2026-09-18, @AmpCode (X): “@sqs @ampcode chatgpt extra usage credits. yes, puck did give me token usage for a thread. it didn't know what that translated to in credits, so i pointed it to openai's pricing table and then it gave me an estimate. my idea is for puck to just have that info from the start. minor qol thing.” [source](https://twitter.com/2937634927/status/2101020248187261199)
- Complaint, 2026-09-17, @AmpCode (X): “@ampcode i'm hitting usage limits and curious how many credits i plausibly would need to buy. could you report number of credits used for threads? puck didn't know how to do this until i pointed it to openai's pricing table at <strict_link>.” [source](https://twitter.com/2937634927/status/2100702502626918718)

### Cline

- Praise, 2026-09-14, @cline (X): “byok. i used gpt models using my codex $200 sub and deepseek flash using my airouter[dot]ch sub. (they suspended my api key two days ago because of overuse lmao) if you're using muse spark 1.3 (which is free) please be very specific of what u want it to do. ideally setup voice typing using their byok/pass else use handy/wisprflow/some other stt tool. i don't hit any limits because i was on byok mostly. but you can track token and cost. they have” [source](https://twitter.com/1654347044503408640/status/2099547364352823600)
- Complaint, 2026-09-25, r/opencodeCLI (Reddit): “their not because cline is not at all transparent when it comes to usage transparent scale opencode > command code > cline” [source](https://www.reddit.com/r/opencodeCLI/comments/1wpp83n/deep_comparison_for_heavy_use_command_code_vs/pbxbzqi/)
- Complaint, 2026-09-21, @cline (X): “@cline could you please explain what you mean by “generous quotas”? other providers like @opencode and @commandcodeai clearly show how much usage is available for each model. why don’t you provide the same level of transparency?” [source](https://twitter.com/912174747416330240/status/2102122288187592917)
- Complaint, 2026-09-21, @cline (X): “@cline any chance we could get a transparent model-by-model usage breakdown for clinepass? how much mimo-v2.6-pro, glm, deepseek, etc. usage do we actually get with the subscription? other similar subscriptions publish this pretty clearly - would be really useful for comparing plans.” [source](https://twitter.com/1770069362608648192/status/2102165413500973469)

### Factory

- Complaint, 2026-09-22, r/FactoryAi (Reddit): “unclear as to why i'm getting a droid core usage limit message when usage settings show 0% have been used.” [source](https://www.reddit.com/r/FactoryAi/comments/1wnh0g4/conflicting_usage_stats/)
- Complaint, 2026-09-20, @droid (X): “@droid also when i see 4x next to fable i'm like im not using that - neeed a less scary way to communicate usage” [source](https://twitter.com/3781517712/status/2101601846183760135)
- Complaint, 2026-09-14, @FactoryAI (X): “ayo @droid homies, ur weekly limits raise % while im afk, can we fix this please? @factoryai i literally went afk 40 min found my 5h limit going up 3 % and weekly 2%? doesn't make any sense.” [source](https://twitter.com/1834314883510226944/status/2099647076481044788)

### Zed

- Complaint, 2026-09-04, r/ZedEditor (Reddit): “check your balance on your openai account. it's probably a missmatched warn message on zed's end.” [source](https://www.reddit.com/r/ZedEditor/comments/1w6jhtb/free_usage_exceeded_error_even_though_i_have_only/p7pesqn/)
- Complaint, 2026-09-04, @zeddotdev (X): “hey @zeddotdev guys, this is probably really far down on the priority list for you guys rn, but please put the usage/balance of the zedvip tokens in the delta app somehow. having to open a browser window to see my balance is a few too many steps. otherwise i’m a huge fan of the app so far 👍👍” [source](https://twitter.com/30111156/status/2095791764284281340)

### Augment Code

- Complaint, 2026-09-26, G2 (G2): “q: what problems is the product solving and how is that benefiting you? a: it completely eliminates the tedious manual effort of navigating unfamiliar codebases, writing repetitive boilerplate, and tracing cross-field dependencies saving me 1-2 hours of grunt work every day. q: what do you like best about the product? a: the context engine is really amazing. in contrast to a simple ai autocomplete function, it creates an index of all of your mult” [source](https://www.g2.com/products/augment-code/reviews/augment-code-review-13618001)
- Complaint, 2026-09-18, @augmentcode (X): “@augmentcode regardless of whether it merges code or offers optimizations, the subscription for augment is very expensive. i ran out of tokens unexpectedly without even realizing it, and the pricing structure is unclear—i wouldn't recommend using it.” [source](https://twitter.com/955391612632252418/status/2100826713957810177)
