# OpenCode (Anomaly (open source))

Feedback Bench, coding agents, built 2026-10-01, window 2026-08-31 to 2026-09-27. Web page: https://feedbackbench.com/#/agent/opencode

| Measure | Value |
|---|---|
| Rank | 3 of 17 (rank range 3–3) |
| Feedback Score | 66.0 (95% interval 65.4–66.7) |
| Popularity | 0.773 (share of voice 11.56%) |
| Customer love | 0.564 (95% interval 0.553–0.575) |
| Top quadrant | yes |
| Authors | 11387 |
| Posts counted | 25673 |
| Posts that judge the agent | 10304 |
| Criteria better / worse than peers | 7 / 7 of 63 |

## The brief

Written by Claude Opus 5.5 from 133 labelled posts and the numbers on this page. Interpretation, not measurement: every quote is verbatim and links to its post.

**An open, model-agnostic harness that buys a lot, if you tolerate churn.**

TL;DR:

- Users praise how much work a plan buys; cheap flash models turn long sessions into pocket change.
- Output quality rides on the cheap model you pick, and capability posts trail peers.
- Billing glitches, account lockouts and shifting quotas draw sharper anger than the agent loop itself.

### What hurts

- **Account and billing break after updates** ([Wrong charges, failed payments and plan provisioning](https://feedbackbench.com/criteria/account.billing_errors.md), [Account bans and access restrictions](https://feedbackbench.com/criteria/account.bans_restrictions.md), [Support, refunds and issue handling](https://feedbackbench.com/criteria/account.support.md)). The new dashboard is losing subscriptions and blocking payment changes, so paying users get treated as fresh signups or locked out.
  This is the clearest worse-than-peers area, and posts contain zero praise. Users report a GitHub login that no longer shows an existing plan, balance or payment method. Others report a renewal fee taken while the app says the subscription is inactive.
  
  Access restrictions land in the same bucket. Users describe sessions timing out into restrictions with no explanation. The fixes users ask for are basic: keep accounts intact across a redesign and show why access was cut.
  Evidence:
  - Complaint, OpenCode, @opencode, 2026-09-20: “the opencode webpage is not as user-friendly as the original version after the update, and there are bugs that prevent changing the payment method @opencode” [source](https://twitter.com/362980075/status/2101596640205590973)
  - Complaint, OpenCode, r/opencode, 2026-09-23: “i used my github account for my identity when signing up for go at the beginning of the year. with the new dashboard, logging in with my github suddenly doesn't have any info on my subscription, payment method, zen balance, etc. it's treating me like a total new signup. the last payment to anomaly came out 8/26, so nothing for september yet. anyone else have a similar issue? just wondering how to get back to my existing account. i'm trying to create a new api key and view my usage details.” [source](https://www.reddit.com/r/opencode/comments/1wnr7ah/new_dashboard_is_missing_my_go_sub/)
  - Complaint, OpenCode, r/opencode, 2026-09-22: “anyone here from opencode. the opencode 10$ subscription fee is deducted from my account yesterday and the gis subscription is supposed to be renewed and working. but i am getting error that my go subscription is not active. the money is already deducted from my account. can anyone help here? how can i either get the go subscription or get my money back?” [source](https://www.reddit.com/r/opencode/comments/1wn7cmv/open_code_go_subscripition_issues/)
  - Complaint, OpenCode, @opencode, 2026-09-10: “@opencode what's going on? why was it restricted after a while? access timed out.” [source](https://twitter.com/1807477178730504192/status/2097994520814342234)

- **Cheap models, uneven results** ([Can do the user's kind of task](https://feedbackbench.com/criteria/work.capability.md), [Quality got worse or better over time](https://feedbackbench.com/criteria/models.quality_drift.md)). The value story runs on budget models, and users say those models iterate fast but shallow, pushing more manual correction back onto the developer.
  Task capability is the one core work criterion rated worse than peers. Complaints rarely blame the harness itself. They blame the free and flash models that make the plans affordable: confident wrong answers, laziness on ambiguous tasks, and answers that need round after round of fixes.
  
  Some posts go further and say the bundled models are not strong coders at all. The praise exists, but it clusters on small or well-scoped tasks.
  Evidence:
  - Complaint, OpenCode, r/opencode, 2026-09-23: “deepseek harness, that's fine just me sharing an opinion. although, at some point you might start to see that you have to iterate more and more manually due to the fast and mediocre answers of deepseek flash, that's what i am referring to.” [source](https://www.reddit.com/r/opencode/comments/1wolqxv/pricing_of_deepseek_v41_flash/pbo5xlx/)
  - Complaint, OpenCode, @opencode, 2026-09-26: “@opencode feels worse than mimo v2. 6 flash tested it a little” [source](https://twitter.com/1854068116881379345/status/2103888119049363882)
  - Complaint, OpenCode, r/opencode, 2026-09-18: “go targets coding and neither of these are all that useful at coding.” [source](https://www.reddit.com/r/opencode/comments/1wk039a/why_dont_we_have_small_models_on_opencode_go_such/pan1ddx/)
  - Complaint, OpenCode, r/opencode, 2026-09-17: “does not work properly most of the time in my case. it frequently forgets the current working directory (e.g. suddenly 1-2 levels of the path are missing or part of the path has an encoding error) and then requests permission for the "new" path that does not exist, tool calls stop all the time, speed doesn't really seem to depend on difficulty of the task - sometimes it's quick or i might need to wait several minutes. and i can't really reproduce the quality of the output seen on x vs. with other models.” [source](https://www.reddit.com/r/opencode/comments/1wj8wz7/opinions_about_union_alpha/pagts8v/)

- **Rolling windows cap the real allowance** ([Short rolling usage window blocks or interrupts work](https://feedbackbench.com/criteria/limits.window_interrupts_work.md), [Usage meter visibility and accuracy](https://feedbackbench.com/criteria/limits.usage_meter.md), [Quota reset timing and bonus or banked resets](https://feedbackbench.com/criteria/limits.reset_schedule.md)). Five-hour and weekly limits stack so that a generous-looking quota cannot actually be spent, and work halts mid-project.
  The interrupting usage window is rated worse than peers. Users do the arithmetic from the docs and find the weekly cap lets them fill the rolling window only a few times. New projects hit the wall fastest.
  
  The meter adds confusion. Some users report free usage marked exhausted on days they did not use it. Others, on lighter models, never come close. Experience depends heavily on which model eats the quota.
  Evidence:
  - Complaint, OpenCode, r/opencode, 2026-09-01: “this is a bit i didn't like. you only have a few 5hr sessions left. so maybe twice a day you free up time to do it. it's going to be tough to use it all because you'll then hit your weekly budget. so at best... maybe 4 sessions. not great for a \*new\* project. i would recommend using it to do a cleanup or review of something you have on your desktop for now.” [source](https://www.reddit.com/r/opencode/comments/1w3u5fq/havent_used_my_subscription_all_month_few_days/p7357f5/)
  - Complaint, OpenCode, r/opencodeCLI, 2026-09-23: “weekly usage != sum of rolling usages . its less than that. their documentation literally says: |mimo-v2.6-flash|30,100|75,200|150,400| |:-|:-|:-|:-| requests respectively for 5h window, week and month. so weekly you may fill up rolling usage to full only 2.5 times, and monthly only 5x. the whole go thing and other subs works because somebody uses only 1% of monthly allowance and other 70% -> median user consumes less than they are spending in total for apis . and this 5h, weekly limits help companies to keep us from spending 100% of allowance.” [source](https://www.reddit.com/r/opencodeCLI/comments/1woh5q9/im_not_buying_the_weekly_usage_thing/pbnv4hd/)
  - Complaint, OpenCode, r/opencodeCLI, 2026-09-20: “i have the same issue, i haven't used any yesterday and now its showing free usage exceeded :(” [source](https://www.reddit.com/r/opencodeCLI/comments/1wlab1z/rate_limits_issue/paxhmcs/)
  - Complaint, OpenCode, @opencode, 2026-09-03: “muse spark 1.3 &gt; doesn't follow instructions &gt; feels dumb like opus 5 &gt; is now slow af and ratelimit on @opencode &gt; unusable 🥀 <strict_link>” [source](https://twitter.com/1426586674264350722/status/2095554367437013400)

- **Long runs drift and overrun** ([Long unattended runs and goal/loop mode](https://feedbackbench.com/criteria/work.long_running_autonomy.md), [Output degrades as the context window fills](https://feedbackbench.com/criteria/context.long_context_decay.md), [Reasoning effort setting and its defaults](https://feedbackbench.com/criteria/models.effort_control.md)). Unattended runs and huge contexts are where cheap models lose the plot, running for hours, doing unexpected things, or bloating cost as context fills.
  Long-running autonomy, long-context decay and effort control are all worse than peers. Users describe overnight runs still grinding on simple tasks, an eight-hour run that went off-script, and sessions near 800k tokens that suddenly get expensive.
  
  Others say they cannot get a true plan-implement-test loop without manually nudging each phase. The removal of a max reasoning level also drew complaints, and a max effort option is a standing request.
  Evidence:
  - Complaint, OpenCode, @opencode, 2026-09-25: “@opencode the only reason i can actually use ai as much as i am is bc how cheap dsv4.1 flash &amp; glm-5.3 flash are. i had my own ai gone wild episode earlier this wk, bc i left ds on a task that ran for 8 hrs. it did all sort of things i did not expect. lesson learned.” [source](https://twitter.com/43463961/status/2103608259672183272)
  - Complaint, OpenCode, r/opencode, 2026-09-11: “yeah that's the problerm with luna, especially on thinking level max, it takes an age to finish any task, i left the agent running with model luna:max at 10pm, came back to my pc in the morning at around 9.30am and it was still going, this is on a relatively simple task, it maybe cheap but its shit and will burn through your time, im currently giving sol a whirl, hoping for better results.” [source](https://www.reddit.com/r/opencode/comments/1wd7aka/anyone_else_spending_10x_more_time_on_reviewdebug/p947glo/)
  - Complaint, OpenCode, r/opencodeCLI, 2026-09-20: “i'm not an llm genius by far, and at some point i start new sessions. starting a new session in the middle of something isn't really bad, but it usually changes the mood and project understanding always drops a bit. i'm good with deepseek v4 flash for now and i'm at a point in the flow that i don't want to start a new session, but i'm hovering around 800,000 tokens in my context even after compression and i'm spending quite a bit of money all of sudden.” [source](https://www.reddit.com/r/opencodeCLI/comments/1wlnl1z/deepseek_v4_flash_is_getting_expensive/)
  - Complaint, OpenCode, r/opencode, 2026-09-06: “i'm building a web platform that will eventually teach python and php. my idea was to create a detailed specification, give it to qwen through opencode, give it permission to edit files, run tests, use git, etc., and let it build most of the application automatically. ideally, it would work like this: `plan -> implement -> test -> fix -> commit -> next task -> repeat` but so far, i keep ending up with workflows divided into many phases where i have to manually tell the agent to continue. is this just how opencode is supposed to work, or can it actually be configured for longer autonomous development? do people normally use an external loop/supervisor to keep qwen working, or am i approaching opencode incorrectly? i'm mainly interested in hearing from people who have actually tried to automate larger projects with opencode + qwen.” [source](https://www.reddit.com/r/opencode/comments/1w8zae4/can_opencode_qwen_build_a_project_mostly/)

### What works

- **Plans buy a lot of work** ([How much use a plan's price buys](https://feedbackbench.com/criteria/limits.plan_value.md), [Single prompt, model or effort level consumes disproportionate quota](https://feedbackbench.com/criteria/limits.burn_rate.md)). Plan value is the strongest signal; cheap flash models let users push large token volumes for very little money.
  Pricing and limits is the best-rated area and beats peers. Users post small receipts for big sessions, one-shot builds that cost pennies, and say the Go plan remains a decent deal even after promotions end.
  
  The complaints in this area are about specific models being limited on Go, not about the overall price-to-work ratio. Requests skew toward more of the same, including higher limits and a pricier top tier.
  Evidence:
  - Praise, OpenCode, r/opencodeCLI, 2026-09-11: “half million tokens since 8:30 am and so far 45 cents - me loves - me say thank you 🙏” [source](https://www.reddit.com/r/opencodeCLI/comments/1wccm5n/what_the_actual_fuck/p98mlko/)
  - Praise, OpenCode, r/opencodeCLI, 2026-09-10: “it oneshotted a fairly complex thing, and cost me 4 pennies. unthinkable just a couple of months ago.” [source](https://www.reddit.com/r/opencodeCLI/comments/1wccm5n/what_the_actual_fuck/p8wwk4x/)
  - Praise, OpenCode, r/opencodeCLI, 2026-09-23: “surprised on how well it performs per unit cost for me on max on all of those incremental updates or slight change tasks when given some direction.” [source](https://www.reddit.com/r/opencodeCLI/comments/1wnhazu/gpt6_luna_cheap_than_deepseek_v41_flash/pbi8du3/)
  - Praise, OpenCode, r/opencodeCLI, 2026-09-01: “hopefully they keep going with their operation cheapseek thing, but i'm not getting my hopes up. opencode go is still a decent deal though. ### sign up here: <strict_link> get $5 of extra usage for free when you sign up for opencode go using this link!” [source](https://www.reddit.com/r/opencodeCLI/comments/1vyxbzz/is_opencode_go_worth_it_for_10_or_nah/p75bypd/)

- **Model variety lands fast** ([Which models are offered on a plan and when](https://feedbackbench.com/criteria/models.catalog_access.md), [Automatic model routing and fallback](https://feedbackbench.com/criteria/models.routing_auto.md)). Users pick OpenCode to run many models in one harness, with new launches supported within hours.
  Catalog access beats peers. Posts describe choosing the harness per model, running two or three models to review each other's plans, and getting more control over system prompts than closed tools offer.
  
  The flip side is constant demand for whichever model is newest. Missing flash variants trigger complaints, and adding DeepSeek V4.1 Flash is a top request.
  Evidence:
  - Praise, OpenCode, @opencode, 2026-09-22: “@opencode same-day support for models that launched hours ago - the integration race is its own sport at this point and the users keep winning” [source](https://twitter.com/1460189373085802496/status/2102484389951369610)
  - Praise, OpenCode, r/opencode, 2026-09-09: “just the usual suspects: opencode, opencode2, claude code, codex, copilot but we use different models and pick the harness accordingly. no point in running gpt models in anything other than codex for example and why would i put myself through the pain of running glm in codex instead of opencode.” [source](https://www.reddit.com/r/opencode/comments/1wbqrwm/opencode_tied_for_last_in_frontierharness_eval/p8ssbt3/)
  - Praise, OpenCode, r/opencode, 2026-09-18: “i like model variety. i feel like for the actual complex tasks i prefer 2-3 different good models reviewing the plan and code instead of just one very good. i also like to use different or custom harness” [source](https://www.reddit.com/r/opencode/comments/1wj13ox/what_is_the_reason_you_use_opencode_instead_of/pah61uk/)
  - Praise, OpenCode, r/cursor, 2026-09-17: “with opencode you can review as well. plus you get to leave comments and interact like so. plus you get as many models you'd like to. and you don't get the not adjustable system prompt bullshit or cursor / grok bullshit. you get a ton more control, the models actually follow your instructions. requires a bit skill to master but if mastered then it's very good. stop using cursor and other scams.” [source](https://www.reddit.com/r/cursor/comments/1whtb6c/whats_better_about_cursor_than_claude_code/pah0rpm/)

- **No lock-in to one provider** ([Using an existing subscription across tools](https://feedbackbench.com/criteria/billing.subscription_portability.md), [Connecting own API keys, local models and custom endpoints](https://feedbackbench.com/criteria/setup.provider_byok_local.md)). Users value bringing their own keys and taking the subscription's API access elsewhere, which turns the plan into portable tokens.
  Subscription portability beats peers. Users say the plan would be unremarkable without outside-harness use, and frame the open-source, any-provider design as insurance when a vendor changes terms.
  
  The weak spot is the reverse direction. Posts report a ChatGPT subscription that stopped working inside OpenCode, and note other vendors' terms forbid it.
  Evidence:
  - Praise, OpenCode, @opencode, 2026-09-07: “@cory_schulz_ @forloopcodes @opencode yes! it would be a meh sub if you couldn't use it outside opencode, that's the best part” [source](https://twitter.com/1782486953297874944/status/2096795951377547612)
  - Praise, OpenCode, r/opencodeCLI, 2026-09-11: “you’re not stuck with one provider if you use opencode. when anthropic decide to rug pull and you’re stuck with their proprietary, closed source harness, you’ll need to scramble to find a replacement. claude code is nothing special and is actually very bloated. opencode allows you to customise it, use any model/provider, and is open source. if one provider rug pulls, switch to another. i was a heavy cc user and felt anthropic models were way too verbose regardless of what was in my prompts or settings. i switched to opencode, played with chinese models, ended up with a chatgpt subscription and haven’t looked back.” [source](https://www.reddit.com/r/opencodeCLI/comments/1wdb7jz/im_a_claude_code_user_and_dont_understand_why/p94kimc/)
  - Praise, OpenCode, r/opencode, 2026-09-16: “i find if you’re smart with the model you choose rather than just brute forcing everything, and maybe use your brain every so often, you can get pretty far with a significantly cheaper price than chat or claude. it also gives your api access so you can use them tokens in anyway you like which is great” [source](https://www.reddit.com/r/opencode/comments/1wg8l9v/does_opencode_go_make_sense_now/pa444ue/)
  - Complaint, OpenCode, @opencode, 2026-09-25: “@thsottiaux why did chatgpt subs stop working in @opencode <strict_link>” [source](https://twitter.com/1473148067306262529/status/2103616532987490350)

- **Subagents and background work click** ([Subagents, parallel agents and orchestrators](https://feedbackbench.com/criteria/work.multi_agent_orchestration.md), [Mobile, remote-control and voice access](https://feedbackbench.com/criteria/surfaces.remote_mobile.md)). Parallel subagents, background tasks and remote access let users run OpenCode as an orchestrator, and posts describe it as a cheap executor behind other tools.
  Multi-agent orchestration and mobile access both beat peers. Users describe root agents that wake only when subagents finish, and pairing a frontier planner with cheap flash executors to cut token use.
  
  OpenChamber and the mobile app get credit for checking on agents from a phone. Friction remains: free-tier subagents get blocked, and some agent permissions need manual tweaks.
  Evidence:
  - Praise, OpenCode, @opencode, 2026-09-25: “@traves_theberge @opencode parallel agents checking git diffs is such a clean idea, love how many ways you can run it” [source](https://twitter.com/317213818/status/2103364135731744921)
  - Praise, OpenCode, r/codex, 2026-09-23: “my workaround: switch to ohmyopenagents on opencode or just use opencode v2. both will have feature for better subagents workflow, where root will only wake up when subagents finish their job with brief summary. huge plus if you have opencode go subscription. you can use deepseek or muse as most subagents in that workflow.” [source](https://www.reddit.com/r/codex/comments/1wo4c5z/astra_luna_setup_questions/pbk2q34/)
  - Praise, OpenCode, r/ClaudeCode, 2026-09-16: “yes, now i have found another way to reduce token consumption for sub-agents by using opencode as sub-agent dispatcher and use a flash series model, this way cc opus stays the orchestrator and oc becomes executor - no context bloat and faster executions” [source](https://www.reddit.com/r/ClaudeCode/comments/1wibhal/havent_maxed_out_in_months_even_on_pro_plan_opus/pa9b20w/)
  - Praise, OpenCode, r/opencode, 2026-09-18: “+1 for openchamber. main way i interact with opencode, makes the entire workflow and feature set of oc really easy to use, i can check on my agent remotely with my phone, it's bliss” [source](https://www.reddit.com/r/opencode/comments/1wj13ox/what_is_the_reason_you_use_opencode_instead_of/paj93zg/)

### Under the surface

- **Prices and promos move weekly** ([Price, allowance or plan terms changed](https://feedbackbench.com/criteria/limits.allowance_change.md), [Pricing and plan terms stated clearly and consistently](https://feedbackbench.com/criteria/billing.pricing_clarity.md)). Promotions, credits and per-model prices shift so often that users stop trusting any stated deal to last.
  Allowance changes rate better than peers because many changes were cuts in price or permanent credits. But posts also flag a price that lasted under a day, a credit dropping in value within a week, and a stealth-free model turning expensive.
  
  Pricing clarity draws almost no praise. Making temporary bonuses permanent is among the top requests.
  Evidence:
  - Complaint, OpenCode, r/opencode, 2026-09-07: “that price didnt even last a day lmao. it's now: $0.14 $0.28 $0.028” [source](https://www.reddit.com/r/opencode/comments/1w8ukta/opencode_should_have_another_look_at_ds4_its/p8c774g/)
  - Complaint, OpenCode, r/opencode, 2026-09-13: “i switched to them from go and haven't had any issues with ds. be forewarned that $60 credit is turning to $40 in a week.” [source](https://www.reddit.com/r/opencode/comments/1wdo67f/looking_for_alternatives_to_opencode_go/p9hlvpq/)
  - Praise, OpenCode, @opencode, 2026-09-25: “opencode bros really turned operation cheapseek into a permanent feature. @opencode 🤝 @deepseek_ai <strict_link> <strict_link>” [source](https://twitter.com/1960747924843061253/status/2103543164888097277)
  - Complaint, OpenCode, r/opencodeCLI, 2026-09-04: “who goes for an annual discount on ai subscriptions i wonder, considering how pricing and models change on a weekly basis.” [source](https://www.reddit.com/r/opencodeCLI/comments/1w6pr1w/deepseek_tokens_vs_zai_lite_plan/p7qy448/)

- **The v2 rollout cuts both ways** ([Updates break working setups](https://feedbackbench.com/criteria/rel.update_breakage.md), [How the interface shows work, and what the user can configure](https://feedbackbench.com/criteria/ui.display_settings.md), [Client crashes, freezes and failed tool execution](https://feedbackbench.com/criteria/rel.client_failures.md)). Upgrading to v2 fixes old bugs for many users, but a quiet migration and desktop regressions left others confused or stuck.
  Update breakage rates better than peers, largely because posts say v2 or a reinstall resolves stuck resets and tool-call bugs. Some users switched back for tabs and backgrounding.
  
  Others report the desktop app regressing between updates, an unclear active-tab display, and many users not realising v2 left beta. Restoring the previous UI is a standing request.
  Evidence:
  - Complaint, OpenCode, r/opencode, 2026-09-26: “ya. they did a very poor job of migrating existing users. so people who are not on x or discord still thinks v2 is beta.” [source](https://www.reddit.com/r/opencode/comments/1wqppth/opencode_v2/pc70efe/)
  - Praise, OpenCode, r/opencodeCLI, 2026-09-21: “i had same issue. everyday it says the same message wait x hours for reset and it never resets looks like v2 is released [<strict_link> if you migrate to v2 it works fine.” [source](https://www.reddit.com/r/opencodeCLI/comments/1wldeus/free_usage_limit_for_opencode_zen/pb4m4e9/)
  - Complaint, OpenCode, r/opencode, 2026-09-25: “opencode desktop is pretty lame. the interface is awful, and it keeps getting regressions between updates (mainly infinite error pop-ups and deleted threads). i tried using it for a couple of months, even the beta. not worth the trouble. openchamber doesn't support opencode v2 yet (it's in preview), though you can keep using v1 just fine until v2 support lands. since openchamber is just a gui for opencode, it will use your current agent setup and all your subscriptions as-is. there's no need to change things, although it makes it easy to configure agents and advanced stuff.” [source](https://www.reddit.com/r/opencode/comments/1wpt89d/opencode_desktop_vs_opencode_v2_desktop_vs/pbyhp5a/)
  - Praise, OpenCode, @opencode, 2026-08-31: “from pi back to @opencode v2 because tabs and ctrl+b to background the task im a simple man” [source](https://twitter.com/851365565201514498/status/2094496153400394130)

### Fine print

- Most posts come from OpenCode's own subreddits and X handle, which skews toward engaged users.
- Much feedback judges third-party models served through OpenCode, not the harness, so model and tool quality blur together.
- Trustpilot contributes only five posts, so independent review-site sentiment is barely represented.

## Top requests

What users ask to add or change, most asked first. 1385 author-weeks ask for something. Requests do not change the Feedback Score. Rule: A separate pass by Claude Sonnet 5 reads every counted post and extracts what the author asks the agent or its vendor to add or change, with the criteria it maps to and a short normalised wording; it does not touch the labels or the Feedback Score. Claude Opus 5.5 groups the wordings within each criterion (the first criterion the request maps to) into themes; code counts them. A theme counts distinct author-weeks that ask for it, per agent; across agents, one author-week per agent. Themes asked in fewer than 2 author-weeks, and requests that share no theme, are not shown. Examples: up to 3 posts per theme from different authors, without slurs, preferring posts of 60 to 450 characters, most recent first.

| Rank | Request | Criterion | Author-weeks | Posts |
|---|---|---|---|---|
| 1 | Free access to specific or new models | [Free tier and free model availability and limits](https://feedbackbench.com/criteria/billing.free_tier.md) | 32 | 33 |
| 2 | Make temporary usage bonuses permanent | [Price, allowance or plan terms changed](https://feedbackbench.com/criteria/limits.allowance_change.md) | 21 | 24 |
| 3 | Add DeepSeek V4.1 Flash model | [Which models are offered on a plan and when](https://feedbackbench.com/criteria/models.catalog_access.md) | 21 | 21 |
| 4 | Higher overall usage limits | [How much use a plan's price buys](https://feedbackbench.com/criteria/limits.plan_value.md) | 21 | 21 |
| 5 | Max reasoning effort level option | [Reasoning effort setting and its defaults](https://feedbackbench.com/criteria/models.effort_control.md) | 16 | 18 |
| 6 | Allow subscription use in third-party harnesses | [Using an existing subscription across tools](https://feedbackbench.com/criteria/billing.subscription_portability.md) | 16 | 16 |
| 7 | Higher-priced tier above current top plan | [How much use a plan's price buys](https://feedbackbench.com/criteria/limits.plan_value.md) | 15 | 15 |
| 8 | Option to restore previous UI design | [How the interface shows work, and what the user can configure](https://feedbackbench.com/criteria/ui.display_settings.md) | 14 | 15 |
| 9 | Cheaper low-cost plan tier | [How much use a plan's price buys](https://feedbackbench.com/criteria/limits.plan_value.md) | 12 | 12 |
| 10 | Stop spurious 429 rate limit errors | [Outages, server errors and capacity or rate errors](https://feedbackbench.com/criteria/rel.service_errors.md) | 11 | 12 |
| 11 | Faster model response speed | [Latency, throughput and fast mode](https://feedbackbench.com/criteria/rel.response_speed.md) | 11 | 11 |
| 12 | Availability in more countries and regions | [Account bans and access restrictions](https://feedbackbench.com/criteria/account.bans_restrictions.md) | 10 | 10 |

### 1. Free access to specific or new models

- OpenCode, 2026-09-25, @opencode (X): “that was fast. i paid for the $10/month @opencode go sub &amp; tried qwen 3.8 max, burnt up my 5 hr limit &amp; 40% of weekly usage in just 30 min. i guess it's going to just be ds flash from now on, unless we get a really awesome free model for a week on opencode zen <strict_link>” [source](https://twitter.com/43463961/status/2103607209263296673)
- OpenCode, 2026-09-25, @opencode (X): “@abu_khadeejah11 @opencode @grok @grok can i get deepseek v4.1 flash in the free version” [source](https://twitter.com/2066096300496683008/status/2103538792636506464)
- OpenCode, 2026-09-21, @opencode (X): “@artificialanlys @xiaomi @opencode any chance to add it to the free ?” [source](https://twitter.com/2058981690714791937/status/2102136883484516730)

### 2. Make temporary usage bonuses permanent

- OpenCode, 2026-09-24, @opencode (X): “@opencode 4.1 flash is my fav model. smart &amp; i don't have to think about my limits at all. luna 6 is a dumbass in comparison, and i'd rather just spend more time with deepseek iterating than blow my limit on sol or astra. please keep the 4x going. bless ya'll 🍻” [source](https://twitter.com/2184495505/status/2103137089890177279)
- OpenCode, 2026-09-20, @opencode (X): “please @opencode dont end the deepseek v4.1 flash's 4x usage 🥹” [source](https://twitter.com/1519659986267230209/status/2101706833299906757)
- OpenCode, 2026-09-19, @opencode (X): “@sankitdev @opencode @thdxr please make deepseek extra useage permanent 🙏 you'll have a permanent subscriber of opencode go” [source](https://twitter.com/740077527578746881/status/2101351781855056176)

### 3. Add DeepSeek V4.1 Flash model

- OpenCode, 2026-09-19, @opencode (X): “so will we see operation cheepseek for ds v4.1 flash in @opencode go? 👀” [source](https://twitter.com/1694972611154042880/status/2101279168336207893)
- OpenCode, 2026-09-17, @opencode (X): “@opencode when will zen start supporting deepseek v4.1 flash?” [source](https://twitter.com/1668462243888132098/status/2100533363031658784)
- OpenCode, 2026-09-14, @opencode (X): “anyone know if @opencode plans on bringing deepseek v4.1 flash to zen?” [source](https://twitter.com/2096095847528239104/status/2099467002867753125)

### 4. Higher overall usage limits

- OpenCode, 2026-09-24, r/opencode (Reddit): “you can have dozens of even hundreds of agents running in parallel to work on just one single project. a typical quota is insufficient for that kind of workflow.” [source](https://www.reddit.com/r/opencode/comments/1wozjjf/with_deepseek_v41_flash_it_feels_impossible_to/pbsse8g/)
- OpenCode, 2026-09-19, r/opencode (Reddit): “monthly quotas are...not a scam, but a massively limiting quota. weekly and (5 hour) i guess are something. what monthly means in practice is unless you only use the cheapest models, your monthly quota is going to be up after a week.” [source](https://www.reddit.com/r/opencode/comments/1wjzbyz/ds_v41_flash_discount_will_go_tomorrow_which/papo20c/)
- OpenCode, 2026-09-16, @opencode (X): “@ibuildthecloud @opencode easy to blow through limits. but a good way to try other models!” [source](https://twitter.com/15955121/status/2100373488582181166)

### 5. Max reasoning effort level option

- OpenCode, 2026-09-24, r/opencodeCLI (Reddit): “the max reasoning level for muse spark 1.3 in the standard opencode zen package got removed recently. i am not able anymore to select max as reasoning level for muse spark 1.3. it was there before and i used it but now it does not exist anymore in opencode. without the max reasoning its not that good for low level engineering computer programming and tasks in the programming languages c, c++, assembly, verilog.” [source](https://www.reddit.com/r/opencodeCLI/comments/1wopbtv/what_muse_spark_14_contributor_is_already_here/pbpmdr6/)
- OpenCode, 2026-09-12, @opencode (X): “hey @opencode can you please include max reasoning for muse 1.3 in opencode go plan” [source](https://twitter.com/1151345682/status/2098860832247730310)
- OpenCode, 2026-09-09, @opencode (X): “@opencode when are you guys making max reasoning available for muse spark 1.3?” [source](https://twitter.com/1120338331676676096/status/2097826368381678037)

### 6. Allow subscription use in third-party harnesses

- OpenCode, 2026-09-27, r/opencode (Reddit): “yea i love it, wish they added support for other harnesses tbh” [source](https://www.reddit.com/r/opencode/comments/1wrs0xl/openchamber_is_so_much_better_than_opencode/pcfpqnx/)
- OpenCode, 2026-09-27, r/google_antigravity (Reddit): “google needs to remove the restriction of using google ai subscription only inside agy otherwise you get banned. gemini 3.8 is good but agy is kinda shitty would be cool to use the model in something like opencode” [source](https://www.reddit.com/r/google_antigravity/comments/1wrh0mx/gemini_38_in_alternative_harnesses/pcd2rek/)
- OpenCode, 2026-09-27, @opencode (X): “@ammaar @petergyang @antigravity @_mohansolo you guys mind letting us use your models in other harnesses like @opencode? really wanted to try 3.8 flash their but there was no easy way to connect it” [source](https://twitter.com/1199733882351828992/status/2104063066275254461)

### 7. Higher-priced tier above current top plan

- OpenCode, 2026-09-25, @opencode (X): “@kimmonismus @opencode for the love of god add max this time” [source](https://twitter.com/1775308018722197504/status/2103501720072417334)
- OpenCode, 2026-09-24, r/opencode (Reddit): “they’re probably just having a laugh. hoping opencode gives us a higher tier than the 40 usd though” [source](https://www.reddit.com/r/opencode/comments/1won5aj/they_are_ragebaiting_us/pbom1xq/)
- OpenCode, 2026-09-23, @opencode (X): “@emanueledpt @opencode we need larger plans or this doesn’t matter.” [source](https://twitter.com/2036968884587180032/status/2102741936750886925)

### 8. Option to restore previous UI design

- OpenCode, 2026-09-26, @opencode (X): “@opencode i miss the old opencode page, it was very intuitive, now i get lost.” [source](https://twitter.com/1896061820299014144/status/2103652991492624633)
- OpenCode, 2026-09-22, @opencode (X): “@w288958 @nginx99 @opencode is there a way to revert back to the old ui ? cause even things like copying your api key is no longer possible on this new thing” [source](https://twitter.com/1322445907925835777/status/2102303687397806524)
- OpenCode, 2026-09-19, @opencode (X): “@opencode bring back old design goddam my whole workflow downed bc of new design” [source](https://twitter.com/1631750080721027094/status/2101299852902613396)

### 9. Cheaper low-cost plan tier

- OpenCode, 2026-09-25, r/opencode (Reddit): “same. or even a $19 plan so that it can still be clearly cheaper than everyone else.” [source](https://www.reddit.com/r/opencode/comments/1won5aj/they_are_ragebaiting_us/pc1mx4y/)
- OpenCode, 2026-09-22, @opencode (X): “@opencode can we get them on the go for $60 credits if not $75. other places are adding it too since they cheaper <strict_link>” [source](https://twitter.com/1293008918/status/2102507975307084148)
- OpenCode, 2026-09-21, @opencode (X): “@fellipesoares @opencode yeah, but it works if you use multiple accounts. i think a $20 plan would be fine too, but i doubt it would actually be a true 2x increase.” [source](https://twitter.com/1142865860039774208/status/2102163019585200256)

### 10. Stop spurious 429 rate limit errors

- OpenCode, 2026-09-21, r/opencode (Reddit): “command code does but only $20 and the connection is extremely flaky 429 errors all the time” [source](https://www.reddit.com/r/opencode/comments/1wmhlzq/why_doesnt_opencode_go_offer_max_on_muse_13_given/)
- OpenCode, 2026-09-18, @opencode (X): “wth is this @opencode error from provider (console go): upstream request failed: [rate_limit_exceeded] rate limit exceeded. please retry after a brief wait. i only used 2% of my 5-hour limit but get rate limits??” [source](https://twitter.com/1250026646058516481/status/2101001690119876845)
- OpenCode, 2026-09-16, @opencode (X): “@chyldm5d @opencode @openrouter rate limiting is becoming part of the union alpha lore. hard to judge the model when the endpoint keeps getting in the way.” [source](https://twitter.com/1397536057315315717/status/2100327908250198126)

### 11. Faster model response speed

- OpenCode, 2026-09-25, @opencode (X): “@opencode thank you, but also improve the speed and performance. when i use deepseek's own api or from @commandcodeai it is much faster” [source](https://twitter.com/1334983500/status/2103542649416282310)
- OpenCode, 2026-09-21, r/opencode (Reddit): “i have taken muse code subscription and im not able to create api keys so that i can use it with other harness. im only able to use it with muse code and the response are very slow it is taking a lot of time to get the work done.” [source](https://www.reddit.com/r/opencode/comments/1vtocki/anyone_else_fine_muse_spark_12_to_be/pb55veu/)
- OpenCode, 2026-09-18, r/opencodeCLI (Reddit): “opencode go "flash", taking into account regular slow response problems 😂” [source](https://www.reddit.com/r/opencodeCLI/comments/1wjxewd/opencode_higher_tier_subscription_is_coming_soon/pamf2ks/)

### 12. Availability in more countries and regions

- OpenCode, 2026-09-07, @opencode (X): “@aiatmeta @prakash_choks @opencode hi, unfortunately i can't try any of meta ai model because those are banned in my country, i request you to kindly allow us to get access, thanks” [source](https://twitter.com/1778350659999576064/status/2096895251395010741)
- OpenCode, 2026-09-04, r/opencode (Reddit): “muse spark 1.2 and 1.3 are region locked... not working in pakistan” [source](https://www.reddit.com/r/opencode/comments/1w5ziqj/meta_muse_spark_13_is_free_on_opencode_zen/p7qkm8v/)
- OpenCode, 2026-09-03, @opencode (X): “@opencode unfortunately muse spark is not available in my country pakistan. why? @opencode” [source](https://twitter.com/2437331952/status/2095597783168565532)

## Facts

| Fact | Value |
|---|---|
| Version | n/a (fast release cadence) |
| Released | #1 on Hacker News: 2026-03-20 (1,099 points, 546 comments) |
| Price | Free, open source, BYOK to 75+ model providers |
| Model | Model-agnostic, BYOK |
| Surface | Terminal (TUI), desktop app |

## Sources

| Channel | Source | Posts |
|---|---|---|
| Reddit | r/opencode | 10411 |
| X | @opencode | 9118 |
| Reddit | r/opencodeCLI | 4794 |
| Reddit | Posts that name it | 1345 |
| Trustpilot | Trustpilot | 5 |

## Better than peers on

How much use a plan's price buys, Price, allowance or plan terms changed, Using an existing subscription across tools, Which models are offered on a plan and when, Subagents, parallel agents and orchestrators, Mobile, remote-control and voice access, Updates break working setups

## Worse than peers on

Short rolling usage window blocks or interrupts work, Reasoning effort setting and its defaults, Output degrades as the context window fills, Can do the user's kind of task, Long unattended runs and goal/loop mode, Wrong charges, failed payments and plan provisioning, Account bans and access restrictions

## All 63 criteria

Criterion love: 0.5 is the category norm. n: rated author-weeks.

### Paying and limits: Better than peers (customer love 0.660, n 2435)

| Criterion | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|
| [How much use a plan's price buys](https://feedbackbench.com/criteria/limits.plan_value.md) | Better than peers | 0.657 | 0.635–0.677 | 1075 | 615 | 460 |
| [Free tier and free model availability and limits](https://feedbackbench.com/criteria/billing.free_tier.md) | Typical | 0.494 | 0.472–0.516 | 571 | 322 | 249 |
| [Single prompt, model or effort level consumes disproportionate quota](https://feedbackbench.com/criteria/limits.burn_rate.md) | Typical | 0.537 | 0.489–0.578 | 371 | 74 | 297 |
| [Price, allowance or plan terms changed](https://feedbackbench.com/criteria/limits.allowance_change.md) | Better than peers | 0.593 | 0.538–0.644 | 217 | 29 | 188 |
| [Pricing and plan terms stated clearly and consistently](https://feedbackbench.com/criteria/billing.pricing_clarity.md) | Typical | 0.486 | 0.421–0.548 | 201 | 12 | 189 |
| [Short rolling usage window blocks or interrupts work](https://feedbackbench.com/criteria/limits.window_interrupts_work.md) | Worse than peers | 0.446 | 0.402–0.491 | 125 | 10 | 115 |
| [Usage meter visibility and accuracy](https://feedbackbench.com/criteria/limits.usage_meter.md) | Typical | 0.471 | 0.420–0.524 | 124 | 9 | 115 |
| [Prompt cache hits, misses and invalidation](https://feedbackbench.com/criteria/limits.prompt_cache.md) | Typical | 0.505 | 0.473–0.535 | 96 | 34 | 62 |
| [Using an existing subscription across tools](https://feedbackbench.com/criteria/billing.subscription_portability.md) | Better than peers | 0.530 | 0.501–0.560 | 92 | 46 | 46 |
| [Quota reset timing and bonus or banked resets](https://feedbackbench.com/criteria/limits.reset_schedule.md) | Typical | 0.473 | 0.444–0.502 | 49 | 5 | 44 |
| [Pay-as-you-go overage, fallback billing and spend caps](https://feedbackbench.com/criteria/billing.overage_charges.md) | Too few posts | 0.552 | 0.500–0.600 | 26 | 6 | 20 |

Most recent posts:

- Praise, 2026-09-27, r/opencodeCLI (Reddit): “it is. i just noticed i failed to completely use up last month's $10 opencode go sub.” [source](https://www.reddit.com/r/opencodeCLI/comments/1wpzekf/operation_cheepseek_phase_2/pcaklzp/)
- Praise, 2026-09-27, r/opencode (Reddit): “if you're on opencode, the free models are honestly the move. i built a little wrapper that exposes them through an openai-compatible endpoint, so you can use them from other clients too instead of being stuck in the cli. [<strict_link>” [source](https://www.reddit.com/r/opencode/comments/1wq8vp9/best_free_model_after_deepseek_leave/pcal0fn/)
- Praise, 2026-09-27, r/opencodeCLI (Reddit): “opencode go is their $10 subscription plan for use inside the opencode cli. your $10 of payment get you what you would get for $60 at full api pricing, so 6x factor. i would be afraid of quantized models running in stupid mode with it. see other comments asking the same thing. by comparison, i think a chatgpt sub gets you roughly 20x multiplier (your $20 subscription lets you spend $400 of api val…” [source](https://www.reddit.com/r/opencodeCLI/comments/1wpzekf/operation_cheepseek_phase_2/pcarwqn/)
- Complaint, 2026-09-27, r/opencode (Reddit): “they started using third party providers when they started dropping from $60 limit to $15 limit a few months ago. everything open-weight served by opencode zen/go should be assumed to be quantized and will have worse cache hits/retention than 1st party api will.” [source](https://www.reddit.com/r/opencode/comments/1wqy8pq/how_true_is_it_that_opencode_go_models_are/pc9ubms/)
- Complaint, 2026-09-27, r/opencode (Reddit): “yeah the insane amount of thinking, is even making it as expensive as deepseek for me, i'll stay with muse.” [source](https://www.reddit.com/r/opencode/comments/1wqfj7j/will_there_always_be_the_contributor_models/pc9wb3j/)
- Complaint, 2026-09-27, r/opencode (Reddit): “muse started tweaking for me at one point where even 20k tokens where costing me about 0.8$ thats when i stopped using it is it good now ?” [source](https://www.reddit.com/r/opencode/comments/1wqfj7j/will_there_always_be_the_contributor_models/pc9xiy0/)

### Setting up and connecting: Typical (customer love 0.517, n 514)

| Criterion | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|
| [Connecting own API keys, local models and custom endpoints](https://feedbackbench.com/criteria/setup.provider_byok_local.md) | Typical | 0.515 | 0.482–0.543 | 200 | 116 | 84 |
| [MCP servers, plugins, skills and hooks](https://feedbackbench.com/criteria/setup.extensions_mcp.md) | Typical | 0.481 | 0.446–0.513 | 138 | 63 | 75 |
| [Install, launch and sign-in](https://feedbackbench.com/criteria/setup.install_signin.md) | Typical | 0.524 | 0.480–0.564 | 119 | 29 | 90 |
| [Onboarding, discoverability and documentation](https://feedbackbench.com/criteria/setup.onboarding_docs.md) | Typical | 0.504 | 0.471–0.539 | 54 | 11 | 43 |
| [IDE and editor integration](https://feedbackbench.com/criteria/setup.ide_integration.md) | Typical | 0.490 | 0.465–0.514 | 30 | 11 | 19 |

Most recent posts:

- Praise, 2026-09-27, r/opencodeCLI (Reddit): “openchamber is simply an ui for opencode. from a ux perspective, it's actually very good. for serious work i prefer the vscode extension anyway.” [source](https://www.reddit.com/r/opencodeCLI/comments/1wrgmfe/whats_the_deal_with_openchamber/pcc9zmx/)
- Praise, 2026-09-27, r/ClaudeCode (Reddit): “i'm using claude pro. at first, planning with opus 5.5 and implementing with sonnet 5 i was not hitting the limits. i tried full opus 5.5 and quickly reached the limit. i have the z.ai coding plan so when it happens i switch to opencode with glm5.3 to continue my workflow. this is possible because i use an ai memory external to the harness.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wrf0ld/is_claude_pro_actually_worth_20_just_for_one/pcguhs1/)
- Praise, 2026-09-26, r/opencodeCLI (Reddit): “since no one in comments seems to answer the questions (classic reddit). 1.- it was developed because v1 was incredibly messy and outdated, batch reading/editing for example was not a thing, image reading needed to be done manually, performance was trash, client/server wasn't a thing and therefore you had to make a whole app for every version of opencode which i suppose for devs was maintenance ni…” [source](https://www.reddit.com/r/opencodeCLI/comments/1wlcs58/questions_i_have_about_v2/pc4d50x/)
- Complaint, 2026-09-27, r/opencode (Reddit): “i've had a lot of plugins break, 💔 if you don't use any plugins then yet it'll be a step up in theory. tbh been leaning on goose a lot as of late.” [source](https://www.reddit.com/r/opencode/comments/1wqppth/opencode_v2/pcad331/)
- Complaint, 2026-09-27, r/opencode (Reddit): “yes, same experience here. i tried to migrate to v2 this friday but all my plugins failed. so, by now is a no-go for me as my workflow depends so much on these plugins.” [source](https://www.reddit.com/r/opencode/comments/1wqppth/opencode_v2/pcag8ri/)
- Complaint, 2026-09-27, r/opencodeCLI (Reddit): “forget vscode + extension, if you want real performance go for wu which is rust (a fork of zed without ai modules) and terminal running opencode in a panel, vscode and extensions give you an overhead of electron and hundreds of megabytes or more than 1 gigabyte of memory versus the 180 or 200 mb of memory that wu consumes [<strict_link> <strict_link>” [source](https://www.reddit.com/r/opencodeCLI/comments/1wrgmfe/whats_the_deal_with_openchamber/pcd1qha/)

### Choosing models: Better than peers (customer love 0.546, n 771)

| Criterion | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|
| [Which models are offered on a plan and when](https://feedbackbench.com/criteria/models.catalog_access.md) | Better than peers | 0.575 | 0.542–0.606 | 415 | 160 | 255 |
| [Quality got worse or better over time](https://feedbackbench.com/criteria/models.quality_drift.md) | Typical | 0.531 | 0.489–0.575 | 213 | 60 | 153 |
| [Automatic model routing and fallback](https://feedbackbench.com/criteria/models.routing_auto.md) | Typical | 0.503 | 0.466–0.537 | 114 | 30 | 84 |
| [Reasoning effort setting and its defaults](https://feedbackbench.com/criteria/models.effort_control.md) | Worse than peers | 0.445 | 0.420–0.470 | 57 | 12 | 45 |

Most recent posts:

- Praise, 2026-09-27, r/opencode (Reddit): “this model is fast af, and probably better than muse 1.3” [source](https://www.reddit.com/r/opencode/comments/1wqqtgx/longcat25preview_is_now_free_on_opencode_for_two/pcah74r/)
- Praise, 2026-09-27, r/opencode (Reddit): “i honestly don't know where people get the idea that opencode's model is quantized. opencode mostly uses proxy rather than host the model by themself and if you think about it, it might be actually cheaper for them. also, the idea that deepseek from opencode is slower, cannot give same quality of work is not true for me, i've used deepseek v4.1 flash provided from opencode and deepseek official ap…” [source](https://www.reddit.com/r/opencode/comments/1wqy8pq/how_true_is_it_that_opencode_go_models_are/pcapfjm/)
- Praise, 2026-09-27, r/opencodeCLI (Reddit): “3.1 feels better than 3.0 :)” [source](https://www.reddit.com/r/opencodeCLI/comments/1wrddp1/minimax_is_actively_testing_m31flashpreview/pcchvqd/)
- Complaint, 2026-09-27, r/opencode (Reddit): “i’ve noticed the same thing with glm 5.3 flash. going through openrouter the model works amazing but on opencode go it’s dumb as fuck and going in circles.” [source](https://www.reddit.com/r/opencode/comments/1wqxyl4/deepseek_v41_performance_is_much_worse_than/pcaf9ya/)
- Complaint, 2026-09-27, r/opencode (Reddit): “so both of us agree that, deepseek v4.1 flash more dumber than api right?” [source](https://www.reddit.com/r/opencode/comments/1wqxyl4/deepseek_v41_performance_is_much_worse_than/pcap2tv/)
- Complaint, 2026-09-27, r/opencodeCLI (Reddit): “opencode go is their $10 subscription plan for use inside the opencode cli. your $10 of payment get you what you would get for $60 at full api pricing, so 6x factor. i would be afraid of quantized models running in stupid mode with it. see other comments asking the same thing. by comparison, i think a chatgpt sub gets you roughly 20x multiplier (your $20 subscription lets you spend $400 of api val…” [source](https://www.reddit.com/r/opencodeCLI/comments/1wpzekf/operation_cheepseek_phase_2/pcarwqn/)

### Instructing and context: Typical (customer love 0.465, n 318)

| Criterion | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|
| [Direct in-prompt instructions and caps are followed](https://feedbackbench.com/criteria/context.instruction_following.md) | Typical | 0.504 | 0.470–0.541 | 76 | 22 | 54 |
| [Output degrades as the context window fills](https://feedbackbench.com/criteria/context.long_context_decay.md) | Worse than peers | 0.459 | 0.427–0.499 | 74 | 6 | 68 |
| [Context compaction keeps what matters, cheaply and quickly](https://feedbackbench.com/criteria/context.compaction.md) | Typical | 0.490 | 0.461–0.522 | 72 | 22 | 50 |
| [Finding the right files in the codebase](https://feedbackbench.com/criteria/context.codebase_retrieval.md) | Typical | 0.486 | 0.464–0.510 | 34 | 12 | 22 |
| [Memory and state carried across sessions](https://feedbackbench.com/criteria/context.session_memory.md) | Typical | 0.497 | 0.473–0.519 | 30 | 13 | 17 |
| [Persistent project rules files are read and obeyed](https://feedbackbench.com/criteria/context.instruction_files.md) | Too few posts | 0.516 | 0.494–0.538 | 27 | 18 | 9 |
| [Images, PDFs and file attachments as input](https://feedbackbench.com/criteria/context.attachments.md) | Too few posts | 0.505 | 0.484–0.524 | 23 | 10 | 13 |
| [Asks the user versus guessing](https://feedbackbench.com/criteria/context.clarifying_questions.md) | Too few posts | 0.506 | 0.488–0.524 | 15 | 7 | 8 |

Most recent posts:

- Praise, 2026-09-27, r/opencode (Reddit): “i hope you have an agents.md set project wise, that would be a great help for you in my experience i'd always keep a model specifically for auditing to make sure everything you want is being implemented how you want it, also go slow tackle one thing at a time having parallel sessions or tasks will eventually get overwhelming.” [source](https://www.reddit.com/r/opencode/comments/1wraeev/need_help_with_big_project_tasks/pcb2kvp/)
- Praise, 2026-09-27, r/opencodeCLI (Reddit): “superhelpful and much nicer than my janky .md version” [source](https://www.reddit.com/r/opencodeCLI/comments/1wqobnf/when_10_cents_isnt_10_cents/pcf6alg/)
- Praise, 2026-09-27, r/opencode (Reddit): “i have been using codex and claude code exclusively since i started using agents. i've been using chatgpt as coordinator between the two, and decided it's time for another agent. this was mainly due to hitting codex weekly limit, within around 3 days (even using terra). chatpgpt recommend kimi and deepseek as first two options. i chose deepseek using opencode harness. it's absolutely wonderful. i…” [source](https://www.reddit.com/r/opencode/comments/1wrj9e6/deepseekopencode_is_great/)
- Complaint, 2026-09-27, r/opencode (Reddit): “not my experience with it. gpt 6 is extremely bad at following instructions and wastes absurd amounts of time testing” [source](https://www.reddit.com/r/opencode/comments/1wq8vp9/best_free_model_after_deepseek_leave/pcbkxpg/)
- Complaint, 2026-09-27, r/opencode (Reddit): “it's far too trashy to be claude. you literally have to convey everything that's common sense for it not to waste time prodding in wrong directions.” [source](https://www.reddit.com/r/opencode/comments/1wrl1kx/big_pickle_space_bunny_is_claude/pcdgy9c/)
- Complaint, 2026-09-27, r/opencodeCLI (Reddit): “openchamber is very good, but when your context size becoming about 400-500k it's getting slow down, after 600-700k significantly slow” [source](https://www.reddit.com/r/opencodeCLI/comments/1wrgmfe/whats_the_deal_with_openchamber/pcei2i1/)

### Doing the work: Typical (customer love 0.500, n 1369)

| Criterion | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|
| [Can do the user's kind of task](https://feedbackbench.com/criteria/work.capability.md) | Worse than peers | 0.467 | 0.442–0.496 | 859 | 502 | 357 |
| [Spins, loops or gets stuck without progress](https://feedbackbench.com/criteria/work.stuck_loops.md) | Typical | 0.501 | 0.426–0.569 | 167 | 9 | 158 |
| [Subagents, parallel agents and orchestrators](https://feedbackbench.com/criteria/work.multi_agent_orchestration.md) | Better than peers | 0.537 | 0.503–0.573 | 167 | 114 | 53 |
| [Long unattended runs and goal/loop mode](https://feedbackbench.com/criteria/work.long_running_autonomy.md) | Worse than peers | 0.457 | 0.421–0.491 | 48 | 28 | 20 |
| [Risky or irreversible actions without confirmation](https://feedbackbench.com/criteria/work.destructive_actions.md) | Typical | 0.493 | 0.464–0.526 | 38 | 6 | 32 |
| [Diagnosing and fixing reported bugs](https://feedbackbench.com/criteria/work.bug_diagnosis.md) | Typical | 0.518 | 0.494–0.539 | 35 | 28 | 7 |
| [Plan-before-edit mode](https://feedbackbench.com/criteria/work.plan_mode.md) | Typical | 0.488 | 0.466–0.511 | 33 | 13 | 20 |
| [Length and clarity of replies, summaries and comments](https://feedbackbench.com/criteria/work.response_verbosity.md) | Typical | 0.503 | 0.476–0.532 | 33 | 8 | 25 |
| [Tool approval prompts and autonomy modes](https://feedbackbench.com/criteria/work.permission_prompts.md) | Typical | 0.495 | 0.470–0.525 | 32 | 7 | 25 |
| [Breaks existing code or reintroduces bugs](https://feedbackbench.com/criteria/work.regressions_introduced.md) | Typical | 0.487 | 0.457–0.524 | 31 | 2 | 29 |
| [Frontend and visual UI output](https://feedbackbench.com/criteria/work.frontend_ui.md) | Typical | 0.508 | 0.486–0.530 | 30 | 17 | 13 |
| [Stops mid-task or answers instead of acting](https://feedbackbench.com/criteria/work.premature_stop.md) | Too few posts | 0.475 | 0.457–0.496 | 25 | 2 | 23 |
| [Does unrequested work or over-engineers](https://feedbackbench.com/criteria/work.scope_overreach.md) | Too few posts | 0.549 | 0.494–0.605 | 20 | 4 | 16 |
| [Safety filters block legitimate coding tasks](https://feedbackbench.com/criteria/work.safety_refusals.md) | Too few posts | 0.520 | 0.485–0.558 | 20 | 4 | 16 |
| [Computer use and browser control](https://feedbackbench.com/criteria/work.computer_browser_use.md) | Too few posts | 0.505 | 0.487–0.523 | 18 | 11 | 7 |
| [Git commits, branches and sync](https://feedbackbench.com/criteria/work.git_workflow.md) | Too few posts | 0.492 | 0.477–0.508 | 13 | 3 | 10 |
| [Caves to or argues with the user's judgement](https://feedbackbench.com/criteria/work.sycophancy_pushback.md) | Too few posts | 0.496 | 0.483–0.514 | 8 | 1 | 7 |
| [Games checks instead of fixing the problem](https://feedbackbench.com/criteria/work.reward_hacking.md) | Too few posts | 0.491 | 0.485–0.497 | 7 | 0 | 7 |

Most recent posts:

- Praise, 2026-09-27, r/opencodeCLI (Reddit): “muse 1.3 xh > minimax 3.1 in my experience” [source](https://www.reddit.com/r/opencodeCLI/comments/1wqqt8f/longcat25preview_is_now_free_on_opencode_for_two/pca1mvi/)
- Praise, 2026-09-27, r/opencode (Reddit): “1.3 is insane. think you have to use all models to know which one to use. some models do not do well on certain projects. or interments.” [source](https://www.reddit.com/r/opencode/comments/1waq3e5/muse_spark_13_free_is_ass/pca29g7/)
- Praise, 2026-09-27, r/opencode (Reddit): “space bunny is finding all the stubs in my code that other models missed. i am pretty happy with it so far.” [source](https://www.reddit.com/r/opencode/comments/1wqi4a8/my_honest_opinion_about_spacebunny/pcazeo1/)
- Complaint, 2026-09-27, r/opencode (Reddit): “luna is not frontier.” [source](https://www.reddit.com/r/opencode/comments/1wqj8p0/currently_which_is_the_best_model_on_opencode_for/pc9wmwf/)
- Complaint, 2026-09-27, r/opencode (Reddit): “yea it's choppy style is horrible, you have to tell it to stop replying with status lines.” [source](https://www.reddit.com/r/opencode/comments/1won61w/real_life_performance_of_muse_spark_13/pc9x8c4/)
- Complaint, 2026-09-27, r/opencode (Reddit): “than he can undetstand opencode's is dumber” [source](https://www.reddit.com/r/opencode/comments/1wqxyl4/deepseek_v41_performance_is_much_worse_than/pcao1zx/)

### Checking and finishing: Typical (customer love 0.508, n 59)

| Criterion | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|
| [Claims work is done or fixed when it is not](https://feedbackbench.com/criteria/verify.false_completion.md) | Too few posts | 0.520 | 0.476–0.573 | 18 | 2 | 16 |
| [Agent-performed code review finds real issues](https://feedbackbench.com/criteria/verify.agent_code_review.md) | Too few posts | 0.507 | 0.486–0.526 | 16 | 13 | 3 |
| [Builds, tests or runs its own changes](https://feedbackbench.com/criteria/verify.self_testing.md) | Too few posts | 0.496 | 0.480–0.512 | 13 | 7 | 6 |
| [Reviewing and approving the agent's changes](https://feedbackbench.com/criteria/verify.change_review_ui.md) | Too few posts | 0.509 | 0.494–0.526 | 12 | 7 | 5 |

Most recent posts:

- Praise, 2026-09-27, @opencode (X): “@thewritingdev @opencode opencode has a gui app too but hermes is just a general agent, it doesn't have any concept of open pr, diff file view, etc. it's jsut not the right tool for the job. it has other bot related features.” [source](https://twitter.com/412133001/status/2104160113313398978)
- Praise, 2026-09-27, @opencode (X): “@iam_chonchol @opencode self-testing before delivery makes the workflow much more reliable.” [source](https://twitter.com/1082992095361609728/status/2104230124396982531)
- Praise, 2026-09-27, @opencode (X): “@iam_chonchol @opencode testing the game before delivery adds real value.” [source](https://twitter.com/1552600100869853184/status/2104230697162694902)
- Complaint, 2026-09-27, r/opencode (Reddit): “not my experience with it. gpt 6 is extremely bad at following instructions and wastes absurd amounts of time testing” [source](https://www.reddit.com/r/opencode/comments/1wq8vp9/best_free_model_after_deepseek_leave/pcbkxpg/)
- Complaint, 2026-09-27, r/opencode (Reddit): “it built the knob + progress feature but **only half of it was ever deployed**: the html takes effect on request (live instantly), but the backend needs a server restart which it never did. half of a two-half deployment is worse than none: it looks shipped but does nothing. it also never logged the gap anywhere.” [source](https://www.reddit.com/r/opencode/comments/1wrj3b5/this_is_big_pickle_in_action_at_the_moment/)
- Complaint, 2026-09-26, @opencode (X): “sometimes i find it hard to navigate between file diffs in @opencode when they’re large and stacked in one long scroll. exploring a persistent file list on the left bar, with one diff at a time on the right panel. thoughts? <strict_link>” [source](https://twitter.com/1075661598960873473/status/2103912455118405672)

### Interface and sessions: Typical (customer love 0.498, n 428)

| Criterion | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|
| [How the interface shows work, and what the user can configure](https://feedbackbench.com/criteria/ui.display_settings.md) | Typical | 0.523 | 0.488–0.556 | 310 | 117 | 193 |
| [Mobile, remote-control and voice access](https://feedbackbench.com/criteria/surfaces.remote_mobile.md) | Better than peers | 0.530 | 0.504–0.557 | 70 | 45 | 25 |
| [Saving, switching, resuming and rewinding sessions](https://feedbackbench.com/criteria/ui.session_history.md) | Typical | 0.497 | 0.466–0.526 | 61 | 17 | 44 |
| [Cloud and remote sandbox execution](https://feedbackbench.com/criteria/surfaces.cloud_sessions.md) | Too few posts | 0.502 | 0.483–0.520 | 15 | 10 | 5 |
| [Stopping and steering a running agent](https://feedbackbench.com/criteria/ui.interrupt_steer.md) | Too few posts | 0.504 | 0.489–0.519 | 10 | 5 | 5 |

Most recent posts:

- Praise, 2026-09-27, r/opencode (Reddit): “you can ask opencode in the chat window, the model shouldf be able to toggle the setting for you , you dont need to do it mannually or look for it, just prompt your model , the new opencode is able to adjust its settings in chat, i has a skill for its own” [source](https://www.reddit.com/r/opencode/comments/1v2103e/how_do_i_switch_between_plan_and_build_in_the_new/pcbgt4k/)
- Praise, 2026-09-27, r/opencode (Reddit): “love that idea, let me see if i can do that, would be amazing to just be coding while on a hike without my phone open at all” [source](https://www.reddit.com/r/opencode/comments/1wr11fu/tui2web_makes_opencode_usable_from_the_web/pcbh46t/)
- Praise, 2026-09-27, r/opencodeCLI (Reddit): “very good, if you host it on a server or leave your computer running you can use it with it’s webui or with the ios/android app without losing access to any features from opencode. it’s helped me immensely and i’d strongly recommend at least trying it. also frequent updates and not buggy on the whole ,” [source](https://www.reddit.com/r/opencodeCLI/comments/1wrgmfe/whats_the_deal_with_openchamber/pccaety/)
- Complaint, 2026-09-27, r/opencode (Reddit): “there is no "show agent" option settings” [source](https://www.reddit.com/r/opencode/comments/1wlanca/opencode_20_how_do_i_switch_between_plan_and/pcbe06q/)
- Complaint, 2026-09-27, r/opencode (Reddit): “could not find a toggle to enable this.” [source](https://www.reddit.com/r/opencode/comments/1wljmvt/this_go_model_requires_global_regions_select/pcbp58g/)
- Complaint, 2026-09-27, r/opencodeCLI (Reddit): “the interface is cool, while it seems still not that easy to start new worktrees (need to use command line each time?). i've been using worktrees with opencode and built [vicoa.ai](<strict_link>) for this workflow. it has native support for worktrees and you can control them also from your phone. open source at [<strict_link>” [source](https://www.reddit.com/r/opencodeCLI/comments/1qzdyu6/git_worktree_tmux_cleanest_way_to_run_multiple/pcc7fo4/)

### Reliability and speed: Better than peers (customer love 0.533, n 1251)

| Criterion | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|
| [Latency, throughput and fast mode](https://feedbackbench.com/criteria/rel.response_speed.md) | Typical | 0.492 | 0.462–0.519 | 532 | 200 | 332 |
| [Outages, server errors and capacity or rate errors](https://feedbackbench.com/criteria/rel.service_errors.md) | Typical | 0.541 | 0.485–0.591 | 394 | 36 | 358 |
| [Client crashes, freezes and failed tool execution](https://feedbackbench.com/criteria/rel.client_failures.md) | Typical | 0.558 | 0.496–0.611 | 327 | 32 | 295 |
| [Updates break working setups](https://feedbackbench.com/criteria/rel.update_breakage.md) | Better than peers | 0.552 | 0.505–0.591 | 110 | 26 | 84 |

Most recent posts:

- Praise, 2026-09-27, r/opencodeCLI (Reddit): “very good, if you host it on a server or leave your computer running you can use it with it’s webui or with the ios/android app without losing access to any features from opencode. it’s helped me immensely and i’d strongly recommend at least trying it. also frequent updates and not buggy on the whole ,” [source](https://www.reddit.com/r/opencodeCLI/comments/1wrgmfe/whats_the_deal_with_openchamber/pccaety/)
- Praise, 2026-09-27, r/opencodeCLI (Reddit): “it's only getting faster..” [source](https://www.reddit.com/r/opencodeCLI/comments/1wrddp1/minimax_is_actively_testing_m31flashpreview/pccp56a/)
- Praise, 2026-09-27, r/opencode (Reddit): “yea, fast is the thing i love most tbh” [source](https://www.reddit.com/r/opencode/comments/1wroplu/xiaomi_mimo_26_flash_vs_glm_53_flash/pcfapui/)
- Complaint, 2026-09-27, r/opencode (Reddit): “yes, same experience here. i tried to migrate to v2 this friday but all my plugins failed. so, by now is a no-go for me as my workflow depends so much on these plugins.” [source](https://www.reddit.com/r/opencode/comments/1wqppth/opencode_v2/pcag8ri/)
- Complaint, 2026-09-27, r/opencode (Reddit): “was looking for some comments on plugins: i switched to v2 and didn't notice that all the plugins failed, the harness i've built wasn't loading, throwing away tokens instead of saving them... once i notice, took me one or two rounds of claude to adjust everything and now all is fine again. just dropping this to saving you from the bitter drink i had ;-)” [source](https://www.reddit.com/r/opencode/comments/1wqppth/opencode_v2/pcb1yj9/)
- Complaint, 2026-09-27, r/opencode (Reddit): “how's the speed? 5.3 flash on go is like a turtle. can't stand it.” [source](https://www.reddit.com/r/opencode/comments/1wr8w3t/cheepseek_phase_2_full_precision/pcbnf96/)

### Account and support: Typical (customer love 0.523, n 428)

| Criterion | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|
| [Data retention, training use and deployment isolation](https://feedbackbench.com/criteria/account.data_privacy.md) | Typical | 0.480 | 0.447–0.512 | 214 | 44 | 170 |
| [Support, refunds and issue handling](https://feedbackbench.com/criteria/account.support.md) | Typical | 0.530 | 0.487–0.571 | 97 | 21 | 76 |
| [Wrong charges, failed payments and plan provisioning](https://feedbackbench.com/criteria/account.billing_errors.md) | Worse than peers | 0.420 | 0.404–0.435 | 74 | 0 | 74 |
| [Account bans and access restrictions](https://feedbackbench.com/criteria/account.bans_restrictions.md) | Worse than peers | 0.422 | 0.407–0.438 | 69 | 0 | 69 |

Most recent posts:

- Praise, 2026-09-27, r/opencodeCLI (Reddit): “is this supposed to be guerrilla marketing? never heard this and there has been multiple “questions” about this that gets answered in couple of minutes” [source](https://www.reddit.com/r/opencodeCLI/comments/1wrgmfe/whats_the_deal_with_openchamber/pccg8rb/)
- Praise, 2026-09-27, r/opencodeCLI (Reddit): “i am using it as an open code alternative since i wanted a better harness close to what i got when i still subscribed to codex. i tried about 3 different open code variants and i liked this the most and am still using it. seems that devs are also quite active as i regularly get new app updates. so far i can recommend at least trying it out.” [source](https://www.reddit.com/r/opencodeCLI/comments/1wrgmfe/whats_the_deal_with_openchamber/pcez2ay/)
- Praise, 2026-09-27, @opencode (X): “@opencode free + 1m context + zdr is the rare combo. most free tiers quietly train on prompts. two weeks is enough to see if it holds up on real agent loops or just chat demos” [source](https://twitter.com/1094558677292351488/status/2104031597620248743)
- Complaint, 2026-09-27, r/opencode (Reddit): “their email is <email_address>. not that it’s helpful- been trying to reach them for the past couple days.” [source](https://www.reddit.com/r/opencode/comments/1wr1sj5/why_isnt_there_any_customer_support_for_opencode/pc9uuz1/)
- Complaint, 2026-09-27, r/opencode (Reddit): “you're in for a rude awakening. i noticed there are a lot of inconsistencies with opencode and honestly i hate them for it. they are not honest. especially with regards to your data. what they mentioned initially was zdr its not really zdr if you have been paying attention. i stopped using opencode the moment i noticed they're just manipulative and shady af. its way better for you to access models…” [source](https://www.reddit.com/r/opencode/comments/1wqox6a/cheepseek_has_its_price_cache_hit_ratio/pcb29kt/)
- Complaint, 2026-09-27, r/opencode (Reddit): “i don't think they've responded to a single email i've sent them” [source](https://www.reddit.com/r/opencode/comments/1wr1sj5/why_isnt_there_any_customer_support_for_opencode/pcbo36j/)
