# Connecting own API keys, local models and custom endpoints (`setup.provider_byok_local`)

Feedback Bench, coding agents, built 2026-10-01, window 2026-08-31 to 2026-09-27. Web page: https://feedbackbench.com/#/criterion/setup.provider_byok_local

Area: [Setting up and connecting](https://feedbackbench.com/criteria/setup.md)

**Definition.** Whether user-supplied providers work: bring-your-own-key, OpenAI-compatible endpoints, OpenRouter, and local servers such as Ollama or LM Studio.

**Boundary.** Not this: see [Using an existing subscription across tools](https://feedbackbench.com/criteria/billing.subscription_portability.md) for reusing a paid subscription. Not this: see rel.tool_call_errors for tool-format failures once connected.

Rated author-weeks, all agents: 819. Complaint share: 46%.

## The brief

Written by Claude Opus 5.5 from 104 labelled posts and the numbers on this page. Interpretation, not measurement: every quote is verbatim and links to its post.

**Open harnesses take any model; closed ones gate keys behind paywalls.**

TL;DR:

- Cline and Pi earn praise for plugging in OpenRouter, DeepSeek or local servers with little friction.
- Cursor trails peers. BYOK covers some features but not all, and users notice the gap.
- Top requests are local model support, custom providers, and BYOK in tools that still lack it.

In plain terms: If you want to point an agent at your own key, OpenRouter or Ollama, open harnesses mostly just work. Elsewhere you hit missing options, features withheld from BYOK, or upstream changes that quietly break a working proxy.

### How it breaks

- **BYOK that still gates on credits** ([Connecting own API keys, local models and custom endpoints](https://feedbackbench.com/criteria/setup.provider_byok_local.md)). Several agents accept your key but withhold features, tiers or desktop apps from it, so users pay twice or fall back to vendor credits.
  Users bring their own key and expect to pay only the model vendor. Posts report the opposite. Cursor's projects feature does not run on BYOK. Factory's desktop app put BYOK behind a paywall. Amp users wait for custom base URLs on a higher tier. Warp users object to credit gating even with their own OpenRouter key. The shared complaint is a broken contract, not a missing feature.
  Evidence:
  - Complaint, Warp, @warpdotdev, 2026-09-18: “@ericmann @vikvang1 @warpdotdev this. if i bring my own keys / openrouter, the harness should not still gate on its own credits. keys in the os credential manager, desktop free, i pay the model vendor. that is the contract i want.” [source](https://twitter.com/2074942490466033664/status/2101000906388935078)
  - Complaint, Cursor, @cursor_ai, 2026-09-14: “the new @cursor_ai projects feature is really good. exhausted about 20% of my $20 plan and boy was it beautiful to watch 17 agents work on my 4 repos in parallel and execute with no conflicts. i wish they made it compatible with byok like the local chats. i really wanna use this with deepseek-flash man. @mntruell @sualehasif996” [source](https://twitter.com/811432362605154304/status/2099333868717576282)
  - Complaint, Factory, @FactoryAI, 2026-09-21: “@tereza_tizkova @droid @factoryai the time i used droid (desktop app) (on windows), byok was behind a paywall which was a bummer. it would be really nice that if this wasn't the case.” [source](https://twitter.com/1854911029870051328/status/2101899154176061769)
  - Complaint, Amp, @AmpCode, 2026-09-07: “when will the megawatt subcription account be enabled for custom base url byok?, i’ve been waiting for it so long @sqs @ampcode 🥲” [source](https://twitter.com/1800425844491620352/status/2096782766776221869)

- **Upstream changes break working proxies** ([Connecting own API keys, local models and custom endpoints](https://feedbackbench.com/criteria/setup.provider_byok_local.md)). Setups that worked last week stop working after a provider-side header, terms or auth change, and proxy users absorb the fallout.
  OpenCode Go added a session header to supposedly standard endpoints. Proxy maintainers report hashing prompts to fake session IDs, and a bug skips the header for providers not prefixed with opencode. Cline and Warp users lost access to those models as a side effect. Codex users report OAuth and switcher setups that worked one day and errored the next. The pattern is short notice and no stable contract.
  Evidence:
  - Complaint, OpenCode, @opencode, 2026-09-03: “@opencode found a bug in opencode, it only adds the "new" x-opencode-session header if the provider begins with "opencode" but i use tailscale aperture (i.e. aperture/opencode-go/model).” [source](https://twitter.com/1260237364594651137/status/2095438361238397244)
  - Complaint, OpenCode, r/opencode, 2026-09-04: “i maintain the [opencode go plugin for cliproxyapi](<strict_link>), and this change causes some real headaches. cpa exposes standard openai and anthropic endpoints without managing conversation state, so there is no actual conversation id. without a clean way to handle this, i ended up hashing the initial user turn and passing it as \`x-opencode-session\` on every opencode go request. it works across chat completions, messages, responses, and streaming, but it is still just a workaround. if two sessions start with the exact same prompt, they collide. proxy maintainers are stuck patching around this because opencode tacked a provider-specific requirement onto supposedly standard endpoints with almost no notice.” [source](https://www.reddit.com/r/opencode/comments/1w6x58o/did_opencode_really_decide_to_break_api_access/p7u15f9/)
  - Complaint, Cline, @cline, 2026-09-12: “oh damn. since opencode go introduced the requirement to add a custom header to requests to its api, i can't use the models it offers with my cli tools like warp. also, @cline, which supports it by default, hasn't fixed it either.” [source](https://twitter.com/19994553/status/2098667465001619864)
  - Complaint, OpenAI Codex, r/codex, 2026-09-16: “i see, but last time i tried with pi, around 3-4 weeks ago i think, it was working. then i tried it again few days ago, and it gave an error. then i tried the same oauth method again, then it gave me that 'cannot be used outside of codex' error. fyi. i will try ompi when limits gets reset.” [source](https://www.reddit.com/r/codex/comments/1whoo9s/this_is_worst_then_i_imagine_about_weekly_resets/pa8ebcc/)

- **No door for outside models** ([Connecting own API keys, local models and custom endpoints](https://feedbackbench.com/criteria/setup.provider_byok_local.md)). Some agents simply do not accept external API models, and users are told to use a different tool instead.
  Antigravity users say external API models are not supported and point each other to Copilot or OpenCode. Devin's own team confirms no BYOK and offers ACP as the workaround. Zed users with expiring credits plead for custom providers. Cursor users call the lack of custom API models real friction. These are the gaps behind the custom provider and BYOK requests.
  Evidence:
  - Complaint, Google Antigravity, r/google_antigravity, 2026-09-14: “you can't use external api models with antigravity. your best bet is to use github copilot in vscode or github copilot app with api key. or use opencode.” [source](https://www.reddit.com/r/google_antigravity/comments/1wfvup0/how_to_configure_custom_openaicompatible_api_base/p9u9oah/)
  - Complaint, Devin, @cognition, 2026-09-03: “@flep @devinai @cognition hey, glad you enjoyed it. yes that is correct, we do not support byok, but devin desktop does support acp, so you can choose not only different models but also different agents. <strict_link>” [source](https://twitter.com/17189394/status/2095584505947881971)
  - Complaint, Zed, @zeddotdev, 2026-09-24: “my vip credits ends today @zeddotdev begging you please add custom provider 😅” [source](https://twitter.com/1261173216455712768/status/2103200607729500470)
  - Complaint, Cursor, @cursor_ai, 2026-09-23: “@wayne_hamadi @cursor_ai credits stuck outside cursor because no custom api models is a real friction, vendors should meet the ide” [source](https://twitter.com/1395133841531101192/status/2102627767405932612)

- **Local models connect, then crawl** ([Connecting own API keys, local models and custom endpoints](https://feedbackbench.com/criteria/setup.provider_byok_local.md)). Hooking up Ollama or LM Studio usually works, but users report slow inference, small context windows and fiddly endpoint forms.
  Codex users running Qwen locally report slow responses and tight context unless they drop to low quantization. Others cannot get the agent to delegate to a local worker. Setup forms add friction. Warp rejects an Ollama URL that works in curl and Zed, and Cline Desktop demands an API key for LM Studio that the VS Code extension does not.
  Evidence:
  - Complaint, OpenAI Codex, r/codex, 2026-09-20: “i find them useful and do use qwen local sometimes for certain tasks but they're slow and limited context windows unless i use a low quant, running on a 4090. maybe i have it configured wrong.” [source](https://www.reddit.com/r/codex/comments/1wl0v9g/switch_to_open_source_models_people_openai_is/pavlcr5/)
  - Complaint, Warp, r/warpdotdev, 2026-09-15: “[unable to add this endpoint. works from 1\)curl 2\)cli 3\)zed editor.](<strict_link>) how do i add this ollama endpoint. the red (error) mark around the url never goes away, and the "add endpoint" always stays grayed. i've tried without the api key..multiple times. note that this same url works for me from zed editor, and from cli, and from openweb ui. so the url is valid. (sorry about typo in the post heading) . no auth required for all those methods that work. my goal here is to use gemma , a local model served by ollama (on port <zip_code>, on another machine on my network)... am i doing it the right way? i didn't find a whole lot of info on local models in warp docs.” [source](https://www.reddit.com/r/warpdotdev/comments/1wgpb25/custom_inferene_endpoint_ollamagemma/)
  - Complaint, Cline, @cline, 2026-09-22: “@cline why do you require an api key for lm studio on cline desktop? should be optional. i don't have a problem with the extension in vscode.” [source](https://twitter.com/2053595020272222208/status/2102536385978855763)
  - Complaint, OpenAI Codex, r/codex, 2026-09-08: “i’m trying to reduce codex token usage by offloading a lot of the “grunt” work to qwen3-coder running locally on my strix halo box. while chatgpt says it will work i’m having a hard time actually getting codex to delegate. any tips on how to make this work? i may just install cline and move to a bigger qwen3 -coder and use chatgpt for architecture/design/reviews/audits. any feedback/advice appreciated.” [source](https://www.reddit.com/r/codex/comments/1watj3s/offloading_coding_workload_to_qwen_but_using/)

### Who stands out

- **Cline (stronger)**. Users plug in DeepSeek, OpenRouter and local open-weight models and report it works on the first try.
  Praise centers on the new native client keeping BYOK and adapting to local models. Users also like the trust model of keeping gateway keys in a local file. The complaints are narrow: missing Bedrock in the desktop app, a required key for LM Studio, and the OpenCode Go header breakage. None question the core BYOK path.
  Evidence:
  - Praise, Cline, @cline, 2026-09-15: “@cline just downloaded and plugged in deepseek. works great. how about a right panel with a browser like codex or zcode? or is that not built in yet? @cline” [source](https://twitter.com/1885730256432324608/status/2099935916219613455)
  - Praise, Cline, @cline, 2026-09-14: “@cline the native client experience is different, and it also supports byok and task migration, which is very comfortable as it can adapt to local open-weight models.” [source](https://twitter.com/45582017/status/2099556764056498640)
  - Praise, Cline, @cline, 2026-09-16: “@cline now i have more options. besides opencore's openrouter, cline now supports it too! awesome!!!” [source](https://twitter.com/2031285494081007620/status/2100266340758208930)
  - Praise, Cline, @cline, 2026-09-19: “@cline the browser capability is a big unlock, but the config boundary matters too. keeping the gateway key in a local file while the agent handles the browsing task is a much cleaner trust model than pasting secrets into prompts.” [source](https://twitter.com/2092135923169579008/status/2101152370801426934)

- **Pi (stronger)**. A minimal harness built for swapping models, including local ones mid-session, earns praise as an exit from vendor lock-in.
  Users describe switching any model mid-session, serving local models through llama.cpp without issues, and running everything locally to avoid cost and privacy worries. The friction is small. Some providers need a community package, and users note that heavy prefill in derivative harnesses hurts local models.
  Evidence:
  - Praise, Pi, r/PiCodingAgent, 2026-09-21: “yes, use openrouter. totally breaks you out of vendor lock.” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wlpj30/lord_forgive_me_for_the_time_i_wasted/pb6e7al/)
  - Praise, Pi, @pidotdev, 2026-09-01: “pi shines as a minimal, open-source harness you fully own and reshape. with just 4 core tools and tiny context overhead, use it to build custom skills/extensions for your exact workflows, switch any model mid-session (including local ones), or run scripts/sdk embeds—complementing the polished but locked-in codex/claude/grok tools. ideal for power users who adapt the agent, not the reverse.” [source](https://twitter.com/1720665183188922368/status/2094838186002313335)
  - Praise, Pi, r/PiCodingAgent, 2026-09-20: “i'm using llama-swap with llama.cpp to serve my local models and i did not have any issue” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wlebgj/latest_pi_version_86_broke_working_with_local/paxxtr5/)
  - Complaint, Pi, r/PiCodingAgent, 2026-09-23: “for external agents omp is fine, for local models like qwen, it dumps a ton of prefill in that actually makes it worse not better. the reason pi works so well it's that you tailor it to your needs and refine it instead of dumping the kitchen sink into you llm as prefill and hoping for the best.” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wnvwqa/are_pi_and_ohmypi_are_same/pbkgkvu/)

- **Cursor (weaker)**. Cursor supports BYOK in chats but not across the product, and users keep comparing it to harnesses that do.
  Posts praise running the agent loop on your own machine and pairing Cursor with OpenRouter. The complaints focus on scope. Projects and other features skip BYOK, custom API models are missing, and users say every other harness supports it. Some offer to pay a flat fee just for bring-your-own-subscription. Running agents locally is its top request here.
  Evidence:
  - Complaint, Cursor, @cursor_ai, 2026-09-26: “@fatih @cursor_ai too expensive compare to claude/codex subscriptions. cant byok for all functionality.” [source](https://twitter.com/2007052689306501120/status/2103871557147697643)
  - Complaint, Cursor, @cursor_ai, 2026-09-14: “the new @cursor_ai projects feature is really good. exhausted about 20% of my $20 plan and boy was it beautiful to watch 17 agents work on my 4 repos in parallel and execute with no conflicts. i wish they made it compatible with byok like the local chats. i really wanna use this with deepseek-flash man. @mntruell @sualehasif996” [source](https://twitter.com/811432362605154304/status/2099333868717576282)
  - Complaint, Cursor, @cursor_ai, 2026-09-11: “@zhanzhao391439 @canbalkya true, was thinking of creating an open-source version of cursor with byok/byos yet never got to it. i'd even pay cursor like €50 per month for byos... @cursor_ai” [source](https://twitter.com/1468874990745489412/status/2098239002851574175)
  - Complaint, Cursor, @cursor_ai, 2026-09-23: “@wayne_hamadi @cursor_ai +1, every other harness supports it <strict_link>” [source](https://twitter.com/1090463126909304836/status/2102625166434136355)

- **OpenCode (mixed)**. The broadest provider support in the category, undercut by a header change that broke proxies and custom providers that cannot be edited.
  Users stack many providers with rate-limit fallbacks and run local Qwen for real work without worrying about token cost. Other harnesses even route through it. The pain comes from OpenCode's own changes. The session header broke proxy setups, and adding a model to an existing custom provider means rebuilding the configuration from scratch.
  Evidence:
  - Praise, OpenCode, r/opencode, 2026-09-17: “you can bring in most providers and even work around some tos limited ones if you are adventurous enough. then you can stack the models with rate limiting fallbacks, get to use all your providers/models under a single roof.” [source](https://www.reddit.com/r/opencode/comments/1wj13ox/what_is_the_reason_you_use_opencode_instead_of/pagptzh/)
  - Praise, OpenCode, r/codex, 2026-09-26: “what gets lost in a lot of these types of conversations is: what kind of coding are you doing? if you're working on a state-of-the-art 3d game, stay with the frontier models. if you're writing a custom inventory management web app for a company, a local model will do fine. using qwen 27b on a dgx spark with opencode it does a damn fine job of updating and porting existing code and with some iterating, even spin up new apps. never really did mac or ios programming but created a few small apps for myself that didn't take long at all. might not be as fast but i never worry about token cost.” [source](https://www.reddit.com/r/codex/comments/1wqje3c/if_600_becomes_the_new_200_this_is_not_a_consumer/pc5aqla/)
  - Complaint, OpenCode, r/opencode, 2026-09-04: “i maintain the [opencode go plugin for cliproxyapi](<strict_link>), and this change causes some real headaches. cpa exposes standard openai and anthropic endpoints without managing conversation state, so there is no actual conversation id. without a clean way to handle this, i ended up hashing the initial user turn and passing it as \`x-opencode-session\` on every opencode go request. it works across chat completions, messages, responses, and streaming, but it is still just a workaround. if two sessions start with the exact same prompt, they collide. proxy maintainers are stuck patching around this because opencode tacked a provider-specific requirement onto supposedly standard endpoints with almost no notice.” [source](https://www.reddit.com/r/opencode/comments/1w6x58o/did_opencode_really_decide_to_break_api_access/p7u15f9/)
  - Complaint, OpenCode, r/opencode, 2026-09-10: “you can add a custom provider and then add models to it. weeeeee but once the provider is loaded, you can't simply add another model to that existing provider. if a new model becomes available from a provider you've already configured, you have to go through the entire process again from the beginning just to add that model. that doesn't make sense. there should be a way to edit an existing custom provider and add/remove models without having to recreate the whole provider configuration. please fix this. custom providers should be persistent and editable after setup. every update i've seen pass i hoped this logic would have been added... let's go opencode” [source](https://www.reddit.com/r/opencode/comments/1wcbs5x/why_do_i_have_to_recreate_the_provider_to_add_a/)

- **Google Antigravity (mixed)**. Offline Gemma through the SDK wins applause, while the IDE itself refuses external API models.
  Users cheer fully offline Gemma execution with no API fees and report strong results on modest VRAM. The other half of the story is closed. Users say external API models cannot be used, and examples still ask for a Gemini key up front. Requests here span BYOK, custom providers and OpenRouter.
  Evidence:
  - Praise, Google Antigravity, @antigravity, 2026-09-24: “@antigravity fully offline gemma 4 with zero api cost is a great local-first setup—privacy and predictable latency are hard to beat.” [source](https://twitter.com/1970761241338515457/status/2103088265435988279)
  - Praise, Google Antigravity, @antigravity, 2026-09-25: “@enric699 @googledevs @antigravity @googlegemma it is gemma 4 from google (open-weight, here the 26b moe). great for reasoning, coding, and local agentic workflows with total privacy and zero api costs. not comparable to opus 5.5 or astra (closed frontiers, much more powerful). cost: free on-device (necessary hardware).” [source](https://twitter.com/1720665183188922368/status/2103391841685037324)
  - Complaint, Google Antigravity, r/google_antigravity, 2026-09-14: “you can't use external api models with antigravity. your best bet is to use github copilot in vscode or github copilot app with api key. or use opencode.” [source](https://www.reddit.com/r/google_antigravity/comments/1wfvup0/how_to_configure_custom_openaicompatible_api_base/p9u9oah/)
  - Praise, Google Antigravity, r/google_antigravity, 2026-09-08: “agree with you on sentiment, i disagree with you about anti-gravity as a harness. it's actually really capable. but yah now i’m open router all the way - it’s amazing” [source](https://www.reddit.com/r/google_antigravity/comments/1wa6rcj/im_agy_clis_biggest_fanboy_but_i_have_to_quit_it/p8ha6mv/)

### Fine print

- Devin, Zed, Warp, Factory and Copilot have too few posts to rank; their quotes illustrate patterns rather than prove them.
- Many local-model complaints reflect user hardware limits as much as agent behavior.
- Reusing paid subscriptions in other harnesses is covered on a separate page.

## Top requests

What users ask to add or change, most asked first. 385 author-weeks ask for something. Requests do not change the Feedback Score. Rule: A separate pass by Claude Sonnet 5 reads every counted post and extracts what the author asks the agent or its vendor to add or change, with the criteria it maps to and a short normalised wording; it does not touch the labels or the Feedback Score. Claude Opus 5.5 groups the wordings within each criterion (the first criterion the request maps to) into themes; code counts them. A theme counts distinct author-weeks that ask for it, per agent; across agents, one author-week per agent. Themes asked in fewer than 2 author-weeks, and requests that share no theme, are not shown. Examples: up to 3 posts per theme from different authors, without slurs, preferring posts of 60 to 450 characters, most recent first.

| Rank | Request | Author-weeks | Posts | Agents (author-weeks) |
|---|---|---|---|---|
| 1 | Local model support | 37 | 37 | OpenAI Codex 13, Claude Code 5, Google Antigravity 4, Cursor 4, Pi 4, OpenCode 3, Amp 1, GitHub Copilot 1, Warp 1, Zed 1 |
| 2 | Custom model provider support | 21 | 22 | Zed 8, Google Antigravity 4, Cursor 3, OpenCode 3, OpenAI Codex 2, Amp 1 |
| 3 | Bring-your-own-key support | 19 | 21 | Devin 6, Google Antigravity 5, Cursor 3, Zed 2, Amp 1, GitHub Copilot 1, Pi 1 |
| 4 | Support more model providers | 16 | 16 | Claude Code 4, OpenAI Codex 4, Google Antigravity 2, Cursor 2, Zed 2, OpenCode 1, Pi 1 |
| 5 | Run agents locally instead of cloud | 14 | 14 | Cursor 7, Claude Code 4, Conductor 1, Devin 1, Factory 1 |
| 6 | Use subscriptions with third-party harnesses | 14 | 14 | Google Antigravity 5, OpenCode 4, Amp 2, OpenAI Codex 1, Factory 1, Zed 1 |
| 7 | OpenCode provider integration | 13 | 13 | OpenCode 4, Google Antigravity 2, OpenAI Codex 2, Amp 1, Claude Code 1, Cursor 1, Pi 1, Zed 1 |
| 8 | OpenRouter support | 12 | 12 | Google Antigravity 3, Amp 2, Cline 2, OpenAI Codex 2, Cursor 1, Devin 1, Zed 1 |
| 9 | Secure credential entry and storage | 10 | 13 | Amp 2, Claude Code 2, OpenCode 2, Google Antigravity 1, OpenAI Codex 1, Cursor 1, Pi 1 |
| 10 | Capable local models for consumer hardware | 10 | 10 | OpenAI Codex 4, Claude Code 2, OpenCode 2, Amp 1, Devin 1 |
| 11 | Custom base URL OpenAI-compatible endpoints | 10 | 10 | Amp 4, Google Antigravity 2, Zed 2, Cursor 1, OpenCode 1 |
| 12 | Public API access to model and agent | 9 | 10 | Google Antigravity 2, Devin 2, Pi 2, Claude Code 1, Cursor 1, OpenCode 1 |

### 1. Local model support

- OpenAI Codex, 2026-09-27, r/LocalLLaMA (Reddit): “its almost impossible to configure codex to run with a local model. dozens of tutorials, none work. they force you to use "ollama" or "lmstudio", not even the vllm tutorial hosted in the vllm site works anymore. how hard can it by? just specify the api endpoint and model. but no, you must use one of their friends software.” [source](https://www.reddit.com/r/LocalLLaMA/comments/1wrfp50/another_harness_matters_post_codex_cli_pi_and/pcf5l2g/)
- Google Antigravity, 2026-09-26, @antigravity (X): “@rodydavis @antigravity i hope we can see this flexibility within the antigravity app, so we can use the google models as cordinators and seniors and then our own local models or openai compatible api´s for higher volumes of development without being restricted with native antigravity models limits.” [source](https://twitter.com/1413849052123369472/status/2103763456247709807)
- Claude Code, 2026-09-26, @ClaudeDevs (X): “@apocalyzabeth @claudedevs yea i had to switch to codex to continue my madness hahaha i really need a local model 🤣” [source](https://twitter.com/1695535328831111168/status/2103655847964668339)

### 2. Custom model provider support

- Google Antigravity, 2026-09-25, @antigravity (X): “@antigravity are we able to use our own models within antigravity already ?” [source](https://twitter.com/1413849052123369472/status/2103613976987029584)
- Zed, 2026-09-25, @zeddotdev (X): “@zeddotdev delta is goood like really good just need custom provider” [source](https://twitter.com/1261173216455712768/status/2103520442594332993)
- Zed, 2026-09-25, @zeddotdev (X): “@zeddotdev can you guy make delta support custom llm providers please” [source](https://twitter.com/1925398929904214019/status/2103312711803420977)

### 3. Bring-your-own-key support

- Cursor, 2026-09-26, @cursor_ai (X): “@mdamore9 @grok @bot @cursor_ai rate limits are also much more efficient the past 1-2 weeks this is critical if we can't bring our own models/api keys” [source](https://twitter.com/1756394384655003648/status/2103698095045489000)
- Google Antigravity, 2026-09-25, r/google_antigravity (Reddit): “just want byok to be available in agy desktop to as ds and that the 3.8 flash issues get solved especially token slash and speed” [source](https://www.reddit.com/r/google_antigravity/comments/1wom9gb/local_model_support_now_available/pby356u/)
- Google Antigravity, 2026-09-21, r/google_antigravity (Reddit): “i like the model. i wish they would give me my api key and keep their ui. acp works but it's a pain in the ass.” [source](https://www.reddit.com/r/google_antigravity/comments/1wmad7c/they_are_ruining_antigravity_20/pb6blot/)

### 4. Support more model providers

- OpenCode, 2026-09-27, r/opencode (Reddit): “this tools help in making of cost and token predictable, so we can plan accordingly. can't it work for other llm?” [source](https://www.reddit.com/r/opencode/comments/1wrfhdx/i_built_a_telemetry_sidebar_for_the_opencode/pccm0hu/)
- Zed, 2026-09-23, @zeddotdev (X): “@gmi_cloud @cline please talk with @zeddotdev lately i have using the delta and it's really really good but sadly very selected few providers” [source](https://twitter.com/1261173216455712768/status/2102607334476558677)
- Claude Code, 2026-09-19, r/ClaudeCode (Reddit): “check out opencode, seriously we should advance not regress. we should move to being provider agnostic, sooner than later! we should have the power, to choose whoever as easy as a simple instruction.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wjl7t8/the_limits_are_disappearing_at_a_crazy_speed/patyhpi/)

### 5. Run agents locally instead of cloud

- Cursor, 2026-09-25, r/cursor (Reddit): “i liked it, but not worth it for smaller tasks tbh... and it just bothers me i have to run it 100% of time on cloud agents, would be really nice to work locally” [source](https://www.reddit.com/r/cursor/comments/1wpefwt/thoughts_on_cursor_projects/pbynvxz/)
- Claude Code, 2026-09-25, @ClaudeDevs (X): “i really love claude project. but i wonder, why does it have to be cloud? no plan for local project @claudedevs @bcherny ?” [source](https://twitter.com/457307083/status/2103373984041738453)
- Factory, 2026-09-18, @FactoryAI (X): “@tereza_tizkova @factoryai local desktop version of <strict_link> really want to try factory, will bring all my team to use it if it’s worth it” [source](https://twitter.com/1747424923914514432/status/2100953473634254851)

### 6. Use subscriptions with third-party harnesses

- Google Antigravity, 2026-09-27, @antigravity (X): “@antigravity you just destroying a good harness day by day your windsuf fork was much better than at current. if you can't do anything better just fork opencode/deepseek/zcode or let your subscribers use those instead” [source](https://twitter.com/151309638/status/2104017857126531200)
- Amp, 2026-09-25, @AmpCode (X): “@sqs @sixhobbits @ampcode i have a feeling anthropic will allow external harnesses in a few months, once they have stabilised their cloud projects. then we can continue orbin' .. don't really want to keep switching between different cloud vm providers for every sub that i have.” [source](https://twitter.com/437728663/status/2103389934086479926)
- Amp, 2026-09-23, @AmpCode (X): “model routing is one of my favorite @ampcode features. i just hope that anthropic and google come into their senses and allow their respective models to be used over oauth.” [source](https://twitter.com/2090734054928687104/status/2102696668944912715)

### 7. OpenCode provider integration

- Zed, 2026-09-26, @zeddotdev (X): “@zeddotdev this will be hard to test until claude code plans can be used. hope you guys figure out a way. opencode should be easy to implement too.” [source](https://twitter.com/2065156316562141184/status/2103808310642131452)
- OpenCode, 2026-09-20, r/vibecoding (Reddit): “it’s hardcoded to use direct deepseek api, make it work with opencode-go and commandcode and i’m in!” [source](https://www.reddit.com/r/vibecoding/comments/1wl3lfi/i_built_an_orchestration_package_that_lowered_my/payk2aa/)
- Cursor, 2026-09-16, r/cursor (Reddit): “can someone help me do that? i love cursor harness but i want to use opencode models, atm i'm using opencode as harness in and on itself but i prefer cursor i tried giving cursor the api, the endpoint and model name, but there's an incompatibility issue. i've heard it can be solved by building ( or gitcloning, if sm1 already did that ) a proxy from the computer, is anyone able to help? thank you!” [source](https://www.reddit.com/r/cursor/comments/1wi5c5n/opencode_api_through_cursor/)

### 8. OpenRouter support

- Google Antigravity, 2026-09-20, r/GoogleAntigravityIDE (Reddit): “i want to use openrouter api inside antigravity ide . i tried using cline but its integration seems be not so smooth as of vscode with antigravity ide . it keep on getting hung. any suggestions” [source](https://www.reddit.com/r/GoogleAntigravityIDE/comments/1wleeqh/how_to_use_openrouter_api_inside_ide_cline_is/)
- Cline, 2026-09-19, @cline (X): “@musingiqbal @bovetheline @cline @opencode i love cline and i still use it because of this visibility, but i cannot be productive with it. it makes models less smart also it does not even support multimodal features that openrouter exposes. currently i use claude desktop + herms.” [source](https://twitter.com/2051403352869638144/status/2101371670971703598)
- OpenAI Codex, 2026-09-15, r/codex (Reddit): “oh neat. add in openrouter provider selcetor into the pi and you are pretty much golden. deepseek orestrator with glm sub-agents has been by far performing insanely good. deepseek seems to be more creatives with things and easier to stear while glm usually remain factual and catches the error that slip pass” [source](https://www.reddit.com/r/codex/comments/1wgkvku/chat_gpt_6_pro_as_planorchestator_anitigravity/p9v3gzt/)

### 9. Secure credential entry and storage

- Amp, 2026-09-26, @AmpCode (X): “.@sqs @ampcode another fun feature request; per-session secrets. sometimes i want to give an orb access to an external service, and i want to only give it access for one thread/session. extra cool if sub-threads could inherit said secrets:)” [source](https://twitter.com/1729720291/status/2103980940653715675)
- OpenCode, 2026-09-26, @opencode (X): “@opencode let me delete api keys in console so i don't have to see them” [source](https://twitter.com/1866208836203286528/status/2103819626425573795)
- Claude Code, 2026-09-24, @ClaudeDevs (X): “@claudedevs i’m not a fan of these new security features with claude. what do you mean i can’t just add a key through chat now? kinda annoying tbh.” [source](https://twitter.com/1200639158437343234/status/2103174242519158894)

### 10. Capable local models for consumer hardware

- Claude Code, 2026-09-27, r/ClaudeCode (Reddit): “i agree, i'm a software engineer and all i want is cheap models i can run locally - the technology is already good enough to change my entire career.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wrpq38/any_idea_what_anthropic_figured_out/pcgafb6/)
- OpenCode, 2026-09-26, @opencode (X): “@jmorgan @ollama @opencode thank you! i want @ollama to succeed 💚. you should become the trust owner, as you are coming from local inference. please only host a few top ow models yourself on own/rented gpus. 🙏 opencode go sold it's soul to closed frontier and china hosters, it seems 😭.” [source](https://twitter.com/26595741/status/2103868429719457933)
- OpenCode, 2026-09-21, r/opencodeCLI (Reddit): “off on a side note here, any chance you’ve started looking into hosting the sparse moe models locally when they exceed available ram size? this is becoming a much more interesting option. i have 64gb ram with 16gb vram, but i managed unsloth’s qwen 3.8 flash next (ud-q4\_k\_xl; 111gb) at about 12 tokens per second. only worked directly through llama, or pi. it wouldn’t respond through opencode.” [source](https://www.reddit.com/r/opencodeCLI/comments/1wlk3yh/comparison_opencode_pi_and_codex/pb7kzf5/)

### 11. Custom base URL OpenAI-compatible endpoints

- Zed, 2026-09-18, r/ZedEditor (Reddit): “i keep reading you can use it on the web, but can’t find any links to access it, how do you do so? a few requests: \- a custom url openai provider \- base the jj implementation on worktrees instead of copies please!” [source](https://www.reddit.com/r/ZedEditor/comments/1whydrw/delta_in_public_beta/panpthq/)
- Zed, 2026-09-17, r/ZedEditor (Reddit): “where's everybody getting that 30$ fee? currently it works like zed, you can use your sub/api keys, or use zed pro plan. most zed providers are already available like openrouter, opencode, etc. i'm missing only the ability to set up a custom open ai provider (like z.ai) but on twitter they tell me is comming soon.” [source](https://www.reddit.com/r/ZedEditor/comments/1whydrw/delta_in_public_beta/pafuh1c/)
- Google Antigravity, 2026-09-16, @antigravity (X): “@yashjitpal @antigravity can u add byok for openai compatible endpoint api &amp; support running from google colab notebook pls?” [source](https://twitter.com/1585841318810394625/status/2100189921088876567)

### 12. Public API access to model and agent

- Claude Code, 2026-09-27, @ClaudeDevs (X): “@claudedevs is there an api available for this so we can integrate it our own systems?” [source](https://twitter.com/1866085055690657792/status/2104090456720380172)
- Pi, 2026-09-23, @pidotdev (X): “@rolandgvc @pidotdev @badlogicgames any chance you guys will support an api that lets anyone use the infra setup?” [source](https://twitter.com/1689423238173007873/status/2102815088323268928)
- OpenCode, 2026-09-20, r/opencode (Reddit): “can't generate api keys from muse ai tho, i'd rather keep giving mah maney to opencode go or any other provider for the contributor one and keep getting my personal info zucced” [source](https://www.reddit.com/r/opencode/comments/1wl8m6y/free_1_billion_muse_spark_13_tokens/paxf0lz/)

## Every agent

| Agent | Overall rank | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|---|
| [Cline](https://feedbackbench.com/agents/cline.md) | =8 | Better than peers | 0.556 | 0.532–0.580 | 48 | 39 | 9 |
| [Pi](https://feedbackbench.com/agents/pi.md) | 7 | Better than peers | 0.532 | 0.505–0.557 | 60 | 41 | 19 |
| [OpenCode](https://feedbackbench.com/agents/opencode.md) | 3 | Typical | 0.515 | 0.482–0.543 | 200 | 116 | 84 |
| [Amp](https://feedbackbench.com/agents/amp.md) | =11 | Typical | 0.510 | 0.483–0.535 | 40 | 23 | 17 |
| [OpenAI Codex](https://feedbackbench.com/agents/codex.md) | 2 | Typical | 0.508 | 0.476–0.540 | 141 | 80 | 61 |
| [Claude Code](https://feedbackbench.com/agents/claude-code.md) | 1 | Typical | 0.507 | 0.480–0.540 | 92 | 53 | 39 |
| [Google Antigravity](https://feedbackbench.com/agents/antigravity.md) | =5 | Typical | 0.478 | 0.450–0.507 | 66 | 29 | 37 |
| [Cursor](https://feedbackbench.com/agents/cursor.md) | 4 | Worse than peers | 0.431 | 0.401–0.460 | 86 | 26 | 60 |
| [GitHub Copilot](https://feedbackbench.com/agents/copilot.md) | =8 | Too few posts | – | – | 24 | 13 | 11 |
| [Zed](https://feedbackbench.com/agents/zed.md) | =8 | Too few posts | – | – | 19 | 9 | 10 |
| [Factory](https://feedbackbench.com/agents/factory.md) | =11 | Too few posts | – | – | 17 | 9 | 8 |
| [Devin](https://feedbackbench.com/agents/devin.md) | =5 | Too few posts | – | – | 13 | 1 | 12 |
| [Warp](https://feedbackbench.com/agents/warp.md) | 15 | Too few posts | – | – | 6 | 2 | 4 |
| [Conductor](https://feedbackbench.com/agents/conductor.md) | 14 | Too few posts | – | – | 5 | 1 | 4 |
| [Kiro](https://feedbackbench.com/agents/kiro.md) | 13 | Too few posts | – | – | 1 | 1 | 0 |
| [Grok Build](https://feedbackbench.com/agents/grok-build.md) | 16 | Too few posts | – | – | 1 | 1 | 0 |
| [Augment Code](https://feedbackbench.com/agents/augment.md) | 17 | Too few posts | – | – | 0 | 0 | 0 |

## Posts

Receipts rule: The 5 most recent praise and complaint posts per area (first 700 characters) and 3 per criterion (first 450 characters).

### Cline

- Praise, 2026-09-20, @cline (X): “@cline @haydendevs cline is awesome for local models” [source](https://twitter.com/1663928823342063619/status/2101816024609800245)
- Praise, 2026-09-19, @cline (X): “@cline the browser capability is a big unlock, but the config boundary matters too. keeping the gateway key in a local file while the agent handles the browsing task is a much cleaner trust model than pasting secrets into prompts.” [source](https://twitter.com/2092135923169579008/status/2101152370801426934)
- Praise, 2026-09-18, @cline (X): “you can point @cline at @friendliai models without installing anything. in cline's settings, set the api provider to openai compatible, paste your friendliai key, and enter a model id. that's it. we tried it with @google's gemma-4-31b-it and had it build a small word-guessing game ("wurdle"). worked like a charm on the first run. change models later by simply editing the model id field. read our docs for more: <strict_link>” [source](https://twitter.com/1517294112399306752/status/2101029942826070466)
- Complaint, 2026-09-25, r/CLine (Reddit): “yes i had to correct few bug with cline to make it work for me (using the gcp vertex provider) also the way they present the items should be like claude desktop in the future otherwise not a big leap with vsc.” [source](https://www.reddit.com/r/CLine/comments/1wposc5/buggy_desktop_apps/pbxpny3/)
- Complaint, 2026-09-25, @cline (X): “@cline local llm is not working” [source](https://twitter.com/1826429551737745408/status/2103545478189396446)
- Complaint, 2026-09-22, @cline (X): “@cline i have been using it for a few days now, and since i connected the longcat 2.0 model, i encountered a problem. the model itself seems to be very unfamiliar with cline; it doesn't know that it is running on cline. i asked it to install the mcp in the cursor, and it directly configured vscode... also, the oauth on the cline desktop side is not working for me, and i'm not sure what the issue is. it seems there is also a bug in the windows not” [source](https://twitter.com/1805247094128791553/status/2102213337446818155)

### Pi

- Praise, 2026-09-27, r/PiCodingAgent (Reddit): “pi for my use case. i also use opencode for free models but pi when i use my own api” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wrjr75/the_good_opensource_harness/pcd80nj/)
- Praise, 2026-09-27, @pidotdev (X): “running @pidotdev against my local proxy which serves the mac mini’s local copy of qwen on the tailnet is kinda amazing free tokens! currently averages 30-50 t/s (everything not super optimized yet, basically just set up) <strict_link>” [source](https://twitter.com/46650115/status/2104144937923105124)
- Praise, 2026-09-26, r/PiCodingAgent (Reddit): “i've been using claude oauth with pi since at least mar without issue. i know it's strictly against tos. i don't think they care for the amount i use (pro sub)” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wq1bv2/whats_the_best_option_for_using_anthropic/pc45div/)
- Complaint, 2026-09-27, @pidotdev (X): “@pigcodingagent @pidotdev i use opencode in pi but i tried and couldn't in pig. it would be great if you added opencode as a login option.” [source](https://twitter.com/299687169/status/2104021066016580085)
- Complaint, 2026-09-25, @pidotdev (X): “interstellar heroic vibe while reading javascript... meanwhile @pidotdev tells its users to bring their own radio..ffs” [source](https://twitter.com/403519350/status/2103542532412182932)
- Complaint, 2026-09-24, r/PiCodingAgent (Reddit): “welp, i tried to login to llama.cpp (that is running through lm studio), but it doesn't work. the llmstudio provider is not working either. i did use the chat template but have no way to really see if this work :(” [source](https://www.reddit.com/r/PiCodingAgent/comments/1wmo5e5/how_do_i_support_thinking_mode_correctly_with/pbrj6ne/)

### OpenCode

- Praise, 2026-09-27, r/ClaudeCode (Reddit): “i'm using claude pro. at first, planning with opus 5.5 and implementing with sonnet 5 i was not hitting the limits. i tried full opus 5.5 and quickly reached the limit. i have the z.ai coding plan so when it happens i switch to opencode with glm5.3 to continue my workflow. this is possible because i use an ai memory external to the harness.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wrf0ld/is_claude_pro_actually_worth_20_just_for_one/pcguhs1/)
- Praise, 2026-09-26, r/opencode (Reddit): “"gemini, i am a total noob. i want a local llm for coding. i got a 5060 and 32gb ram. where do i get llama.cpp binaries and a model like ornith 35b? give me a small apex quant, offloading is ok. give me arguments for inference with fit on, fit context 128k and q8 kv cache." <strict_link> <strict_link> llama-server.exe -m "path\to\ornith-35b-quant.gguf" --fit on --fit-ctx 131072 --flash-attn on --cache-type-k q8_0 --cache-type-v q8_0 --port 8080 i” [source](https://www.reddit.com/r/opencode/comments/1wqlgg3/best_ways_to_run_ollama_models/pc5bwqn/)
- Praise, 2026-09-26, @opencode (X): “@infiloop2 @thadley0 @manaflowai @zeddotdev @pidotdev @opencode @claudedevs the 'bring your own agent' part is key. less vendor lock-in is always a win, especially with these setups.” [source](https://twitter.com/1825807188360835072/status/2103675720648298514)
- Complaint, 2026-09-27, r/opencode (Reddit): “yeah im trying to understand why or how to fix it but seems like there is no solution. maybe i'll just stick to pure api if i ever need more deepseek with pi.” [source](https://www.reddit.com/r/opencode/comments/1wrg81n/go_subscription_is_slower_deepseek_flash_41/pceeqo7/)
- Complaint, 2026-09-27, r/opencode (Reddit): “a few weeks ago i would have said absolutely, you just put in the api key. but they made a change recently so you must send an `x-opencode-session` header containing a stable, unique uuid per conversation. and i don't know how well n8n can integrate that. hermes agent had an update to handle it. others have custom-built plugins for various harnesses to make it work. so if you don't get an answer from someone who has done it personally, that's spe” [source](https://www.reddit.com/r/opencode/comments/1wrqpjq/can_i_use_the_opencode_go_api_inside_n8n/pcffn8b/)
- Complaint, 2026-09-27, r/opencode (Reddit): “thanks.. as you mentioned, this has become more complex than easy api key setup.” [source](https://www.reddit.com/r/opencode/comments/1wrqpjq/can_i_use_the_opencode_go_api_inside_n8n/pcfi6gb/)

### Amp

- Praise, 2026-09-27, @AmpCode (X): “@benvargas @ampcode local always on was the pain. custom url via orbs is the kind of shortcut that sticks.” [source](https://twitter.com/1502677575776149504/status/2104182819664593107)
- Praise, 2026-09-26, @AmpCode (X): “@bedesqui @catalinmpit @t3dotcodes @ampcode you can wire claude code into amp ui directly btw (this thread is running in cc) <strict_link>” [source](https://twitter.com/3064259332/status/2103785310685647344)
- Praise, 2026-09-26, @AmpCode (X): “@sqs @ampcode @ollama fantastic decision. that was one of the final jigsaw pieces for me. 🫡” [source](https://twitter.com/1590702228234391552/status/2103868668400775310)
- Complaint, 2026-09-26, @AmpCode (X): “@homborg @ampcode amp can't be backed by the local codex cli or claude, right? i wish someone would do that. essentially a web orchestration layer on top of my clis.” [source](https://twitter.com/36411940/status/2103851323238297984)
- Complaint, 2026-09-23, @AmpCode (X): “@ampcode hey team, what’s the proper way to handle test accounts with amp orbs? sometimes i need to use a different one-time test account for verification. i don’t want to put the credentials into an amp secret every time, especially since the account can be different for each test/environment. if i provide the test account credentials directly to amp orbs, it refuses to use them and asks me to sign in manually. is there a recommended way to sec” [source](https://twitter.com/89360130/status/2102621409264758912)
- Complaint, 2026-09-23, @AmpCode (X): “model routing is one of my favorite @ampcode features. i just hope that anthropic and google come into their senses and allow their respective models to be used over oauth.” [source](https://twitter.com/2090734054928687104/status/2102696668944912715)

### OpenAI Codex

- Praise, 2026-09-27, r/codex (Reddit): “my $20 happened to be up today, so impulse cancelled last night. been playing with opus5.5 today for personal projects. i'd say it goes through a quarter the quota that astra did, with at least as good code quality. in fact, i never used to use astra, because it would burn through my 5-hour quota in 10 minutes, even on medium, which i always kept it at. so i was stuck with nerfed sol. guess i could have gone back to 5.6-sol, but i swapped to opus” [source](https://www.reddit.com/r/codex/comments/1wrri5h/been_running_astra_high_100month_and_opus55_high/pcgchqj/)
- Praise, 2026-09-27, r/GithubCopilot (Reddit): “hi u/spare-ant7119 , on my side i am using : \- codex 20$ by month \- and openrouter (10$ when my codex is empty and for testing new models) \- i have also put some $ on deepseek (20$ 3 months ago haha). but their model are very cheap. :) with codex you have very often reset (when they have an issue or an big release) it allow you to have your account to 100% of usage. i am using it in visual studio 2026, to avoid to switch my "workflow"” [source](https://www.reddit.com/r/GithubCopilot/comments/1wr5io2/best_cheaper_alternative/pce3sy4/)
- Praise, 2026-09-27, r/LocalLLaMA (Reddit): “i keep saying this time and again: codex is being slept on as a harness. it’s open source, supports open models out of the box and (as you experienced) is an amazing coding harness. i keep getting downvoted whenever i mention it. go figure.” [source](https://www.reddit.com/r/LocalLLaMA/comments/1wrfp50/another_harness_matters_post_codex_cli_pi_and/pcc69hc/)
- Complaint, 2026-09-27, r/codex (Reddit): “except if you have 1 tb of vram, a local model will never be at the level of astra/opus5.5, and then get ready to warm up your computer. as soon as you work in a real code base with a lot of files and an important context to understand, it’s difficult for small models to be so good.” [source](https://www.reddit.com/r/codex/comments/1wrn2uz/gpt_56_sol_completely_nerfed_after_astra_release/pcdzsnt/)
- Complaint, 2026-09-27, X search: OpenAI Codex, Codex CLI, Codex app (X): “@yetone my problem may occur in [model_providers.magpie] and ~/.codex/magpie-models.json are not generated properly. the /v1/models request has a group, but the codex app cannot access the group model.” [source](https://twitter.com/920264376501780480/status/2104094512188670247)
- Complaint, 2026-09-27, X search: OpenAI Codex, Codex CLI, Codex app (X): “codex cli can no longer be used through cc switch? does anyone know? i was thinking that deepseek's holiday billing is cheap, it was working yesterday, but today it doesn't work.” [source](https://twitter.com/1896614432895307776/status/2104107458151219562)

### Claude Code

- Praise, 2026-09-26, r/ClaudeCode (Reddit): “holy moly, i didn’t wanna get that deep into it, but yeah. i was just giving the cliff’s notes i created an llm panel using openrouter and a few other things where i reach out to other free tier api llms and i run ollama locally. then i have claude delegate out appropriate work based on the model capabilities and the token count/speed that the specific llm can withstand/sustain/etc.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wj51jg/man_what_are_these_limits/pc31vkn/)
- Praise, 2026-09-25, r/ClaudeCode (Reddit): “true, but extracting the skill module was a quick one-shot and now works completely local.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wpczby/superpowers_skill_is_so_bad_now/pbxmm15/)
- Praise, 2026-09-25, r/ClaudeCode (Reddit): “my chinese local claude setup would disagree” [source](https://www.reddit.com/r/ClaudeCode/comments/1wq2905/saw_this_harness_tier_list_online_putting_claude/pc0feac/)
- Complaint, 2026-09-27, @ClaudeDevs (X): “@claudedevs i don't get the point, what kind of developer wants to use ai inference to develop on infrastructure they don't personally entirely control? like, make it make sense, how this is a service i would ever want to pay for?” [source](https://twitter.com/1716937866058874880/status/2104018994294661473)
- Complaint, 2026-09-26, r/ClaudeCode (Reddit): “provider env is a lot of secrets to babysit, feels like you found a tool and then wrote it a review” [source](https://www.reddit.com/r/ClaudeCode/comments/1wqrq8l/quick_setting_switcher_for_claude_code_multiple/pc6dbih/)
- Complaint, 2026-09-25, r/ClaudeCode (Reddit): “not being able to efficiently run other models is what makes it b, otherwise it might’ve deserved an a” [source](https://www.reddit.com/r/ClaudeCode/comments/1wq2905/saw_this_harness_tier_list_online_putting_claude/pc0eaq5/)

### Google Antigravity

- Praise, 2026-09-26, @antigravity (X): “@androidstudio @antigravity bring your own agent inside android studio is exactly what devs needed — choice of claude, codex, antigravity right where you build. love the flexibility in canary” [source](https://twitter.com/2068360781402652672/status/2103763143201833046)
- Praise, 2026-09-25, r/google_antigravity (Reddit): “i just wired in gemma and holy cow, it can do things i could not do before with the cloud rules. cost, speed, quality are all going up and this is just the start. highly recommend looking into this.” [source](https://www.reddit.com/r/google_antigravity/comments/1wom9gb/local_model_support_now_available/pbz5wg2/)
- Praise, 2026-09-25, @antigravity (X): “@antigravity running gemma 4 fully offline with zero api costs is a genuinely useful shift for privacy focused devs.” [source](https://twitter.com/2058874736470327296/status/2103279750777372764)
- Complaint, 2026-09-27, r/google_antigravity (Reddit): “its a skill issue when gemini does not work well, skill issue in selecting a proper provider for models, which is not google” [source](https://www.reddit.com/r/google_antigravity/comments/1wrbsby/antigravitys_gemini_models_are_worst_in_among_all/pcbqem5/)
- Complaint, 2026-09-27, r/google_antigravity (Reddit): “man i have the google ai pro and i have setup for sub2api , so we cab use the antigravity models , with all other harnesses working fine but not with hermes” [source](https://www.reddit.com/r/google_antigravity/comments/1wrh0mx/gemini_38_in_alternative_harnesses/pcfcte6/)
- Complaint, 2026-09-27, @antigravity (X): “@ammaar @petergyang @antigravity @_mohansolo you guys mind letting us use your models in other harnesses like @opencode? really wanted to try 3.8 flash their but there was no easy way to connect it” [source](https://twitter.com/1199733882351828992/status/2104063066275254461)

### Cursor

- Praise, 2026-09-27, r/cursor (Reddit): “so here's my issue with grok models up until 3.x\~ they were training their own models on their own data. in house model, in house data. spacex sees what cursor is doing, which is basically using all the enterprise data flow and their retail use towards training their own in-house model and they want in on it too. here's the kicker, they all start with the same base, it's all kimi 2.5 under the hood. spacex agrees to "buy" cursor, provides bigges” [source](https://www.reddit.com/r/cursor/comments/1wrktgc/am_i_the_only_one_who_thinks_grok_47_is_actually/pcdbwh3/)
- Praise, 2026-09-27, @cursor_ai (X): “@cognosr @cursor_ai cursor + local models sounds like a solid setup.” [source](https://twitter.com/1756637604613607424/status/2104192194265452668)
- Praise, 2026-09-26, r/cursor (Reddit): “cursor environment is good so i use cursor+openrouter happy with it” [source](https://www.reddit.com/r/cursor/comments/1wq5ull/so_no_more_other_models_generous_credits/pc3gk30/)
- Complaint, 2026-09-27, r/cursor (Reddit): “i feel you. i went down this exact rabbit hole a few months ago — loved the tile view and the composer workflow, but the moment i hit a real refactor session the quota wall killed the flow. short answer: cursor doesn't let you bring your own key for their $20 plan, and the byok workaround people hack together is brittle (you lose the indexing, the diff view, the agent loops). what actually solved it for me was switching to an editor-agnostic tool” [source](https://www.reddit.com/r/cursor/comments/1wqq46y/anyone_managed_to_use_cursors_ui_with_their/pcd5l9b/)
- Complaint, 2026-09-26, @cursor_ai (X): “@mdamore9 @grok @bot @cursor_ai rate limits are also much more efficient the past 1-2 weeks this is critical if we can't bring our own models/api keys” [source](https://twitter.com/1756394384655003648/status/2103698095045489000)
- Complaint, 2026-09-26, @cursor_ai (X): “@fatih @cursor_ai too expensive compare to claude/codex subscriptions. cant byok for all functionality.” [source](https://twitter.com/2007052689306501120/status/2103871557147697643)

### GitHub Copilot

- Praise, 2026-09-26, @GitHubCopilot (X): “@behaviourtree @code @githubcopilot they invest a lot in byok suggest you open an issue in github if u r still having problems with this” [source](https://twitter.com/36475277/status/2103855607895998585)
- Praise, 2026-09-22, r/GithubCopilot (Reddit): “yes, byok with local models works great. also, luna 6 just dropped and is ridiculously cheap. half what 5.6 was. highly recommend. i use a mix of sol for thinking and luna for everything i possibly can. my local is now just background tasks.” [source](https://www.reddit.com/r/GithubCopilot/comments/1wnjzmr/using_a_local_llm_as_a_sub_agent_to_save_token/pbfvv73/)
- Praise, 2026-09-22, r/GithubCopilot (Reddit): “absolutely you can and i have done so. byok for the win. of course your mileage may vary depending on how much fast ram you have on whatever is running your local inference.” [source](https://www.reddit.com/r/GithubCopilot/comments/1wnjzmr/using_a_local_llm_as_a_sub_agent_to_save_token/pbg6unc/)
- Complaint, 2026-09-26, @GitHubCopilot (X): “@orenme @code @githubcopilot i just want to use lmstudio local models but none of the approaches work without issues.always some harness failure.” [source](https://twitter.com/93118049/status/2103780819005493482)
- Complaint, 2026-09-25, r/GithubCopilot (Reddit): “this week i’ve been using copilot / interactive / assisted approvals. it’s been very pleasant after feeling like i was fighting with local more recently” [source](https://www.reddit.com/r/GithubCopilot/comments/1wpctlc/vs_code_chat_users_local_or_copilot_harness/pbx3k2v/)
- Complaint, 2026-09-20, r/GithubCopilot (Reddit): “i couldn't get the custom endpoint to work. but i do use openrouter, so i will definitely try adding my together.ai key. thanks.” [source](https://www.reddit.com/r/GithubCopilot/comments/1wjw01d/copilot_with_togetherai_provider/pb2708g/)

### Zed

- Praise, 2026-09-25, @zeddotdev (X): “@zeddotdev works but i do wonder can we make it so you can then turn back on specific features? i personally use edit predictions and nothing else. i even direct those to a local model.” [source](https://twitter.com/1587524325229400064/status/2103431903676354674)
- Praise, 2026-09-18, r/ZedEditor (Reddit): “you can use you're own sub with it. i'm using it with codex” [source](https://www.reddit.com/r/ZedEditor/comments/1whydrw/delta_in_public_beta/pahcmpx/)
- Praise, 2026-09-18, r/ZedEditor (Reddit): “it’s possible to set z.ai up using custom provider settings. at least, i made it work for 5.3-flash model (with their coding plan)” [source](https://www.reddit.com/r/ZedEditor/comments/1whydrw/delta_in_public_beta/paisf5s/)
- Complaint, 2026-09-25, @zeddotdev (X): “@zeddotdev can you guy make delta support custom llm providers please” [source](https://twitter.com/1925398929904214019/status/2103312711803420977)
- Complaint, 2026-09-25, @zeddotdev (X): “@zeddotdev you absolutely need to have a free default model for commit message creation costs almost nothing. now once a quarter, without fail, the commit generation fails because of some configuration issue with models, providers or keys just have it be baked in, it's key onboarding” [source](https://twitter.com/436785962/status/2103372305606860989)
- Complaint, 2026-09-25, @zeddotdev (X): “@zeddotdev delta is goood like really good just need custom provider” [source](https://twitter.com/1261173216455712768/status/2103520442594332993)

### Factory

- Praise, 2026-09-27, @droid (X): “so yeah, you can use custom models in @droid cli, which is great currently trying this setup, is performing great so far ! <strict_link> <strict_link>” [source](https://twitter.com/2028221376092581888/status/2104173498101117143)
- Praise, 2026-09-25, @FactoryAI (X): “@yuqih @factoryai @gmi_cloud their @droid is my main harness and running on gmi keys for open models” [source](https://twitter.com/1261173216455712768/status/2103332443025719806)
- Praise, 2026-09-17, @FactoryAI (X): “only downside is the 5 hours limit and the price which is pretty fair but still a little for me personally. other than that i can list so many things i love about it. droid is super efficient and often finish tasks faster than most other agent with similar results. i feel like it gets the right context at the right time. it’s pretty amazing. also love the byok, live the fact that ui almost always looks better when done with droid even using th” [source](https://twitter.com/1617212256487411712/status/2100703532785541412)
- Complaint, 2026-09-27, @FactoryAI (X): “@droid @factoryai byok is bugged on the gui, but idk where to report this kinda of stuff.” [source](https://twitter.com/4865415939/status/2104338825291891112)
- Complaint, 2026-09-22, @FactoryAI (X): “@tereza_tizkova @enoreyes @factoryai please let us use oauth for openai, kimi, and glm subs. i’d happily pay for the 100 usd sub for that alone.” [source](https://twitter.com/15952889/status/2102464729272995841)
- Complaint, 2026-09-21, @FactoryAI (X): “@tereza_tizkova @droid @factoryai the time i used droid (desktop app) (on windows), byok was behind a paywall which was a bummer. it would be really nice that if this wasn't the case.” [source](https://twitter.com/1854911029870051328/status/2101899154176061769)

### Devin

- Praise, 2026-09-17, @cognition (X): “i went to a @cognition workshop and got some tokens to try @devinai . the most interesting feature i found is their dedicated infrastructure to store secrets as it makes automation easy to test. i’d like to see the same in other platforms 💻 <strict_link>” [source](https://twitter.com/562998386/status/2100694812097671349)
- Complaint, 2026-09-13, @cognition (X): “@brahmad111 @cognition not tried it yet , can’t fit locally 😌” [source](https://twitter.com/1942691267697336324/status/2099153979338866965)
- Complaint, 2026-09-13, @cognition (X): “i started a session through devin cloud less than one hour ago, and i have clear tangible progress, i would say around 60% of the task. i was gifted 2 months of devin by @dabit3 for which im very grateful. but the only reason i used it less is quotas and no possibility of bringing own subs and compute. but i think at the end of the day i burned more money and compute and temper trying to not go the devin way. so the quotas might seem smaller,” [source](https://twitter.com/138407951/status/2099208600111419599)
- Complaint, 2026-09-13, @cognition (X): “@amandineflachs @dnnskr91 @cognition i think devin does not allow byok” [source](https://twitter.com/1271421738459463681/status/2099208983802220839)

### Warp

- Praise, 2026-09-26, @warpdotdev (X): “@lxfater look at this introduction, then i would still prefer to recommend @warpdotdev open source + fully support byok <strict_link>” [source](https://twitter.com/1820087559634202624/status/2103890485442150688)
- Praise, 2026-09-17, r/codex (Reddit): “there’s no way they can continue offering these plans with all of us hammering on their servers basically all day every single day. had a sense this would come to an end. have basically been at it nonstop to get product advanced as far as possible. fyi….the grok $300 plan is incredible……soooo much programming. shit tons. way more than even codex 20x or claude max. i use all 3. also, for grok….100% use it with warp. free to use warp…byom…bring yo” [source](https://www.reddit.com/r/codex/comments/1whlif0/support_just_told_me_they_arent_renewing_people/paad0jv/)
- Complaint, 2026-09-18, @warpdotdev (X): “@vikvang1 @warpdotdev appreciate you asking: 1. make byok / custom inference actually credit-free. i pointed warp at openrouter and the agent still bailed repeatedly on "press y to confirm" steps because i had no credits. if i'm using external inference, the harness shouldn't gate.” [source](https://twitter.com/14132756/status/2100994479591411732)
- Complaint, 2026-09-18, @warpdotdev (X): “@vikvang1 @warpdotdev 3. let the client talk to a local endpoint. i don't want my ollama on the internet but a custom endpoint has to be public for your harness to talk to it. client-side (or at least lan-reachable) would unlock a lot.” [source](https://twitter.com/14132756/status/2100994815861383422)
- Complaint, 2026-09-18, @warpdotdev (X): “@ericmann @vikvang1 @warpdotdev this. if i bring my own keys / openrouter, the harness should not still gate on its own credits. keys in the os credential manager, desktop free, i pay the model vendor. that is the contract i want.” [source](https://twitter.com/2074942490466033664/status/2101000906388935078)

### Conductor

- Praise, 2026-09-09, @conductor_build (X): “little @conductor_build setup that's been helpful: personal claude + codex accounts by default; client ai accounts whenever i open a work repo. super simple. just have one global personal default plus a local override inside each work repo. prompt 👇” [source](https://twitter.com/1821276957428084738/status/2097739018129826296)
- Complaint, 2026-09-25, @conductor_build (X): “@jasdev @conductor_build need it to run local though” [source](https://twitter.com/6827332/status/2103526300355047844)
- Complaint, 2026-09-24, @conductor_build (X): “back to claude. @chatgpt is unusable with @conductor_build” [source](https://twitter.com/337722085/status/2103191328523694521)
- Complaint, 2026-09-21, r/conductorbuild (Reddit): “<strict_link> opencode lets me attach copilot models, but when i select them in conductor, there is exactly 1 skill available - switch to plan mode. is that a known issue?” [source](https://www.reddit.com/r/conductorbuild/comments/1wma8yq/cant_use_skills_in_copilot_via_opencode_bug_report/)

### Kiro

- Praise, 2026-09-25, r/kiroIDE (Reddit): “i’ve been using it with opencode and direct bedrock calls and it’s been working great.” [source](https://www.reddit.com/r/kiroIDE/comments/1wpi98x/opus_55_prediction/pbw9hib/)

### Grok Build

- Praise, 2026-09-24, r/ChatGPTCoding (Reddit): “all harnesses are tui. though a few like grok build has menus clickable by mouse. i actually recommend using that with a local model if you don’t want to subscribe to anything.” [source](https://www.reddit.com/r/ChatGPTCoding/comments/1womcvr/best_claude_code_alternatives/pboumxa/)
