# Computer use and browser control (`work.computer_browser_use`)

Feedback Bench, coding agents, built 2026-10-01, window 2026-08-31 to 2026-09-27. Web page: https://feedbackbench.com/#/criterion/work.computer_browser_use

Area: [Doing the work](https://feedbackbench.com/criteria/work.md)

**Definition.** Whether the agent operates desktop apps and browsers well, without taking over the user's windows.

**Boundary.** Not this: see [Images, PDFs and file attachments as input](https://feedbackbench.com/criteria/context.attachments.md) for reading screenshots.

Rated author-weeks, all agents: 556. Complaint share: 45%.

## The brief

Written by Claude Opus 5.5 from 67 labelled posts and the numbers on this page. Interpretation, not measurement: every quote is verbatim and links to its post.

**Codex sets the computer-use bar; rivals get measured against it.**

TL;DR:

- Codex draws the most praise, and other agents' users name it as the benchmark to match.
- Antigravity users call its built-in browser broken and route browser work to Codex instead.
- Top asks are built-in computer use, background operation that leaves windows alone, and Linux support.

In plain terms: When it works, the agent clicks through your app faster than you and closes the testing loop. When it fails, it stalls on the wrong OS, a login, or a dropdown, or quietly leaves browsers running.

### How it breaks

- **Works on one OS, not yours** ([Computer use and browser control](https://feedbackbench.com/criteria/work.computer_browser_use.md)). Computer use often ships for one platform first, and users on Linux, WSL, or the Windows app hit a wall or a weaker version.
  Codex users describe a feature that works in native Windows but not under WSL, does not run on Linux, and feels behind the Mac app on Windows. One user says a third-party tool beats the native Windows app. Another reports an update that limited computer use to Chrome. Claude Code's background mode draws the same question, with users asking when it stops being Mac-only. Linux support is a recurring request.
  Evidence:
  - Complaint, OpenAI Codex, r/codex, 2026-09-07: “i always grant full access to any project. i’ve always used wsl, but lately i’ve been wanting to test projects more and more using computer use, especially with the release of gpt-astra, but it doesn’t work in wsl, only in native windows.” [source](https://www.reddit.com/r/codex/comments/1w9tk34/does_the_codex_app_work_well_with_native_windows/p8d9ivk/)
  - Complaint, OpenAI Codex, r/codex, 2026-09-22: “yeah i tried opencode as well, then there was the whole debacle with axing of cheap models and 15$ usage stuff, can't be asked to deal with that, codex just works only issue is computer use doesn't work on linux yet” [source](https://www.reddit.com/r/codex/comments/1wnh014/absolutely_no_fking_way_the_pricing_is_wowww/pbgsbqr/)
  - Complaint, OpenAI Codex, X search: OpenAI Codex, Codex CLI, Codex app, 2026-09-11: “i finally went back to using the codex cli on windows and i'm already noticing improvements in terms of speed and quota burn rate. trycua is also better on windows vs the native codex app computer use. hopefully the windows app reaches mac level soon.” [source](https://twitter.com/717586332/status/2098358921522192839)
  - Complaint, OpenAI Codex, r/codex, 2026-09-06: “really took a dive today. not sure what's going on, same repo, same [agents.md](http://agents.md), same instructions. astra is now returning a message every single turn without continuing its work; when tests are red, it literally just looks at the test, then creates a commit and says "tests are still red! i'm not going to lie and say they're green", and if i tell it to continue to make the tests green, it literally just looks at the tests again and then does it over again without changing a single line of code. also, my codex desktop, after the most recent update, somehow can no longer do computer use outside of chrome browser, and none of my agents can figure out why. tried with astra medium and sol high, and both of them come up blank.” [source](https://www.reddit.com/r/codex/comments/1w3i0mm/codex_usage_and_operation_discussion_last_updated/p88zgr7/)

- **Built-in browsers that cannot see** ([Computer use and browser control](https://feedbackbench.com/criteria/work.computer_browser_use.md)). Several agents ship a browser panel the agent cannot visually drive, so users fall back to MCP add-ons or switch tools for front-end checks.
  Devin's built-in browser is called too rudimentary, with a hard-to-use element selector and no visual operation by the agent. Cursor users say the agent barely controls its integrated browser. A Conductor user calls its in-app browser lacking and names Codex king there. Antigravity users wish for an isolated browser the agent can click around in. The fix users describe is usually a different tool, not a setting.
  Evidence:
  - Complaint, Devin, @cognition, 2026-09-15: “the built-in browser function of @cognition is too rudimentary compared to cursor and codex. the element selector is difficult to use, it cannot reference area screenshots, and the agent cannot visually operate the built-in browser, which greatly limits the convenience of front-end development and the capabilities of agent automated testing.” [source](https://twitter.com/1236527033435312128/status/2099904812749844666)
  - Complaint, Cursor, r/cursor, 2026-09-06: “is there any way to make computer use efficient on cursor, like the codex app, because on cursor it barely controls the integrated browser right ?rather the system or the external browser if so, please let me know. thanks . i tried codex computer user great stuff , but the token are flying and actually im on pro+ cursor” [source](https://www.reddit.com/r/cursor/comments/1w9al8a/computer_use_on_cursor_like_codex/)
  - Complaint, Conductor, @conductor_build, 2026-09-13: “for managing worktrees and complex setups it’s fantastic. it feels super quick and intuitive. i love having multiple chats within one worktree too. they feel more involved and first class citizens rather than codex “side chats” that feel mainly for asking questions. it uses then codex cli so it can still access computer use etc… the in-app browser is lacking though. codex is still king there and the plugin ecosystem feels seamless within codex too. i prefer it for all models outside of codex.” [source](https://twitter.com/50570112/status/2099046528161567197)
  - Complaint, Google Antigravity, r/google_antigravity, 2026-09-24: “been doing this for a while now, with occasional usage of vscode just to look through certain things (very rarely). wish antigravity had a built in browser, similar to codex, to let the agents click around in an isolated browser. looking forward to trying agy on googlebooks tho 👀” [source](https://www.reddit.com/r/google_antigravity/comments/1wp6mpo/poll_how_do_you_code_in_late_2026/pbtewsr/)

- **Foreground agents hijack the desktop** ([Computer use and browser control](https://feedbackbench.com/criteria/work.computer_browser_use.md)). Users want computer use that runs in the background, because foreground control turns the user into a babysitter.
  Claude Code's background computer use beta drew praise for exactly this reason. Users say overnight runs now get a whole desktop instead of one terminal window. The open questions are what happens when the laptop sleeps and when it leaves macOS. Background operation without taking over windows is a named request for both Claude Code and Codex.
  Evidence:
  - Praise, Claude Code, @ClaudeDevs, 2026-09-02: “@claudedevs background computer use is the real unlock. foreground agents always felt like you were babysitting a coworker curious how long til this stops being mac-only.” [source](https://twitter.com/2094945695757475840/status/2095232067818651923)
  - Praise, Claude Code, @ClaudeDevs, 2026-09-02: “@claudedevs background computer use means my overnight agent runs just got a whole desktop instead of one terminal window” [source](https://twitter.com/1998401809665425408/status/2095228499586175021)
  - Praise, Claude Code, @ClaudeDevs, 2026-09-03: “@claudedevs background computer use on pro/max, macos-only is the right beta shape. real question: does overnight agent work survive when the laptop sleeps? that's the product gap.” [source](https://twitter.com/2788825408/status/2095456720684274172)

- **Silent failures on real-world flows** ([Computer use and browser control](https://feedbackbench.com/criteria/work.computer_browser_use.md)). Agents that look capable in demos stumble on dynamic dropdowns, auth flows, captchas, and leftover processes, often without telling the user.
  A Claude Code user reports the agent launching a debug Chrome in the background, summarising before checking its output, and leaving the browser burning CPU for hours. A Cursor user saw models trigger a reload that restarted Cursor itself instead of the page. Others flag captchas and login flows as hard stops. A Warp post puts the worry plainly: who owns a silent failure days in.
  Evidence:
  - Complaint, Pi, @pidotdev, 2026-09-08: “@pidotdev @openai frontier leading for computer use is a meaningful claim right up until you hand it a dynamic dropdown or an auth flow.” [source](https://twitter.com/1910204476973096960/status/2097329557804081326)
  - Complaint, Claude Code, r/ClaudeCode, 2026-09-02: “yeah this is a real bug, seen it a lot with the browser tasks. claude will launch a `chrome --remote-debugging-port` or a `run_in_background: true` bash job and then summarize before it ever calls `taskoutput` to wait for it, so it tells you gates are clean while `ps aux` still shows chrome burning cpu for hours. what fixed it for me was adding a rule in `claude.md` to never use background without an await and to always run `ps aux | grep chrome` at the end, and when it happens just `pkill -f "chrome --remote-debugging"` then kill the leftover background id. `claude --verbose` also shows the orphaned task id it forgot to check.” [source](https://www.reddit.com/r/ClaudeCode/comments/1w506l1/orphaned_processes/p7bst8u/)
  - Complaint, Cursor, @cursor_ai, 2026-09-10: “just noticed something @cursor_ai, some models seem to be thinking that cdp page.reload reloads a browser page (gemini 3.8 and kimi k3) but that reloads cursor instead 🤔” [source](https://twitter.com/4209730385/status/2098006730936406046)
  - Complaint, Warp, @warpdotdev, 2026-09-15: “@warpdotdev computer-use that needs a human every step is labor with extra latency. who owns silent failure on day 3?” [source](https://twitter.com/1977033941514072064/status/2099852678754975934)

- **Slow and token-hungry runs** ([Computer use and browser control](https://feedbackbench.com/criteria/work.computer_browser_use.md)). Even when computer use succeeds, users say it is slow enough to watch cartoons through and expensive enough to drain plans.
  A Cursor user says cloud agents are so slow at computer use that doing it by hand would have been faster. Another tried Codex computer use, liked it, but says the tokens were flying. Faster, cheaper computer use is a request across Codex, Claude Code, and Cursor. Speed and cost now matter as much as whether the click lands.
  Evidence:
  - Complaint, Cursor, @cursor_ai, 2026-09-12: “@cursor_ai cloud agents are so slow at computer use on their own machine that i’m just watching cartoons with my 4 yo. we need @grok 4.7 asap! would have been faster for me to do it but i’m naturally lazy. <strict_link>” [source](https://twitter.com/1215028704/status/2098845440804635034)
  - Complaint, Cursor, r/cursor, 2026-09-06: “is there any way to make computer use efficient on cursor, like the codex app, because on cursor it barely controls the integrated browser right ?rather the system or the external browser if so, please let me know. thanks . i tried codex computer user great stuff , but the token are flying and actually im on pro+ cursor” [source](https://www.reddit.com/r/cursor/comments/1w9al8a/computer_use_on_cursor_like_codex/)

### Who stands out

- **OpenAI Codex (stronger)**. Codex is the reference point, praised for reliable control across apps and sessions, and named by rival users as the standard to match.
  Users say it works every time once enabled, drives desktop apps beyond the browser, and lets computer control carry across chats and the iOS app. Claude Code users concede computer use is the edge Codex holds. Complaints cluster on platform reach, not quality: WSL, Linux, and a Windows app that trails the Mac build.
  Evidence:
  - Praise, OpenAI Codex, X search: OpenAI Codex, Codex CLI, Codex app, 2026-09-04: “@brick295 @mvanhorn @mvanhorn just install the chatgpt/codex app on your computer, enable computer use, and enable codex computer use in openclaw. that’s it. or even just use codex directly. it works every time.” [source](https://twitter.com/2841774497/status/2095686076518023474)
  - Praise, OpenAI Codex, X search: OpenAI Codex, Codex CLI, Codex app, 2026-09-27: “i regret to inform you - the codex app is miles better than claude app - gpt6 is the best computer use agent and it’s not even close this is after spamming opus 5.5 for the last few days (still the best model)” [source](https://twitter.com/1328913688892346370/status/2104109109549367435)
  - Praise, OpenAI Codex, r/ClaudeCode, 2026-09-06: “one thing i love codex more than claude is how computer control works from different chat; and i can access sessions from my ios app, compared to claude where i need to /rc session to get access or use dispatch which is just a single thread so i can’t separate sessions.” [source](https://www.reddit.com/r/ClaudeCode/comments/1w8kgah/just_moved_to_codex_and_wow/p83i5a6/)
  - Complaint, Claude Code, r/codex, 2026-09-27: “yah that is what i'm planning really, going back and forth between the two, sometimes anthropic models are ahead and sometimes openai are leading, also i find that computer use is superior in codex really regardless of how now slow slo 6 or luna 6. that is maybe the only edge thay have over claude code.” [source](https://www.reddit.com/r/codex/comments/1wrn697/new_openai_product_o_but_not_for_plus_users/pcefpfh/)

- **Google Antigravity (weaker)**. Antigravity's browser tooling draws the sharpest complaints, and users describe routing browser work to Codex instead.
  Posts call its browser command garbage and its MCP integration a disaster. One user advises against setting up computer use at all and suggests Playwright via MCP. Another says the IDE version handles the browser better than the agent build. Fans exist: some users run it as a free browser tester or pair it with Codex to do the clicking. Built-in computer use is its most requested fix.
  Evidence:
  - Complaint, Google Antigravity, r/google_antigravity, 2026-09-20: “well i cant use /browser in antigravity because its garbage, so i use codex browser-use when im developing browser extensions and when i need to modify a third party website ui, the only way the ai can know what to target is if codex uses browser-use to see the dom and all html.” [source](https://www.reddit.com/r/google_antigravity/comments/1wluoqe/which_antigravity_surface_do_you_use_the_most/pb26ive/)
  - Complaint, Google Antigravity, @antigravity, 2026-09-26: “@ivanleomk @antigravity open-source it pls the mcp integration inside antigravity is an absolute disaster to use. third-party solutions like agent browser and browser use just don't hold a candle to codex's browser control.” [source](https://twitter.com/2080328280927055875/status/2103687543564927358)
  - Complaint, Google Antigravity, r/google_antigravity, 2026-09-16: “kind of short answer: yes it can but you have to set it up, i don't advice setting up computer use in agy for two main reasons first i don't really trust gemini models and second is there is really no need to instead set the mcp for he particular app you want it to use, like playwright for browser use but that one already comes preinstalled with antigravity” [source](https://www.reddit.com/r/google_antigravity/comments/1wi8ker/antigravity_control_the_computer/pa8v51y/)
  - Praise, Google Antigravity, r/google_antigravity, 2026-09-26: “for what you get for free it's crazy good a dependable. i get alot of use from just having it use the browser and try to make the crap i've built fail. the flash models are good for that and you get some opus and sonnet 4.6 as well. also i have it as basically my google hands to control and use my gemini notebooks and all the other free google stuff like stitch. don't sleep on using it with google vids for some easy editing and video creation. i kind of just have agy running that kind of stuff and bunch of other harnesses in a herdr setup. not really a set up i just find it easier to organize if seperate work this way” [source](https://www.reddit.com/r/google_antigravity/comments/1wqmd4n/why_is_the_antigravitycli_so_underrated/pc89ome/)

- **Claude Code (mixed)**. Claude Code wins praise for background computer use and app click-throughs, but users still rank Codex ahead and report browser rough edges.
  Users say it closes the loop by clicking through apps faster than a human and surfacing weak selectors. The background beta is called the real unlock. On the other side, posts describe orphaned browser processes, passkeys that need work in its browser, and a plain verdict that Codex computer use is better.
  Evidence:
  - Complaint, Claude Code, @ClaudeDevs, 2026-09-20: “@claudedevs thanks for the cookies import. passkeys in claude's browser still need some tlc. cmux has a really cool implementation. any attention to this would really help the browser use experience.” [source](https://twitter.com/186789799/status/2101670343949582647)
  - Praise, Claude Code, r/ClaudeCode, 2026-09-15: “are you saying that claude actually said "you good to poke around yourself" to you? are you speaking to it using super casual slang all the time? also there is nothing wrong with this, it literally closes the loop and will click through the app way faster than you, will surface issues where you have poor accessibility due to bad selectors since it is optimized to use that, etc” [source](https://www.reddit.com/r/ClaudeCode/comments/1whdo6u/today_i_lost_any_shred_of_self_respect_that_i_had/pa1x7go/)
  - Complaint, Claude Code, r/ClaudeCode, 2026-09-04: “+ computer use is so much better and offers so much usage compared to claude.” [source](https://www.reddit.com/r/ClaudeCode/comments/1w6x360/todays_comparison_between_claude_and_codex_20/p7qwlhs/)
  - Complaint, Claude Code, r/ClaudeCode, 2026-09-02: “yeah this is a real bug, seen it a lot with the browser tasks. claude will launch a `chrome --remote-debugging-port` or a `run_in_background: true` bash job and then summarize before it ever calls `taskoutput` to wait for it, so it tells you gates are clean while `ps aux` still shows chrome burning cpu for hours. what fixed it for me was adding a rule in `claude.md` to never use background without an await and to always run `ps aux | grep chrome` at the end, and when it happens just `pkill -f "chrome --remote-debugging"` then kill the leftover background id. `claude --verbose` also shows the orphaned task id it forgot to check.” [source](https://www.reddit.com/r/ClaudeCode/comments/1w506l1/orphaned_processes/p7bst8u/)

- **Cursor (mixed)**. Cursor's cloud agents with their own VMs impress users, but the local experience lacks control of the user's own browser and runs slowly.
  Users like that cloud agents get a VM and send back screenshots and videos, and one describes long multi-step browser jobs done persistently. The in-app browser wins fans. Complaints say there is no remote control and no use of the user's logged-in browser, and computer use on its own machine crawls.
  Evidence:
  - Praise, Cursor, @cursor_ai, 2026-09-06: “@nick_kango @cursor_ai the in-app browser is what got me. i used to bounce tabs all night” [source](https://twitter.com/1773662247023468545/status/2096497192450031640)
  - Praise, Cursor, r/cursor, 2026-09-12: “hey there. i am currently trying out cursor and i really like it. the local agent feel quite similar to other agent harnesses like copilot but what really stands out for me are the cloud agents. is it possible to get cloud agents with computer use that work with other harnesses like opencode, for example? i especially like that the agent has its own vm and and sends screenshots and videos of the result. thanks!” [source](https://www.reddit.com/r/cursor/comments/1we5n68/can_i_achieve_cursorlike_cloud_agents_with_other/)
  - Complaint, Cursor, @cursor_ai, 2026-09-22: “ok @cursor_ai is lagging behind now. no remote control (cloud sessions doesnt count), no browser control (user's own one with logins), i use these 2 so much in other tools but missing in cursor” [source](https://twitter.com/1306443936194596865/status/2102290226945331335)
  - Praise, Cursor, r/cursor, 2026-09-04: “i’m pretty sure it’s doing it all through browser automation - you can open computer and watch it navigate the pa site. i’ve given it a bunch of resources and had it build 25 item listings (that were like 8 pages with attachments and generated images, cropped packaging shots, local files. etc.). it’s revised a compliance label for me and made it look way better than what chat gpt produced. it’s pretty wild - troubleshoots issues and is very persistent on getting the job done. it’s currently building a welcome and account approval email flow in shopify and building some cursor projects for me.” [source](https://www.reddit.com/r/cursor/comments/1vzg2wd/any_one_tried_grok_bot_is_it_garbage_or_just_me/p7p12o5/)

### Fine print

- Most agents have too few posts here to separate from the pack; Devin, OpenCode, and others are not ranked.
- Several posts praise a model rather than the harness, which blurs credit between agent and model.
- Codex dominates the sample, so its complaints are better documented than smaller agents' complaints.

## Top requests

What users ask to add or change, most asked first. 218 author-weeks ask for something. Requests do not change the Feedback Score. Rule: A separate pass by Claude Sonnet 5 reads every counted post and extracts what the author asks the agent or its vendor to add or change, with the criteria it maps to and a short normalised wording; it does not touch the labels or the Feedback Score. Claude Opus 5.5 groups the wordings within each criterion (the first criterion the request maps to) into themes; code counts them. A theme counts distinct author-weeks that ask for it, per agent; across agents, one author-week per agent. Themes asked in fewer than 2 author-weeks, and requests that share no theme, are not shown. Examples: up to 3 posts per theme from different authors, without slurs, preferring posts of 60 to 450 characters, most recent first.

| Rank | Request | Author-weeks | Posts | Agents (author-weeks) |
|---|---|---|---|---|
| 1 | Built-in computer use capability | 45 | 46 | Google Antigravity 13, OpenAI Codex 12, Devin 7, Cursor 6, Cline 3, Amp 1, Claude Code 1, GitHub Copilot 1, OpenCode 1 |
| 2 | Better browser automation capability | 16 | 16 | Google Antigravity 5, Claude Code 5, OpenAI Codex 2, Devin 2, Cursor 1, OpenCode 1 |
| 3 | Integrated in-app browser panel | 15 | 18 | Google Antigravity 6, Cline 2, OpenCode 2, Amp 1, Claude Code 1, OpenAI Codex 1, Conductor 1, Pi 1 |
| 4 | Computer use support on Linux | 8 | 9 | OpenAI Codex 6, Google Antigravity 1, Claude Code 1 |
| 5 | Faster, cheaper computer use | 8 | 8 | OpenAI Codex 3, Claude Code 2, Cursor 2, Google Antigravity 1 |
| 6 | Fix browser automation bugs | 8 | 8 | OpenAI Codex 4, Cursor 2, Claude Code 1, OpenCode 1 |
| 7 | Mac VM with iOS simulator | 8 | 8 | Devin 8 |
| 8 | Remote control of other computers | 8 | 8 | OpenAI Codex 3, Claude Code 2, Google Antigravity 1, GitHub Copilot 1, Cursor 1 |
| 9 | Background computer use without taking over windows | 7 | 7 | Claude Code 4, OpenAI Codex 3 |
| 10 | Computer and browser use in CLI | 7 | 7 | OpenAI Codex 5, Google Antigravity 1, Cursor 1 |
| 11 | Control of native desktop applications | 7 | 7 | OpenAI Codex 2, Amp 1, Google Antigravity 1, Claude Code 1, Factory 1, Zed 1 |
| 12 | Mobile emulator access for agents | 6 | 6 | Amp 1, Google Antigravity 1, Claude Code 1, Cursor 1, Devin 1, Factory 1 |

### 1. Built-in computer use capability

- OpenAI Codex, 2026-09-24, r/google_antigravity (Reddit): “is there any skill, extension or plugin like codex computer use? like getting it to control pc at its own and complete the job. i'm surprise this feature isn't in antigravity yet since been focus on multi-model and interaction? <strict_link>” [source](https://www.reddit.com/r/google_antigravity/comments/1wotyua/computer_use_in_antigravity/)
- Cline, 2026-09-24, @cline (X): “@cline browser automation with a built-in browser? computer use? when can we expect that” [source](https://twitter.com/1696542879735222272/status/2103056296488407298)
- Devin, 2026-09-23, @cognition (X): “i've been an @cursor_ai user since feb 2024, but after grok 4.7, i'm looking for alternatives. @droid @cognition are in the lead for me, but what i really need is 1. cloud agents/desktop 2. agnostic harness &amp; computer use 3. mobile 4. good connectors anyone have suggestions?” [source](https://twitter.com/1902193987244408832/status/2102596587046531556)

### 2. Better browser automation capability

- Google Antigravity, 2026-09-26, @antigravity (X): “@ivanleomk @antigravity open-source it pls the mcp integration inside antigravity is an absolute disaster to use. third-party solutions like agent browser and browser use just don't hold a candle to codex's browser control.” [source](https://twitter.com/2080328280927055875/status/2103687543564927358)
- Claude Code, 2026-09-23, @ClaudeDevs (X): “@trq212 loving 5.5! any plans to add browser access when i run in cloud? @claudedevs” [source](https://twitter.com/1332447750/status/2102735069312090503)
- Google Antigravity, 2026-09-23, @antigravity (X): “@rodydavis @antigravity as a harness, i think it should be far better than what it is right now. i know you are a huge developer and i am nobody but z code is also better than antigravity multiple things it needs plug-in and browser” [source](https://twitter.com/1405851706848530432/status/2102604582119789031)

### 3. Integrated in-app browser panel

- OpenCode, 2026-09-23, @opencode (X): “@opencode @opencode , could you add an integrated browser feature to the desktop app, similar to how it works in codex?” [source](https://twitter.com/384597968/status/2102774203707502639)
- Pi, 2026-09-22, r/PiCodingAgent (Reddit): “this is clean as af, it's everything i've been looking for and it being a fork of openchamber on top of everything else is icing on the cake. it's crazy how much of us are living the same lives, thanks much for this. does this include browser functionality / previews? with original openchamber those features don't work on web so if that worked here this would be a dream.” [source](https://www.reddit.com/r/PiCodingAgent/comments/1whrhm5/pichamber_same_pi_session_on_desktop_browser_and/pbaiyss/)
- Google Antigravity, 2026-09-15, r/google_antigravity (Reddit): “yeah, tried that. i think we’re talking about two different things. generative ui seems focused on artifact/html previews; i’m referring to a persistent browser window for the actual dev server, with screenshots/annotations alongside the agent, more like codex’s integrated browser workflow. for full-stack development that removes a lot of context switching.” [source](https://www.reddit.com/r/google_antigravity/comments/1wh9uyh/antigravity_2_release_v2140/pa10h8p/)

### 4. Computer use support on Linux

- Google Antigravity, 2026-09-24, r/google_antigravity (Reddit): “use browser mcp, i also need general computer use for linux but there isn't any stable...” [source](https://www.reddit.com/r/google_antigravity/comments/1wotyua/computer_use_in_antigravity/pbpv9na/)
- OpenAI Codex, 2026-09-22, r/codex (Reddit): “yeah i tried opencode as well, then there was the whole debacle with axing of cheap models and 15$ usage stuff, can't be asked to deal with that, codex just works only issue is computer use doesn't work on linux yet” [source](https://www.reddit.com/r/codex/comments/1wnh014/absolutely_no_fking_way_the_pricing_is_wowww/pbgsbqr/)
- OpenAI Codex, 2026-09-20, X search: OpenAI Codex, Codex CLI, Codex app (X): “tibo aka @thsottiaux ... dropper of resets turn dropper of **hints**? could it be that openai will soon release better computer use support in the chatgpt/codex app for linux desktop? or a whole new agentic distro? his previous post alluded to how much new stuff they're going to announce soon ... so "i'm monitoring the situation."” [source](https://twitter.com/10667142/status/2101652767362240605)

### 5. Faster, cheaper computer use

- Cursor, 2026-09-23, @cursor_ai (X): “.@cursor_ai please make fast cua we have jev and other stuff now, sonnet 4.5 is super slow and expensive - makes me stop using cloud agents <strict_link>” [source](https://twitter.com/373274900/status/2102668343681618156)
- Claude Code, 2026-09-22, r/ClaudeCode (Reddit): “can it open a web browser and do research, fill in info and chat with customer service reps without hitting the limit like it has done in the past? that's all i care about” [source](https://www.reddit.com/r/ClaudeCode/comments/1wnecru/introducing_claude_opus_55_the_first_model_in_our/pbesdli/)
- OpenAI Codex, 2026-09-20, r/codex (Reddit): “personally the computer control is cumbersome and drains tokens too fast. i find it better to pilot and send screenshot to sol (or astra if you have the usage).” [source](https://www.reddit.com/r/codex/comments/1wlbrd0/how_to_ai_assisted_cad_design/paxm9rx/)

### 6. Fix browser automation bugs

- OpenCode, 2026-09-26, @opencode (X): “please @thdxr help fix @opencode so we can use the <strict_link> drivers for computer use without these errors every time. computer use is a critical feature these days and this issue is holding opencode back. <strict_link> <strict_link> <strict_link>” [source](https://twitter.com/20052949/status/2103796383450825001)
- Cursor, 2026-09-25, @cursor_ai (X): “browser use by subagents in @cursor_ai is broken <strict_link>” [source](https://twitter.com/467130927/status/2103491442882830418)
- Cursor, 2026-09-22, @cursor_ai (X): “@bot is soooo lame at connecting to any site!!! times out constantly - i am turinig in to a babysitter! @cursor_ai @bot you have to fix this” [source](https://twitter.com/48466418/status/2102412074337141154)

### 7. Mac VM with iOS simulator

- Devin, 2026-09-17, @cognition (X): “@ptbthefirst @cognition a dedicated mac vm with ios simulator does close a real gap for mobile automation if it actually works reliably in practice.” [source](https://twitter.com/1963715144392863744/status/2100395669034852542)
- Devin, 2026-09-17, @cognition (X): “@ptbthefirst @cognition giving devin its own mac vm with an ios simulator actually solves a real bottleneck for mobile testing.” [source](https://twitter.com/2058874736470327296/status/2100391330106974685)
- Devin, 2026-09-17, @cognition (X): “@ptbthefirst @cognition devin getting a full mac vm with an ios simulator finally closes the mobile dev gap.” [source](https://twitter.com/2061512950322548736/status/2100388341631861217)

### 8. Remote control of other computers

- Claude Code, 2026-09-16, @ClaudeDevs (X): “@raroque agreed on computer use. hope @claudedevs steps up their game on remote!” [source](https://twitter.com/1628155650621714432/status/2100047683113385988)
- Google Antigravity, 2026-09-14, r/google_antigravity (Reddit): “ah im new to antigravity, so previously it had a terminal option? i would love to drop into a shell from remote control sometimes damn” [source](https://www.reddit.com/r/google_antigravity/comments/1wf7eu7/agy_cli_122_no_more_terminal_in_remote_mode/p9o9yta/)
- OpenAI Codex, 2026-09-14, X search: OpenAI Codex, Codex CLI, Codex app (X): “@thsottiaux i need remote control other computer inside the codex app linux please 👏👏👏” [source](https://twitter.com/1544380649519435780/status/2099452559962345813)

### 9. Background computer use without taking over windows

- OpenAI Codex, 2026-09-25, r/codex (Reddit): “honestly the way claude does this is kinda weird. sometimes it just uses the computer normally, like a regular person would, super organic. then other times the whole screen turns orange and it's like "claude is doing this, this and this" and it looks so strange. idk why they even built it like that” [source](https://www.reddit.com/r/codex/comments/1wpvp4o/for_people_with_both_codex_and_claude_code_what/pc08ub2/)
- OpenAI Codex, 2026-09-09, r/codex (Reddit): “i get the computer use alure on macos. my macbook uses it beautifuly without disrupting what i'm doing, since it can run entirely on background. but it's so shitty on my main machine (windows). sadly it can't run in the background there, so i literally have to let the ai use my computer for me while it does stuff with computer use.” [source](https://www.reddit.com/r/codex/comments/1wb6msb/shots_were_fired/p8qqyxg/)
- Claude Code, 2026-09-02, @ClaudeDevs (X): “@claudedevs does it need the app in focus or can it work behind a fullscreen window” [source](https://twitter.com/1649303522930749442/status/2095240400378142901)

### 10. Computer and browser use in CLI

- Google Antigravity, 2026-09-26, @antigravity (X): “@ivanleomk @antigravity need that cli browser for concur filing” [source](https://twitter.com/1811332417099055105/status/2103792711467946295)
- OpenAI Codex, 2026-09-15, X search: OpenAI Codex, Codex CLI, Codex app (X): “theory: because codex cli is opensource, it's inferior to the codex desktop. so it still doesn't have browser support yet.” [source](https://twitter.com/1252559794080268288/status/2099745400274264440)
- OpenAI Codex, 2026-09-08, X search: OpenAI Codex, Codex CLI, Codex app (X): “@voxyz_ai i wish codex cli allowed browser use and computer use. how do you get around that?” [source](https://twitter.com/2551067293/status/2097141437380870194)

### 11. Control of native desktop applications

- Factory, 2026-09-27, @FactoryAI (X): “great to hear! as a solo founder, it is a massively important part of my stack. i don't have any specific or intensive needs other than following up on emails, having computer use and secure password/payment sharing for login and forms, and automatically creating tickets from user feedback” [source](https://twitter.com/2070908287978246144/status/2104338878144667993)
- Google Antigravity, 2026-09-24, r/google_antigravity (Reddit): “trying to use something like computer use on desktop app instead of browser, but thanks” [source](https://www.reddit.com/r/google_antigravity/comments/1wotyua/computer_use_in_antigravity/pbpxkg9/)
- Claude Code, 2026-09-24, @ClaudeDevs (X): “@claudedevs it's great! could you add computer use like grok bot so i can update a local document in the same session?” [source](https://twitter.com/1920079071284965376/status/2102946551438221695)

### 12. Mobile emulator access for agents

- Devin, 2026-09-17, @cognition (X): “@ptbthefirst @cognition mobile dev automation has been stuck without device level testing so this is a meaningful unlock.” [source](https://twitter.com/2067687598571569152/status/2100393725713076709)
- Google Antigravity, 2026-09-16, @antigravity (X): “@antigravity when will you link an emulator and a browser for my tests?” [source](https://twitter.com/2079811715273809920/status/2100239754390372797)
- Claude Code, 2026-09-09, r/ClaudeCode (Reddit): “can it work for mobile apps? for example does the whole demo thing but by interacting with iphone simulator” [source](https://www.reddit.com/r/ClaudeCode/comments/1wbg30i/i_made_an_mcp_app_so_claude_code_can_record_edit/p8ql3wc/)

## Every agent

| Agent | Overall rank | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|---|
| [OpenAI Codex](https://feedbackbench.com/agents/codex.md) | 2 | Better than peers | 0.534 | 0.511–0.561 | 292 | 176 | 116 |
| [Claude Code](https://feedbackbench.com/agents/claude-code.md) | 1 | Typical | 0.486 | 0.455–0.516 | 100 | 50 | 50 |
| [Cursor](https://feedbackbench.com/agents/cursor.md) | 4 | Typical | 0.482 | 0.458–0.505 | 38 | 17 | 21 |
| [Google Antigravity](https://feedbackbench.com/agents/antigravity.md) | =5 | Worse than peers | 0.461 | 0.435–0.486 | 47 | 16 | 31 |
| [Devin](https://feedbackbench.com/agents/devin.md) | =5 | Too few posts | – | – | 27 | 16 | 11 |
| [OpenCode](https://feedbackbench.com/agents/opencode.md) | 3 | Too few posts | – | – | 18 | 11 | 7 |
| [Amp](https://feedbackbench.com/agents/amp.md) | =11 | Too few posts | – | – | 9 | 5 | 4 |
| [Pi](https://feedbackbench.com/agents/pi.md) | 7 | Too few posts | – | – | 7 | 2 | 5 |
| [Cline](https://feedbackbench.com/agents/cline.md) | =8 | Too few posts | – | – | 7 | 5 | 2 |
| [GitHub Copilot](https://feedbackbench.com/agents/copilot.md) | =8 | Too few posts | – | – | 6 | 5 | 1 |
| [Conductor](https://feedbackbench.com/agents/conductor.md) | 14 | Too few posts | – | – | 2 | 1 | 1 |
| [Zed](https://feedbackbench.com/agents/zed.md) | =8 | Too few posts | – | – | 1 | 0 | 1 |
| [Factory](https://feedbackbench.com/agents/factory.md) | =11 | Too few posts | – | – | 1 | 1 | 0 |
| [Warp](https://feedbackbench.com/agents/warp.md) | 15 | Too few posts | – | – | 1 | 0 | 1 |
| [Kiro](https://feedbackbench.com/agents/kiro.md) | 13 | Too few posts | – | – | 0 | 0 | 0 |
| [Grok Build](https://feedbackbench.com/agents/grok-build.md) | 16 | Too few posts | – | – | 0 | 0 | 0 |
| [Augment Code](https://feedbackbench.com/agents/augment.md) | 17 | Too few posts | – | – | 0 | 0 | 0 |

## Posts

Receipts rule: The 5 most recent praise and complaint posts per area (first 700 characters) and 3 per criterion (first 450 characters).

### OpenAI Codex

- Praise, 2026-09-27, r/codex (Reddit): “not only did this work but astra found this post and did the steps itself using computer use lmao. tysm!” [source](https://www.reddit.com/r/codex/comments/1wdfzef/workaround_for_codex_computer_use_timing_out_on/pca61g1/)
- Praise, 2026-09-27, r/codex (Reddit): “yeah this is pretty much the case for me. with the release of opus 5.5, the use of any models from openai becomes trivial because you get such a high level intelligence at a discount. the only things i truly use codex for at the moment are: \- computer use (still miles ahead of claude) \- generative images (for some of my workflows) \- astra for bug finding (still ahead of claude on this according to benchmarks) i’m going to way until devday on” [source](https://www.reddit.com/r/codex/comments/1wrazdi/i_cannot_take_this_anymore/pcb96ez/)
- Praise, 2026-09-27, r/codex (Reddit): “yah that is what i'm planning really, going back and forth between the two, sometimes anthropic models are ahead and sometimes openai are leading, also i find that computer use is superior in codex really regardless of how now slow slo 6 or luna 6. that is maybe the only edge thay have over claude code.” [source](https://www.reddit.com/r/codex/comments/1wrn697/new_openai_product_o_but_not_for_plus_users/pcefpfh/)
- Complaint, 2026-09-27, r/codex (Reddit): “shopping sites block bots so no” [source](https://www.reddit.com/r/codex/comments/1wrn697/new_openai_product_o_but_not_for_plus_users/pcf97vq/)
- Complaint, 2026-09-27, r/codex (Reddit): “not with computer use.” [source](https://www.reddit.com/r/codex/comments/1wrn697/new_openai_product_o_but_not_for_plus_users/pcf9luf/)
- Complaint, 2026-09-27, r/codex (Reddit): “sharing the fix that worked on my mac today, after extension reinstalls, app restarts, new chats, a different project, and rebuilding the browser/chrome plugin cache had all failed. i worked through the diagnosis with codex. the error was: “browser use cannot access [<strict_link> because saved browser permissions could not be verified.” chrome and its tabs were visible, but actual page access failed. the built-in browser failed too. the cause on” [source](https://www.reddit.com/r/codex/comments/1wrtr68/fixed_saved_browser_permissions_could_not_be/)

### Claude Code

- Praise, 2026-09-27, r/ClaudeCode (Reddit): “true, it's good at computer use, graphics etc” [source](https://www.reddit.com/r/ClaudeCode/comments/1wqvwry/just_resubscribed_to_max_after_a_long_time_and/pcaf8cp/)
- Praise, 2026-09-26, r/ClaudeCode (Reddit): “claude can't hear, but it produced my edmish house remix anyway: \~40 iterations of me listening and claude measuring i'd never used strudel (a live-coding music tool in the browser) before this week. i got obsessed with making a bootleg of scissor sisters – it can't come quickly enough, and did the whole thing with claude code. my part: listening and describing what i wanted. claude's part: everything technical. the interesting bit is how it wor” [source](https://www.reddit.com/r/ClaudeCode/comments/1wqhfzz/scissor_sisters_it_cant_come_quickly_enough/)
- Praise, 2026-09-26, @ClaudeDevs (X): “@claudedevs that is awesome. i use remote control a lot, adding a safe way to execute commands on the box (mini terminal ) and communicate secrets would rock. all part of the chat now. @bcherny” [source](https://twitter.com/901081378682531841/status/2103766323436179797)
- Complaint, 2026-09-27, r/ClaudeCode (Reddit): “mine is small but i have used it constantly since i made it, a command that reads my iphone's screen with apple's vision ocr on the mac and gives claude the text plus where to tap, instead of claude looking at a screenshot every single step which was the slowest part of the whole loop for the school emails one, does it add the events straight to the calendar or ask you first?” [source](https://www.reddit.com/r/ClaudeCode/comments/1wrcmjn/what_tool_have_you_built_for_yourself_with_claude/pcch8pz/)
- Complaint, 2026-09-27, @ClaudeDevs (X): “@claudedevs i’m sorry but opus / fable cannot use the computer like astra. astra is absolutely worth using.” [source](https://twitter.com/2067376675432583169/status/2104262447393890595)
- Complaint, 2026-09-27, r/codex (Reddit): “yah that is what i'm planning really, going back and forth between the two, sometimes anthropic models are ahead and sometimes openai are leading, also i find that computer use is superior in codex really regardless of how now slow slo 6 or luna 6. that is maybe the only edge thay have over claude code.” [source](https://www.reddit.com/r/codex/comments/1wrn697/new_openai_product_o_but_not_for_plus_users/pcefpfh/)

### Cursor

- Praise, 2026-09-22, @cursor_ai (X): “@itseduvieira @cursor_ai a browser inside the agent kills the screenshot relay (and a lot of blind css guesses)” [source](https://twitter.com/2007194879781474305/status/2102411755422949506)
- Praise, 2026-09-22, @cursor_ai (X): “@itseduvieira @cursor_ai having the browser built in sounds like a big advantage for ui work. it cuts down on screenshot relays and guesswork.” [source](https://twitter.com/2160573998/status/2102464451354271902)
- Praise, 2026-09-16, @cursor_ai (X): “😱 <strict_link> @cursor_ai @bot down i really am impressed with how good it has been for me the past month. ✅ quick answers - not like claude 😴 ✅ true computer usage - bash, browser etc all in it's own sandbox or mine <strict_link>” [source](https://twitter.com/63228936/status/2100272839114821995)
- Complaint, 2026-09-25, @cursor_ai (X): “browser use by subagents in @cursor_ai is broken <strict_link>” [source](https://twitter.com/467130927/status/2103491442882830418)
- Complaint, 2026-09-25, @cursor_ai (X): “@fatih @timekeepur @cursor_ai my biggest issue with this is i can’t open the terminal for my local agents. it somehow still opens a cloud terminal. this sucks because i want to run the mobile simulator from it to test the changes but have to open a separate terminal app to start it.” [source](https://twitter.com/48140152/status/2103623815297175555)
- Complaint, 2026-09-22, @cursor_ai (X): “ok @cursor_ai is lagging behind now. no remote control (cloud sessions doesnt count), no browser control (user's own one with logins), i use these 2 so much in other tools but missing in cursor” [source](https://twitter.com/1306443936194596865/status/2102290226945331335)

### Google Antigravity

- Praise, 2026-09-27, @antigravity (X): “people shit on @antigravity a lot - but its very usable free student plan gives a ton of usage its one of the fastest models and it's great when you need smt simple just done. its also pretty good for browser use <strict_link>” [source](https://twitter.com/2031861244118945792/status/2104295241935143042)
- Praise, 2026-09-26, r/google_antigravity (Reddit): “for what you get for free it's crazy good a dependable. i get alot of use from just having it use the browser and try to make the crap i've built fail. the flash models are good for that and you get some opus and sonnet 4.6 as well. also i have it as basically my google hands to control and use my gemini notebooks and all the other free google stuff like stitch. don't sleep on using it with google vids for some easy editing and video creation. i” [source](https://www.reddit.com/r/google_antigravity/comments/1wqmd4n/why_is_the_antigravitycli_so_underrated/pc89ome/)
- Praise, 2026-09-26, @antigravity (X): “@ivanleomk @antigravity browser access changes the ceiling” [source](https://twitter.com/1513567206352764929/status/2103784906656526397)
- Complaint, 2026-09-27, @antigravity (X): “talking shit while your own clientbase is starting to revolt is a good one. i like how you talked shit about claudes computer use when it's actually ahead of openai's computer use. you only care about what works on your shitty little macbook, boasting about a feature that 0 of your customers using windows can even use. you are so incredibly out of touch and should not be commenting on the bleeding edge of consumer facing products. you constantly” [source](https://twitter.com/1642487601449029632/status/2104018673463882156)
- Complaint, 2026-09-27, @antigravity (X): “@antigravity we need in app browser” [source](https://twitter.com/229365187/status/2104053240841093320)
- Complaint, 2026-09-26, @antigravity (X): “@ivanleomk @antigravity open-source it pls the mcp integration inside antigravity is an absolute disaster to use. third-party solutions like agent browser and browser use just don't hold a candle to codex's browser control.” [source](https://twitter.com/2080328280927055875/status/2103687543564927358)

### Devin

- Praise, 2026-09-16, @cognition (X): “yoooo devin has access to a mac now it can build a mobile app and send you a recording of the app working end-to-end and you don't even need your laptop open sending you a testflight link to test is crazy as well amazing day for app builders @cognition is cooking <strict_link>” [source](https://twitter.com/1762175022527877120/status/2100056831548932272)
- Praise, 2026-09-16, @cognition (X): “@cognition a mac vm is what makes ios work for an agent. linux-only setups could write swift, they just couldn’t prove the app launched.” [source](https://twitter.com/1127587680001437697/status/2100058272606900397)
- Praise, 2026-09-16, @cognition (X): “@cognition neat, apple scripts better for computer use than linux, so smart move.” [source](https://twitter.com/2946409799/status/2100166759919521980)
- Complaint, 2026-09-23, @cognition (X): “@fei2411 @devindesktop @cognition @devinai the blogger wants the official website to properly improve its desktop version, as there is no computer use. browser automation. multiple agent sessions are still lagging.” [source](https://twitter.com/1178669733572423680/status/2102558550576992283)
- Complaint, 2026-09-16, @cognition (X): “hey @jkelleyrtp this is great! qq, on desktop app in macos environment, should i be able to click around in the desktop on things? because i can't. trying to figure out what the issue issue tldr: devin started my app in the macos, and it works but i can't click on anything in the window like i can in an ubuntu env or previously in namespace. can dm video example if you want.” [source](https://twitter.com/414497508/status/2100243449471488188)
- Complaint, 2026-09-15, @cognition (X): “the built-in browser function of @cognition is too rudimentary compared to cursor and codex. the element selector is difficult to use, it cannot reference area screenshots, and the agent cannot visually operate the built-in browser, which greatly limits the convenience of front-end development and the capabilities of agent automated testing.” [source](https://twitter.com/1236527033435312128/status/2099904812749844666)

### OpenCode

- Praise, 2026-09-27, r/opencodeCLI (Reddit): “depending on what you want to do, open code gui can be quite useful occasionally. i've been debugging a server in cli with qwen 3.8 and when it came to testing the web frontend, i switched to gui and opened the same chat there, so it kept the context, but had access to an internal webbrowser it could capture through vision and interact with, while reading js logs. although i like tui (i mean... i'm a linux guy, i love that terminal), there are so” [source](https://www.reddit.com/r/opencodeCLI/comments/1wrgmfe/whats_the_deal_with_openchamber/pcgrtgm/)
- Praise, 2026-09-25, r/opencode (Reddit): “just dont bro, i wouldnt build this if you could just use playwright and have a smooth sail. try it out, or just read the readme if you dont wanna commit and see the difference yourself. its far better as an agent browser driver than playwright” [source](https://www.reddit.com/r/opencode/comments/1wpr9sx/deepseek_v41_flash_using_a_browser_to_get_content/pbxxcx3/)
- Praise, 2026-09-25, r/opencode (Reddit): “reloading the site should fix it, i also do rarely get blocked in reddit, a reload usually fixes it, if that dosent, tell me and i will fix the stealth problem of reddit too edit: i am working on a fix for those who encounter this problem, and even testing if i get blocked if do many task with reddit in the current version, my agent has done like 30+ task and it has not yet been blocked at all, so its probably your ip or something, i will add a s” [source](https://www.reddit.com/r/opencode/comments/1wpr9sx/deepseek_v41_flash_using_a_browser_to_get_content/pby045b/)
- Complaint, 2026-09-26, @opencode (X): “please @thdxr help fix @opencode so we can use the <strict_link> drivers for computer use without these errors every time. computer use is a critical feature these days and this issue is holding opencode back. <strict_link> <strict_link> <strict_link>” [source](https://twitter.com/20052949/status/2103796383450825001)
- Complaint, 2026-09-25, r/opencode (Reddit): “yeah i got blocked from reddit after opening like 6 pages lol” [source](https://www.reddit.com/r/opencode/comments/1wpr9sx/deepseek_v41_flash_using_a_browser_to_get_content/pbxy14s/)
- Complaint, 2026-09-22, r/codex (Reddit): “yeah i tried opencode as well, then there was the whole debacle with axing of cheap models and 15$ usage stuff, can't be asked to deal with that, codex just works only issue is computer use doesn't work on linux yet” [source](https://www.reddit.com/r/codex/comments/1wnh014/absolutely_no_fking_way_the_pricing_is_wowww/pbgsbqr/)

### Amp

- Praise, 2026-09-25, @AmpCode (X): “@iannuttall @ampcode running directly inside your active chrome session is the real unlock. remote sandboxes choke on oauth, 2fa, and stored state. we found the exact same thing building dassi—local cdp control with a clean human handoff for passwords/2fa cuts out 95% of the auth friction.” [source](https://twitter.com/2098583078805573632/status/2103475225669308566)
- Praise, 2026-09-25, @AmpCode (X): “@iannuttall @ampcode computer use on your own machine is huge, way better than a sandboxed one” [source](https://twitter.com/1619563255621627904/status/2103507618073657647)
- Praise, 2026-09-17, @AmpCode (X): “well. i just made an @ampcode plugin using cua driver + typesafe and now i have a pretty fast computer use and i am using deepseek for it working wonders. already testing it to do motion capture and rotoscoping in after effects :)” [source](https://twitter.com/2931128860/status/2100708556320219145)
- Complaint, 2026-09-25, @AmpCode (X): “i love @ampcode and have been using it for 90% of my coding tasks in the last few weeks. but i really wanted to use my mac mini for computer use because it's already logged in to chrome and has all my apps. turns out you can just ask amp to build that for you! amp does have it's own computer use but for my use at least it was a bit slow, i couldn't paste passwords properly, and it wasn't easy to get it logged in to all my chrome tabs. so i chatte” [source](https://twitter.com/9111552/status/2103429597496778899)
- Complaint, 2026-09-14, @AmpCode (X): “the only thing missing now from @ampcode is a bit of a nice in-app browser experience. but overall, this is pretty top notch way to split my openai and grok super heavy subs (i like to have different views of the world)” [source](https://twitter.com/279111065/status/2099491434562756926)
- Complaint, 2026-09-08, @AmpCode (X): “@ampcode by any chance can we have smt like computer use inside amp? thats prob the only thing making me go to codex app these days (use it to control after effects, editting software, photoshop and things like thar)” [source](https://twitter.com/2931128860/status/2097289514125336919)

### Pi

- Praise, 2026-09-22, @pidotdev (X): “@pidotdev <strict_link> not just a gui solution - but expandeture of capabilities. such as new sub-agent system, browser annotation etc.” [source](https://twitter.com/2059310670131195904/status/2102406216131756254)
- Praise, 2026-09-08, @pidotdev (X): “@pidotdev @openai 这轮更看 computer use。真能点、键、拖完一整套流程，比聊天窗口里嘴上聪明管用。pi 里先跑一遍，比干看榜单有用。” [source](https://twitter.com/2088999223241101312/status/2097293014473617541)
- Complaint, 2026-09-16, r/PiCodingAgent (Reddit): “what the best way to to "computer use" with the pi agent. of course it makes sense to use api/cli/mcp servers, playwright browser use and all these things, but sometimes computer use makes all the difference and it is the only thing that keeps me in codex app.” [source](https://www.reddit.com/r/PiCodingAgent/comments/1whxzpc/computer_use_with_pi_agent/)
- Complaint, 2026-09-08, @pidotdev (X): “currently working with it inside pi now, but i had a lot of issues of astra not being able to properly control my utm virtual machine. will try using the official codex app to see if it works better there but hopefully it will work well in both pi and the official chatgpt/codex app soon.” [source](https://twitter.com/174263209/status/2097301443547971874)
- Complaint, 2026-09-08, @pidotdev (X): “@pidotdev @openai frontier leading for computer use is a meaningful claim right up until you hand it a dynamic dropdown or an auth flow.” [source](https://twitter.com/1910204476973096960/status/2097329557804081326)

### Cline

- Praise, 2026-09-20, @cline (X): “nice move putting jev-driven browser work inside the desktop app. background chrome for web tasks is a strong default. the gap i still hit is everything that is not a browser: native installers, ide dialogs, apps with no dom. for those i want the same agent loop on the os window, with shell when a command exists and ui only when it does not. shipping an early windows desktop agent in that lane (chat + terminal + operator). browser plugins cover a” [source](https://twitter.com/1416221221432172550/status/2101494793075363880)
- Praise, 2026-09-19, @cline (X): “@cline the browser capability is a big unlock, but the config boundary matters too. keeping the gateway key in a local file while the agent handles the browsing task is a much cleaner trust model than pasting secrets into prompts.” [source](https://twitter.com/2092135923169579008/status/2101152370801426934)
- Praise, 2026-09-19, @cline (X): “@cline browser powered jev in cline sounds like a huge productivity boost!” [source](https://twitter.com/2008812694628175872/status/2101385019348627924)
- Complaint, 2026-09-18, @cline (X): “@cline also bring computer use feature in it like codex thats very usefull” [source](https://twitter.com/1846962321442091009/status/2101026989276860880)
- Complaint, 2026-09-16, r/CLine (Reddit): “godd job!!! you need to add browser inside!!” [source](https://www.reddit.com/r/CLine/comments/1wg8pqc/introducing_cline_desktop_a_native_interface_for/pa7dtt0/)

### GitHub Copilot

- Praise, 2026-09-23, r/GithubCopilot (Reddit): “you don’t have a build pipeline that deploys the server? you should be able to code both locally and just have the ai commit/deploy the web server and wait until it’s deployed to run tests. ask it to setup a github action that deploys the server. or scripting if it’s a local server. also the cloud github copilot can easily run tools like playwright which is like running a command line chrome to run and verify websites.” [source](https://www.reddit.com/r/GithubCopilot/comments/1wo80xe/how_to_pair_two_copilots_working_on_different/pbkrz35/)
- Praise, 2026-09-09, r/GithubCopilot (Reddit): “are you using the vscode agent? copilot cli does not lack computer use or remote control, and in my experience is much better than opencode go (apart from cost to consumer)” [source](https://www.reddit.com/r/GithubCopilot/comments/1wbbhzs/open_code_vs_github_cp/p8qwjcn/)
- Praise, 2026-09-09, r/ClaudeCode (Reddit): “this is awesome, i had been using playwright in vscode <> claude <> copilot at work and see it do the work in the vscode shared browser was shocked when i couldn't do the same with the claude chat plugin in vscode (how i run claude for home projects)” [source](https://www.reddit.com/r/ClaudeCode/comments/1wbg30i/i_made_an_mcp_app_so_claude_code_can_record_edit/p8rtv46/)
- Complaint, 2026-09-09, r/GithubCopilot (Reddit): “where else can i use my gh copilot tokens? github copilot is cool but not for genetic workflows, lacks computer use, remote control, etc etc. seems like there should be better alternatives” [source](https://www.reddit.com/r/GithubCopilot/comments/1wbbhzs/open_code_vs_github_cp/)

### Conductor

- Praise, 2026-09-27, @conductor_build (X): “@conductor_build now records stuff from computer mode for testing, it's awesome.” [source](https://twitter.com/1060104763885477888/status/2104227093194461548)
- Complaint, 2026-09-13, @conductor_build (X): “for managing worktrees and complex setups it’s fantastic. it feels super quick and intuitive. i love having multiple chats within one worktree too. they feel more involved and first class citizens rather than codex “side chats” that feel mainly for asking questions. it uses then codex cli so it can still access computer use etc… the in-app browser is lacking though. codex is still king there and the plugin ecosystem feels seamless within codex to” [source](https://twitter.com/50570112/status/2099046528161567197)

### Zed

- Complaint, 2026-09-02, r/ZedEditor (Reddit): “like directly access and intract with native applications like cad, eda (electronic design automation) and others” [source](https://www.reddit.com/r/ZedEditor/comments/1w5eiox/is_zed_support_ai_native_computer_use_tool_api/)

### Factory

- Praise, 2026-09-27, @FactoryAI (X): “@factoryai @enoreyes agent foxxy just aced its browser test by autonomously uploading a video to youtube—acting just like a human! 🤖🔥 <strict_link> <strict_link>” [source](https://twitter.com/2093686029186396160/status/2104206958949810353)

### Warp

- Complaint, 2026-09-15, @warpdotdev (X): “@warpdotdev computer-use that needs a human every step is labor with extra latency. who owns silent failure on day 3?” [source](https://twitter.com/1977033941514072064/status/2099852678754975934)
