Specialist automation
Which AI agent platforms can drive a browser or a desktop?
This is the most specialized axis in the taxonomy and the numbers say so. Of 946 vendors, 105 document full browser, desktop or remote computer control for workflows that cannot be done through stable APIs. 89 document partial coverage and 752 document none.
Every vendor in the index is assessed against the same 14 point taxonomy from public documentation, and no vendor pays for placement. Counts on this page were measured across all 946 public vendors on September 30, 2026.
How the 946 vendors split
No public evidence means the reviewed sources did not document the capability. On this index that is a statement about the evidence, not proof that the capability is absent. See methodology.
What counts as full coverage
Full coverage means documented control of a real interface: a hosted or local browser, a desktop session, or remote computer control the agent drives itself. Partial usually means web scraping, a headless fetch or a third party browser engine wired in as an integration rather than owned by the platform. A terminal, sandbox or code execution environment is autonomous execution but not this axis, an ingestion crawler that builds the knowledge corpus belongs on knowledge grounding, and an agent operating the vendor's own admin screens does not count. Where a vendor documents actions inside a platform that publishes no API but does not state the mechanism, the cell is Partial.
How to read these numbers
Read the 786 as scope rather than as failure. Most agents in the index act through APIs by design, which is faster, more reliable and easier to audit, so an absent score here is usually the correct engineering choice. The browser lane itself documents full coverage for 47 of its 49 members, as it should. The number that carries information is the spread outside that lane: coding at 19 percent, agent builders at 17, multi agent platforms at 15 and agent infrastructure at 14, then a long tail in single digits, with SRE at zero. Screen driving remains a specialist tool bought for systems that never got an API, not a capability spreading across the market.
Leading platform for browser & computer use in each use case
Picked mechanically: the highest total coverage vendor in each lane that documents full evidence on this axis, one vendor per row. Scores are out of 14.
-
1. Appian, for enterprise operations agents
14 / 14Full enterprise operations agents ranking · Compare the whole lane on all 14 axes
-
2. FLOWX.AI, for multi-agent platforms
14 / 14Full multi-agent platforms ranking · Compare the whole lane on all 14 axes
-
3. Gumloop, for GTM and revenue agents
14 / 14Full GTM and revenue agents ranking · Compare the whole lane on all 14 axes
-
4. Mastra, for agent infrastructure platforms
14 / 14Full agent infrastructure platforms ranking · Compare the whole lane on all 14 axes
-
5. ServiceNow, for customer support agents
14 / 14Full customer support agents ranking · Compare the whole lane on all 14 axes
-
6. UiPath, for agent builders
14 / 14Full agent builders ranking · Compare the whole lane on all 14 axes
-
7. GitHub Copilot, for coding agents
13.5 / 14Full coding agents ranking · Compare the whole lane on all 14 axes
-
8. HappyRobot, for voice agents
13 / 14Full voice agents ranking · Compare the whole lane on all 14 axes
-
9. Adopt AI, for browser and computer-use agents
12.5 / 14Full browser and computer-use agents ranking · Compare the whole lane on all 14 axes
-
10. Redbird, for data analyst agents
11.5 / 14Full data analyst agents ranking · Compare the whole lane on all 14 axes
-
11. CoOrdio, for healthcare agents
8.5 / 14Full healthcare agents ranking · Compare the whole lane on all 14 axes
-
12. Assail, for security and SOC agents
8 / 14Full security and SOC agents ranking · Compare the whole lane on all 14 axes
Documented coverage by use case
Share of each lane documenting full coverage on this axis. Vendors that sit in two lanes count in both, the same rule the rankings and matrices use.
| Use case | Full coverage | Share |
|---|---|---|
| browser and computer-use agents | 42 of 42 | 100% |
| coding agents | 17 of 64 | 27% |
| agent builders | 23 of 115 | 20% |
| multi-agent platforms | 9 of 52 | 17% |
| agent infrastructure platforms | 27 of 186 | 15% |
| enterprise operations agents | 27 of 297 | 9% |
| data analyst agents | 4 of 59 | 7% |
| healthcare agents | 4 of 69 | 6% |
| GTM and revenue agents | 5 of 138 | 4% |
| voice agents | 3 of 83 | 4% |
| customer support agents | 4 of 113 | 4% |
| security and SOC agents | 1 of 85 | 1% |
| SRE and DevOps agents | 0 of 37 | 0% |
Recent verified changes from the vendors named above
Capability coverage is not a static picture. These are the most recent sourced change log entries for the platforms listed above, newest first, one per vendor. Scores on this page update as entries like these are verified.
-
GitHub Copilot browser/computer use
High impactGitHub Copilot can now operate desktop apps on macOS and Windows, in public preview in Copilot CLI and the Copilot app. It reads what is on screen, clicks, types, scrolls and drags through workflows, including software with no API or MCP integration.
October 1, 2026 · Verified · All GitHub Copilot changes
-
Mastra deployment / data residency
Medium impactAny environment on Mastra's hosted platform can now get a Postgres database that accepts connections only from inside the same private network. It is created with one CLI command or at deploy time, placed in the environment's region and wired in automatically.
October 1, 2026 · Verified · All Mastra changes
-
UiPath human approval / guardrails
High impactAdministrators can now bring their own safety vendor into UiPath's AI Trust Layer, in preview, so agent guardrails run under the customer's own vendor agreement and region. Azure AI Language, Azure Content Safety and Noma are supported, checks can be enforced across the organization, and each evaluation is recorded in the agent trace.
October 1, 2026 · Partially Verified · All UiPath changes
-
Gumloop agent capability
Medium impactUsers can now text Gumball, Gumloop's personal agent, over iMessage, RCS or SMS to start a task, get replies in the same thread and approve steps by text. It is a public beta that admins can switch off by role.
September 29, 2026 · Partially Verified · All Gumloop changes
-
FLOWX.AI agent capability
High impactFlowX.AI 5.13.0 adds an AI agent to its Designer that surveys an existing app, drafts a plan from a plain language request, then builds or edits processes, screens, workflows and data types. It runs only after the builder confirms the plan and reviews each step, works on its own branch, and refuses destructive changes such as deleting a process with live instances.
September 28, 2026 · Verified · All FLOWX.AI changes
Full change log · updated weekly across the whole index
Questions buyers ask
Should I expect this from a general platform?
No. If your target systems have APIs, an agent that uses them is the better architecture. This axis matters when a critical system has no API and the only path is the interface a person would use.
How is this different from the browser agent category?
This page measures the capability across every vendor in the index. The browser and computer use lane is the set of vendors whose whole product is that capability, ranked and compared on its own pages.
Why do some infrastructure vendors score full?
Because they sell hosted browsers and control interfaces for other people's agents. They are suppliers to the capability rather than agents that use it, and the grid records the documented capability either way.
The other 13 axes
No single axis decides a shortlist. Buyers who care about this one usually check testing, debugging & optimization and integrations & tool calling next, or open the full taxonomy to see how the 14 axes fit together.