First question
What can it touch?
The practical difference is not only whether the agent can browse or click. It is which websites, files, accounts, and review steps sit inside its reach.
AI Resources
A focused map for agents, models, and tooling that point AI systems at browsers, desktops, or visual interfaces.
Computer-use tools can touch accounts, files, websites, and private data. Use this page to narrow what to explore next, then check permissions, terms, and review steps before trying anything important.
Questions to check
These checks frame the source-linked Resources below. They do not rank products or cover every option.
First question
The practical difference is not only whether the agent can browse or click. It is which websites, files, accounts, and review steps sit inside its reach.
Human review
Look for permissions, confirmations, logs, screenshots, and ways to stop or correct the agent before it changes something important.
Testing path
Use throwaway tasks, test accounts, and manual review before connecting anything private, paid, or hard to undo.
Coverage and freshness
These groups are selective starting points, not a complete directory. The date reflects the newest included Resource’s LifeHubber added date, not a recheck of every linked source. Check the original source for current setup, terms, limits, privacy, access, costs, and behaviour.
Fresh in this topic
Recently added Resources from the groups below.
Browser and desktop control
Use this group when the agent needs to browse, click, type, observe a screen, or automate a visual workflow.
browser-use/browser-use
Browser-first navigation, clicking, typing, and custom tools keep the comparison on website control rather than a broader desktop runtime.
iFurySt/open-codex-computer-use
An MCP service with setup paths across Windows, macOS, Linux, Codex, and other clients shows a computer-control layer that plugs into an existing MCP-capable agent.
trycua/cua
Cua pairs a background desktop driver and agent sandboxes with benchmark and SDK layers for isolated computer environments that also need an evaluation path.
Hcompany/holo31
Model sizes from 0.8B to 35B-A3B and local quantized checkpoints let web, desktop, or mobile control be matched to available hardware.
bytedance/UI-TARS-desktop
A desktop application with vision-language control and local or remote operation adds a visible GUI-agent workspace beyond a browser library.
nico-martin/gemma4-browser-extension
Running through Transformers.js and WebGPU keeps this browser-agent experiment on-device when extension constraints and page-local models shape the choice.
alibaba/page-agent
In-page JavaScript control and text-based DOM interaction show natural-language operation embedded directly into a web interface.
allenai/molmoweb
Multimodal navigation from natural-language instructions makes visual page understanding the main execution path.
lightpanda-io/browser
A headless browser built for automation brings infrastructure footprint and agent-oriented execution into the comparison.
Also in AI
Keep the thread going with AI Guides for decision habits for messy AI choices, AI Access for free and low-cost ways to compare AI model access, AI Ballot for a clearer view of what readers are leaning toward.