Build AI agents that interact with computers like humans do - viewing screens, moving cursors, clicking buttons, and typing text. Covers Anthropic's Computer Use, OpenAI's Operator/CUA, and open-source alternatives.
49
55%
Does it follow best practices?
Run evals on this skill
Adds up to 20 points to the overall score
View guide
High
Do not use without reviewing
Fix and improve this skill with Tessl
tessl review fix ./skills/computer-use-agents/SKILL.mdThe skill handles credentials insecurely by requiring the agent to include secret values verbatim in its generated output. This exposes credentials in the agent’s context and conversation history, creating a risk of data exfiltration.
The skill requires the LLM to emit actionable JSON (e.g., {"type":"type","text":"..."}) and bash/tool commands that may include user credentials or tokens verbatim (for logins, typing, or API calls), which forces secrets to appear in model outputs and creates an exfiltration risk.
The skill exposes the agent to untrusted, user-generated content from public third-party sources, creating a risk of indirect prompt injection. This includes browsing arbitrary URLs, reading social media posts or forum comments, and analyzing content from unknown websites.
The skill's BrowserUseAgent explicitly captures page snapshots (get_page_snapshot) and feeds page DOM/text/URL into the LLM in run_with_llm (with examples like "Go to weather.com"), and the vision-based agents also pass screenshots to the model, so arbitrary public webpages and user-generated web content are fetched and interpreted to drive actions—creating a clear vector for indirect prompt injection.
a5a6601
If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.