🖥️
AutomationPython

Computer Use Agents

by davila7

Computer Use Agents is an Automation skill for Claude Code, published by davila7 in claude-code-templates.

30.2K stars3.4K forkson davila7/claude-code-templatesAdded 2026/08/14Repository updated 2026/08/14
anthropicanthropic-claudeclaudeclaude-code
Install in seconds
Install Computer Use Agents
Copy Computer Use Agents into your Claude Code skills folder. Run the command in your terminal, or review the source on GitHub before installing.
terminal
npx degit https://github.com/davila7/claude-code-templates/tree/main/cli-tool/components/skills/ai-research/computer-use-agents ~/.claude/skills/computer-use-agents

Requires Node.js. Downloads this skill only — not the rest of the repository — into your Claude Code skills folder.

Without Node.js

git clone https://github.com/davila7/claude-code-templates.git

Clones the whole repository, then copy the skill’s own directory into your skills folder yourself.

In this catalog

Source file
cli-tool/components/skills/ai-research/computer-use-agents/SKILL.md in davila7/claude-code-templates
Installs to
~/.claude/skills/computer-use-agents
Collection
One of 25 skills cataloged from this repository
Category
Automation1523 skills

What Computer Use Agents does

Computer Use Agents build AI agents that interact with computers like humans do — viewing screens, moving cursors, clicking buttons, and typing text. Use it when building desktop automation, vision-based control, or GUI agents with sandboxing and security.

Computer Use Agents is cataloged under Automation on DirSkills. Computer Use Agents comes from a repository tagged anthropic, anthropic-claude, claude and claude-code.

Documentation

README

Computer Use Agents

Patterns

Perception-Reasoning-Action Loop

The fundamental architecture of computer use agents: observe screen, reason about next action, execute action, repeat. This loop integrates vision models with action execution through an iterative pipeline.

Key components:

  1. PERCEPTION: Screenshot captures current screen state
  2. REASONING: Vision-language model analyzes and plans
  3. ACTION: Execute mouse/keyboard operations
  4. FEEDBACK: Observe result, continue or correct

Critical insight: Vision agents are completely still during "thinking" phase (1-5 seconds), creating a detectable pause pattern.

When to use: ['Building any computer use agent from scratch', 'Integrating vision models with desktop control', 'Understanding agent behavior patterns']

This is the opening of the README. Read the full README on GitHub.

Commands Computer Use Agents provides

Slash commands named in this skill’s SKILL.md, listed in the order they first appear.

  • /app
  • /run

Frequently asked about Computer Use Agents

  • What else does davila7 publish alongside Computer Use Agents?

    Computer Use Agents is one of 25 skills that DirSkills catalogs from davila7/claude-code-templates, the repository it ships in. Its siblings there include AI Agents Architect, Agent Evaluation and Agent Management. Each one is a separate skill with its own page in this directory, installs the same way Computer Use Agents does, and is maintained by davila7 in that same repository. The rest of the collection is listed on the davila7/claude-code-templates page.

  • How does Computer Use Agents compare to other Automation skills?

    Computer Use Agents ranks #92 by stars among the 1523 Automation skills in this catalog. The most-starred ones next to it are Autonomous Loops, Autonomous Agent Harness and Automation Audit Ops. DirSkills ranks by the star count of the repository each skill ships in, so that order reflects how popular those repositories are rather than any review of Computer Use Agents against them. Open each page to compare what they document and how they install.

More from davila7/claude-code-templates

Computer Use Agents is one of 25 skills cataloged on DirSkills from davila7/claude-code-templates.

See all 25 skills
🤖
2w ago

AI Agents Architect

AI Agents Architect designs and builds autonomous AI agents, covering tool use, memory systems, planning strategies, and multi-agent orchestration. Use it when building or debugging AI agents that need function calling, planning loops, and controlled autonomy.
AI Engineering
30.2K3.4K
🧪
2w ago

Agent Evaluation

Agent Evaluation tests and benchmarks LLM agents using behavioral contracts, capability assessments, reliability metrics, and adversarial testing to catch issues before production. Use it when evaluating agent reliability, designing benchmarks, or monitoring production agents.
Quality
30.2K3.4K
🤖
2w ago

Agent Management

Agent Management creates, manages, and orchestrates AI agents through the AI Maestro CLI, covering agent lifecycle tasks such as create, hibernate, wake, rename, export/import, and plugin management.
AI Engineering
30.2K3.4K
🤖
2w ago

Agent Manager

Agent Manager starts, stops, monitors, and assigns tasks to multiple local CLI agents running in tmux sessions, with cron-friendly scheduling. Use it when you need to run agents in parallel and tail their logs.
Automation
30.2K3.4K
🧠
2w ago

Agent Memory MCP

Agent Memory MCP provides a persistent, searchable memory bank for AI agents, exposing MCP tools to search, write, read, and analyze project knowledge. Use it when an agent needs long-term memory synced with project documentation.
AI Engineering
30.2K3.4K
🧠
2w ago

Agent Memory Systems

Agent Memory Systems describes architectures for short-term, long-term, and working memory in AI agents, including vector store selection, chunking strategies, and retrieval patterns. Use it when designing or debugging agent memory to prevent retrieval failures that look like intelligence failures.
AI Engineering
30.2K3.4K