Desktop Click

SkillWeb & browsing

Lets your agent click buttons, icons, and menus in desktop applications by name or screen coordinates.

Instructions available. Your AI can read the instructions. Execution depends on the setup they require.

Add ahel to your AI once: Claude, ChatGPT, Cursor, Claude Code or Codex. Then ask it to use this.

Then ask your AI: use the Desktop Click skill

About this skill

Click desktop elements or coordinates. When you need to click buttons/icons in applications, select menu items, or interact with desktop UI. Supports element description, name prefix, or coordinates. For browser webpage elements, use browser_click instead.

What this skill tells your AI

The instructions your AI receives, as published by openakita/openakita in skills/system/desktop-click/SKILL.md and read by ahel’s review.

点击桌面上的 UI 元素或指定坐标。

Parameters

参数类型必填说明
targetstring是元素描述或坐标(如 '确定按钮' 或 '100,200')
buttonstring否鼠标按钮:left, right, middle,默认 left
doubleboolean否是否双击,默认 false
methodstring否查找方法:auto, uia, vision,默认 auto

Target Formats

  • 元素描述:"保存按钮"、"name:确定"
  • 坐标:"100,200"

Find Methods

  • auto: 自动选择(推荐)
  • uia: 只用 UIAutomation
  • vision: 只用视觉识别

Examples

点击按钮(元素描述):

{"target": "确定按钮"}

点击坐标:

{"target": "100,200"}

右键点击:

{"target": "文件图标", "button": "right"}

双击打开:

{"target": "文档.txt", "double": true}

Notes

  • 如果点击的是浏览器内的网页元素,请使用 browser_click
  • 优先使用 UIAutomation(快速准确),失败时用视觉识别

Related Skills

  • browser-click: 点击浏览器网页元素
  • desktop-type: 输入文本
  • desktop-find-element: 先查找元素

Signals

GitHub stars
2k
Forks
280
Last commit
Sep 2026
Advanced
Item type
skill
Key
desktop-click
Source
github.com/openakita/openakita