{"repo":"yangzhuxinyzx/auto-openclaw","free":true,"listed":false,"github":"https://github.com/yangzhuxinyzx/auto-openclaw","clone":"git clone https://github.com/yangzhuxinyzx/auto-openclaw.git","description":"Give OpenClaw the power to control your desktop — UI-TARS VLM-driven GUI automation","language":"TypeScript","stars":18,"topics":["ai-agent","bytedance","computer-use","desktop-automation","electron","gui-agent","gui-automation","natural-language","nut-js","openclaw"],"license":"Apache-2.0","category":"ai-agents","readme_excerpt":"auto-openclaw 让 OpenClaw 拥有操作电脑的能力 — 基于 UI-TARS 视觉语言模型的桌面自动化集成。 本项目基于 UI-TARS Desktop（v0.3.0-beta.11）修改，将其 CLI 作为 OpenClaw 的 skill 接入，让 OpenClaw 能够像人一样看屏幕、点鼠标、打键盘。 工作原理 前置要求 - Node.js = 20 - pnpm 9.10.0 - OpenClaw 已安装 - VLM API（推荐火山引擎 doubao-seed-2-0-pro-260215） 安装 1. 克隆并构建 验证： 2. 配置 VLM 模型 创建 /.ui-tars-cli.json ： 支持任何 OpenAI 兼容的视觉语言模型 API。 3. 接入 OpenClaw 将 skill 文件复制到 OpenClaw workspace： OpenClaw 会自动发现并加载，无需额外配置。 使用 在 OpenClaw 对话中直接下达 GUI 任务： 手动测试 输出格式 JSON Lines，每行一个事件： event 含义 关键字段 ------- ------ ---------- screenshot 截屏 loop , width , height prediction 模型决策 action type , thought , action inputs error 出错 message done 结束 status , summary , screenshotPath 退出码： 0 成功 / 1 出错 / 2 需人工 / 3 用户中止 相对原版的改动 基于 UI-TARS Desktop v0.3.0-beta.11： - packages/ui-tars/cli/ — 新增 --output json 结构化输出、退出码、操作摘要、最终截图保存 - packages/ui-tars/operators/nut-js/ — type 操作剪贴板 fallback、scroll 改进、wait 缩短 致谢 - UI-TARS Desktop — ByteDance 开源的 GUI Agent 桌面应用（Apache-2.0） - OpenClaw — 开源自托管 AI Agent License Apache-2.0（继承自 UI-TARS Desktop）","default_branch":null,"files":null,"tree":[],"storefront":"/r/yangzhuxinyzx","claimed":false,"request_supported":{"post":"https://gitbuyer.com/r/yangzhuxinyzx/auto-openclaw/request-supported","requests":0},"note":"indexed from public GitHub; nothing is for sale on this page. Clone it from GitHub. Paid listings live at /search."}