When you need a browser, read this Skill by default. Use it to open and operate websites, fill forms, click buttons, take screenshots, extract page data, sign…
Updated 2026-09-11 21:03:55 +08:00
browser-automation: Vision-driven browser automation using Midscene. Operates from screenshots — no DOM or accessibility labels needed. Runs in headless Puppeteer — does NOT take…; android-device-automation: Vision-driven Android device automation using Midscene. Operates entirely from screenshots — no DOM or accessibility labels required. Can interact with all…; desktop-computer-automation: Vision-driven desktop automation using Midscene. Control your local desktop (macOS, Windows, Linux) or…
Updated 2026-09-08 14:12:42 +08:00
opencli-usage: Use at the start of any OpenCLI session — this is the top-level map of what `opencli` can do, how to discover adapters, what flags and output formats are…; opencli-browser: Use when an agent needs to drive a real Chrome window via opencli — inspect a page, fill forms, click through logged-in flows, or extract data ad-hoc. Covers…; opencli-autofix: Automatically fix broken OpenCLI adapters when commands fail. Load this skill when an opencli command fails — it guides you through co…
Updated 2026-08-31 01:35:37 +08:00