Browser
Windows PowerShell toolkit for AI agents: Microsoft 365 (mail, calendar, Teams, SharePoint, OneDrive) via Work IQ + Microsoft Graph, Outlook COM, Edge browser (CDP), desktop, and system control - all returning structured JSON.
npx -y skills add aloth/PowerSkills --skill browserAssembled from the repository path, not quoted from the project. Check it against their README if it does not work.
What its author says it does
Copied from the file, not written here
Edge browser automation via Chrome DevTools Protocol (CDP). List tabs, navigate, take screenshots, extract page content/HTML, execute JavaScript, click elements, type text, fill forms, scroll. Use when needing to control Edge browser, scrape web content, automate web forms, or take browser screenshots on Windows. Requires Edge with --remote-debugging-port=9222.
The file declares its own license as MIT. That is the author’s claim about this one file, and it is not the same thing as the license GitHub reports for the repository, which is listed with the other numbers below.
SKILL.md
2.7 KB, as published. Nobody here has run it
PowerSkills — Browser
Edge browser automation via CDP (Chrome DevTools Protocol).
Requirements
- Microsoft Edge running with remote debugging:
Start-Process "msedge" -ArgumentList "--remote-debugging-port=9222" - Default port configurable in
config.json(edge_debug_port)
Actions
.\powerskills.ps1 browser <action> [--params]
| Action | Params | Description |
|---|---|---|
tabs | List open browser tabs | |
navigate | --url URL | Navigate to URL |
screenshot | --out-file path.png [--target-id id] | Capture page as PNG |
content | [--target-id id] | Get page text content |
html | [--target-id id] | Get full page HTML |
evaluate | --expression "js" | Execute JavaScript expression |
click | --selector "#btn" | Click element by CSS selector |
type | --selector "#input" --text "hello" | Type into element |
new-tab | --url URL | Open new tab |
close-tab | --target-id id | Close tab by ID |
scroll | --scroll-target top|bottom|selector | Scroll page |
fill | --fields-json '[{"selector":"#a","value":"b"}]' | Fill multiple form fields |
wait | --seconds N | Wait N seconds (default: 3) |
Examples
# List open tabs
.\powerskills.ps1 browser tabs
# Navigate and screenshot
.\powerskills.ps1 browser navigate --url "https://example.com"
.\powerskills.ps1 browser screenshot --out-file page.png
# Extract page text
.\powerskills.ps1 browser content
# Run JavaScript
.\powerskills.ps1 browser evaluate --expression "document.title"
# Fill a login form
.\powerskills.ps1 browser fill --fields-json '[{"selector":"#user","value":"alex"},{"selector":"#pass","value":"secret","submit":"#login"}]'
Multi-Tab Support
Pass --target-id (from tabs output) to operate on a specific tab. Without it, actions target the first page.
Fill Fields Format
JSON array of objects with selector, value, and optional submit:
[
{"selector": "#search-input", "value": "PowerShell automation"},
{"selector": "#filter-type", "value": "recent", "submit": "#apply-btn"}
]
Supports text inputs, selects, and checkboxes. Last field can include submit to click a button.