Getting it into your agent
One page per mod, every tool's command on it. A separate URL per tool would split the same page into five that compete with each other.
npx skills add Syysean/web-ai-skills --skill gemini-workergit clone --depth 1 https://github.com/Syysean/web-ai-skillsWrote this? Show the measurements
A badge with what this costs and how it scanned, read live from this page, so it follows the numbers instead of freezing them. Markdown for a README, HTML for a documentation site or a project page.
[](https://agentmods.dev/skills/syysean/web-ai-skills/gemini-worker)<a href="https://agentmods.dev/skills/syysean/web-ai-skills/gemini-worker"><img src="https://agentmods.dev/badge/skills/syysean/web-ai-skills/gemini-worker/github.svg" alt="Measured on agentmods" height="20"></a>Or the 80×15 button, for a site that already has a row of RSS and ATOM ones. Only the verdict fits; the numbers stay here.
<a href="https://agentmods.dev/skills/syysean/web-ai-skills/gemini-worker"><img src="https://agentmods.dev/badge/skills/syysean/web-ai-skills/gemini-worker.svg" alt="Reviewed on agentmods" width="80" height="20"></a>What it costs to keep this loaded
Counted locally with the o200k_base tokenizer, which is exact for GPT models; Claude uses its own tokenizer and its counts differ. Treat this as one consistent yardstick across the catalogue rather than a bill. Prices are per million input tokens.
| Model | Per session | Once invoked |
|---|---|---|
| Fable 5.1 | $0.00063 | $0.01475 |
| Opus 5 | $0.00032 | $0.00737 |
| Sonnet 5 | $0.00013 | $0.00295 |
| Haiku 4.5 | $0.00006 | $0.00147 |
Grade A, and why
gemini-worker scanned grade A with 0 findings against 26 rules in 11 categories — prompt injection, anti-refusal, data exfiltration, privilege escalation, supply chain, agent snooping, system-prompt leakage, SSRF and excessive agency — measured 11d ago.
A static scan of the body, not an audit. Every finding is printed with the line that produced it so you can judge whether it matters here. A mod is markdown that instructs an agent; that is exactly why what it instructs is worth reading.
Nothing flagged
None of the 26 patterns this scan looks for appear in this file: no shell pipes, no recursive deletes, no credential paths, no hidden text, no instruction-override or anti-refusal phrasing, no agent-config snooping. That is not a guarantee, it is the absence of the things that are checkable.
How it starts
The opening of the file, as written. The whole thing — 215 lines — stays where its author put it; the contents beside it link to each section on GitHub.
Gemini Worker
Send a prompt to Gemini via browser and return the response.
Environment
- Browser: Chrome on localhost:9222 (managed by local-browser-tools)
- Login: Not required — Gemini guest mode is confirmed working
- Tools:
browser-start.js,browser-nav.js,browser-eval.js,browser-screenshot.js
Input
prompt: string # The text to send to Gemini
task_type: string # qa | summary | translation | rewrite | extraction | comparison | brainstorming | analysis
Output
Always output this JSON block first:
{
"site": "gemini",
"task_type": "<task_type>",
"status": "success | partial | error",
"answer": "<extracted response text>",
"retry_count": 0,
"elapsed_ms": 0,
"error_reason": null
}
Then follow with a short human-readable summary (2–3 sentences max).
Workflow
Step 1 — Start browser
node ~/.pi/agent/skills/local-browser-tools/browser-start.js
Expected: ✓ Chrome already running on :9222 or ✓ Chrome started on :9222
If fails: return error status, stop.
Step 2 — Navigate to Gemini
node ~/.pi/agent/skills/local-browser-tools/browser-nav.js https://gemini.google.com/app
After navigation, take a screenshot and verify:
- Input box is visible
- Page is NOT a login wall (a "Sign in" button in the top-right corner is normal and acceptable — guest mode works)
If a full-page login form appears with no input box: return error status, stop.
Step 3 — Locate input box
Try selectors in this order until one works:
[aria-label="Enter a prompt for Gemini"][contenteditable="true"][role="textbox"]div.ql-editor
Use browser-eval.js to check:
!!document.querySelector('[aria-label="Enter a prompt for Gemini"]')
If no selector works after 3 attempts: trigger retry.
Step 4 — Insert prompt
(function() {
const input = document.querySelector('[aria-label="Enter a prompt for Gemini"]')
|| document.querySelector('[contenteditable="true"][role="textbox"]')
|| document.querySelector('div.ql-editor');
if (!input) return "NOT_FOUND";
input.focus();
input.textContent = "";
input.textContent = PROMPT_TEXT;
input.dispatchEvent(new Event("input", { bubbles: true }));
return input.textContent;
})()
What this file has done since we first saw it
Hashed on every crawl. A supply-chain change to an agent config is a question of when, not whether, so the history is kept rather than the latest state alone.
- 11d ago First seen · 215 lines · 63 tokens per session scan A 56ce92075e3b
gemini-worker is a skill published in the GitHub repository Syysean/web-ai-skills (5 stars, last pushed 3mo ago), licensed MIT. It adds 63 tokens to every session and 1,475 once invoked, about $0.0003 per session on Opus 5. A static security scan graded it A with 0 findings. No closer match exists in the catalogue, so it is treated as the original; first seen 2026-08-31.
Other skills, from other repositories
edgeone skill scanner
A local static scanner that checks agent-skill files for security risks before they are installed or used. Static analysis examines files without running them.
sandbase
Access 2,000+ AI models and API tools through one MCP interface for inference, media generation, search, scraping, embeddings, social data, and structured retrieval. Use sandbasediscover before building custom integrations or declaring external data inaccessible; prefer an existing dedicated tool or API key when the…
dev-workflow
The complete development workflow for SkillHub contributors including local dev, staging validation, testing, and PR creation. Ensures agents follow the correct sequence of steps.
video-frames
Extract a single frame from a local video at the first frame, a timestamp, or a zero-based frame index using FFmpeg.
convex-performance-audit
Audits Convex performance for reads, subscriptions, write contention, and function limits. Use for slow features, insights findings, OCC conflicts, or read amplification.
convex-insights
Query a running Convex app's logs + health in natural language (official MCP): failures, slow/expensive functions, deploy causality — scoped, evidence-backed, with a dashboard deep link.