Gemini Browser Image
SkillWeb & browsingDrive the logged-in Gemini web app in a real browser to generate or edit images without API keys. Use when browser persistence, existing login state, uploads, or real Gemini page behavior matter more than direct API access.
Available today. Use it from your connected AI after setup.
No other account needed.
Connect ahel once, and every AI you use reads what you have installed.
Then ask your AI: use the Gemini Browser Image skill
What this skill tells your AI
The instructions your AI receives, as published by qianleigood/crawclaw in skills-optional/gemini-browser-image/SKILL.md and read by ahel’s review.
Use this skill for Gemini website image workflows, not API-based image generation.
Use this skill for
- text-to-image in Gemini Web
- edit or reference-image generation
- browser-login-dependent image work
- saving real local outputs from the Gemini UI
Mandatory workflow
- Reuse the logged-in browser profile.
- Normalize Gemini into a stable image-generation state before acting.
- Collect only minimum prompt/edit constraints.
- Submit, wait, and verify a real local file exists before declaring success.
Working rules
- One Gemini tab per job.
- Stop on CAPTCHA, login friction, or other human-verification walls.
- Do not claim success from a visible button alone; confirm a saved local artifact.
Read references as needed
references/upload-and-attach.mdFor uploads and attachment behavior.references/output-capture.mdFor save, verification, export, and fallback capture rules.
Signals
- GitHub stars
- 30
- Forks
- 1
- Last commit
- Sep 2026
Advanced
- Catalog kind
- skill
- Gateway key
gemini-browser-image- Source
- github.com/qianleigood/crawclaw