‹ 首页

gemini-browser-image

@qianleigood · 收录于 昨天 · 上游提交 1 个月前

Drive the logged-in Gemini web app in a real browser to generate or edit images without API keys. Use when browser persistence, existing login state, uploads, or real Gemini page behavior matter more than direct API access.

适合你,如果不想设置 API 密钥,直接使用 Gemini 网页生成图像

/ 通过 npx 安装 校验哈希
npx oh-my-skill add qianleigood/crawclaw/gemini-browser-image
/ 通过 bash 安装
curl -fsSL https://oh-my-skill.com/install.sh | bash -s -- qianleigood/crawclaw/gemini-browser-image
/ 已经装过?验证本机副本,不用重装
npx oh-my-skill verify qianleigood/crawclaw/gemini-browser-image
安装目标可用 --agent / --scope 或 --to 明确指定;省略时只会在唯一已存在的 agent 目录上自动选择,零命中或多命中会停止并提示。content_hash 缺失或不一致均拒装。
30GitHub stars
~251最小装载
~1.1K含声明引用
~5.9K文本包总量
索引托管

怎么用

商店整理自技能原文 · 版本 f27e175 · 表述以原文为准
它做什么

装上后,Claude会驱动已登录的Gemini网页,在真实浏览器中根据你的描述生成或编辑图像,无需API密钥。完成生成后,它会将图像保存到本地。

什么时候触发

当你需要生成图像或基于已有图像进行编辑,且更依赖浏览器登录状态和真实页面行为时触发。

装好后可以这样说
这会让Gemini生成图像并保存到本地。
需要你先上传参考图到Gemini。
会利用Gemini的编辑功能。
技能原文 SKILL.md作者撰写 · MIT · f27e175

Gemini Browser Image

Use this skill for Gemini website image workflows, not API-based image generation.

Use this skill for
  • text-to-image in Gemini Web
  • edit or reference-image generation
  • browser-login-dependent image work
  • saving real local outputs from the Gemini UI
Mandatory workflow
  1. Reuse the logged-in browser profile.
  2. Normalize Gemini into a stable image-generation state before acting.
  3. Collect only minimum prompt/edit constraints.
  4. Submit, wait, and verify a real local file exists before declaring success.
Working rules
  • One Gemini tab per job.
  • Stop on CAPTCHA, login friction, or other human-verification walls.
  • Do not claim success from a visible button alone; confirm a saved local artifact.
Read references as needed
  • references/upload-and-attach.md For uploads and attachment behavior.
  • references/output-capture.md For save, verification, export, and fallback capture rules.
按 MIT 许可原样转载,未经改动 · 在 GitHub 查看 →

评论

登录即可评论;带「已验证安装」的,是发布者名下有本店的安装或持有记录。