Original Summary

I built AI Vision around one small loop: select part of a normal webpage, ask Gemini what it means, and follow up while the screenshot stays in context. It is useful for explaining a chart or error, copying text from an image, summarizing a supported page, or comparing tabs. To use it, simply open AI Vision on a regular HTTP/HTTPS page, drag a region, ask Gemini, then use the answer field for a follow-up. Setup: install, paste your own Gemini API key from Google AI Studio. I’m looking for specific feedback on the first screenshot action and the key requirement: what feels unclear when you try it? Store: https://chromewebstore.google.com/detail/ai-vision-gemini-screensh/ghmmlbclopoakmjjbkkmoefjldgjimgk Demo and setup: https://stiwarilbj.github.io/AI_Vision/   submitted by   /u/Traditional_Yam9806 [link]   [comments]


  • 情报分类:技术学习与提效
  • 分类依据:内容涉及技术、AI、软件工具或工程实践
  • 信息来源:Reddit · SideProject
  • 发布时间:2026/9/13 22:50:37