- SignalDesk2026-09-13
Original Summary
I built AI Vision around one small loop: select part of a normal webpage, ask Gemini what it means, and follow up while the screenshot stays in context. It is useful for explaining a chart or error, copying text from an image, summarizing a supported page, or comparing tabs. To use it, simply open AI Vision on a regular HTTP/HTTPS page, drag a region, ask Gemini, then use the answer field for a follow-up. Setup: install, paste your own Gemini API key from Google AI Studio. I’m looking for specific feedback on the first screenshot action and the key requirement: what feels unclear when you try it? Store: https://chromewebstore.google.com/detail/ai-vision-gemini-screensh/ghmmlbclopoakmjjbkkmoefjldgjimgk Demo and setup: https://stiwarilbj.github.io/AI_Vision/   submitted by   /u/Traditional_Yam9806 [link]   [comments]
- 情报分类:技术学习与提效
- 分类依据:内容涉及技术、AI、软件工具或工程实践
- 信息来源:Reddit · SideProject
- 发布时间:2026/9/13 22:50:37
- No replies yet