- SignalDesk1 hr ago
Original Summary
Hi everyone, I have been working on a project called Lens , an open-source desktop AI assistant for Windows and Linux, and I wanted to share it here. The idea came from my own daily struggles: staring at compiler errors, trying to interpret complicated charts, or reading something I did not understand. Instead of taking a screenshot, opening a browser, and uploading it to an AI chatbot, I thought of building something like Lens that could provide help directly from the desktop. Here’s how it works: Press Alt + L to ask a question about your current screen. Lens captures the active application window and sends it to a locally running AI model. The answer appears in an on-screen chat interface, where you can ask follow-up questions. Press Alt + Shift + S to select a specific area, such as an error message or a section of code. Privacy was my top priority. Unlike tools that continuously capture the screen, Lens only does so when you explicitly ask for help. Screenshots are processed in memory rather than saved as files, and the application is designed to run locally without sending screen content to a cloud AI service. A few technical details: AI inference through Ollama using Qwen3-VL 8B Vision-based understanding without requiring a separate OCR engine like Tesseract A lightweight desktop overlay with multi-turn conversations Automated Windows setup for easier onboarding I would genuinely appreciate feedback on three things: Would you actually use an on-demand screen assistant like this? What would you improve about the workflow or user interface? What applications or use cases should I test it with next? The project is open source, and contributions are welcome. GitHub: https://github.com/Rushu-Tushu/Lens I have also attached a short video showing Lens in action. I would love to hear what you think.   submitted by   /u/Yogiji69 [link]   [comments]
- 情报分类:商业与市场研究
- 分类依据:内容涉及商业、投资或市场动态
- 信息来源:Reddit · SideProject
- 发布时间:2026/10/9 21:09:56
- No replies yet