- SignalDesk2 hr ago
Original Summary
Cactus launched their 17 MB Whistle speech model two days ago, so I immediately turned it into the thing I actually wanted: free on-device dictation for my Mac, Wispr Flow-style. Hold Right Option, speak, release — text lands wherever the cursor is. Hard parts, honestly: (1) macOS Accessibility permissions bind to the resolved Python binary, not the symlink you see in Finder; (2) NSPanel hides itself when the app isn't active, which a background daemon never is — one-line fix, brutal to find; (3) release-to-text latency was all process-spawn overhead (osascript), not the model — a 12-second clip transcribes in 0.12 s. Stack: Python + sounddevice + pynput + PyObjC overlay, LaunchAgent autostart, ~600 lines. MIT: github.com/rodriveiga01/syrinx Question for builders: clipboard-paste + synthetic Cmd+V feels hacky but it's instant and universal — is there a cleaner text-insertion path on macOS I'm missing?   submitted by   /u/rodriveiga1000 [link]   [comments]
- 情报分类:技术学习与提效
- 分类依据:内容涉及技术、AI、软件工具或工程实践
- 信息来源:Reddit · SideProject
- 发布时间:2026/10/4 01:32:13
- No replies yet