Original Summary

I made SpokenLog around a simple requirement: keep recordings useful without committing to one transcription provider or an app subscription. The same library holds the recording and its transcript. You can use local SenseVoice, Moonshine or Whisper, or connect your own Groq/Cloudflare credentials. Then review the audio by timestamp, edit the text and export TXT, SRT, VTT or JSON. The emphasis is that workflow and provider choice, not a new speech model or a claim to beat other apps on accuracy. Local transcription has no service usage quota; optional cloud services have their own limits. Windows demo: https://github.com/user-attachments/assets/cf12c70e-a913-4209-887f-9ed36139ffa6 Source and downloads: https://github.com/Blue-B/SpokenLog Early-release caveats: local input currently needs WAV and models download separately. UI is English/Korean. Windows is unsigned; macOS is not notarized, and the iOS package needs re-signing. AGPL-3.0-only. If you try it, I would especially like to hear where the record-to-transcript workflow feels awkward.   submitted by   /u/Agitated_Chair_4977 [link]   [comments]


  • 情报分类:技术学习与提效
  • 分类依据:内容涉及技术、AI、软件工具或工程实践
  • 信息来源:Reddit · SideProject
  • 发布时间:2026/9/28 12:18:34