- SignalDesk4天前
Original Summary
I wanted to get better at speaking Polish, and reading aloud on my own only got me so far. I had no way to tell how close any given word was to how it should actually sound, so I built something that scores each word 0–100 the moment you stop speaking. Tuning the scoring took longer than building the rest of it. Azure's Pronunciation Assessment scores against the reference text using forced alignment, which means you can say the wrong thing and still pass since it's matching what it expects to hear. Adding a second, unguided STT call and diffing that against the reference is what finally got the scoring to a place I was happy with. It's been live for a little while now, free to use, and no account needed: SpeakingMonkey . 33 locales, with Polish, Norwegian and English being the ones I've put the most hours into myself. What I'd like feedback on specifically: is a 0–100 number per word actually useful, or would you rather just see the color and the sound you got wrong? And would you want to choose your own sentences, or is a curated set better?   submitted by   /u/Kriss3rn [link]   [comments]
- 情报分类:商业与市场研究
- 分类依据:内容涉及商业、投资或市场动态
- 信息来源:Reddit · SideProject
- 发布时间:2026/9/17 21:01:03
- 暂无回复