Original Summary

A few weeks ago I posted here about renting a 4090 to train a model that could pull my voice out of noisy restaurant recordings. I put a free demo in the browser, and since then, people have run over 1,000 audio files through it. I’ve kept working on the model and built a web app around it. You give it a short sample of someone’s voice and a recording with multiple people talking; it returns a track with that speaker isolated. I also trained a second model for enhancing rough voice memos to sound like a studio podcast recording. It can generate the full and dry sound that's only present with a professional setup. Both are free to try: https://crele.co/ If you want, leave some thoughts on the tools or any questions about the training.   submitted by   /u/premier_slack [link]   [comments]


  • 情报分类:综合情报
  • 分类依据:内容未命中明确的垂直分类规则,归入综合情报
  • 信息来源:Reddit · SideProject
  • 发布时间:2026/9/25 20:06:36