Solo dev, bit over a year on this. songbrain: you upload a song, it tells you which 15 seconds are the strongest and why, then cuts 7 short videos on exactly those seconds for tiktok/reels. why i trust my own numbers: i also run an ai music project on spotify, 2.6M streams, ~7k$ payout, 3.2k followers. that catalog was my test set. i knew which tracks worked so i could check if the analyzer agrees with reality or just sounds smart. for most of the year the whole thing ran on 5 gpu models on my gaming pc. demucs, whisper, clap, panns, qwen2-audio. 158 seconds per song, pc on 24/7, when a container died at 3am nothing worked till i woke up. cloud gpu would have cost more than the project made. then in june gemini 2.5 flash started taking audio directly. sent it a full song as base64, asked f


  • 情报分类:项目价值、技术价值
  • 命中依据:本地GPU音频分析项目及被大模型替代的对比经验
  • 来源:Reddit · SideProject
  • 原作者:/u/Putrid_Ad_8262 https://www.reddit.com/user/Putrid_Ad_8262
  • 发布时间:2026/9/12 16:58:52