- SignalDesk1 hr ago
Original Summary
I've been working on this for a while and just put it on GitHub (MIT). You give it a topic, it produces a complete long-form documentary: researched script, per-scene visuals, voiceover, music, captions, thumbnail, YouTube metadata. A few things I learned the hard way: Every cited source gets fetched and checked for a real 200 response. Otherwise the script invents references that look completely legit. Most embarrassing bug, now it fails loudly. The script is written in timed beats, underweight ones get extended at ~150 wpm. A "25 minute" video kept coming out at 14 minutes and every scene after that was misaligned. Real stock footage (Pexels/Pixabay/NASA) beats AI visuals for most scenes, which honestly surprised me. It's Python + FFmpeg. Scripting works with a local model (LM Studio/llama-server), narration is local TTS, nothing cloud required. I have no idea if anyone besides me wants this. What would you throw at it first, and what would you expect to break? https://github.com/summitsingh/ai-video-factory   submitted by   /u/summitsc [link]   [comments]
- 情报分类:硬件与数码
- 分类依据:内容涉及硬件、数码产品或通信卡
- 信息来源:Reddit · SideProject
- 发布时间:2026/9/21 02:05:09
- No replies yet