- SignalDesk6小时前
Original Summary
Hi Reddit, I built Airy TTS, a text-to-speech API focused on minimizing end-to-end latency — the time to receive the entire audio, not just the first chunk. Most TTS APIs optimize for time-to-first-audio (TTFA) in streaming setups. But for longer text, there can still be a significant gap between the first chunk arriving and the full audio being ready. Airy TTS is built around the full round-trip instead. Fast end-to-end response • A 495-character input (27s of audio) is fully generated and received in ~0.2s (US-West) • 2,000+ characters/sec throughput for English TTS Low cost • Pricing scales down to $2 / 1M characters at higher volume Built this for anyone from indie devs who want cheap TTS to services that need to generate large volumes of speech quickly. Hope it’s useful: https://airy.so   submitted by   /u/ANLGBOY [link]   [comments]
- 情报分类:技术学习与提效
- 分类依据:内容涉及技术、AI、软件工具或工程实践
- 信息来源:Reddit · SideProject
- 发布时间:2026/9/19 09:21:55
- 暂无回复