Text-to-speech and transcription via AI SDK 7 with low-latency streaming and word-level timestamps; billing begins September 19 unless you use the `-free` suffix.
Summary
Removes vendor lock-in friction for audio features—test production models at zero cost before committing to per-character/per-hour rates. Speech-to-text returns word-level timing, unblocking real-time captioning workflows.
Why it matters
Removes vendor lock-in friction for audio features—test production models at zero cost before committing to per-character/per-hour rates. Speech-to-text returns word-level timing, unblocking real-time captioning workflows.
Implementation verdict
Replaces Fish Audio SDK calls with unified AI SDK functions (`generateSpeech`, `transcribe`). Requires Node.js 18+, one npm install, and model-name awareness for post-trial billing. Worth trying now if you're evaluating audio infrastructure; use `-free` suffix to auto-cutoff on September 19.
Sources
Dev Signal
Get briefs like this in your inbox — free, every weekday.
100+ sources compressed into one 4-minute read. Ranked, cited, implementation-ready.