Use `streamTranscribe()` to get transcript deltas in real-time as audio arrives, replacing batch-file workflows for live captioning and voice agents.
Summary
Developers building voice interfaces no longer wait for full audio upload before text arrives—critical for responsive live captioning and agent voice modes. Enables voice input to text-based agents without changing agent logic.
Why it matters
Developers building voice interfaces no longer wait for full audio upload before text arrives—critical for responsive live captioning and agent voice modes. Enables voice input to text-based agents without changing agent logic.
Implementation verdict
Replaces previous full-file batch transcription. Requires audio stream input and compatible model (OpenAI's gpt-realtime-whisper confirmed, cross-provider API via AI SDK). Beta status—ready to integrate now for production voice features, but expect API surface changes.
Sources
Dev Signal
Get briefs like this in your inbox — free, every weekday.
100+ sources compressed into one 4-minute read. Ranked, cited, implementation-ready.