Two speech-to-text models from Google, one for transcribing recordings and one for streaming live audio.
Turning recorded calls, interviews, and meetings into accurate text, or building voice input into a product.
You need transcription that gets names and jargon right, or you need text appearing as someone speaks.
Anyone processing recorded audio at volume, and developers building voice-first interfaces.
Pick a topic and leave with something you created.
Get new module drops and weekly AI strategy.