Skip to content

Topic

Speech-to-speech models

Models that take audio in and return audio out without a separate text stage, built for live spoken interaction rather than transcription.

Current clusters