Choose a model for the work you need to doMe
← Use cases
Speech to text: how to choose a model
Speech recognition transcribes recordings. File transcription, live recognition, speaker separation, and timestamps have different requirements.
Before integrating
- 01Check audio format, sample rate, and duration limits.
- 02Distinguish uploaded files from live streaming connections.
- 03Verify language, timestamp, and speaker separation support.
- 04Review names, terminology, and low-quality audio.
A formal integration contract is not yet verified for this scenario. The catalog supports discovery only.
Example request: “Transcribe a meeting recording into text”Find models for this scenario →
Related integrations
Browse the catalog →No complete integration is available for this scenario yet. Explore the provider and model directories and check official API documentation.
