Choose a model for the work you need to doMe
← Use cases

Speech to text: how to choose a model

Speech recognition transcribes recordings. File transcription, live recognition, speaker separation, and timestamps have different requirements.

Before integrating

  • 01Check audio format, sample rate, and duration limits.
  • 02Distinguish uploaded files from live streaming connections.
  • 03Verify language, timestamp, and speaker separation support.
  • 04Review names, terminology, and low-quality audio.

A formal integration contract is not yet verified for this scenario. The catalog supports discovery only.

Example request: “Transcribe a meeting recording into text
Find models for this scenario →

Related integrations

Browse the catalog →
No complete integration is available for this scenario yet. Explore the provider and model directories and check official API documentation.

Learn about sources and verification →