Home / Speech to Text
Speech to Text
Transcribe audio and video files to text with AI. Supports 99 languages, multiple models, and automatic language detection.
Upload Audio or Video
Drag & drop your file here, or browse
Supports MP3, WAV, FLAC, OGG, M4A, WEBM, MP4. Max 50MB.file.mp3
0 MBTranscription Result
Credit Cost
2 credits per minute of audio
Billed based on audio duration, rounded up to the nearest minute.
Available Models
OpenAI Whisper
OpenAI's robust speech recognition API supporting 57 languages.
OpenAI GPT-4o Transcribe
OpenAI's latest transcription model with improved accuracy and structured output.
Deepgram Nova-3
Deepgram's fastest and most accurate STT with real-time streaming and diarization.
Google Cloud Chirp 3
Google's latest speech recognition with 100+ language support and auto-punctuation.
Microsoft Azure STT
Azure speech recognition with 100+ languages, custom models, and real-time streaming.
ElevenLabs Scribe
ElevenLabs' speech-to-text with speaker diarization and 99 language support.
Best For
Meetings, interviews, podcasts, lectures, support calls, and searchable audio archives.
Frequently Asked Questions
Need a specific TTS workflow?
Compare providers, test voices, then run it through one brokered API.