
Speech Text
AI speech to text for audio, video and live recordings
- Pricing
- Freemium
- Category
- Voice AI Agents
- Access
- Closed Source
- Industry
- Horizontal
What is Speech Text?
Speech Text is an online AI speech to text workspace that converts audio, video, live browser recordings, and media URLs into editable transcripts. Powered by OpenAI Whisper technology, it supports over 100 languages with automatic detection or manual selection, speaker labels, and timestamps. Users upload files, record in the browser, or paste a media link, then search, edit, and export transcripts as TXT, SRT, VTT, DOCX, JSON, or PDF. Pro AI tools add summaries, transcript chat, and translation into 100+ languages.
Key features
- Transcribe uploaded audio and video files
- Record speech live in the browser
- Transcribe from a media URL
- Automatic language detection across 100+ languages
- Export as TXT, SRT, VTT, DOCX, JSON, and PDF
Use cases
- Meeting notes and documentation
- Interview transcription for articles
- Podcast repurposing into show notes
- Subtitle files for video captions
- Searchable archives of calls and lectures
Speech Text alternatives
Other voice ai agents agents worth comparing.