AI speech to text workspace turning audio and video into timestamped, speaker-labelled transcripts.
GPT Transcribe Online: Speech to Text with Subtitles and Speaker Labels · OpenAI's new gpt-transcribe API returns plain text only. This browser tool adds what's missing: SRT/VTT subtitles, speaker labels, timestamps. 5 free minutes. · What gpt-transcribe Is — and What It Leaves Out · Why Use This Instead of Calling the API · Speech to Text AI Features · gpt-transcribe vs Whisper: What Each One Actually Does
GPT Transcribe is an online AI speech to text workspace that converts audio and video recordings into searchable, timestamped transcripts. Every job runs on OpenAI Whisper, so regional accents, industry jargon and background noise hold up far better than in phone dictation tools. Upload an MP3 or MP4, record live in the browser, or paste a media link; GPT Transcribe detects the language from 100+ options, labels each speaker, and drops the transcript into an editor where you can search, fix a misheard name, and export as TXT, SRT, VTT, JSON, PDF or DOCX. Every account starts with free transcription minutes.
| Type | Agent |
| Section | AI agents |
| Pricing | has a free tier (от $4.9/mo) |
| Platform | API only |
| Systems | api, web |
| Site language | en |
| Rating | 0.00 (0 reviews) |
| Views | 11 |
| Launched | 2026-07-30 |