Speech Text

speechtext.co
Visit site

AI speech to text for audio, video and live recordings

Description

Speech Text — AI Speech to Text for Audio and Video · Convert audio, video, live recordings, or media URLs into editable speech to text transcripts with search, speaker options, and flexible exports. · What Speech Text Does for Speech to Text · Why Use Speech Text for Speech to Text · Speech Text Features for Speech to Text

Speech Text is an online AI speech to text workspace that converts audio, video, live browser recordings, and media URLs into editable transcripts. Powered by OpenAI Whisper technology, it supports over 100 languages with automatic detection or manual selection, speaker labels, and timestamps. Users upload files, record in the browser, or paste a media link, then search, edit, and export transcripts as TXT, SRT, VTT, DOCX, JSON, or PDF. Pro AI tools add summaries, transcript chat, and translation into 100+ languages.

Features

Transcribe uploaded audio and video files
Record speech live in the browser
Transcribe from a media URL
Automatic language detection across 100+ languages
Export as TXT, SRT, VTT, DOCX, JSON, and PDF

Use cases

Meeting notes and documentation
Interview transcription for articles
Podcast repurposing into show notes
Subtitle files for video captions
Searchable archives of calls and lectures

FAQ

Speech Text is an online AI-powered workspace that transcribes audio and video files, browser recordings, or media URLs into editable and searchable text transcripts.

Yes, new users receive starter transcription credits to test the service before subscribing for more capacity.

Accuracy varies with audio quality, speaker clarity, and background noise. Clear source audio yields the most reliable results.

It supports common audio and video formats including MP3, WAV, M4A, MP4, MOV, and WEBM for upload.

Yes, transcripts include a built-in editor for making corrections and fixing wording before export.

You can export transcripts as TXT, SRT, VTT, DOCX, JSON, and PDF files for different use cases.

Yes, Pro and higher plans include speaker labeling to distinguish between different voices in a recording.

Yes, you can paste a supported media URL to transcribe online content without downloading it first.

The core transcription uses OpenAI Whisper technology, and the platform also lists other models like Nemotron 3.5 ASR.

It's used by journalists, podcasters, students, researchers, and businesses for meeting notes, interviews, captions, and archives.

Specs

Type Agent
SectionAI agents
Pricing has a free tier (от $4.9/mo)
Platform API only
Systems api, web
Who forсоздатели контента
Site languageen
Rating0.00 (0 reviews)
Views7
Launched2026-07-25

Integrations

Platforms

Similar in «AI agents»

Submit a site to the catalog

Just send the link — we will work out the rest.

We will review what you send and add it to the catalog if it fits.

Not sure how to implement it? We can help

Tell us about your task — we will pick the tools and suggest where to start.

0 / 5000
Verification code

Fields marked with an asterisk are required. Your data is used only to reply.