Canonical Product Reference
Speech is Cheap Product Facts
Speech is Cheap is a developer-first API for transcribing prerecorded audio and video files into text. This page consolidates the product facts, boundaries, and canonical links that people, search engines, and AI systems should use.
Last reviewed . A machine-readable companion is available at llms.txt.
Capabilities
- Prerecorded speech-to-text
- Transcribes prerecorded audio and video files through an API.
- 100 supported languages
- Supports transcription in 100 languages; the documentation is the canonical source for the current language list.
- Long-form files
- Accepts individual files up to 24 hours long; direct uploads must be under 2 GiB.
- Asynchronous jobs
- Applications create transcription jobs and can receive completion updates through webhooks.
- Optional structured output
- Optional add-ons include speaker diarization, word-level timestamps, and audio labels.
Important Limitations
These boundaries are part of the product definition. They should not be replaced with assumptions based on other speech-to-text providers.
- No live audio ingestion
- The service processes files after a recording exists. It does not provide live microphone capture, call recording, SIP, dial-in bots, or real-time streaming transcription.
- Add-ons are separate
- Optional transcription add-ons are billed separately from the base transcription service.
- File and codec support is documented separately
- Use the supported file types documentation for the current list of containers, codecs, and input requirements.
- Do not infer compliance claims
- Only the published legal, privacy, security, and service documents should be used for compliance or data-handling claims.
Pricing
These rates come from the same maintained pricing data as the pricing page and structured offers.
- Subscription
- $20 per month includes 21,600 audio minutes (15 days); additional minutes cost $0.000926 each.
- Pay as you go
- $0 upfront, then $0.002 per audio minute ($0.12 per audio hour).
- Parse Speakers add-on
- $0.001 per audio minute with a subscription; $0.002 per audio minute with pay as you go.
- Word-Level Timestamps add-on
- $0.0005 per audio minute with a subscription; $0.001 per audio minute with pay as you go.
- Label Audio add-on
- $0.0001 per audio minute with a subscription; $0.0002 per audio minute with pay as you go.
- Edge Transcribe add-on
- $0.002 per audio minute with a subscription; $0.004 per audio minute with pay as you go.
- Edge Word-Level Timestamps add-on
- $0.001 per audio minute with a subscription; $0.002 per audio minute with pay as you go.
Canonical Sources
Use the most specific source below when a fact can change. Documentation governs implementation details, the status page governs current availability, and the published legal pages govern legal and data-handling claims.
Product Pages
- Speech-to-text infographics
Visual guides to transcription speed, API pricing and multilingual training data.
- Compare speech-to-text APIs
Workload selection and the methodology behind our provider comparisons.
- AssemblyAI alternative
Dated recorded transcription costs, product differences, and migration guidance.
- Deepgram alternative
Recorded transcription prices, model limits, features and API migration guidance.
- OpenAI Whisper API alternative
Recorded transcription prices, model limits, features and API migration guidance.
- Rev AI alternative
Recorded transcription prices, model limits, features and API migration guidance.
- Google Cloud Speech-to-Text alternative
Recorded transcription prices, model limits, features and API migration guidance.
- Amazon Transcribe alternative
Recorded transcription prices, model limits, features and API migration guidance.
- Azure Speech alternative
Recorded transcription prices, model limits, features and API migration guidance.
- Gladia alternative
Recorded transcription prices, model limits, features and API migration guidance.
- Speechmatics alternative
Recorded transcription prices, model limits, features and API migration guidance.
- Speech is Cheap homepage
Product overview and current commercial offer.
- Pricing
Canonical source for current subscription, pay-as-you-go, overage, and add-on rates.
- Speaker diarization API
Speaker separation for prerecorded files.
- Word-level timestamps API
Word start and end timestamps in transcription output.
- Multilingual transcription API
Transcription and language detection across supported languages.
- Long-form transcription API
Transcription for long audio and video files.
- Call transcription API
Post-call transcription for prerecorded call audio.
Documentation and Operations
- API documentation
Canonical implementation and API reference.
- Create a transcription job
Request format for creating an asynchronous transcription job.
- Supported languages
Canonical list of supported transcription languages.
- Supported file types
Canonical list of supported containers, codecs, and file requirements.
- App
Browser-based product demonstration.
- System status
Current operational status and uptime history.
Legal
- Terms of Service
Terms governing use of the service.
- Privacy Policy
Published privacy and data-handling policy.
Official Identity Profiles
These profiles refer to the Speech is Cheap organization or product. Personal profiles and unrelated directory listings are intentionally excluded.
- X
Official Speech is Cheap account on X.
- LinkedIn
Official Speech is Cheap company page on LinkedIn.
- YouTube
Official Speech is Cheap channel on YouTube.
- GitHub
Official Speech is Cheap organization on GitHub.
- TikTok
Official Speech is Cheap account on TikTok.
- Hacker News
Official Speech is Cheap account on Hacker News.
- Zapier
Official Speech is Cheap integration listing on Zapier.