Tool overview
Deepgram is listed under Audio AI tools.
What is Deepgram?
Deepgram provides speech-to-text, text-to-speech, audio intelligence, and a real-time Voice Agent API. Developers can use public APIs and SDKs for streaming or recorded audio and can deploy selected models in private cloud environments.
Best for
Developers and enterprises building production speech and voice applications
Who is it for?
Decision note
Rebuilt from the original export under the complete V412/V411 factual-source-verification workflow. Preview only. Apply remains blocked until Failures = 0, Warnings = 0, Unmapped = 0, Missing = 0; explicit clears are reviewed; image import is disabled or V380 accepts the asset; a representative WordPress edit screen is compared with the export and proposed row; and a post-Apply zero-change Preview succeeds.
Key features
Streaming and prerecorded speech recognition
Aura text-to-speech models
Real-time Voice Agent API
Language, diarization, and audio intelligence
Use cases
Transcribe calls and meetings
Build low-latency voice agents
Generate speech from application text
Analyze recorded audio at scale
Pros
- No minimum pay-as-you-go contract
- $200 signup credit without a card
- Enterprise and private-cloud deployment options
Limitations
The signup credit is not an ongoing free plan or time-limited software trial.
Language and model compatibility must be checked for each endpoint.
Pricing details
Billing options
Pricing note
Pay As You Go includes a one-time $200 credit with no card required. Nova-3 prerecorded speech-to-text starts around $0.0043 per minute, Aura text-to-speech starts at $0.015 per 1,000 characters, and the full Voice Agent stack is listed at $4.50 per hour. The credit is not a permanent free plan.
Supported languages
- English
Integrations
Amazon Connect
AWS SageMaker
AWS Bedrock
Twilio
LiveKit
Pipecat
Please log in to join the discussion.