Supported Products
Follow this AWS Marketplace link to see the Deepgram products that are supported on the SageMaker AI platform. No login to your AWS account is required to view this public AWS Marketplace website.
Product listings
Section titled “Product listings”For Speech-to-Text (STT), Deepgram publishes a separate product listing for each combination of:
- Model family — such as Nova-3 or Flux
- Language coverage — monolingual or multilingual
- Processing mode — streaming or batch
For example, Deepgram Voice AI- Nova-3 Monolingual Speech-to-Text (STT) Streaming is one listing.
For Text-to-Speech (TTS), Deepgram publishes a single product listing per model family (such as Aura-2), with no separate listings for language coverage or processing mode. Subscribe to and deploy a SageMaker Endpoint for each product you wish to utilize. Your application code will need to route requests to the SageMaker Endpoint for the product you wish to run inference against.
Within a listing, individual languages are delivered as versions of the model package. A monolingual listing may offer one version covering English and French, and another covering Vietnamese and Thai. Read the version name and its release notes to understand the set of languages each version provides, and select the version that matches the languages you need when deploying.
Language Requests: If there is a transcription language that is not currently available on the AWS Marketplace, please work with your account manager to request additional language models to be added. For a full list of the Deepgram supported transcription languages, check out this document. You can also view the Changelog to see recent product announcements.
Instance types
Section titled “Instance types”Every Deepgram SageMaker product requires a GPU-accelerated instance. Choose an instance type from the table for the product you are deploying, and request SageMaker quota for it before you create an endpoint.
| Product | Recommended | Also supported | Not supported |
|---|---|---|---|
| Nova-3 STT | ml.g6.2xlarge |
ml.g7.2xlarge, ml.g7e.2xlarge, ml.g6e.2xlarge, ml.g5.2xlarge, ml.g4dn.2xlarge |
— |
| Flux STT | ml.g6.2xlarge |
ml.g7.2xlarge, ml.g7e.2xlarge, ml.g6e.2xlarge, ml.g5.2xlarge |
ml.g4dn.* (no sm_75 kernel) |
| Aura-2 TTS | ml.g6.12xlarge |
ml.g7.12xlarge, ml.g7e.12xlarge, ml.g5.12xlarge, ml.g6e.12xlarge, ml.g4dn.12xlarge |
Single-GPU types (Aura-2 needs 2+ GPUs) |
| Flux TTS (Aura-3) | ml.g6e.2xlarge |
ml.g7.2xlarge, ml.g7e.2xlarge, ml.g6.2xlarge |
ml.g5.*, ml.g4dn.* |
The host driver your instances boot with is set separately from the instance type. Current Deepgram model packages require a recent inference AMI version — see Inference AMI Versions.
Related resources
Section titled “Related resources”references/products.jsonin the dg-sagemaker repository — a machine-readable equivalent of this page (product IDs, invocation modes, supported instance types, and required parameters)- Subscribe on AWS Marketplace
- Deploy Deepgram on Amazon SageMaker
- Requesting SageMaker Quota
- Deployment Environments