Deepgram Release Notes

Follow

148 release notes curated from 226 sources by the Releasebot Team. Last updated: Oct 2, 2026

Get this feed:
  • Oct 2, 2026
    • Date parsed from source:
      Oct 2, 2026
    • First seen by Releasebot:
      Oct 2, 2026
    Deepgram logo

    Deepgram

    October 2, 2026

    Deepgram updates Nova-3 keyterms mid-stream speech-to-text.

    Update Nova-3 Keyterms Mid-Stream

    Speech-to-Text

    Original source
  • Sep 30, 2026
    • Date parsed from source:
      Sep 30, 2026
    • First seen by Releasebot:
      Sep 30, 2026
    Deepgram logo

    Deepgram

    September 30, 2026

    Deepgram adds Flux TTS inline controls for pause and pronunciation in text-to-speech.

    Flux TTS Inline Controls: Pause and Pronunciation

    Text-to-Speech

    Original source
  • All of your release notes in one feed

    Join Releasebot and get updates from Deepgram and hundreds of other software products.

    Create account
  • Sep 25, 2026
    • Date parsed from source:
      Sep 25, 2026
    • First seen by Releasebot:
      Sep 28, 2026
    Deepgram logo

    Deepgram

    September 25, 2026

    Deepgram adds mid-stream numeral toggling on Flux STT Speech-to-Text.

    Toggle Numerals Mid-Stream on Flux STT

    Speech-to-Text

    Original source
  • Sep 24, 2026
    • Date parsed from source:
      Sep 24, 2026
    • First seen by Releasebot:
      Sep 25, 2026
    Deepgram logo

    Deepgram

    September 24, 2026

    Deepgram improves Nova-3 speech-to-text models for Flemish, German (Switzerland), Lithuanian, and Portuguese.

    Nova-3 Improved Models for Flemish, German (Switzerland), Lithuanian, and Portuguese

    Speech-to-Text

    Original source
  • Sep 22, 2026
    • Date parsed from source:
      Sep 22, 2026
    • First seen by Releasebot:
      Sep 22, 2026
    Deepgram logo

    Deepgram

    September 22, 2026

    Deepgram improves Nova-3 speech-to-text models for nine languages, including Danish, Italian, Polish, Urdu, and Vietnamese.

    Nova-3 Improved Models for Danish, Estonian, Flemish, Italian, Lithuanian, Macedonian, Polish, Urdu, and Vietnamese

    Speech-to-Text

    Original source
  • Similar to Deepgram with recent updates:

  • Sep 18, 2026
    • Date parsed from source:
      Sep 18, 2026
    • First seen by Releasebot:
      Sep 18, 2026
    Deepgram logo

    Deepgram

    SEP 18, 2026

    Deepgram corrects Browser Agent SDK, client-side VAD, and Voice Agent token minting TTL usage.

    Voice Agent

    Correction: Browser Agent SDK has no client-side VAD

    Voice Agent

    Correction: token minting uses ttl_seconds

    Original source
  • Sep 17, 2026
    • Date parsed from source:
      Sep 17, 2026
    • First seen by Releasebot:
      Sep 17, 2026
    Deepgram logo

    Deepgram

    September 17, 2026

    Deepgram launches Nova-3 Pharma speech-to-text for English pharmaceutical use cases.

    Nova-3 Pharma: Speech-to-Text for Pharmaceutical Use Cases (English)

    Speech-to-Text

    Original source
  • Sep 17, 2026
    • Date parsed from source:
      Sep 17, 2026
    • First seen by Releasebot:
      Sep 17, 2026
    Deepgram logo

    Deepgram

    Introducing Nova-3 Pharma: The First Speech-to-Text Model Purpose-Built for the Pharmaceutical Industry

    Deepgram introduces Nova-3 Pharma, a pharma-specific speech-to-text model built for accurate drug name and medical vocabulary transcription. It ranks #1 in drug name recognition and overall accuracy across batch and streaming, and is available now via hosted API and self-hosted deployments.

    Deepgram sets a new standard for pharmaceutical speech recognition with the first pharma-specific speech-to-text model, ranking #1 in both drug name recognition and overall transcription accuracy across batch and streaming.

    In pharmaceutical workflows, getting a medication name wrong isn’t just another transcription error. It can introduce incorrect information into critical workflows, from prescription refill automation and pharmacy IVRs to clinical documentation. A transcript can score highly on overall accuracy, but an error on a medication name can carry far greater consequences than errors elsewhere in the transcript.

    That’s why we built Nova-3 Pharma, the first speech-to-text model purpose-built for the pharmaceutical industry.

    Nova-3 Pharma is specifically designed to accurately transcribe drug names, pharmaceutical terminology, and medical vocabulary while maintaining the same high-quality transcription performance Deepgram customers expect from Nova-3. In our latest competitive benchmarks, Nova-3 Pharma ranks #1 in drug name recognition and #1 in overall transcription accuracy across both batch and streaming.

    With Nova-3 Pharma, developers can build pharmacy IVRs, prescription pickup and refill automation, member services and contact center voice agents, and clinical documentation applications where drug name accuracy is critical.

    It’s available now for both batch and streaming transcription through Deepgram’s hosted API and self-hosted deployments.

    High accuracy for pharmaceutical terminology

    General-purpose speech-to-text models are designed to perform well across a broad range of speech. Medical models add specialized clinical vocabulary. But pharmaceutical speech introduces another challenge: thousands of medication names that can be uncommon, phonetically similar, and difficult for general-purpose and medical STT systems to recognize reliably.

    For pharmaceutical applications, aggregate accuracy alone doesn’t tell the full story. Critical terminology like medication names requires a higher level of precision.For example, if a patient says a medication name like “Benlysta” or “Dofetilide” and the model transcribes a different medication, a strong overall accuracy score can obscure a critical error.

    Nova-3 Pharma is specifically designed for this level of specialized accuracy. The model combines Deepgram’s medical speech foundation with pharma-specific training to improve recognition of drug names and pharmaceutical terminology while maintaining strong general transcription performance.

    The result is more accurate transcription where pharmaceutical terminology matters most.

    Nova-3 Pharma is #1 in drug name recognition

    Overall transcription accuracy and drug name recognition measure two different dimensions of performance. Word Error Rate (WER) measures accuracy across the full transcript, while Keyword Recognition Rate (KRR) measures whether specific terms spoken in the audio are correctly captured. For pharmaceutical speech, we evaluate both, using KRR to measure drug name recognition directly.

    Nova-3 Pharma is trained on more than 5,000 drug names to improve recognition of pharmaceutical terminology. We evaluated the model against leading speech-to-text and medical speech-to-text models using real English medical audio across therapy, telehealth, clinical dictation, triage, and other medical use cases.

    Nova-3 Pharma ranked #1 in drug name KRR across every model evaluated in both batch and streaming:

    • 91.59% drug name KRR in batch
    • 91.28% drug name KRR in streaming

    Nova-3 Pharma is #1 in overall transcription accuracy

    Nova-3 Pharma pairs industry-leading drug name recognition with industry-leading overall transcription accuracy. On the same benchmark, it delivered the lowest Word Error Rate of every model evaluated in both batch and streaming:

    • 10.17% WER in batch
    • 11.69% WER in streaming

    Together, the results show that Nova-3 Pharma leads on both drug name recognition and overall transcription accuracy.

    Power pharmaceutical applications with Nova-3 Pharma

    Nova-3 Pharma gives developers a specialized speech-to-text foundation for voice applications across pharmaceutical workflows.

    Build pharmacy IVRs and prescription refill automation that accurately capture medication names. Power payer and member-services voice agents handling medication-related requests. Transcribe clinical documentation where drug name accuracy is critical.

    Deploy Nova-3 Pharma now

    Nova-3 Pharma uses the same architecture and compute footprint as Nova-3 Medical and is available through Deepgram’s hosted API and self-hosted deployment.

    It supports both streaming and batch transcription in eight English variants:
    en, en-US, en-GB, en-AU, en-CA, en-IE, en-IN, en-NZ.

    Get started with:
    model=nova-3-pharma

    Try Nova-3 Pharma in the API Playground
    Read the docs
    Contact Sales

    Original source
  • Sep 15, 2026
    • Date parsed from source:
      Sep 15, 2026
    • First seen by Releasebot:
      Sep 15, 2026
    Deepgram logo

    Deepgram

    Deepgram's India Endpoint Is Now Generally Available

    Deepgram launches its India endpoint as generally available, bringing fully managed voice AI with in-country storage and inference for speech-to-text, text-to-speech, and Voice Agent workloads. Customers can use the same API, models, and SDKs with a simple base URL change.

    Deepgram's global expansion began with Deepgram Dedicated and our EU endpoint, then Australia. India is the next region.

    The Deepgram India endpoint is now generally available, giving teams a fully managed way to run voice AI with storage and inference inside India.

    Today, the Deepgram India endpoint is generally available to every customer. Any organization with Indian data-residency requirements, whether based in India or serving Indian users from abroad, can now run speech-to-text, text-to-speech, and Voice Agent workloads on Deepgram, with storage and inference inside India, using the same API, models, and SDKs available on our Global, EU, and Australia endpoints.

    India has been among the most-requested regions from our customers, with demand concentrated in banking, payments, insurance, and the software vendors that serve them. This launch is a direct response to that demand.

    For regulated industries in India, where voice data is stored and processed is a procurement requirement rather than a preference. Data protection law applies broadly, and the binding localization obligations sit in sector regulation covering payments and insurance. Teams in those sectors consistently tell us they go further than the rules strictly require, because auditors scrutinize any customer data that leaves the country. For them, keeping audio and transcripts in country is what separates a deployment that passes review from one that stalls.

    Until now, teams building on Deepgram in India could route audio to our Global or regional endpoints, which introduced cross-border exposure, or deploy Deepgram self-hosted within their own Indian infrastructure, which requires the engineering capacity to operate it in production. The India endpoint provides a fully managed, fully onshore option that uses the same Deepgram Voice AI API as our Global, EU, or Australia endpoints.

    What GA Means for Your Workloads

    The India endpoint is production-ready and available to all customers, with no waitlist and no enterprise-only restriction. It runs in AWS's ap-south-2 region (Hyderabad), and pricing matches our Global, EU, and Australia rates at launch.

    Adoption requires a single change. Point your integration at api.in.deepgram.com, and your existing API keys and SDK integrations continue to work. There is no separate account, no migration project, and no change to how you call the API.

    Where Your Data Is Processed and Stored

    By default, storage and inference for your audio, transcripts, and speech output happen in India. Opting out of model improvement (mip_opt_out=true) extends that to all processing and access.

    The residency commitment is specific. For speech-to-text and text-to-speech on Deepgram models, storage and inference happen on Indian infrastructure (AWS ap-south-2, Hyderabad). Two boundaries are explicit, so compliance teams can review the full picture:

    • Voice Agent LLM. If a Voice Agent uses a managed third-party LLM (the reasoning step), that provider processes the request outside India. Confirm the provider's residency and processing guarantees independently; Deepgram runs listen and speak in-country.
    • Operational metadata and billing are processed in the US.

    For the strictest requirements, where regulated entities avoid any customer data leaving the country, opt out of model improvement (mip_opt_out=true). That keeps customer content out of any training workflow, so no content is accessed or processed across Indian borders for any purpose.

    Supported APIs and Models

    The India endpoint provides the same API surface as our Global, EU, and Australia regions. Speech-to-text is available at /v1/listen and /v2/listen, text-to-speech at /v1/speak and /v2/speak, Voice Agent at /v1/agent/converse, and text intelligence at /v1/read.

    Deepgram model availability is at parity with our EU and Australia regions, so the speech-to-text and text-to-speech models teams already run on Deepgram are available in India without changes.

    Running Voice Agent on the India endpoint keeps Deepgram's speech-to-text and text-to-speech onshore. The agent's LLM step runs on the model provider you select or provide, and that provider determines where the reasoning is processed. Confirm the provider's residency and processing guarantees independently. Teams with the strictest requirements typically bring their own model, hosted in India, and point the agent at it.

    How to Use the India Endpoint

    Set the base URL to the India endpoint in your existing client. The rest of your code stays the same.

    Voice Agent connections use api.in.deepgram.com/v1/agent/converse. All customer payloads route through Indian infrastructure.

    Built for Regulated Industries

    Indian teams evaluating voice AI have real in-country options, and the differences between them are worth understanding. Some global providers offer India as a storage region while inference continues to run elsewhere, so processing still crosses borders. Others make an in-country deployment available only through infrastructure the customer deploys and operates. Deepgram runs storage and inference in India as a fully managed service, with no infrastructure for the customer to run, and without asking teams to trade model quality for a local footprint.

    That distinction matters most in banking, payments, insurance, and the vendors building for them, where an onshore guarantee separates a deployment that passes review from one that stalls. Deepgram holds SOC 2 Type I and Type II certifications, and the India endpoint extends that compliance posture with in-country processing and storage.

    Start Building in India Today

    The India endpoint is live now and available to every Deepgram customer.

    • Read how regional endpoints work →
    • India endpoint quickstart →

    To move a workload to India, update your base URL to api.in.deepgram.com, or contact your account team to plan the transition.

    The India endpoint is part of a broader expansion. Over the past year, Deepgram has extended where voice AI runs, from Dedicated deployments to our EU endpoint, then Australia, and now India, while continuing to expand the languages our models support. The goal is consistent: Deepgram should run where your business runs, in the language your customers speak. India is the next step, and more regions will follow.

    Original source
  • Sep 14, 2026
    • Date parsed from source:
      Sep 14, 2026
    • First seen by Releasebot:
      Sep 14, 2026
    Deepgram logo

    Deepgram

    September 14, 2026

    Deepgram highlights SDK releases for developer tools.

    SDK releases

    Developer Tools

    Original source
  • Sep 10, 2026
    • Date parsed from source:
      Sep 10, 2026
    • First seen by Releasebot:
      Sep 11, 2026
    Deepgram logo

    Deepgram

    September 10, 2026

    Deepgram releases @deepgram/react 0.2.0 for Voice Agent.

    @deepgram/react 0.2.0

    Voice Agent

    Original source
  • Sep 8, 2026
    • Date parsed from source:
      Sep 8, 2026
    • First seen by Releasebot:
      Sep 8, 2026
    Deepgram logo

    Deepgram

    SEP 8, 2026

    Deepgram adds Hold a function call until the user's turn is confirmed Voice Agent

    Hold a function call until the user's turn is confirmed

    Voice Agent

    Original source
  • Sep 3, 2026
    • Date parsed from source:
      Sep 3, 2026
    • First seen by Releasebot:
      Sep 3, 2026
    Deepgram logo

    Deepgram

    September 3, 2026

    Deepgram adds Kazakh to Nova-3 and improves speech-to-text models for Estonian, Hebrew, Latvian, Lithuanian, Macedonian, Malay, and Polish.

    Nova-3 Adds Kazakh, Plus Improved Models for Estonian, Hebrew, Latvian, Lithuanian, Macedonian, Malay, and Polish

    Speech-to-Text

    Original source
  • Aug 31, 2026
    • Date parsed from source:
      Aug 31, 2026
    • First seen by Releasebot:
      Sep 2, 2026
    Deepgram logo

    Deepgram

    August 31, 2026

    Deepgram expands Flux TTS Voice Agent with a wider speed range.

    Wider speed range for Flux TTS

    Text-to-Speech

    Voice Agent

    Original source
  • Aug 28, 2026
    • Date parsed from source:
      Aug 28, 2026
    • First seen by Releasebot:
      Aug 28, 2026
    Deepgram logo

    Deepgram

    August 28, 2026

    Deepgram adds flexible turn-taking control for Flux Speech-to-Text and improves Nova-3 for 10 languages.

    Flexible turn-taking control for Flux

    Speech-to-Text
    Voice Agent

    Nova-3 Model Improvements for Bulgarian, Croatian, Estonian, Georgian, Italian, Latvian, Lithuanian, Malay, Marathi, and Telugu

    Speech-to-Text
    Voice Agent

    Original source
Releasebot

Curated by the Releasebot team

Releasebot is an aggregator of official release notes from hundreds of software vendors and thousands of sources.

Our editorial process involves the manual review and audit of release notes procured with the help of automated systems.