Volcengine Release Notes

Follow

46 release notes curated from 1 source by the Releasebot Team. Last updated: Sep 1, 2026

Get this feed:
  • Jun 1, 2026
    • Date parsed from source:
      Jun 1, 2026
    • First seen by Releasebot:
      Sep 1, 2026
    Volcengine logo

    Volcengine

    【2026年6月】

    Volcengine 发布豆包音频生成模型1.0和实时语音模型3.0,支持单条Prompt生成影视级音频与实时对话办事。

    豆包音频生成模型1.0

    业内首个可通过单条 Prompt 直接生成影视级音频的模型,支持文本、音频等多模态输入,显著降低声音创作门槛,兼顾高表现力与全要素覆盖。

    豆包实时语音模型3.0

    原生全双工端到端语音大模型(Seeduplex),具备精准遵循、抗干扰、动态判断三大优势,不仅更懂何时听、何时说,还能在对话中直接调用工具完成任务,实现“边听边说边办事”的实时交互体验。

    Original source
  • Apr 1, 2026
    • Date parsed from source:
      Apr 1, 2026
    • First seen by Releasebot:
      Sep 1, 2026
    Volcengine logo

    Volcengine

    【2026年4月】

    Volcengine adds the Doubao machine translation model with 32-language translation, context awareness, and custom terminology.

    豆包机器翻译模型

    支持 32 种语言互译,提供高质量、上下文感知的翻译能力,并支持术语定制,满足专业领域翻译需求。

    Original source
  • All of your release notes in one feed

    Join Releasebot and get updates from Volcengine and hundreds of other software products.

    Create account
  • Mar 1, 2026
    • Date parsed from source:
      Mar 1, 2026
    • First seen by Releasebot:
      Sep 1, 2026
    Volcengine logo

    Volcengine

    【2026年3月】

    Volcengine adds a Doubao voice design model with text and image input for custom character voices.

    豆包音色设计模型

    支持文本与图片双模态输入,为每个角色定制更贴合需求的专属音色。覆盖故事片、AI 电影、AI 短剧等场景,解决音色版权与人物声音同质化问题,让角色表达更丰富、更生动。

    Original source
  • Jan 1, 2026
    • Date parsed from source:
      Jan 1, 2026
    • First seen by Releasebot:
      Sep 1, 2026
    Volcengine logo

    Volcengine

    【2026年1月】

    Volcengine adds Doubao Real-time Speech Model 2.0 with more controllable synthesis, flexible voices, better cloning, and upgraded singing.

    豆包实时语音模型2.0

    在延续 1.0 低时延、高质量体验的基础上,合成效果更可控、音色更灵活、复刻更出色,歌唱表现也全面升级。

    Original source
  • Dec 1, 2025
    • Date parsed from source:
      Dec 1, 2025
    • First seen by Releasebot:
      Sep 1, 2026
    Volcengine logo

    Volcengine

    【2025年12月】

    Volcengine adds Doubao Speech Recognition 2.0 with visual context to improve transcription by understanding images and video.

    豆包语音识别模型2.0

    将语音识别从“听懂文字”升级为“看懂场景”。在语音识别过程中,将上下文理解的范围从纯文本扩展到视觉层面,通过理解图像和视频内容,帮助模型更精准地完成语音转录。

    Original source
  • Similar to Volcengine with recent updates:

  • Nov 12, 2025
    • Date parsed from source:
      Nov 12, 2025
    • First seen by Releasebot:
      Dec 23, 2025
    • Modified by Releasebot:
      Mar 20, 2026
    Volcengine logo

    Volcengine

    【2025.11】

    Volcengine adds new TTS voices, including one audiobook voice and 18 roleplay and emotional voices.

    1. TTS 2.0音色上新 | 新音色*1,新增有声阅读音色:1个。

    2. TTS 1.0音色上新 | 新音色*18,新增角色扮演、多情感音色:18个。

    Original source
  • Nov 1, 2025
    • Date parsed from source:
      Nov 1, 2025
    • First seen by Releasebot:
      Feb 13, 2026
    • Modified by Releasebot:
      May 22, 2026
    Volcengine logo

    Volcengine

    【2025.11】

    Volcengine adds new TTS voices, expanding TTS 2.0 with an audiobook voice and TTS 1.0 with 18 role-play and multi-emotion Chinese voices for richer speech experiences.

    1. TTS 2.0音色上新 | 新音色*1,新增有声阅读音色:1个。

    语种 类别 名称 Speaker

    中文 有声阅读 儿童绘本 zh_female_xueayi_saturn_bigtts

    2. TTS 1.0音色上新 | 新音色*18,新增角色扮演、多情感音色:18个。

    语种 类别 名称 Speaker

    中文 角色扮演 寡言小哥 ICL_zh_male_xiaoge_v1_tob
    中文 角色扮演 清朗温润 ICL_zh_male_renyuwangzi_v1_tob
    中文 角色扮演 潇洒随性 ICL_zh_male_xiaosha_v1_tob
    中文 角色扮演 清冷矜贵 ICL_zh_male_liyisheng_v1_tob
    中文 角色扮演 沉稳优雅 ICL_zh_male_qinglen_v1_tob
    中文 角色扮演 清逸苏感 ICL_zh_male_chongqingzhanzhan_v1_tob
    中文 角色扮演 温柔内敛 ICL_zh_male_xingjiwangzi_v1_tob
    中文 角色扮演 低沉缱绻 ICL_zh_male_sigeshiye_v1_tob
    中文 角色扮演 蓝银草魂师 ICL_zh_male_lanyingcaohunshi_v1_tob
    中文 角色扮演 清冷高雅 ICL_zh_female_liumengdie_v1_tob
    中文 角色扮演 甜美娇俏 ICL_zh_female_linxueying_v1_tob
    中文 角色扮演 柔骨魂师 ICL_zh_female_rouguhunshi_v1_tob
    中文 角色扮演 甜美活泼 ICL_zh_female_tianmei_v1_tob
    中文 角色扮演 成熟温柔 ICL_zh_female_chengshu_v1_tob
    中文 角色扮演 贴心闺蜜 ICL_zh_female_xnx_v1_tob
    中文 角色扮演 温柔白月光 ICL_zh_female_yry_v1_tob
    中文 角色扮演 高冷沉稳 zh_male_bv139_audiobook_ummv3_bigtts
    中文 多情感 深夜播客 zh_male_shenyeboke_emo_v2_mars_bigtts

    Original source
  • Oct 1, 2025
    • Date parsed from source:
      Oct 1, 2025
    • First seen by Releasebot:
      Sep 1, 2026
    Volcengine logo

    Volcengine

    【2025年10月】

    Volcengine adds Doubao Voice Clone Model 2.0 with contextual reference and voice instruction tags for more natural speech.

    豆包声音复刻模型2.0

    支持灵活引用对话上下文,并结合语音指令标签,精准调控语气、节奏与表达,让合成效果更自然、更贴合场景。

    Original source
  • Oct 1, 2025
    • Date parsed from source:
      Oct 1, 2025
    • First seen by Releasebot:
      Mar 8, 2026
    • Modified by Releasebot:
      Jun 10, 2026
    Volcengine logo

    Volcengine

    【2025.10】

    Volcengine adds new TTS voices, including one Cantonese fun-accent voice for TTS 1.0 and 11 fresh voices for TTS 2.0 across general use, video dubbing, and role-play scenarios.

    TTS 1.0音色上新 | 新音色*1,新增趣味口音音色:1个。

    语种 | 类别 | 名称 | Speaker

    中文 | 趣味口音 | 粤语小溏 | zh_female_yueyunv_mars_bigtts

    TTS 2.0音色上新 | 新音色*11,新增通用场景、视频配音、角色扮演音色:11个。

    语种 类别 名称 Speaker

    中文、英语 通用场景 vivi 2.0 zh_female_vv_uranus_bigtts
    中文 视频配音 大壹 zh_male_dayi_saturn_bigtts
    中文 视频配音 黑猫侦探社咪仔 zh_female_mizai_saturn_bigtts
    中文 视频配音 鸡汤女 zh_female_jitangnv_saturn_bigtts
    中文 视频配音 魅力女友 zh_female_meilinvyou_saturn_bigtts
    中文 视频配音 流畅女声 zh_female_santongyongns_saturn_bigtts
    中文 视频配音 儒雅逸辰 zh_male_ruyayichen_saturn_bigtts
    中文 角色扮演 可爱女生 ICL_zh_female_keainvsheng_tob
    中文 角色扮演 调皮公主 ICL_zh_female_tiaopigongzhu_tob
    中文 角色扮演 爽朗少年 ICL_zh_male_shuanglangshaonian_tob
    中文 角色扮演 天才同桌 ICL_zh_male_tiancaitongzhuo_tob

    Original source
  • Oct 1, 2025
    • Date parsed from source:
      Oct 1, 2025
    • First seen by Releasebot:
      Dec 23, 2025
    • Modified by Releasebot:
      Mar 9, 2026
    Volcengine logo

    Volcengine

    【2025.10】

    Volcengine announces new TTS voices in 1.0 and 2.0, adding 12 new voices for playful accents and general video narration roles.

    TTS 1.0音色上新

    • 新音色*1,新增趣味口音音色:1个。

    TTS 2.0音色上新

    • 新音色*11,新增通用场景、视频配音、角色扮演音色:11个。
    Original source
  • Sep 1, 2025
    • Date parsed from source:
      Sep 1, 2025
    • First seen by Releasebot:
      Sep 1, 2026
    Volcengine logo

    Volcengine

    【2025年9月】

    Volcengine launches Doubao Speech Synthesis Model 2.0 for more natural, richer, and more expressive voice output.

    豆包语音合成模型2.0

    供更加自然、更丰富情感、更具有表现力的语音合成效果

    Original source
  • Sep 1, 2025
    • Date parsed from source:
      Sep 1, 2025
    • First seen by Releasebot:
      Apr 23, 2026
    Volcengine logo

    Volcengine

    【2025.09】

    Volcengine adds real-time voice model updates with customizable end detection, longer system prompts, and SC support for web, RAG summaries, and rewrites.

    豆包端到端实时语音大模型:

    通用能力:

    • 用户判停时间支持自定义,end_smooth_window_ms字段用于客户调整判断用户停止说话的时间,默认1500ms,取值范围[500ms, 50s];
    • System prompt(O版本system_role & speaking_style字段;SC版本character_manifest字段)放开字数限制,允许更多信息的输入;
    • SC版本:已对齐O版本支持联网、外部RAG总结和口语化改写接口。
    Original source
  • Sep 1, 2025
    • Date parsed from source:
      Sep 1, 2025
    • First seen by Releasebot:
      Apr 23, 2026
    Volcengine logo

    Volcengine

    【2025.09】

    Volcengine升级Strong Character和Omni语音模型,前者支持声音复刻,后者支持RAG总结改写与低延时对话。

    产品能力升级

    • Strong Character版本(S2S-SC版本):强人格版本,主要for角色扮演、情感陪聊场景,支持声音复刻;暂不支持联网、RAG总结改写接口。
      • 目前支持20+公版音色,发音人ID以“ICL”开头,如:ICL_zh_female_aojiaonvyou_tob、ICL_zh_female_bingjiaojiejie_tob、ICL_zh_female_chengshujiejie_tob...(具体音色list详见接口文档)
    • Omni版本(S2S-O版本):定位是一个低延时语音端到端助手模型,覆盖闲聊、客服、车载等ToB多场景的低延时端到端模型;支持外部RAG总结和口语化改写接口,可支持客户外接RAG搜索内容,通过接口传入,S2S模型会按照人设进行口语化总结改写,并进行播报。

    SC版和O版本的功能差异详见接口文档。

    Original source
  • Sep 1, 2025
    • Date parsed from source:
      Sep 1, 2025
    • First seen by Releasebot:
      Jan 8, 2026
    • Modified by Releasebot:
      Jun 10, 2026
    Volcengine logo

    Volcengine

    【2025.09】

    Volcengine新增隐式meta水印写入,支持大模型语音合成、声音复刻和语音播客v3协议接口,覆盖mp3/wav/ogg_opus音频格式。

    已支持隐式 meta 水印写入,当前仅大模型语音合成、声音复刻和 语音播客v3 协议接口支持,音频格式支持mp3/wav/ogg_opus。

    官网接口文档→链接,搜索 “aigc_metadata”。

    Original source
  • Sep 1, 2025
    • Date parsed from source:
      Sep 1, 2025
    • First seen by Releasebot:
      Jan 8, 2026
    • Modified by Releasebot:
      Apr 23, 2026
    Volcengine logo

    Volcengine

    【2025.09】

    Volcengine releases Doubao Voice Synthesis 2.0 with dialogue-style TTS and async long-text jobs for richer, more natural audio.

    大模型语音合成2.0版本上新:

    • 推出豆包语音合成模型2.0,支持TTS对话式合成新范式(Query-Response),提供更加自然、更丰富情感、更具有表现力的语音合成效果。
    • 新上线异步执行长文本任务接口:最大单次可执行的文本长度为10万字符,合成音频数据在服务端可保存7天。适用于批量进行音频内容生产(如有声小说等),但对时效性要求不高的场景;调用的价格跟大模型语音合成/声音复刻短文本定价保持一致;
    Original source
Releasebot

Curated by the Releasebot team

Releasebot is an aggregator of official release notes from hundreds of software vendors and thousands of sources.

Our editorial process involves the manual review and audit of release notes procured with the help of automated systems.