EN
返回档案库

档案库 · AI 与模型 · 战略决策 · 2019–2026

Resemble AI 押注 MIT 开源语音克隆成为默认 TTS

语音初创公司以 MIT 许可发布 Chatterbox,登顶 GitHub 趋势榜并获 26.2k 星标;销售托管、水印和检测服务。

Resemble AI

它在赌什么MIT 许可的顶尖语音模型将使开源成为实时智能体的默认 TTS,而 Resemble 通过托管、水印和检测服务盈利。已上线

做的是什么生意

Resemble AI is a generative voice company that clones and synthesizes voices with consent controls, PerTh audio watermarking and its Resemble Detect authenticity model, and since 2025 publishes the open-source Chatterbox text-to-speech family.

启动资金$8M Series A led by Javelin Venture Partners with Craft and Ubiquity, announced 2023-07-12, bringing total raised to $12M at the time (TechCrunch).

起因

Resemble AI was founded in 2019 by Zohaib Ahmed, a former Magic Leap lead software engineer, and Saqib Muhammad, after they noticed video-game voice lines could not keep up with frequent game updates. The company required consent clips before cloning a voice, enforced strict usage guidelines, and built PerTh watermarking plus Resemble Detect early, positioning itself as the responsible player in a market it acknowledged was controversial.

经过

In April 2025 Resemble shipped Chatterbox, its first production-grade open-source TTS model, under an MIT license. The repo grew 10.4x in Q3 2025 per the ROSS Index, and star-history shows it reached GitHub trending #1 on December 17, 2025, with repeated top-10 trending days into March 2026. On December 26, 2025 the company released Chatterbox Turbo: first audio in under 150 milliseconds, zero-shot cloning from five seconds of audio, built-in PerTh watermarking for verifying AI-generated speech, MIT-licensed, and available on Hugging Face, RunPod, Modal, Replicate and Fal.

结果

Live and scaling its bet: as of September 2026 the MIT-licensed Chatterbox family spans multilingual and Turbo models, the main repo holds 26.2k stars, and Resemble continues to sell hosted voice plus AI-generated-content detection to enterprises and regulated industries.

背景

Resemble AI 是一家生成式语音公司,由 Zohaib Ahmed 和 Saqib Muhammad 于 2019 年创立,起源是一个简单观察:游戏配音跟不上频繁更新。公司从游戏配音扩展到配音、个性化语音消息和实时对话智能体,并早期在同意和安全上差异化——克隆前要求同意录音,并构建 PerTh 水印和 Resemble Detect,一个评估音频真伪的模型。

2025 年 Resemble 反转常规语音 AI 玩法。它以宽松的 MIT 许可发布了生产级文本转语音模型 Chatterbox,随后在 2025 年 12 月发布 Chatterbox Turbo——150 毫秒内首段音频,五秒素材零样本克隆,内置 PerTh 水印,自由商用。The Decoder 和 PingWest 均将此次发布视为对 ElevenLabs 等商业语音平台的直接挑战,目标是为构建实时智能体、客户支持、游戏、头像和社交产品的开发者。

开源赌注带来了封闭 API 无法产生的流量。resemble-ai/chatterbox 仓库在 2025 年第三季度增长 10.4 倍(据 Runa 的 ROSS 指数),2025 年 12 月 17 日达到 GitHub 趋势榜首,截至 2026 年 9 月 1 日持有 26.2k 星标,共 21 天进入趋势榜。模型在 Hugging Face、RunPod、Modal、Replicate 和 Fal 上分发,开发者可在不接触 Resemble 付费平台的情况下测试和部署。

Resemble 的策略是让开源俘获开发者,向企业盈利:提供低延迟托管服务、对其自身生成语音有效的水印,以及面向担忧深度伪造的受监管行业的检测工具。截至 2026 年 9 月,公司仍在运营,开放 Chatterbox 系列涵盖多语言和 Turbo 变体,并围绕 AI 生成内容建立商业安全业务。

这件事要成立,得有什么

  • 开源翻转了信任问题:在 MIT 许可下展示模型代码给开发者提供质量和控制的证明,将语音克隆供应商的怀疑者转化为采用者。
  • 发布时机契合实时智能体浪潮:150 毫秒内首段音频和五秒克隆直接瞄准语音智能体、支持机器人、游戏和头像——不仅是离线配音。
  • 安全功能捆绑在赠品中:内置 PerTh 水印使开放模型可用于受监管行业,否则他们会拒绝无标记的合成语音。
  • 分发倍增触及:在 Hugging Face、RunPod、Modal、Replicate 和 Fal 上发布使模型面向从不访问供应商网站的开发者。

可借鉴之处

赠送皇冠明珠可以扩大你销售的市场:一个 MIT 许可的顶尖模型建立了开发者信任和心智份额,这是封闭 API 永远无法做到的。

后续进展

截至 2026 年 9 月 2 日,Resemble AI 以开源加商业语音策略运行:MIT 许可的 Chatterbox 系列包括多语言和 Turbo 模型,GitHub 仓库持有 26.2k 星标并曾在 2025 年 12 月 17 日登顶趋势榜,模型通过 Hugging Face、RunPod、Modal、Replicate 和 Fal 服务。它仍向企业和受监管行业销售托管语音、PerTh 水印和 Resemble Detect。悬而未决的问题是:放弃其最佳模型是否比信任工具和企业支持更快蚕食托管 TTS 收入。

资料来源

发现哪里写错了?告诉我们。

轮到你了

你刚读完一家。说说你在做什么,看看谁在赌同一件事。

免费账号 · 3 次免费提问 · 不用绑卡

相关案例