档案库 · 硬件与设备 · 技术决策 · 2023–2026
Taalas将模型权重蚀刻进硅片;融资2.19亿美元,2026年被AMD收购
Taalas押注将模型硬连线到硅片上——无需HBM读取——赢得AI推理:演示速度达16,960 token/秒,融资2.19亿美元,随后被AMD收购。
Taalas
做的是什么生意
Taalas builds model-specific AI chips ('Hardcore Models') that etch an LLM's weights directly into silicon, so inference skips the memory fetches that slow general-purpose GPUs.
启动资金:$219M total raised by Feb 2026 — a $169M round from Quiet Capital, Fidelity and Pierre Lamond plus earlier funding; AMD did not disclose deal terms.
起因
Ljubisa Bajic, founder of AI-chip maker Tenstorrent, started Taalas in Toronto in 2023 with COO Lejla Bajic on the premise that 'the model is the computer': simulate nothing, hardwire the model. Its first TSMC 6nm test chip ran Llama 3.1 8B at a claimed 16,960 tokens/sec, about 48x an Nvidia GPU of the era.
经过
Taalas emerged from stealth and in Feb 2026 raised $169M (about $219M total from Fidelity, Quiet Capital and Pierre Lamond), unveiling HC1, a chip optimized for Llama 3.1 8B, with a second-generation chip for roughly 20B-parameter models planned for 2026. The tradeoff — one model per chip, frozen at manufacturing — was the entire point.
结果
On 2026-08-06 AMD announced a definitive agreement to acquire Taalas; terms were undisclosed. AMD plans to pair Taalas silicon with Instinct GPUs and Helios rackscale systems under its ROCm software, expecting the deal to close in Q4 2026 — framed as a business acquisition, not an acqui-hire.
背景
Taalas是一家多伦多的芯片初创公司,制造“硬核模型”:一种AI推理芯片,在制造时将特定模型的权重蚀刻进硅片作为掩膜ROM。由于权重从不从HBM移动到处理器,token生成绕过了主导现代推理成本的内存瓶颈。
创始人Ljubisa Bajic,此前创立了AI芯片公司Tenstorrent,于2023年与COO Lejla Bajic共同创立Taalas,基于“模型即计算机”的理念。首款TSMC 6nm测试芯片运行Llama 3.1 8B,声称达16,960 token/秒——大约是同期NVIDIA GPU的48倍,以及Cerebras加速器的8.5倍。Taalas于2026年2月融资1.69亿美元(总融资约2.19亿美元),来自Fidelity、Quiet Capital和Pierre Lamond,并推出了针对Llama 3.1 8B优化的HC1芯片。
其权衡是有意为之:每片芯片在制造时冻结为单一模型,因此它只在单一模型服务海量实时流量时获胜。2026年8月6日,AMD宣布最终协议收购Taalas(条款未披露),计划将技术集成到Instinct GPU和Helios机架级系统中,在ROCm下运行,预计于2026年第四季度完成。
这件事要成立,得有什么
- 将权重蚀刻进掩膜ROM消除了限制GPU推理速度的HBM权重读取往返——这是一个数量级的架构差异,而非优化。
- 演示优先策略——一片芯片对应一个模型,Llama 3.1 8B运行于16,960 token/秒——使价值可视化,并在HN线程引发724条评论的讨论。
- 在硬件发货前融资2.19亿美元,使公司在押注成熟期间能挺过流片。
- 每芯片单模型的限制将市场限定在少数巨大工作负载,使平台所有者的收购成为自然退出。
可借鉴之处
押注行业的标准瓶颈——内存而非计算——可以定义初创公司,但固定单一模型的芯片只有在平台所有者收购时才值得。
后续进展
截至2026年9月2日,Taalas正被纳入AMD:收购预计于2026年第四季度完成,待监管批准,AMD计划在Helios机架级系统中,通过Instinct GPU处理提示,通过Taalas硅片生成token。初创公司的独立路线图——面向约200亿参数模型的第二代芯片——在AMD加速器路线图中继续。交易条款未披露。
资料来源
- AMD Acquires Taalas to Advance Compute Solutions for Rapidly Growing AI Inference Market
- AMDが新興Taalas買収、AIモデルをシリコンに焼き付ける技術を獲得──AI推論でHBM依存を崩すか
发现哪里写错了?告诉我们。
轮到你了
你刚读完一家。说说你在做什么,看看谁在赌同一件事。
免费账号 · 3 次免费提问 · 不用绑卡