EN
返回档案库

档案库 · AI 与模型 · 产品决策 · 2026

Weave 押注大部分编程提示不需要前沿模型;3.8k 星路由器

将 Claude Code、Codex 和 Cursor 的提示在 50 毫秒内发送给最便宜的合适模型;声称节省 40-70%,GitHub 星标 3.8k。

Weave

它在赌什么大部分编程提示不需要前沿模型,因此将每个请求路由到满足质量的最便宜模型,在 50 毫秒内完成,可以节省 40-70% 的 token 支出,而不损失质量。已上线

做的是什么生意

Weave sells a model router for AI coding agents: teams point Claude Code, Codex or Cursor at a proxy that classifies every request in under 50ms and sends routine prompts to cheaper models (Kimi, DeepSeek, Llama, Gemini, GPT variants) while reserving frontier models for hard work, with one policy across harnesses, automatic failover, and per-team quality-versus-cost settings.

起因

Weave — which markets itself on weaveos.com as an engineering-intelligence platform that tracks code quality, keep rate and cost per merged PR — created the weave-os/router repository on 2026-04-27 and sells a hosted router. The pitch is the freight-truck analogy: most coding prompts don't need a frontier model, so the router 'reads each request, sends it to the smallest model that finishes the job at full quality', with classification in under 50ms at Weave's edge.

经过

By the 2026-09-04 crawl the repo showed 3,849 stars. The product page, crawled the same day, promised 40-70% cost cuts with an endpoint change, an example calculator showing 53% lower monthly spend, 74% of requests routed to the cheaper pool, 99.9% completion with automatic failover and a 91/100 average code-quality score across routed traffic — all self-reported — and offered three policies (quality-first, balanced, cost-first) so teams set the quality floor the router may not cross.

结果

Still live and early as of 2026-09-05: weave-os/router has 3,849 stars, the hosted sign-up and demo flow are open, and the source is available; the 40-70% savings and quality numbers remain Weave's own claims, with no independent revenue, funding or customer data in the record.

背景

Weave 为代理式编码工具销售模型路由器:一个即插即用的代理,位于 Claude Code、Codex 或 Cursor 与 LLM 提供商之间,在 50 毫秒内对每个请求进行分类,并将常规提示发送给更便宜的模型,而保留前沿模型用于有难度的工作。营销口号是“不要派卡车送明信片”,页面声称只需更改端点即可节省 40-70% 的成本。

其赌注是大多数编程提示不需要前沿模型,因此按质量每 token 路由 —— 以完整质量完成工作的最便宜模型 —— 可以大致减半团队的 token 账单,而不会造成明显质量损失。Weave 将路由器定位在一个已跟踪代码质量、保留率和合并 PR 成本的产品中,因此节省基于工程团队已关注的指标报告,扣除缓存未命中。

路由器于 2026 年推出:weave-os/router 仓库于 2026-04-27 创建,截至 2026-09-04 抓取达到 3,849 星。产品页面提供三种策略(质量优先、平衡和成本优先),跨每个工具统一策略,99.9% 的完成率自动故障转移,示例计算器显示在 12,000 美元账单上每月支出降低 53% —— 这些数字目前是 Weave 自身的声明。

采用机制注重低摩擦:npx 安装程序修改每个提供商的环境变量,客户端在不感知代理的情况下工作,源代码开放供团队检查或自托管。截至 2026-09-05,记录中没有融资、收入、客户或独立基准数据。

这件事要成立,得有什么

  • 可验证的关注度:weave-os/router 从 2026-04-27 创建到 2026-09-04 抓取达到 3,849 星 —— 对于一个四个月大的开发者工具来说增长迅速。
  • 赌注在经济上是合理且明确的:在模型价格差异大的情况下,为常规提示支付前沿价格是浪费,路由可以消除这种浪费。
  • 设计消除了采用摩擦:一键安装、掉线 API 兼容性、跨 Claude Code、Codex 和 Cursor 的统一策略,无需修改提示。
  • 声明结构可验证 —— 开源核心、节省扣除缓存未命中、基于保留率和合并 PR 成本的质量度量 —— 但每个数字仍是自报。

可借鉴之处

当模型价格差异巨大时,路由成为产品:出售用户设定质量下限的可衡量的每 token 质量权衡,并开源核心,使成本声明可被验证。

后续进展

截至 2026-09-05,Weave 路由器活跃且处于早期:weave-os/router 在 2026-09-04 抓取中为 3,849 星,weaveos.com 页面邀请注册或演示,源代码公开供自托管。节省和质量数字(40-70% 成本削减,示例计算器 53%,91/100 质量评分)是 Weave 自身的营销申请,如抓取所示;材料中没有融资、收入、客户列表或独立评估。赌注 —— 将常规编程提示路由到更便宜的模型,成本降低一半而不损失质量 —— 仍未被公司自己的测量之外的数据证实。

资料来源

发现哪里写错了?告诉我们。

轮到你了

你刚读完一家。说说你在做什么,看看谁在赌同一件事。

免费账号 · 3 次免费提问 · 不用绑卡

相关案例