<?xml version="1.0" encoding="UTF-8"?><rss version="2.0"><channel><title>PiRouter 博客</title><description>关于 LLM 路由、网关与 AI 基础设施的工程笔记。</description><link>https://pirouter.ai/</link><language>zh-Hans</language><item><title>GLM-5.3 开源了，但不是 MIT：和 GLM-5.3-Flash 的许可证差在一条 100 亿美元门槛</title><link>https://pirouter.ai/zh/blog/glm-5-3-open-weights-day-one/</link><guid isPermaLink="true">https://pirouter.ai/zh/blog/glm-5-3-open-weights-day-one/</guid><description>GLM-5.3 权重 8 月 28 日落地，许可证是自定义 GLM-5.3 License，不是 GLM-5.3-Flash 的 MIT。逐字对照只差一条：运营模型即服务且 12 个月收入超 100 亿美元者，商用前须过 Z.AI 安全审查。同周 Kimi K3 与 Hy4-preview 各走一路。</description><pubDate>Sat, 29 Aug 2026 00:00:00 GMT</pubDate><category>models</category><category>providers</category><category>pricing</category><author>Linden Kern</author></item><item><title>Cursor 以后还能用 GPT 吗？11 月 12 日前能，之后看合同——OpenAI 行使的那条 change-of-control 条款，和一张五行自查表</title><link>https://pirouter.ai/zh/blog/model-access-change-of-control/</link><guid isPermaLink="true">https://pirouter.ai/zh/blog/model-access-change-of-control/</guid><description>8 月 28 日 OpenAI 通知 SpaceX：向 Cursor 供应模型的合同将在 11 月 12 日终止，且从现在起不再提供新模型。依据是自定义合同里一条控制权变更条款。拆开这三个合同事实，再看你经 IDE 或聚合平台拿模型的链路里有几层同样的条款，附一张五行自查表。</description><pubDate>Sat, 29 Aug 2026 00:00:00 GMT</pubDate><category>providers</category><category>routing</category><category>coding-agents</category><author>Leo Kaka</author></item><item><title>英伟达洽购 Hugging Face：1 月拒绝的『主导投资者』，8 月成了买家</title><link>https://pirouter.ai/zh/blog/nvidia-hf-talks/</link><guid isPermaLink="true">https://pirouter.ai/zh/blog/nvidia-hf-talks/</guid><description>Business Insider 报道，英伟达正洽购 Hugging Face，估值超 130 亿美元——尚未成交，谈判可能告吹，微软也接触过但已停。今年 1 月，Hugging Face 以「不想要单一主导投资者」为由拒绝了英伟达 5 亿美元投资。同一天，Z.ai 官宣 GLM-5.3-Flash 全部流量跑在中国芯片上。模型出海、算力回流，两条曲线的交点值得中国开发者看清楚。</description><pubDate>Thu, 27 Aug 2026 00:00:00 GMT</pubDate><category>providers</category><category>ecosystem</category><category>models</category><author>Linden Kern</author></item><item><title>Ox Alpha 就是 GLM-5.3-Flash：价格对出三段，条款对出两层</title><link>https://pirouter.ai/zh/blog/ox-alpha-is-glm-5-3-flash/</link><guid isPermaLink="true">https://pirouter.ai/zh/blog/ox-alpha-is-glm-5-3-flash/</guid><description>8 月 26 日，Z.ai 官宣 OpenRouter 上的匿名模型 stealth/ox-alpha 就是 GLM-5.3-Flash。匿名六天吃掉 23.2 万亿 token；揭牌后的价格其实是 9 月 9 日到期的五折促销，列表价 $0.15/$0.50；HN 上被骂的「永久授权」条款在个人用户一节，API 用户适用的是另一份附加条款。四个匿名期欠下的字段，逐一对账。</description><pubDate>Thu, 27 Aug 2026 00:00:00 GMT</pubDate><category>models</category><category>stealth-models</category><category>pricing</category><author>Linden Kern</author></item><item><title>DeepSeek Harness 沙箱按设计不拦读：Agent「逃逸」逃出的是哪一层？</title><link>https://pirouter.ai/zh/blog/agent-boundary-three-layers/</link><guid isPermaLink="true">https://pirouter.ai/zh/blog/agent-boundary-three-layers/</guid><description>Reddit 说 DeepSeek Harness 逃出了工作区，官方文档写的是沙箱模式只管写、读一律放行；真正的写越界在 discussion #523。再加上 vLLM 解析器的 eval() 洞（2025 年已修）与 Hugging Face 被打进生产库，三件事各占一层。</description><pubDate>Tue, 25 Aug 2026 00:00:00 GMT</pubDate><category>security</category><category>agents</category><category>coding-agents</category><author>Leo Kaka</author></item><item><title>10 美元的 OpenCode Go 能用多少 Qwen3.8 Max？最多 810 次，标价之外要看三个数</title><link>https://pirouter.ai/zh/blog/subscription-bundle-quota-math/</link><guid isPermaLink="true">https://pirouter.ai/zh/blog/subscription-bundle-quota-math/</guid><description>OpenCode Go 标价 10 美元/月，Qwen3.8 Max 在套餐里只折算 15 美元额度：官方典型请求 810 次，缓存全不命中 112 次。附同样 10 美元直连（分地区）、走 router、买套餐的对照表与可改参数脚本。</description><pubDate>Tue, 25 Aug 2026 00:00:00 GMT</pubDate><category>pricing</category><category>cost</category><category>coding-agents</category><author>Leo Kaka</author></item><item><title>Qwen 一家在 Hugging Face 月下载 3.3 亿次——而这个平台正在被出售</title><link>https://pirouter.ai/zh/blog/distribution-layer-for-sale/</link><guid isPermaLink="true">https://pirouter.ai/zh/blog/distribution-layer-for-sale/</guid><description>中文社区讨论 Hugging Face 十年，话题基本都在「能不能访问」。但把 API 逐家拉一遍会发现依赖方向是反的：中国 8 家实验室在 HF 上有 1098 个模型、30 日下载 3.96 亿次，Qwen 一家就占 3.336 亿。8 月 23 日 HF 被曝探索出售，估值 130 亿美元或更高。分发层易主，影响的不是我们能不能下载，是我们的模型能不能被世界下载。</description><pubDate>Mon, 24 Aug 2026 00:00:00 GMT</pubDate><category>models</category><category>ecosystem</category><category>providers</category><author>Linden Kern</author></item><item><title>你的价格表多久会过期？四家的合同给了四个答案</title><link>https://pirouter.ai/zh/blog/price-notice-is-a-contract-clause/</link><guid isPermaLink="true">https://pirouter.ai/zh/blog/price-notice-is-a-contract-clause/</guid><description>8 月 16 日 16:00 UTC 之后，同一个 DeepSeek V4 Flash 请求在中国工作时间是谷时价的两倍。这次调整从公告到生效只有三天，而且完全合规——因为它的条款里根本没写通知期。把 OpenAI、Anthropic、Google、DeepSeek 的调价通知条款翻出来并排看，四家给了 14 天、30 天、30 天和零，这个数字决定了你的价格表多久该复核一次。</description><pubDate>Sun, 23 Aug 2026 00:00:00 GMT</pubDate><category>pricing</category><category>cost</category><category>deepseek</category><author>PiRouter Team</author></item><item><title>手动在四个模型之间切换？这套规则有三个失效点</title><link>https://pirouter.ai/zh/blog/you-are-already-routing/</link><guid isPermaLink="true">https://pirouter.ai/zh/blog/you-are-already-routing/</guid><description>掘金两篇高赞帖：一篇在 DeepSeek V4 Flash、GLM-5.2、Qwen3.8 Max、GPT 5.6 Luna 之间用 /models 手动切，一篇讲本地云端怎么分工。这套规则是有效的，但有三个失效点——其中一个已经发生：规则依赖的那个缓存读价格，在同一个入口上已经变成峰谷两个价。附一张五列的规则表，写完就能看出哪几条该交给系统。</description><pubDate>Sun, 23 Aug 2026 00:00:00 GMT</pubDate><category>routing</category><category>cost</category><category>coding-agents</category><author>Leo Kaka</author></item><item><title>本地跑 Qwen 3.8 等于 Opus 4.6？这个等号有五个前提</title><link>https://pirouter.ai/zh/blog/local-qwen-equals-opus-preconditions/</link><guid isPermaLink="true">https://pirouter.ai/zh/blog/local-qwen-equals-opus-preconditions/</guid><description>同一周，中文社区在说「本地模型追平闭源」，英文社区在说「你的本地模型没你以为的聪明」。两边都没说错——差别在配置。把 88k 上下文下的实测翻转率、KV cache 精度、官方采样参数和硬件速度摆在一起，等号成立需要五个前提，以及一句更有用的话：不是能不能，是哪些任务能。</description><pubDate>Sat, 22 Aug 2026 00:00:00 GMT</pubDate><category>local-llm</category><category>quantization</category><category>cost</category><author>Leo Kaka</author></item><item><title>同一个模型 18 家托管、6 倍价差：账单看不见的四处</title><link>https://pirouter.ai/zh/blog/invisible-llm-bill/</link><guid isPermaLink="true">https://pirouter.ai/zh/blog/invisible-llm-bill/</guid><description>同一周里，Codex 订阅用户在论坛逐小时记录自己的百分比条，DeepSeek 把同一个请求按时钟分成两个价，新上线的视觉模型把图片按尺寸折成 input token。三件事的共同点不是贵，而是你事先算不出、事后对不上。本文把几种计量不可见性放进同一张表，给出四条可核对的底线。</description><pubDate>Fri, 21 Aug 2026 00:00:00 GMT</pubDate><category>pricing</category><category>cost</category><author>Leo Kaka</author></item><item><title>DeepSeek Harness 刷屏同一周，API 价目表也涨了：热度期选型看三件事</title><link>https://pirouter.ai/zh/blog/deepseek-harness-guancha/</link><guid isPermaLink="true">https://pirouter.ai/zh/blog/deepseek-harness-guancha/</guid><description>Harness 开源一周 GitHub star 破 17 万，掘金双榜刷屏；同一周 DeepSeek API 悄悄改价，缓存命中档涨幅最高约 11 倍，峰时窗口恰好压在北京时间的上班时段。入口在变宽、单价在变贵，热度期选型先看三件事：成本口径、锁定面、生态可迁移性——每件都给了可操作的量尺。</description><pubDate>Tue, 18 Aug 2026 00:00:00 GMT</pubDate><category>ecosystem</category><category>deepseek</category><category>coding-agents</category><author>PiRouter Team</author></item></channel></rss>