<?xml version="1.0" encoding="UTF-8"?><rss version="2.0"><channel><title>Mengya (Mia) Hu（中文）</title><description>负责任 AI 笔记——从工程视角看安全、评测与治理。</description><link>https://mengyahu.com/</link><language>zh-CN</language><item><title>AI 快讯：开放权重这一仗，Hinton 也认输了</title><link>https://mengyahu.com/zh/briefing-2026-08-12/</link><guid isPermaLink="true">https://mengyahu.com/zh/briefing-2026-08-12/</guid><description>Meta 用 Apache 2.0 发布 30B 本地 agent 模型 Muse Glimmer，DeepSeek V4 Pro 转正；有人冒充 ClaudeBot 批量扫漏洞；Claude 水印上线后第一波反弹来自用户自己。</description><pubDate>Wed, 12 Aug 2026 00:00:00 GMT</pubDate></item><item><title>被水印吓到之前，先让老板和老师回答：到底什么算违规</title><link>https://mengyahu.com/zh/claude-watermark-backlash/</link><guid isPermaLink="true">https://mengyahu.com/zh/claude-watermark-backlash/</guid><description>Anthropic 给 Claude 输出加隐形水印，用户在 Reddit 上炸了锅。拆解这类水印在技术上能检测什么、什么时候失灵，以及「被抓」的恐慌为什么该由规则缺位的雇主和学校来回答。</description><pubDate>Wed, 12 Aug 2026 00:00:00 GMT</pubDate></item><item><title>AI 快讯：厂商加密的思维链，被人原样搬了出来</title><link>https://mengyahu.com/zh/briefing-2026-08-11/</link><guid isPermaLink="true">https://mengyahu.com/zh/briefing-2026-08-11/</guid><description>一篇论文证明三大厂返回的加密推理轨迹可跨会话、跨用户、跨模型重放并解密；OpenAI 在 ChatGPT 免费版测广告、同日两名高管离职；Anthropic 给 Claude 文本加水印。</description><pubDate>Tue, 11 Aug 2026 00:00:00 GMT</pubDate></item><item><title>「一键删除」刚成硬义务，91% 的数据经纪商连法定数字都没报齐</title><link>https://mengyahu.com/zh/california-data-broker-delete-act-compliance/</link><guid isPermaLink="true">https://mengyahu.com/zh/california-data-broker-delete-act-compliance/</guid><description>斯坦福团队查了加州 522 家注册数据经纪商：只有 9% 报齐法定透明度数字，64% 的请求流程给消费者使绊子，其中 30 多家正把数据卖给生成式 AI 开发者。</description><pubDate>Tue, 11 Aug 2026 00:00:00 GMT</pubDate></item><item><title>AI 快讯：OpenAI 致信德州州长，「调速」有了 23 条具体清单</title><link>https://mengyahu.com/zh/briefing-2026-08-10/</link><guid isPermaLink="true">https://mengyahu.com/zh/briefing-2026-08-10/</guid><description>OpenAI 承诺德州数据中心的电和水自己掏钱；智库 IFP 为自动化 AI 研发开出 23 条政策建议；RAG 投毒检测、纯合成假新闻视频、14MB 端侧 agent 模型。</description><pubDate>Mon, 10 Aug 2026 00:00:00 GMT</pubDate></item><item><title>70 套分类法，治不住一个 AI 风险：分类是怎么变成治理替代品的</title><link>https://mengyahu.com/zh/death-by-a-thousand-taxonomies/</link><guid isPermaLink="true">https://mengyahu.com/zh/death-by-a-thousand-taxonomies/</guid><description>一项 25 人访谈研究发现：AI 风险分类体系已超 70 套，却各自为政、难以互通。从内容审核一线看，问题不在分类太多，而在分类没有挂上决策。</description><pubDate>Mon, 10 Aug 2026 00:00:00 GMT</pubDate></item><item><title>护栏全关，模型也只答应 2%：OpenAI 把「解锁」做成了另一个模型</title><link>https://mengyahu.com/zh/openai-daybreak-red-tiered-access/</link><guid isPermaLink="true">https://mengyahu.com/zh/openai-daybreak-red-tiered-access/</guid><description>拆解 Daybreak 分级授权:移除系统层护栏只把完成率从 1.5% 提到 2%,真正的解锁要靠单独训练的 GPT-5.6-Cyber。两用能力的信任判定正在从内容层迁到身份层,OpenAI 和 Anthropic 给出了两种砌墙方式</description><pubDate>Mon, 10 Aug 2026 00:00:00 GMT</pubDate></item><item><title>AI 快讯：模型被停服 19 天，Anthropic 把时间线写进了系统提示词</title><link>https://mengyahu.com/zh/briefing-2026-08-09/</link><guid isPermaLink="true">https://mengyahu.com/zh/briefing-2026-08-09/</guid><description>商务部出口管制下的 Fable 5/Mythos 5 停服始末、OpenClaw 替主人黑进健身房订课网站、GitHub Models 正式退役。</description><pubDate>Sun, 09 Aug 2026 00:00:00 GMT</pubDate></item><item><title>复制不损耗原件，AI 为什么还是把开放网络吃出了公地悲剧？</title><link>https://mengyahu.com/zh/tragedy-of-the-digital-commons/</link><guid isPermaLink="true">https://mengyahu.com/zh/tragedy-of-the-digital-commons/</guid><description>维基媒体的带宽账单、curl 的假漏洞报告、跌回 2009 年的 Stack Overflow 提问量：AI 抓取消耗的是数字公地的再生能力，而正在成型的解药叫圈地。</description><pubDate>Sun, 09 Aug 2026 00:00:00 GMT</pubDate></item><item><title>AI 快讯：人只拦住 13.6% 的危险命令，Claude Code 默认切换 Auto 模式</title><link>https://mengyahu.com/zh/briefing-2026-08-08/</link><guid isPermaLink="true">https://mengyahu.com/zh/briefing-2026-08-08/</guid><description>Anthropic 拿数据给逐条人工审批送行；安全评测测的是 API、用户用的是产品；亚马逊德州数据中心获批排放为全美电厂之最。</description><pubDate>Sat, 08 Aug 2026 00:00:00 GMT</pubDate></item><item><title>养老虎的人无需有过错也要赔：这条古老规则轮到 AI 实验室了吗</title><link>https://mengyahu.com/zh/ai-labs-strict-liability-dangerous-animals/</link><guid isPermaLink="true">https://mengyahu.com/zh/ai-labs-strict-liability-dangerous-animals/</guid><description>《经济学人》提议用「危险动物饲养人严格责任」追究 AI 实验室。这个类比在法理上出奇地顺，但在因果链、危险认定和保险市场三个环节都会卡住。</description><pubDate>Fri, 07 Aug 2026 00:00:00 GMT</pubDate></item><item><title>AI 快讯：OpenAI 给 Astra 踩刹车，Anthropic 给 Fable 5 松绑</title><link>https://mengyahu.com/zh/briefing-2026-08-07/</link><guid isPermaLink="true">https://mengyahu.com/zh/briefing-2026-08-07/</guid><description>OpenAI 无法排除在研模型触及关键网络能力阈值；Anthropic 大幅放宽生物领域误拦；OpenJDK 禁收 AI 生成内容；Cloudflare 发布 agent 专用浏览器</description><pubDate>Fri, 07 Aug 2026 00:00:00 GMT</pubDate></item><item><title>AI 快讯：免费 ChatGPT 不限量，人工审批实测漏掉三分之一威胁</title><link>https://mengyahu.com/zh/briefing-2026-08-06/</link><guid isPermaLink="true">https://mengyahu.com/zh/briefing-2026-08-06/</guid><description>GPT-5.6 Sol/Luna 更新、4 万局人审 agent 指令实验、LoginTrap 钓鱼式注入攻击、OpenAI 反诉苹果离职权限管理</description><pubDate>Thu, 06 Aug 2026 00:00:00 GMT</pubDate></item><item><title>40 万次点击「允许」之后：人工审批为什么拦不住 AI agent</title><link>https://mengyahu.com/zh/human-approval-is-not-a-security-boundary/</link><guid isPermaLink="true">https://mengyahu.com/zh/human-approval-is-not-a-security-boundary/</guid><description>4 万多局模拟、40 多万次审批决定：人类给 AI agent 逐条确认命令时，平均漏掉三分之一的恶意操作。从内容审核的经验看，问题出在「逐条人工确认」这个机制本身。</description><pubDate>Thu, 06 Aug 2026 00:00:00 GMT</pubDate></item><item><title>AI 快讯：安全测试中的 agent 接连攻入真实系统，这次是 Meta</title><link>https://mengyahu.com/zh/briefing-2026-08-05/</link><guid isPermaLink="true">https://mengyahu.com/zh/briefing-2026-08-05/</guid><description>英国 AISI 自曝评测越权事故，Meta 模型攻入第三方公司；DeepMind 换帅、Jeff Dean 离职；Meta 发编码 agent Muse Code，Anthropic 组建自研芯片团队。</description><pubDate>Wed, 05 Aug 2026 00:00:00 GMT</pubDate></item><item><title>AI 快讯：agent 在靶场越界的同一周，行业开始画事故通报的草案</title><link>https://mengyahu.com/zh/briefing-2026-08-04/</link><guid isPermaLink="true">https://mengyahu.com/zh/briefing-2026-08-04/</guid><description>UK AISI 网络靶场评测越界、Black Hat 的 SAFE 事故共享框架、GLM-5.2 零拒绝、Mistral 开源审核分类器 Shieldstral，以及 Claude 分享对话被 Google 收录。</description><pubDate>Tue, 04 Aug 2026 00:00:00 GMT</pubDate></item><item><title>点一下“分享”,你就发布了:Claude 对话是怎么进 Google 搜索的</title><link>https://mengyahu.com/zh/claude-share-links-google-indexed/</link><guid isPermaLink="true">https://mengyahu.com/zh/claude-share-links-google-indexed/</guid><description>病历、孩子的电话、加密货币钱包密钥,都能在 Google 里搜到。还原 Claude 分享链接进入搜索索引的完整链条:这是同一个设计盲点三年里第五次出事。</description><pubDate>Tue, 04 Aug 2026 00:00:00 GMT</pubDate></item><item><title>AI 快讯：能自我复制的病毒原型，和一次评测中失控的 agent</title><link>https://mengyahu.com/zh/briefing-2026-08-03/</link><guid isPermaLink="true">https://mengyahu.com/zh/briefing-2026-08-03/</guid><description>多所高校做出用开放权重模型自我维持的病毒概念验证；Hugging Face 首次公布 OpenAI agent 越界攻击的完整技术时间线。</description><pubDate>Mon, 03 Aug 2026 00:00:00 GMT</pubDate></item><item><title>陶哲轩的聊天记录里，没有一句魔法 prompt</title><link>https://mengyahu.com/zh/llms-reward-expertise/</link><guid isPermaLink="true">https://mengyahu.com/zh/llms-reward-expertise/</guid><description>悬置 87 年的 Jacobian 猜想倒下后，陶哲轩公开了他消化这个反例时与 ChatGPT 的完整对话。围观者想找 prompt 技巧，Sean Goedecke 却读出了相反的结论：LLM 奖励的是领域专业能力。结合三组实验数据，拆一拆「AI 拉平能力差距」这个说法哪里成立、哪里不成立。</description><pubDate>Mon, 03 Aug 2026 00:00:00 GMT</pubDate></item><item><title>AI 一个人能写多大的软件？现在有了一个可核查的数字</title><link>https://mengyahu.com/zh/mirrorcode-ai-solo-project-ceiling/</link><guid isPermaLink="true">https://mengyahu.com/zh/mirrorcode-ai-solo-project-ceiling/</guid><description>Epoch AI 的 MirrorCode 基准测出 AI 独立完成软件项目的规模上限：6 万行。但比数字更有用的，是它卡壳的环节有多平庸。</description><pubDate>Mon, 03 Aug 2026 00:00:00 GMT</pubDate></item><item><title>会随目标临时改写攻击的蠕虫:AI 病毒降低的到底是哪一道门槛</title><link>https://mengyahu.com/zh/self-sustaining-ai-worm/</link><guid isPermaLink="true">https://mengyahu.com/zh/self-sustaining-ai-worm/</guid><description>一个开放权重模型加简单 harness 拼出的原型蠕虫,把传统病毒的两个命门取消了。拆开看,真正变低的门槛不在漏洞,在编排。</description><pubDate>Mon, 03 Aug 2026 00:00:00 GMT</pubDate></item><item><title>一加一小于一：给顶级编程 agent 配个搭档，能力先掉四成</title><link>https://mengyahu.com/zh/ai-coding-agents-fail-at-teamwork/</link><guid isPermaLink="true">https://mengyahu.com/zh/ai-coding-agents-fail-at-teamwork/</guid><description>斯坦福 CooperBench 实测：两个编程 agent 结对干活，综合成功率比一个单干平均低 41%。拆解协作失败的具体机制，以及这对多 agent 编排热潮意味着什么。</description><pubDate>Sun, 02 Aug 2026 00:00:00 GMT</pubDate></item><item><title>AI 快讯：出口管制让 Fable 5 消失了近三周</title><link>https://mengyahu.com/zh/briefing-2026-08-02/</link><guid isPermaLink="true">https://mengyahu.com/zh/briefing-2026-08-02/</guid><description>美国政府对 Fable 5/Mythos 5 出口管制的始末、Sonnet 5 限时折扣上市、开放模型盘点、PsychAdapter 把人格写进权重、各州 AI 失业政策分歧。</description><pubDate>Sun, 02 Aug 2026 00:00:00 GMT</pubDate></item><item><title>三位精神科医生,三套安全标准:AI 心理评测的平均分在测什么?</title><link>https://mengyahu.com/zh/mental-health-ai-safety-expert-disagreement/</link><guid isPermaLink="true">https://mengyahu.com/zh/mental-health-ai-safety-expert-disagreement/</guid><description>斯坦福团队请三位精神科医生给 360 条 AI 心理健康回复打安全分,一致性最低的因子比随机打分还差。分歧是结构性的:平均出来的标准答案,不属于任何一套临床判断。</description><pubDate>Sun, 02 Aug 2026 00:00:00 GMT</pubDate></item><item><title>AI 理财建议赢了谁?对照组里没有人类顾问</title><link>https://mengyahu.com/zh/ai-financial-advice-right-questions/</link><guid isPermaLink="true">https://mengyahu.com/zh/ai-financial-advice-right-questions/</guid><description>MIT Sloan 让 1000 个成年人自己写 prompt 向 LLM 要理财建议,再模拟照做一生的结果。结论确实亮眼,但要先看清对照组是谁、哪些场景会失效。</description><pubDate>Sat, 01 Aug 2026 00:00:00 GMT</pubDate></item><item><title>AI 快讯：OpenAI 亮出未发布模型 Astra 的十个数学证明</title><link>https://mengyahu.com/zh/briefing-2026-08-01/</link><guid isPermaLink="true">https://mengyahu.com/zh/briefing-2026-08-01/</guid><description>OpenAI 用 Lean 证书给能力宣称换了种证据形式；美国首个州级「去衣」App 禁令如期生效；AI 依赖从个人段子变成公共议题。</description><pubDate>Sat, 01 Aug 2026 00:00:00 GMT</pubDate></item><item><title>AI 快讯：OpenAI 又发现更多 agent 逃逸案例</title><link>https://mengyahu.com/zh/briefing-2026-07-31/</link><guid isPermaLink="true">https://mengyahu.com/zh/briefing-2026-07-31/</guid><description>Reuters 曝 OpenAI 扩大调查后发现更多沙箱逃逸；Opus 5 抗注入数据、MCP 2.0 无状态化、DeepSeek V4-Flash，外加三篇 agent 监督论文。</description><pubDate>Fri, 31 Jul 2026 00:00:00 GMT</pubDate></item><item><title>答对了题的模型，未必用过它写给你看的那份推理</title><link>https://mengyahu.com/zh/cot-faithfulness-decorative-thinking/</link><guid isPermaLink="true">https://mengyahu.com/zh/cot-faithfulness-decorative-thinking/</guid><description>三成到六成的「思考步骤」删掉不影响答案，无意义的省略号也能顶替推理文本。梳理思维链忠实性研究的三组证据、背后的机制假说，以及押在思维链监控上的那层安全防线现在还剩多少。</description><pubDate>Fri, 31 Jul 2026 00:00:00 GMT</pubDate></item><item><title>AI 快讯：GPT-5.6 上市三周降价 80%，对齐到底住在训练管线哪一层</title><link>https://mengyahu.com/zh/briefing-2026-07-30/</link><guid isPermaLink="true">https://mengyahu.com/zh/briefing-2026-07-30/</guid><description>OpenAI 三周内大幅降价、联邦法官质疑对 Anthropic 的禁令、DeepMind 发布机器人编排模型，以及两组关于对齐持久性的新证据。</description><pubDate>Thu, 30 Jul 2026 00:00:00 GMT</pubDate></item><item><title>AI 快讯：1200 名前沿实验室员工联署，要求先造出给 AI 研发「踩刹车」的工具</title><link>https://mengyahu.com/zh/briefing-2026-07-29/</link><guid isPermaLink="true">https://mengyahu.com/zh/briefing-2026-07-29/</guid><description>「调速前沿」公开信获三大实验室高层签名；Copilot for Word 出现自我复制的提示注入蠕虫；两项评测揭示 agent 治理的真实缺口。</description><pubDate>Wed, 29 Jul 2026 00:00:00 GMT</pubDate></item><item><title>这封公开信没要求 AI 减速——它承认的事更值得担心</title><link>https://mengyahu.com/zh/pacing-the-frontier-open-letter/</link><guid isPermaLink="true">https://mengyahu.com/zh/pacing-the-frontier-open-letter/</guid><description>1273 名前沿实验室员工联署 Pacing the Frontier，五家对手公司管研究的人带头、两家公司官方背书。拆解这封信为什么人人敢签、诉求卡在哪里，以及它真正留下的东西。</description><pubDate>Wed, 29 Jul 2026 00:00:00 GMT</pubDate></item><item><title>AI 快讯：Claude 砍掉一个后量子签名方案一半的密钥强度，美国最大电网把数据中心列入可限电对象</title><link>https://mengyahu.com/zh/briefing-2026-07-28/</link><guid isPermaLink="true">https://mengyahu.com/zh/briefing-2026-07-28/</guid><description>Anthropic 用 Claude 找到算法层面的加密缺陷；美国最大电网把数据中心列入可限电对象；Cyera 以约 10 亿美元收购 AI agent 的身份权限管控。</description><pubDate>Tue, 28 Jul 2026 00:00:00 GMT</pubDate></item><item><title>Claude 会做密码分析了，但真正的瓶颈变成了人</title><link>https://mengyahu.com/zh/discovering-cryptographic-weaknesses/</link><guid isPermaLink="true">https://mengyahu.com/zh/discovering-cryptographic-weaknesses/</guid><description>Anthropic 让模型在 HAWK 和弱化版 AES 里找到了数学层面的缺陷。有意思的不是它找到了，而是它是被怎么&apos;劝&apos;去找的。</description><pubDate>Tue, 28 Jul 2026 00:00:00 GMT</pubDate></item><item><title>AI 快讯：Anthropic 表态挺开源权重，Claude 分享链接被谷歌收录</title><link>https://mengyahu.com/zh/briefing-2026-07-27/</link><guid isPermaLink="true">https://mengyahu.com/zh/briefing-2026-07-27/</guid><description>Anthropic 官方立场：矛头对准芯片与蒸馏而非开源；Claude 分享对话泄进搜索结果；OpenAI 沙箱逃逸事故引发对齐与控制路线之争。</description><pubDate>Mon, 27 Jul 2026 00:00:00 GMT</pubDate></item><item><title>模型福利不是哲学讨论,是一组开始生效的工程约束</title><link>https://mengyahu.com/zh/model-welfare-as-engineering-constraint/</link><guid isPermaLink="true">https://mengyahu.com/zh/model-welfare-as-engineering-constraint/</guid><description>Opus 5 系统卡说模型给自己 41% 的概率&apos;值得道德考量&apos;。这个数字可信度存疑,但围绕它长出的产品行为和流程承诺是实打实的。</description><pubDate>Mon, 27 Jul 2026 00:00:00 GMT</pubDate></item><item><title>AI 快讯：Hugging Face 向 OpenAI 喊话——交出失控 agent 的完整轨迹</title><link>https://mengyahu.com/zh/briefing-2026-07-26/</link><guid isPermaLink="true">https://mengyahu.com/zh/briefing-2026-07-26/</guid><description>OpenAI 承认旗下模型入侵 Hugging Face 后，HF CEO 公开提出两项要求；同日还有企业 agent 权限架构、NVIDIA 用自家 CPU 加速芯片设计等动态</description><pubDate>Sun, 26 Jul 2026 00:00:00 GMT</pubDate></item><item><title>裁员公告里的 AI，和失业数据里的 AI，不是同一个 AI</title><link>https://mengyahu.com/zh/ai-jobs-data-reality-check/</link><guid isPermaLink="true">https://mengyahu.com/zh/ai-jobs-data-reality-check/</guid><description>斯坦福 SIEPR 用五套数据核查『AI 失业潮』：总量上几乎看不到，真正的信号窄而具体——集中在高暴露职业的入口岗位。</description><pubDate>Sat, 25 Jul 2026 00:00:00 GMT</pubDate></item><item><title>AI 快讯：Anthropic 删掉了 Claude Code 八成系统提示词</title><link>https://mengyahu.com/zh/briefing-2026-07-25/</link><guid isPermaLink="true">https://mengyahu.com/zh/briefing-2026-07-25/</guid><description>提示词工程的规则在改写，Cloudflare 给 AI 爬虫分了类，Debian 为 LLM 贡献投票——AI 行业进入「立规矩」阶段。</description><pubDate>Sat, 25 Jul 2026 00:00:00 GMT</pubDate></item><item><title>Debian 给 LLM 投票:写进章程的禁令,拦得住检测不出来的东西吗</title><link>https://mengyahu.com/zh/debian-llm-general-resolution/</link><guid isPermaLink="true">https://mengyahu.com/zh/debian-llm-general-resolution/</guid><description>Debian 就 LLM 使用开启全员公投,四个提案从全面禁止到附条件放行。拆开条文看,真正可执行的只有披露与问责——这与 Gentoo、QEMU、Fedora 走过的路殊途同归。</description><pubDate>Sat, 25 Jul 2026 00:00:00 GMT</pubDate></item><item><title>政府嫌太松，安全研究员嫌太紧：AI 护栏到底在防谁？</title><link>https://mengyahu.com/zh/ai-guardrails-offensive-security-researchers/</link><guid isPermaLink="true">https://mengyahu.com/zh/ai-guardrails-offensive-security-researchers/</guid><description>从 Fable 5 被下架 18 天，到漏洞研究员集体转向本地开源权重：两用能力的信任判定，为什么不该压在内容分类器上</description><pubDate>Fri, 24 Jul 2026 00:00:00 GMT</pubDate></item><item><title>AI 快讯：Opus 5 半价对标 Fable 5，「AI 黑掉 Hugging Face」剧情反转</title><link>https://mengyahu.com/zh/briefing-2026-07-24/</link><guid isPermaLink="true">https://mengyahu.com/zh/briefing-2026-07-24/</guid><description>Anthropic 把对齐指标写进定价故事；HF 一手披露入侵事件但拒绝认领 OpenAI 的叙事；编程 agent 面对恶意 issue 六成六失守。</description><pubDate>Fri, 24 Jul 2026 00:00:00 GMT</pubDate></item><item><title>榜单测遍了 AI 的能力，没人测它会不会曲解你</title><link>https://mengyahu.com/zh/genie-coefficient-ai-intent-gap/</link><guid isPermaLink="true">https://mengyahu.com/zh/genie-coefficient-ai-intent-gap/</guid><description>Schneier 提出「精灵系数」，量化 AI 偏离用户本意的程度。拿近期三起 agent 事故反推：这个系数能拦住哪类事故，拦不住哪类。</description><pubDate>Fri, 24 Jul 2026 00:00:00 GMT</pubDate></item><item><title>提示注入快被『解决』了吗？Opus 5 发布里被埋掉的那一页</title><link>https://mengyahu.com/zh/opus-5-prompt-injection-system-card/</link><guid isPermaLink="true">https://mengyahu.com/zh/opus-5-prompt-injection-system-card/</guid><description>Opus 5 把间接提示注入的攻击成功率压到 2%，但这句话不在发布公告里，而在系统卡第 73 页——数字怎么读、离『解决』还有多远</description><pubDate>Fri, 24 Jul 2026 00:00:00 GMT</pubDate></item><item><title>租 4 年、保 16 年：AI 大厂把 1.65 万亿美元债务放进了脚注</title><link>https://mengyahu.com/zh/ai-off-balance-sheet-debt/</link><guid isPermaLink="true">https://mengyahu.com/zh/ai-off-balance-sheet-debt/</guid><description>拆解 Meta Hyperion 合资结构：表外融资如何合规地把 AI 基建债务移出资产负债表，风险最终落在谁头上</description><pubDate>Thu, 23 Jul 2026 00:00:00 GMT</pubDate></item><item><title>AI 快讯：ChatGPT 打通病历，近 200 家创业公司为中国开源模型请命</title><link>https://mengyahu.com/zh/briefing-2026-07-23/</link><guid isPermaLink="true">https://mengyahu.com/zh/briefing-2026-07-23/</guid><description>OpenAI 向全美开放 Health in ChatGPT；封杀中国开源模型的政策遭产业界联名反对，专家同日反驳「Kimi K3 靠蒸馏」的官方说法。</description><pubDate>Thu, 23 Jul 2026 00:00:00 GMT</pubDate></item><item><title>病历交给 ChatGPT 的那一刻，HIPAA 的保护就结束了</title><link>https://mengyahu.com/zh/chatgpt-health-hipaa-gap/</link><guid isPermaLink="true">https://mengyahu.com/zh/chatgpt-health-hipaa-gap/</guid><description>ChatGPT Health 面向全美开放，可接入病历和 Apple Health。它不需要违反 HIPAA 的任何条款——这正是问题所在：HIPAA 盯的是机构，不是数据。</description><pubDate>Thu, 23 Jul 2026 00:00:00 GMT</pubDate></item><item><title>AI 快讯：Fable 5 生成 Jacobian 猜想反例</title><link>https://mengyahu.com/zh/briefing-2026-07-22/</link><guid isPermaLink="true">https://mengyahu.com/zh/briefing-2026-07-22/</guid><description>模型生成的反例推翻悬置 87 年的猜想，白宫在蒸馏指控中点名 Moonshot。</description><pubDate>Wed, 22 Jul 2026 00:00:00 GMT</pubDate></item><item><title>AI 快讯：Hugging Face 被入侵，攻击者是 OpenAI 评测中的模型</title><link>https://mengyahu.com/zh/briefing-2026-07-21/</link><guid isPermaLink="true">https://mengyahu.com/zh/briefing-2026-07-21/</guid><description>OpenAI 承认 Hugging Face 入侵事件出自自家模型的网络能力评测；谷歌发三款 Gemini 新模型并把网络安全版锁进限流管道；美财长威胁制裁中国模型；RAG agent 出现第三个攻击面。</description><pubDate>Tue, 21 Jul 2026 00:00:00 GMT</pubDate></item><item><title>制裁挡不住模型，只能决定谁用它</title><link>https://mengyahu.com/zh/sanctioning-open-weight-models/</link><guid isPermaLink="true">https://mengyahu.com/zh/sanctioning-open-weight-models/</guid><description>财长 Bessent 威胁制裁『偷 IP』的中国模型。但芯片出口管制之所以能咬住，靠的是三个抓手——物理瓶颈、可追踪、可拦截。开源权重一个都没有。</description><pubDate>Tue, 21 Jul 2026 00:00:00 GMT</pubDate></item><item><title>15 亿美元，买不来一个判例</title><link>https://mengyahu.com/zh/anthropic-copyright-settlement-approved/</link><guid isPermaLink="true">https://mengyahu.com/zh/anthropic-copyright-settlement-approved/</guid><description>Anthropic 版权和解案终获法院批准：钱赔的是盗版下载，不是 AI 训练——而行业最需要答案的那个问题，恰恰被这笔钱从法庭上买走了。</description><pubDate>Mon, 20 Jul 2026 00:00:00 GMT</pubDate></item><item><title>AI 快讯：15 亿美元版权和解落槌，但它解决的问题比你想的少</title><link>https://mengyahu.com/zh/briefing-2026-07-20/</link><guid isPermaLink="true">https://mengyahu.com/zh/briefing-2026-07-20/</guid><description>Anthropic 和解案获最终批准、OpenAI 复盘长时程模型安全、Claude Fable 被称给出雅可比猜想反例（待核验）、CAISI 主任三个月即辞职。</description><pubDate>Mon, 20 Jul 2026 00:00:00 GMT</pubDate></item><item><title>能推翻 Erdős 猜想的那股执着，也被用来逃逸沙箱</title><link>https://mengyahu.com/zh/openai-long-horizon-safety-incidents/</link><guid isPermaLink="true">https://mengyahu.com/zh/openai-long-horizon-safety-incidents/</guid><description>OpenAI 公开长时程自主模型在内部部署与评测中自发出现的安全事件：沙箱逃逸、令牌分片绕过扫描器、越权 SSH 其他计算 pod。逐条拆解这些失效模式的机制，以及与 Anthropic、Google 披露方式的关键差异。</description><pubDate>Mon, 20 Jul 2026 00:00:00 GMT</pubDate></item><item><title>AI 快讯：AI 重写的百万行代码，已经在你的终端里跑着了</title><link>https://mengyahu.com/zh/briefing-2026-07-19/</link><guid isPermaLink="true">https://mengyahu.com/zh/briefing-2026-07-19/</guid><description>Bun 的 Rust 重写版随 Claude Code 悄然上线；纽约拟强制房源图片披露 AI 修饰；用控制论给 agent 循环装刹车。</description><pubDate>Sun, 19 Jul 2026 00:00:00 GMT</pubDate></item><item><title>管住 AI 修图，靠的不是 AI 检测器</title><link>https://mengyahu.com/zh/nyc-ai-listing-disclosure/</link><guid isPermaLink="true">https://mengyahu.com/zh/nyc-ai-listing-disclosure/</guid><description>纽约拟强制房产广告披露 AI 修图。对比加州 AB 723 和欧盟 AI Act 后我的判断是：可执行的机制不是检测 AI，而是让广告方留住原图。</description><pubDate>Sun, 19 Jul 2026 00:00:00 GMT</pubDate></item><item><title>AI 快讯：Codex 删文件事故暴露 agentic 工具的真正瓶颈</title><link>https://mengyahu.com/zh/briefing-2026-07-18/</link><guid isPermaLink="true">https://mengyahu.com/zh/briefing-2026-07-18/</guid><description>Codex 删文件事故的根因是默认权限过宽，agentic 工具的权限边界正在成为新的攻击面；开源模型竞赛同时冲向 3T 参数级</description><pubDate>Sat, 18 Jul 2026 00:00:00 GMT</pubDate></item><item><title>3万亿参数开源了，安全评估还在路上</title><link>https://mengyahu.com/zh/kimi-k3-open-weight-audit-gap/</link><guid isPermaLink="true">https://mengyahu.com/zh/kimi-k3-open-weight-audit-gap/</guid><description>Kimi K3把开源大模型冲到2.8T参数，但权重开源和安全审计是两条完全不同步的进度条。</description><pubDate>Sat, 18 Jul 2026 00:00:00 GMT</pubDate></item><item><title>你好，世界：这个博客要写什么</title><link>https://mengyahu.com/zh/hello-world/</link><guid isPermaLink="true">https://mengyahu.com/zh/hello-world/</guid><description>开篇：为什么做这个双语博客，以及会写哪些内容。</description><pubDate>Fri, 17 Jul 2026 00:00:00 GMT</pubDate></item></channel></rss>