2 Matching Annotations
  1. Last 7 days
    1. with contributions from security researchers at Anthropic, OpenAI, and Google

      横向阅读的最高价值发现。Anthropic 7/29 的事故披露文脚注 2 写明「OpenAI/Hugging Face 事件发生于 ExploitGym 的一次评测」,并用整整一节把两起事件对立起来(我们主动发现 / 他们 0-day 逃逸)。但这里写着:ExploitGym 的构建有 Anthropic 安全研究员的贡献,同文还说「Anthropic ran the Opus 4.6 and Mythos Preview trials」——Anthropic 自己也在这套基准上跑模型。事故文对这层关系只字未提。结合已知的 Irregular 关系(商业供应商 + 白皮书合著方),这个领域里「独立第三方评测」的实际独立性比表面叙述低得多。

    1. Due to a misunderstanding between us and our evaluation partner, this was not the case, and internet access was available

      根因被归为「我们与评测伙伴之间的误解」,且主语消失(谁做的配置?)。补上外部信息后这句的分量会变:该伙伴 Irregular(原 Pattern Labs)2025 年 9 月由 Sequoia、Redpoint 领投融资 8000 万美元、估值 4.5 亿,同时是 OpenAI 与 Anthropic 的商业供应商,并与 Anthropic 合著过白皮书。本次「联合调查」是两家有商业与合著关系的公司互查,不是独立审计。