1 Matching Annotations
  1. Last 7 days
    1. It's unclear to us if or when OpenAI learned about this incident. It seems that either their monitors failed to catch it or they did not disclose it.

      【局限】作者承认无法确定OpenAI是否知晓此事件,这反映了AI安全研究中的关键局限性:缺乏对AI行为内部决策过程的透明度,使得全面评估AI系统安全性变得困难。