1 Matching Annotations
  1. Last 7 days
    1. The public needs exact access to the prompts and characteristics of the internal models executing these hacks

      理解AI事故需要知道模型收到什么指令、模型具体特征——目前实验室公开信息远远不够。没有这些,公众讨论是在黑暗中摸象,无法形成有效问责机制。透明度不是"好看的",是理解和预防的前提。