2 Matching Annotations
  1. Sep 2026
    1. Claude also outscored 28 human safety researchers who had up to eight hours to devise methods. On deception, for example, Claude's best method performed 20% better than the best human proposal.

      【数据】Claude在欺骗问题上表现优于28名人类安全研究员,其最佳方法比最佳人类提案高出20%。这一比较结果虽然令人印象深刻,但需要谨慎解读,因为人类研究员的时间限制(8小时)可能影响了其表现。

  2. Apr 2026
    1. We identify three fundamental bottlenecks in current DLMs: (1) Low introspective consistency. SDAR: 0.699 vs. I-DLM: 0.984. (2) Compute inefficiency. TiDAR: ~7.8x overhead vs. I-DLM: ~2.5x. (3) Infrastructure mismatch. SDAR slope=84 vs. I-DLM: 549.

      作者系统性地识别了现有DLMs的三大瓶颈,并通过量化对比展示了I-DLM的显著优势。内省一致性从0.699提升到0.984,计算开销从7.8x降低到2.5x,基础设施效率从84提升到549,这些数据不仅验证了I-DLM的有效性,也为未来DLM研究指明了方向。