3 Matching Annotations
  1. Last 7 days
    1. It also provides strong reasoning capabilities, scoring 97.7% on Big Bench Audio, while maintaining a highly competitive price point compared to other frontier models.

      【数据】97.7%的推理能力得分是一个极高的数字,需要了解Big Bench Audio的具体测试方法和评分标准。同时,文章提到'具有竞争力的价格点'但没有提供具体的价格信息或与其他模型的比较数据,这使得这一说法难以验证。

    2. Gemini 3.8 Live Extended Thinking provides enterprise-grade task completion and intelligence, capturing the #1 overall spot on Artificial Analysis' Speech to Speech Quality Index (82.6), and leads in agentic task completion with 68.6% on τ-Voice and 35.1% on Sierra's τ-Voice-banking benchmark.

      【数据】这些具体的性能指标需要独立验证,特别是第三方基准测试的排名和分数。Gemini 3.8 Live Extended Thinking在多个基准测试中声称领先,但没有提供完整的测试方法和比较对象的信息,这些数据的可信度需要进一步确认。

  2. Feb 2020