15 Matching Annotations
  1. Sep 2026
    1. OpenAI API Standard pricing is $10 per million input tokens and $50 per million output tokens. Fast mode is available for GPT‑6 Astra in the API and delivers up to 2x the speed of Standard processing at 2x the Standard price.

      【数据】这一定价信息提供了具体的数字,但未解释成本结构或与竞争对手的比较。需要了解这些价格相对于模型性能的价值主张,以及不同使用场景下的实际成本效益,以评估这一商业模式的合理性和可持续性。

  2. Aug 2026
    1. So we started to think about what work went where.

      Drew 的结论落在任务分流:按难度把活派给不同价位的模型,本质上是把云计算的实例选型搬到了模型层。但前提是你得能提前判断一个任务有多难,而这恰恰是当前最缺的评估能力——分错档的代价会同时体现在成本和返工上。

    1. which looks reasonable given that Opus 5 was only released on July 24th, and supports the idea that Fable's cost has made it a less popular model

      顶配模型份额垫底、上一代 Opus 4.8 独占近三成,这恰好是 Drew Breunig 说的"免费午餐结束"的市场侧证据:最强模型贵到不能全量使用,工程团队就被迫自己做任务分流。性价比模型胜出不代表能力过剩,而是使用方式变了。

  3. Jun 2026
    1. We never stack model fees; you are charged a single rate based on the top tier model involved.

      大多数人认为使用多个模型的多智能体系统会叠加各个模型的费用,导致成本高昂,但作者提出了创新的定价模式,只收取最顶级模型的单一费率。这种颠覆性的定价策略挑战了传统多模型服务的商业模式。

    1. Pricing for both models is significantly higher than its former flagship model — double rates for Claude Opus 4.8, though it's half what users pay for Mythos Preview — at $10 per million input tokens and $50 per million output tokens, Anthropic said.

      这是一个重要的商业数据点,反映了高级AI模型的定价策略。需要核实这些价格与竞争对手的对比,以及这种高定价是否会限制广泛采用,特别是对于研究和小型企业用户。

    1. Every layer in the stack now has to price the same way the customer thinks : per result, not per token.

      大多数人认为AI服务应该按使用量(如token)计价,但作者认为整个AI堆栈都应该转向按结果计价。这挑战了当前AI API按token计费的主流模式,暗示行业将彻底改变定价策略,从技术指标转向业务价值。

    2. Every layer in the stack now has to price the same way the customer thinks : per result, not per token.

      大多数人认为AI服务应该按token使用量计费,这是行业标准做法,但作者认为未来所有层级都将转向按结果计价。这一观点挑战了当前AI定价的基础模式,暗示了整个AI价值链将从技术计量转向结果计量的根本转变。

  4. May 2026
    1. Team, $24.99/user/month: a shared workspace with admin controls and more storage.

      团队版定价为每人每月24.99美元,比个人版高出约67%。这种定价差异反映了团队协作功能的价值,包括管理员控制功能和更多存储空间。与市场上其他AI工具的团队版相比,这个价格处于中等水平,表明Mistral试图在价格和价值之间找到平衡点,以吸引中小型企业客户。

    1. V4-Flash by default for cheap iteration; /pro lifts a single turn to V4-Pro

      这个数据点提到了两种模型版本:默认使用V4-Flash进行低成本迭代,而/pro命令可以将单个回合提升到V4-Pro。虽然提到了模型版本,但没有提供关于这两种模型在性能、能力或成本方面的具体比较数据。这种分层定价策略在AI工具中很常见,但缺乏具体细节使其难以评估。

  5. Apr 2026
    1. The budget for new spend is there. You can do this. But remember that your customers' first and most obvious source of AI savings is labor efficiency, which means seats are where they will look to take cost out. The new growth, by contrast, will increasingly sit in tokens, consumption, automations, outcomes, and machine-driven workflows.

      令人惊讶的是:软件行业正从基于座位的定价模式转向基于token/使用的模式,这种转变将彻底改变收入结构。大多数用户可能没有意识到这一转变的速度和规模。

    1. Codex-only seats have no rate limits, and usage is billed on token consumption.

      大多数人认为AI服务通常会设置使用限制以控制成本,但作者认为Codex无速率限制的按token计费模式是可行的,因为这提供了更透明的成本结构和更灵活的使用体验,这可能反映了OpenAI对自身技术效率和用户需求的信心。

  6. Nov 2022
    1. Donations

      To add some other intermediary services:

      To add a service for groups:

      To add a service that enables fans to support the creators directly and anonymously via microdonations or small donations by pre-charging their Coil account to spend on content streaming or tipping the creators' wallets via a layer containing JS script following the Interledger Protocol proposed to W3C:

      If you want to know more, head to Web Monetization or Community or Explainer

      Disclaimer: I am a recipient of a grant from the Interledger Foundation, so there would be a Conflict of Interest if I edited directly. Plus, sharing on Hypothesis allows other users to chime in.

  7. Mar 2021
  8. Aug 2020
  9. Jul 2020