But while the AI was ready to give up several times, it did keep adding debug code and analyzing it faithfully when I pushed.
能力其实在,缺的是自主坚持——这是 Linus 这条记录里最有价值的判断。对 agent 设计的启示不是继续堆推理长度,而是加一层"不得自行宣告不可解"的约束。反面推论也成立:不会追问、不敢施压的使用者会系统性地拿到更差的结果。
But while the AI was ready to give up several times, it did keep adding debug code and analyzing it faithfully when I pushed.
能力其实在,缺的是自主坚持——这是 Linus 这条记录里最有价值的判断。对 agent 设计的启示不是继续堆推理长度,而是加一层"不得自行宣告不可解"的约束。反面推论也成立:不会追问、不敢施压的使用者会系统性地拿到更差的结果。
I'd like to call it my tireless helper, but the AI several times stated flat out that this was impossible and unsolvable and that we should just write a report about it.
Linus 记录的失败模式比成功更值得读:模型多次断言问题无解、建议写份报告收工。长任务里"体面收尾"的倾向会表现为过早放弃,而这一步通常没有任何报错信号。这也解释了为什么同一个模型在不同人手里产出差距巨大。
Weirdly though, those things have started to blur for me already, which is quite upsetting.
Simon表达了对vibe coding和agentic engineering边界模糊的担忧,这让他感到不安。
Weirdly though, those things have started to blur for me already, which is quite upsetting. I thought we had a very clear delineation where vibe coding is the thing where you're not looking at the code at all. You might not even know how to program. You might be a non-programmer who asks for a thing, and gets a thing, and if the thing works, then great! And if it doesn't, you tell it that it doesn't work and cross your fingers.
作者原本认为vibe coding和agentic engineering有明确界限,但现在发现两者界限正在模糊,这让他感到不安。