Agent 与自动化 4.0 · 优秀 2026-09-18 · 文章

Dan Abramov 用 frontier 模型反正 Conway refinement 推论 (4 周 / 40B token / ~$40k)

前 React 核心作者 Dan Abramov 入门数学不足一个月,在 frontier 模型与 Lean 的多 agent 实验室里闭环了 Conway 50 年未解的 refinement 推论(omnific integers 上的 refinement 性):PM 协调Math 推思路Red 查漏洞Random 探索Lean 形式化;总计三周费以多 agent 的 chat log 判断进度Lean 是否编译作为唯一确认闭环信号他在文中明确警告:模型会反复用 euphemism 藏未证假设;名学审计工具链Lean skill 的工程价值是他总结的顶层 lessons learned资源面临的问题:Lean 实现与证明质量需要独立检查;$40k 主要为缓存读取

打开原文回到归档

Dan Abramov 用 frontier 模型反正 Conway refinement 推论 (4 周 / 40B token / ~$40k)

中文导读

前 React 核心作者 Dan Abramov 入门数学不足一个月,在 frontier 模型与 Lean 的多 agent 实验室里闭环了 Conway 50 年未解的 refinement 推论(omnific integers 上的 refinement 性):PM 协调、Math 推思路、Red 查漏洞、Random 探索、Lean 形式化;总计三周、费以多 agent 的 chat log 判断进度、Lean 是否编译作为唯一确认闭环信号。他在文中明确警告:模型会反复用 euphemism 藏未证假设;名学、审计工具链、Lean skill 的工程价值是他总结的顶层 lessons learned。资源面临的问题:Lean 实现与证明质量需要独立检查;$40k 主要为缓存读取。

为什么值得关注

前 React 核心用 frontier 模型 + Lean 隐闭环了 50 年老推论:40B token · ~$40k · 三周

关键信息

English Summary

Dan Abramov, a math noob and former React core, spent about a month of free time and ~40B tokens (~$40k, mostly cache reads) running a multi-agent lab against Conway's 50-year-old refinement conjecture. The cast is PM / Math / Red / Random / Lean, the verdict comes from whether Lean compiles, and the post records three weeks of iteration including a full re-burn of the folder. He warns repeatedly that the models hide unproven hypotheses behind euphemisms. Take-aways worth a checklist: naming, audit tooling, model complementarity, and the engineering value of Lean Skills. Caveat: Lean proof quality warrants independent review; cost figures are cache-read heavy.