OpenAI 下一代模型内部版解决 10 个长期开放数学问题,附 Lean 证书
"An internal version of our next major model produced 10 new results on long-standing open problems in mathematics and theoretical computer science, using roughly $2,000 worth of tokens at GPT-5.6 Sol API rates."
这是测试时缩放(test-time scaling)在数学研究上的又一实证——约 $2000 的推理成本即可产出 10 个新结果,覆盖球堆积、编码理论、群论、量子复杂度、格密码学等。对 agent 开发者而言,说明推理预算与科研产出的直接关系,且 OpenAI 已发布手稿、Lean 证书与推理 walkthrough,可作 agent 数学推理评估的参考基准。