ChatGPT 5.5 Pro 数学能力获顶尖数学家认可
推荐指数 70.0 NO. 012 · 2026.05.10
发布2026/05/09Score522Comments382
为什么值得看
菲尔兹奖得主 Timothy Gowers 测试 ChatGPT 5.5 Pro 后大幅上调了对大模型数学能力的评估。这意味着前沿 LLM 可能已突破高等数学推理的关键门槛,对自动化形式证明和数学研究辅助有直接影响。
编辑判断
Gowers 不是普通用户,他是组合数学领域的权威,曾主导 Polymath 众包数学项目,对机器辅助证明一直持审慎态度。他公开"大幅上调"评估的分量,相当于给 LLM 数学能力做了学术背书。
这会对 Lean、Coq 等交互式定理证明社区产生直接冲击——如果自然语言模型能可靠地生成可验证的形式化证明,数学家对专用证明助手的依赖模式将改变。做数学教育、科研工具或自动形式化的团队,需要重新评估产品定位,OpenAI 可能正在吃掉这个垂直领域。
社区反馈
意见分歧 256 条评论
核心争论:AI数学能力突破是否削弱人类研究者价值,教育评估体系如何应对
I saw Tim Gowers give a talk at the AMS-MAA joint meeting in Seattle about ten years ago where he predicted that in 100 years humans would no longer be doing research mathematics. I wonder if he’s adjusted his timeline. At the time I thought the key missing tool was a natural language search that ac
And Teichmüller thought that Germany would win WW2 and volunteered for the Eastern Front. Being a gifted mathematician does not make you right. In fact, mathematicians have a lot of bizarre theories.
Is the assessment system of undergraduate mathematics education no longer effective?