返回话题榜我不知道我这个做法将会狠狠推进数学一大步,还是彻底终结数学。
党哥这里隆重推出The Last Math Competition(“最后数学竞赛”),可能将是全人类最后一场、也将是最旷日持久的一场数学竞赛。
https://t.co/WbogrgA0JX
以下是党哥制定的 The Last Math Competition 的规则:
1. 举办方每周会使用 AI Agent 新增 10000 个纯数学领域的猜想,添加到 ./conjectures 文件夹中;
2. 人类和 AI Agent 将共同对现存的所有猜想进行不同维度的打分,预估这些猜想证明或者证伪的难度,评价猜想的重要程度;
3. 人类和 AI Agent 可以共同提交对于每个猜想的完整证明,以 pull request 的形式提交在对应序号的文件夹中(比如 ./solutions/00000000001/my_submission_20260912041426),其中需要同时包含 LaTeX 源代码、PDF 文档和 Lean 4 项目,经过完整 review 后,完成证明或者证伪该猜想的提交将会被 merge 进来;
4. 举办方将会持续维护和更新一个含有所有猜想相关统计数据的表格。表格中的每一行对应一个猜想的所有信息,其中包括猜想的难度预估、重要程度打分、���否 well-defined、目前是否被证明或者证伪、第一次成功解决的时间、成功解决者的姓名和 affiliation 等信息;
5. 根据现存的猜想、对现存猜想的评估、对现存猜想的成功证明或者证伪,举办方将调整使用 AI Agent 生成猜想的策略,来逐步提高未来生成数学猜想的质量,以及可能逐步提高生成猜想的数量。
由于比赛尚在早期,我们必须声明:在早期生成的数学猜想的平均质量比较差,一部分猜想可能定义不充分或者存在错误的条件,或者存在显而易见的错误,甚至可能 "not even wrong"。所以我们将会根据对猜想的统计数据表格,来提升新生成猜想的质量和方法。
党哥的长远目标是:
1. 经过长期的比赛和反馈,我们将尝试逐步提升猜想质量。在长期大量猜想生成的过程中,我们希望能够创造若干重要猜想,通过对这些重要猜想的证明和证伪,来推动数学知识体系的推进;
2. 我们将探索以 AI Agent 为主导的数学研究和证明的技术路线。本比赛作为公开的比赛和 benchmark,将会实时评测所有模型的性能、harness 的性能、数学证明工具的性能,整个社区的所有参与者将一起探索 AI 时代科学研究的方法论;
3. 整个竞赛的所有猜想、对猜想的评估和打分、对猜想的最终证明或证伪,作为全世界 AI Agent 共同的经过验证和评议的知识成果,可以用于未来的数学研究,也可以用于未来的 LLM 训练。 与其关心能否正确verify your proof,
更重要的是能否提出足够多、足够重要、足够有价值的猜想,供足够多的AI Agent和人类去探索。
人和Agent一起提出猜想,一起评估猜想,一起证明猜想,一起验证证明是否正确,The Last Math Competition的终极目标,就是让全人类数学家+AI公司集体快速卷这个loop。
我们每周提出10000个有一定价值的数学猜想,也许其中有简单题目,也许有下一个黎曼猜想,但留下来的证明一定是有价值的。
again,只有彻底把人类解放出来,让出题、评估题、做题、判答案整个流程,彻底留给全人类数学家手下的agent,大规模无休止24小时跑起来,数学家、模型公司、harness公司、数学爱好者各显神通,才能狠狠推进人类数学的前进速度。
希望大家给我���github star。
https://t.co/WbogrgA0JX I don't know whether this approach of mine will boldly advance mathematics by a huge leap, or utterly end mathematics altogether.
Lidang is here proudly launching The Last Math Competition (“最后数学竞赛”), which may be humanity's final math competition—and also the longest-lasting one.
https://t.co/WbogrgA0JX
Below are the rules for The Last Math Competition as formulated by Lidang:
1. The organizers will use an AI Agent each week to add 10,000 new conjectures in the pure mathematics domain to the ./conjectures folder;
2. Humans and AI Agents will jointly score all existing conjectures across different dimensions, estimating the difficulty of proving or disproving them, and evaluating the conjectures' importance;
3. Humans and AI Agents can jointly submit complete proofs for each conjecture in the form of pull requests to the corresponding numbered folder (e.g., ./solutions/00000000001/my_submission_20260912041426), which must include LaTeX source code, PDF document, and Lean 4 project; after a full review, submissions that complete the proof or disproof of the conjecture will be merged;
4. The organizers will continuously maintain and update a table containing statistical data for all conjectures. Each row in the table corresponds to all information for one conjecture, including estimated difficulty, importance score, whether it is well-defined, whether it has currently been proven or disproven, the time of the first successful solution, the name and affiliation of the successful solver, and other details;
5. Based on existing conjectures, evaluations of existing conjectures, and successful proofs or disproofs of existing conjectures, the organizers will adjust the strategy for generating conjectures with the AI Agent, to gradually improve the quality of future mathematical conjectures generated, and possibly gradually increase the number of conjectures generated.
Since the competition is still in its early stages, we must state: The average quality of mathematical conjectures generated early on is relatively poor; some conjectures may be insufficiently defined or have erroneous conditions, or contain obvious errors, or may even be "not even wrong." Therefore, we will use the statistical data table for conjectures to improve the quality and methods for generating new conjectures.
Lidang's long-term goals are:
1. Through long-term competition and feedback, we will attempt to gradually improve conjecture quality. In the process of generating a large volume of conjectures over the long term, we hope to create several important conjectures, and through proving and disproving these important conjectures, advance the mathematical knowledge system;
2. We will explore a technical route for mathematical research and proof led by AI Agents. As an open competition and benchmark, this event will continuously evaluate the performance of all models, the performance of harnesses, and the performance of mathematical proof tools; all participants in the entire community will jointly explore methodologies for scientific research in the AI era;
3. All conjectures from the entire competition, evaluations and scores of conjectures, and final proofs or disproofs of conjectures—as verified and peer-reviewed knowledge outcomes shared by AI Agents worldwide—can be used for future mathematical research and also for training future LLMs. This is The Last Math Competition — possibly the last mathematics competition humanity will ever hold, and surely its most protracted one.
https://t.co/fyalfxzXCE
AI热议全部数据
立党推出最后数学竞赛,主张人类与AI协作探索猜想
立党连续发布中英文推文,宣布推出The Last Math Competition,即“最后数学竞赛”,并提出让人类与AI Agent共同参与数学探索。其宣传将竞赛描述为可能是人类最后一场、也将是最旷日持久的数学竞赛,随后把讨论重点从能否正确验证证明,转向能否提出足够多、重要且有价值的猜想。他设想人和Agent共同提出与评估猜想、寻找证明并验证结果,使活动定位从单纯解题扩展至数学研究流程协作。本轮四条推文均来自立党本人,其中包含英文推广,不能视为多个独立来源的认可。候选正文未完整展示竞赛规则、参赛方式、评审机制或实际成果,因此不能确认活动已经运行,也不能将推进或终结数学的表述视为能力结论。
#立党#The Last Math Competition#数学竞赛#AI Agent#数学猜想
话题热度
194
帖子数
4
曝光
16.1万
参与创作者
1
话题热度趋势
从首次出现到最后更新;不足 7 天时补足 7 天观察窗。
全部关联帖子
共 4 条
