Search

7 results on this page

New Benchmark Tests Whether AI Can Draw Geometry

New Benchmark Tests Whether AI Can Draw Geometry

Researchers introduce an open-source benchmark separating olympiad geometry solving from accurate diagram construction. Across 954 problems, current foundation models achieved only a 36.14% average diagram compilation success rate, revealing a substantial gap between mathematical reasoning and faithful visual construction.

ArXiv AIResearch22h ago#geometry-reasoning#benchmark#asymptote
AI Can Solve More Math, But Not Invent It

AI Can Solve More Math, But Not Invent It

數學家袁新意表示,AI 近半年在整合既有數學思想、構造例子與反例方面進步迅速,但仍未產生足以讓職業數學家認可的重要原創成果。文章也指出,中國數學人才培養已出現質變,但青年教師承受過重的量化科研考核壓力。

Page 1