ملف الباحث
Xiaoshuai Song
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
We-Math: Does Your Large Multimodal Model Achieve Human-like Mathematical Reasoning?
2024 · arXiv (Cornell University)
Visual mathematical reasoning, as a fundamental visual reasoning ability, has received widespread attention from the Large Multimodal Models (LMMs) community. Existing benchmarks, such as MathVista and MathVerse, focus more on the result-oriented performance but neglect …