Synapse
⌘+K
Synapse
PulseExploreClubsResearchersJournals
Instagram
HomeClubsExplore
March 12, 2024Open Access

FineMath: A Fine-Grained Mathematical Evaluation Benchmark for Chinese Large Language Models

View Full Paper
Ask AI
Bookmark
Share

Authors

YLYan LiuRJRenren JinLSLin Shi

Discussion

Loading...

Member takes

Overview

Key Points

Key points are not available for this paper at this time.

Cite This Study

Liu et al. (2024) studied this question.

synapsesocial.com/papers/68e747e6b6db6435876c0d1fhttps://doi.org/10.48550/arxiv.2403.07747
View Full Paper
Ask AI
Bookmark
Share

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1ConceptMath: A Bilingual Concept-wise Benchmark for Measuring Mathematical Reasoning of Large Language Models2024 · 1 citations
  2. 2CMMaTH: A Chinese Multi-modal Math Skill Evaluation Benchmark for Foundation Models2024 · 2 citations
  3. 3Mathify: Evaluating Large Language Models on Mathematical Problem Solving Tasks2024 · 5 citations
  4. 4Mathfish: Evaluating Language Model Math Reasoning via Grounding in Educational Curricula2024
  5. 5MathBench: Evaluating the Theory and Application Proficiency of LLMs with a Hierarchical Mathematics Benchmark2024 · 1 citations