Multilingual Mathematical Reasoning: Advancing Open-Source LLMs in Hindi and English

Anand, Avinash; Prasad, Kritarth; Kirtani, Chhavi; Nair, Ashwin R; Nema, Manvendra Kumar; Jaiswal, Raj; Shah, Rajiv Ratn

Computer Science > Computation and Language

arXiv:2412.18415 (cs)

[Submitted on 24 Dec 2024]

Title:Multilingual Mathematical Reasoning: Advancing Open-Source LLMs in Hindi and English

Authors:Avinash Anand, Kritarth Prasad, Chhavi Kirtani, Ashwin R Nair, Manvendra Kumar Nema, Raj Jaiswal, Rajiv Ratn Shah

View PDF HTML (experimental)

Abstract:Large Language Models (LLMs) excel in linguistic tasks but struggle with mathematical reasoning, particularly in non English languages like Hindi. This research aims to enhance the mathematical reasoning skills of smaller, resource efficient open-source LLMs in both Hindi and English. We evaluate models like OpenHathi 7B, LLaMA-2 7B, WizardMath 7B, Mistral 7B, LLeMMa 7B, MAmmoTH 7B, Gemini Pro, and GPT-4 using zero-shot, few-shot chain-of-thought (CoT) methods, and supervised fine-tuning. Our approach incorporates curriculum learning, progressively training models on increasingly difficult problems, a novel Decomposition Strategy to simplify complex arithmetic operations, and a Structured Solution Design that divides solutions into phases. Our experiments result in notable performance enhancements. WizardMath 7B exceeds Gemini's accuracy on English datasets by +6% and matches Gemini's performance on Hindi datasets. Adopting a bilingual approach that combines English and Hindi samples achieves results comparable to individual language models, demonstrating the capability to learn mathematical reasoning in both languages. This research highlights the potential for improving mathematical reasoning in open-source LLMs.

Comments:	Accepted at AAAI 2025
Subjects:	Computation and Language (cs.CL); Artificial Intelligence (cs.AI)
Cite as:	arXiv:2412.18415 [cs.CL]
	(or arXiv:2412.18415v1 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2412.18415

Submission history

From: Chhavi Kirtani [view email]
[v1] Tue, 24 Dec 2024 13:07:29 UTC (2,136 KB)

Computer Science > Computation and Language

Title:Multilingual Mathematical Reasoning: Advancing Open-Source LLMs in Hindi and English

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Multilingual Mathematical Reasoning: Advancing Open-Source LLMs in Hindi and English

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators