SemiHMER: Semi-supervised Handwritten Mathematical Expression Recognition using pseudo-labels

Chen, Kehua; Shen, Haoyang

Computer Science > Computer Vision and Pattern Recognition

arXiv:2502.07172 (cs)

[Submitted on 11 Feb 2025 (v1), last revised 20 Feb 2025 (this version, v3)]

Title:SemiHMER: Semi-supervised Handwritten Mathematical Expression Recognition using pseudo-labels

Authors:Kehua Chen, Haoyang Shen

View PDF HTML (experimental)

Abstract:In this paper, we study semi-supervised Handwritten Mathematical Expression Recognition (HMER) via exploring both labeled data and extra unlabeled data. We propose a novel consistency regularization framework, termed SemiHMER, which introduces dual-branch semi-supervised learning. Specifically, we enforce consistency between the two networks for the same input image. The pseudo-label, generated by one perturbed recognition network, is utilized to supervise the other network using the standard cross-entropy loss. The SemiHMER consistency encourages high similarity between the predictions of the two perturbed networks for the same input image and expands the training data by leveraging unlabeled data with pseudo-labels. We further introduce a weak-to-strong strategy by applying different levels of augmentation to each branch, effectively expanding the training data and enhancing the quality of network training. Additionally, we propose a novel module, the Global Dynamic Counting Module (GDCM), to enhance the performance of the HMER decoder by alleviating recognition inaccuracies in long-distance formula recognition and reducing the occurrence of repeated characters. The experimental results demonstrate that our work achieves significant performance improvements, with an average accuracy increase of 5.47% on CROHME14, 4.87% on CROHME16, and 5.25% on CROHME19, compared to our baselines.

Comments:	17 pages,3 figures
Subjects:	Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI)
Cite as:	arXiv:2502.07172 [cs.CV]
	(or arXiv:2502.07172v3 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2502.07172

Submission history

From: Kehua Chen [view email]
[v1] Tue, 11 Feb 2025 01:39:11 UTC (238 KB)
[v2] Wed, 19 Feb 2025 02:52:10 UTC (1,052 KB)
[v3] Thu, 20 Feb 2025 01:17:01 UTC (1,052 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:SemiHMER: Semi-supervised Handwritten Mathematical Expression Recognition using pseudo-labels

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:SemiHMER: Semi-supervised Handwritten Mathematical Expression Recognition using pseudo-labels

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators