ViANLI: Adversarial Natural Language Inference for Vietnamese

Van Huynh, Tin; Van Nguyen, Kiet; Nguyen, Ngan Luu-Thuy

Computer Science > Computation and Language

arXiv:2406.17716v1 (cs)

[Submitted on 25 Jun 2024 (this version), latest version 1 Jul 2024 (v2)]

Title:ViANLI: Adversarial Natural Language Inference for Vietnamese

Authors:Tin Van Huynh, Kiet Van Nguyen, Ngan Luu-Thuy Nguyen

View PDF HTML (experimental)

Abstract:The development of Natural Language Processing (NLI) datasets and models has been inspired by innovations in annotation design. With the rapid development of machine learning models today, the performance of existing machine learning models has quickly reached state-of-the-art results on a variety of tasks related to natural language processing, including natural language inference tasks. By using a pre-trained model during the annotation process, it is possible to challenge current NLI models by having humans produce premise-hypothesis combinations that the machine model cannot correctly predict. To remain attractive and challenging in the research of natural language inference for Vietnamese, in this paper, we introduce the adversarial NLI dataset to the NLP research community with the name ViANLI. This data set contains more than 10K premise-hypothesis pairs and is built by a continuously adjusting process to obtain the most out of the patterns generated by the annotators. ViANLI dataset has brought many difficulties to many current SOTA models when the accuracy of the most powerful model on the test set only reached 48.4%. Additionally, the experimental results show that the models trained on our dataset have significantly improved the results on other Vietnamese NLI datasets.

Subjects:	Computation and Language (cs.CL)
Cite as:	arXiv:2406.17716 [cs.CL]
	(or arXiv:2406.17716v1 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2406.17716

Submission history

From: Tin Van Huynh [view email]
[v1] Tue, 25 Jun 2024 16:58:19 UTC (491 KB)
[v2] Mon, 1 Jul 2024 15:19:51 UTC (503 KB)

Computer Science > Computation and Language

Title:ViANLI: Adversarial Natural Language Inference for Vietnamese

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:ViANLI: Adversarial Natural Language Inference for Vietnamese

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators