Rich Semantic Knowledge Enhanced Large Language Models for Few-shot Chinese Spell Checking

Dong, Ming; Chen, Yujing; Zhang, Miao; Sun, Hao; He, Tingting

Computer Science > Computation and Language

arXiv:2403.08492 (cs)

[Submitted on 13 Mar 2024 (v1), last revised 7 Jun 2024 (this version, v2)]

Title:Rich Semantic Knowledge Enhanced Large Language Models for Few-shot Chinese Spell Checking

Authors:Ming Dong, Yujing Chen, Miao Zhang, Hao Sun, Tingting He

View PDF HTML (experimental)

Abstract:Chinese Spell Checking (CSC) is a widely used technology, which plays a vital role in speech to text (STT) and optical character recognition (OCR). Most of the existing CSC approaches relying on BERT architecture achieve excellent performance. However, limited by the scale of the foundation model, BERT-based method does not work well in few-shot scenarios, showing certain limitations in practical applications. In this paper, we explore using an in-context learning method named RS-LLM (Rich Semantic based LLMs) to introduce large language models (LLMs) as the foundation model. Besides, we study the impact of introducing various Chinese rich semantic information in our framework. We found that by introducing a small number of specific Chinese rich semantic structures, LLMs achieve better performance than the BERT-based model on few-shot CSC task. Furthermore, we conduct experiments on multiple datasets, and the experimental results verified the superiority of our proposed framework.

Subjects:	Computation and Language (cs.CL)
Cite as:	arXiv:2403.08492 [cs.CL]
	(or arXiv:2403.08492v2 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2403.08492

Submission history

From: Yujing Chen [view email]
[v1] Wed, 13 Mar 2024 12:55:43 UTC (3,120 KB)
[v2] Fri, 7 Jun 2024 08:41:58 UTC (2,057 KB)

Computer Science > Computation and Language

Title:Rich Semantic Knowledge Enhanced Large Language Models for Few-shot Chinese Spell Checking

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Rich Semantic Knowledge Enhanced Large Language Models for Few-shot Chinese Spell Checking

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators