APrompt4EM: Augmented Prompt Tuning for Generalized Entity Matching

Xia, Yikuan; Chen, Jiazun; Li, Xinchi; Gao, Jun

Full-text links:

Download:

Current browse context:

cs.CL

< prev | next >

new | recent | 2405

Computer Science > Computation and Language

Title: APrompt4EM: Augmented Prompt Tuning for Generalized Entity Matching

Authors: Yikuan Xia, Jiazun Chen, Xinchi Li, Jun Gao

(Submitted on 8 May 2024)

Abstract: Generalized Entity Matching (GEM), which aims at judging whether two records represented in different formats refer to the same real-world entity, is an essential task in data management. The prompt tuning paradigm for pre-trained language models (PLMs), including the recent PromptEM model, effectively addresses the challenges of low-resource GEM in practical applications, offering a robust solution when labeled data is scarce. However, existing prompt tuning models for GEM face the challenges of prompt design and information gap. This paper introduces an augmented prompt tuning framework for the challenges, which consists of two main improvements. The first is an augmented contextualized soft token-based prompt tuning method that extracts a guiding soft token benefit for the PLMs' prompt tuning, and the second is a cost-effective information augmentation strategy leveraging large language models (LLMs). Our approach performs well on the low-resource GEM challenges. Extensive experiments show promising advancements of our basic model without information augmentation over existing methods based on moderate-size PLMs (average 5.24%+), and our model with information augmentation achieves comparable performance compared with fine-tuned LLMs, using less than 14% of the API fee.

Subjects:	Computation and Language (cs.CL); Artificial Intelligence (cs.AI)
Cite as:	arXiv:2405.04820 [cs.CL]
	(or arXiv:2405.04820v1 [cs.CL] for this version)

Submission history

From: Yikuan Xia [view email]
[v1] Wed, 8 May 2024 05:38:56 GMT (3180kb,D)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> cs > arXiv:2405.04820

Download:

Current browse context:

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

Computer Science > Computation and Language

Title: APrompt4EM: Augmented Prompt Tuning for Generalized Entity Matching

Submission history