Search engine for discovering works of Art, research articles, and books related to Art and Culture
ShareThis
Javascript must be enabled to continue!

Paraphrase-Sense-Tagged Sentences

View through CrossRef
Many natural language processing tasks require discriminating the particular meaning of a word in context, but building corpora for developing sense-aware models can be a challenge. We present a large resource of example usages for words having a particular meaning, called Paraphrase-Sense-Tagged Sentences (PSTS). Built on the premise that a word’s paraphrases instantiate its fine-grained meanings (i.e., bug has different meanings corresponding to its paraphrases fly and microbe) the resource contains up to 10,000 sentences for each of 3 million target-paraphrase pairs where the target word takes on the meaning of the paraphrase. We describe an automatic method based on bilingual pivoting used to enumerate sentences for PSTS, and present two models for ranking PSTS sentences based on their quality. Finally, we demonstrate the utility of PSTS by using it to build a dataset for the task of hypernym prediction in context. Training a model on this automatically generated dataset produces accuracy that is competitive with a model trained on smaller datasets crafted with some manual effort.
Title: Paraphrase-Sense-Tagged Sentences
Description:
Many natural language processing tasks require discriminating the particular meaning of a word in context, but building corpora for developing sense-aware models can be a challenge.
We present a large resource of example usages for words having a particular meaning, called Paraphrase-Sense-Tagged Sentences (PSTS).
Built on the premise that a word’s paraphrases instantiate its fine-grained meanings (i.
e.
, bug has different meanings corresponding to its paraphrases fly and microbe) the resource contains up to 10,000 sentences for each of 3 million target-paraphrase pairs where the target word takes on the meaning of the paraphrase.
We describe an automatic method based on bilingual pivoting used to enumerate sentences for PSTS, and present two models for ranking PSTS sentences based on their quality.
Finally, we demonstrate the utility of PSTS by using it to build a dataset for the task of hypernym prediction in context.
Training a model on this automatically generated dataset produces accuracy that is competitive with a model trained on smaller datasets crafted with some manual effort.

Related Results

Chinese Medical Paraphrase Generation: Based on Neural Machine Translation
Chinese Medical Paraphrase Generation: Based on Neural Machine Translation
Abstract Background: As people prefer to obtain medical knowledge online, medical intelligence question-answer systems based on question matching have attracted more and mo...
Pemerolehan Kalimat Bahasa Indonesia Anak Usia 4;0–5;0 Tahun
Pemerolehan Kalimat Bahasa Indonesia Anak Usia 4;0–5;0 Tahun
The purpose of this research is to describe the acquisition of Indonesian sentences for children aged 4;0-5;0 years old. Specifically, this study aims to explain the types of sente...
Analisis Jenis Kalimat Berdasarkan Tujuan pada Teks Drama Buku Bahasa dan Bersastra Indonesia Kelas XI Kurikulum Merdeka
Analisis Jenis Kalimat Berdasarkan Tujuan pada Teks Drama Buku Bahasa dan Bersastra Indonesia Kelas XI Kurikulum Merdeka
The Indonesian language is the national and state language, but it is also a compulsory subject at all levels of education. The research entitled Analysis Types of Sentences Based ...
ANALISIS PENGGUNAAN KALIMAT IMPERATIF DALAM TERJEMAHAN QS. 2 (AL-BAQARAH)
ANALISIS PENGGUNAAN KALIMAT IMPERATIF DALAM TERJEMAHAN QS. 2 (AL-BAQARAH)
This research aims to describe the forms of use and types of imperative sentences found in QS translations. 2 (Al-Baqarah). This type of research is descriptive qualitative. The qu...
The syntactic structure of the initial and final sentences
The syntactic structure of the initial and final sentences
The article analyzes the key role of syntactic features in the composition of the text. It was determined that the first and last sentences of the text are syntactically diverse. S...
ANALISIS KEEFEKTIFAN KALIMAT DALAM MAJALAH WARTA USK
ANALISIS KEEFEKTIFAN KALIMAT DALAM MAJALAH WARTA USK
ABSTRAK Penelitian analisis keefektifan kalimat dalam majalah Warta USK ini bertujuan untuk mengetahui bentuk-bentuk kalimat tidak efektif yang terdapat di dalam majalah Warta USK....
Comparison of Four Types of Interrogative Sentences in Chinese and Arabic Language (A Grammatical Study) Grammar
Comparison of Four Types of Interrogative Sentences in Chinese and Arabic Language (A Grammatical Study) Grammar
An interrogative sentence is a common sentence used in daily life and communication. The main purpose of an interrogative sentence is to ask questions that the inquirer does not ye...
Urdu Short Paraphrase Detection at Sentence Level
Urdu Short Paraphrase Detection at Sentence Level
Paraphrase detection systems uncover the relationship between two text fragments and classify them as paraphrased when they convey the same idea; otherwise non-paraphrased. Previou...

Back to Top