C 2022

Evaluation of Automatically Constructed Word Meaning Explanations

STARÁ, Marie, Pavel RYCHLÝ and Aleš HORÁK

Basic information

Original name

Evaluation of Automatically Constructed Word Meaning Explanations

Authors

STARÁ, Marie, Pavel RYCHLÝ and Aleš HORÁK

Edition

Rickmansworth, UK, Logically Speaking: A Festschrift for Marie Duží, p. 99-112, 14 pp. Tributes, Volume 49, 2022

Publisher

College Publications

Other information

Language

English

Type of outcome

Kapitola resp. kapitoly v odborné knize

Field of Study

10200 1.2 Computer and information sciences

Country of publisher

United Kingdom of Great Britain and Northern Ireland

Confidentiality degree

není předmětem státního či obchodního tajemství

Publication form

printed version "print"

Organization unit

Faculty of Informatics

ISBN

978-1-84890-419-4

Keywords in English

explanations; word sketches; explanation construction

Tags

International impact, Reviewed
Změněno: 29/3/2023 14:33, RNDr. Pavel Šmerk, Ph.D.

Abstract

V originále

Preparing exact and comprehensive word meaning explanations is one of the key steps in the process of monolingual dictionary writing. In standard methodology, the explanations need an expert lexicographer who spends a substantial amount of time checking the consistency between the descriptive text and corpus evidence. In the following text, we present a new tool that derives explanations automatically based on collective information from very large corpora, particularly on word sketches. We also propose a quantitative evaluation of the constructed explanations, concentrating on explanations of nouns. The methodology is to a certain extent language independent; however, the presented verification is limited to Czech and English. We show that the presented approach allows to create explanations that contain data useful for understanding the word meaning in approximately 90% of cases. However, in many cases, the result requires post-editing to remove redundant information.

Links

LM2018101, research and development project
Name: Digitální výzkumná infrastruktura pro jazykové technologie, umění a humanitní vědy (Acronym: LINDAT/CLARIAH-CZ)
Investor: Ministry of Education, Youth and Sports of the CR