OYXOY : A Modern NLP Test Suite for Modern Greek

Konstantinos Kogkalidis*, Stergios Chatzikyriakidis*, Eirini Chrysovalantou Giannikouri, Christina Klironomou, Christina Koula, Thelka Pasparaki, Efthymia Sakellariou, Vasiliki Katsouli, Dimitris Papadakis, Erofili Psaltaki, Hara Soupiona

*Corresponding author for this work

Research output: Chapter in Book/Report/Conference proceedingConference article in proceedingsScientificpeer-review

Abstract

This paper serves as a foundational step towards the development of a linguistically motivated and technically relevant evaluation suite for Greek NLP. We initiate this endeavor by introducing four expert-verified evaluation tasks, specifically targeted at natural language inference, word sense disambiguation (through example comparison or sense selection) and metaphor detection. More than language-adapted replicas of existing tasks, we contribute two innovations which will resonate with the broader resource and evaluation community. Firstly, our inference dataset is the first of its kind, marking not just one, but rather all possible inference labels, accounting for possible shifts due to e.g. ambiguity or polysemy. Secondly, we demonstrate a cost-efficient method to obtain datasets for under-resourced languages. Using ChatGPT as a language-neutral parser, we transform the Dictionary of Standard Modern Greek into a structured format, from which we derive the other three tasks through simple projections. Alongside each task, we conduct experiments using currently available state of the art machinery. Our experimental baselines affirm the challenging nature of our tasks and highlight the need for expedited progress in order for the Greek NLP ecosystem to keep pace with contemporary mainstream research.

Original languageEnglish
Title of host publicationEACL 2024 - 18th Conference of the European Chapter of the Association for Computational Linguistics, Findings of EACL 2024
EditorsYvette Graham, Matthew Purver, Matthew Purver
PublisherAssociation for Computational Linguistics
Pages311-322
Number of pages12
ISBN (Electronic)979-8-89176-093-6
Publication statusPublished - 2024
MoE publication typeA4 Conference publication
EventConference of the European Chapter of the Association for Computational Linguistics - St. Julian's, Malta
Duration: 17 Mar 202422 Mar 2024
Conference number: 18

Conference

ConferenceConference of the European Chapter of the Association for Computational Linguistics
Abbreviated titleEACL
Country/TerritoryMalta
CitySt. Julian's
Period17/03/202422/03/2024

Fingerprint

Dive into the research topics of 'OYXOY : A Modern NLP Test Suite for Modern Greek'. Together they form a unique fingerprint.

Cite this