COMFRE: A Visualization for Comparing Word Frequencies in Linguistic Tasks

Research output: Chapter in Book/Report/Conference proceedingConference contributionScientificpeer-review

Researchers

Research units

  • Trinity College Dublin
  • University of Edinburgh

Abstract

Comparing frequency distributions is a basic task in statistics and in disciplines that rely on statistical analysis, such as corpus linguistics. However, support for comparing word frequencies between different corpora in corpus linguistics tasks such as lexical analysis and corpus-based translation studies, is often limited to fairly basic techniques like tabular word lists. While other visualizations such as word clouds do exist, they are not widely used in linguistic analysis tasks due to their lack of precision, unsuitability to dealing with the high frequencies of common words, and lack of effective mechanisms for direct manipulation. In this paper, we propose a visualization for comparing word frequencies across two corpora using a combination of slope charts and histogram contours. An interactive implementation of this visualization is also presented. The design of visualization, and the development of the prototype, have been guided through the involvement of expert linguist users.

Details

Original languageEnglish
Title of host publicationAVI 2018 - Proceedings of the 2018 International Conference on Advanced Visual Interfaces
Publication statusPublished - 29 May 2018
MoE publication typeA4 Article in a conference publication
EventInternational Working Conference on Advanced Visual Interfaces - Grosseto, Italy
Duration: 29 May 20181 Jun 2018
Conference number: 14
https://sites.google.com/dis.uniroma1.it/avi2018/home

Conference

ConferenceInternational Working Conference on Advanced Visual Interfaces
Abbreviated titleAVI
CountryItaly
CityGrosseto
Period29/05/201801/06/2018
Internet address

    Research areas

  • frequency comparisons, histograms, linguistics, set frequencies, slope charts, word clouds, word lists

ID: 27182868