Skip to main navigation Skip to search Skip to main content

Application and analysis of a multi-layered scheme for irony on the Italian twitter corpus TWITTIRÒ

  • Alessandra Teresa Cignarella
  • , Cristina Bosco
  • , Viviana Patti
  • , Mirko Lai

Research output: Chapter in Book/Report/Conference proceedingConference contributionpeer-review

Abstract

In this paper we describe the main issues emerged within the application of a multi-layered scheme for the fine-grained annotation of irony (Karoui et al., 2017) on an Italian Twitter corpus, i.e. TWITTIRÒ, which is composed of about 1,500 tweets with various provenance. A discussion is proposed about the limits and advantages of the application of the scheme to Italian messages, supported by an analysis of the outcome of the annotation carried on by native Italian speakers in the development of the corpus. We present a quantitative and qualitative study both of the distribution of the labels for the different layers involved in the scheme which can shed some light on the process of human annotation for a validation of the annotation scheme on Italian irony-laden social media contents collected in the last years. This results in a novel gold standard for irony detection in Italian, enriched with fine-grained annotations, and in a language resource available to the community and exploitable in the cross- and multi-lingual perspective which characterizes the work that inspired this research.

Original languageEnglish
Title of host publicationLREC 2018 - 11th International Conference on Language Resources and Evaluation
EditorsNicoletta Calzolari, Khalid Choukri, Christopher Cieri, Thierry Declerck, Sara Goggi, Koiti Hasida, Hitoshi Isahara, Bente Maegaard, Joseph Mariani, Helene Mazo, Asuncion Moreno, Jan Odijk, Stelios Piperidis, Takenobu Tokunaga
PublisherEuropean Language Resources Association (ELRA)
Pages4204-4211
Number of pages8
ISBN (Electronic)9791095546009
Publication statusPublished - 2018
Externally publishedYes
Event11th International Conference on Language Resources and Evaluation, LREC 2018 - Miyazaki, Japan
Duration: 7 May 201812 May 2018

Publication series

NameLREC 2018 - 11th International Conference on Language Resources and Evaluation

Conference

Conference11th International Conference on Language Resources and Evaluation, LREC 2018
Country/TerritoryJapan
CityMiyazaki
Period7/05/1812/05/18

Keywords

  • Corpora
  • Figurative language processing
  • Irony
  • Italian
  • Social media

Fingerprint

Dive into the research topics of 'Application and analysis of a multi-layered scheme for irony on the Italian twitter corpus TWITTIRÒ'. Together they form a unique fingerprint.

Cite this