Advanced search
1 file | 772.53 KB Add to list

TRIC : a confidence-aware multilingual dataset for irony detection and rationales in English, Dutch and Italian

Aaron Maladry (UGent) , Alessandra Teresa Cignarella (UGent) , Cynthia Van Hee (UGent) , Els Lefever (UGent) and Veronique Hoste (UGent)
Author
Organization
Project
Abstract
This article presents the novel manually annotated Trilingual Recognition of Irony with Confidence (TRIC) dataset for English, Dutch and Italian, publicly available on Hugging Face as Amala3/TRIC. The annotations for this dataset include irony likelihood labels, indicating how likely the annotators believe a text is ironic, as well as trigger words, indicating which words in a sentence are essential for understanding the irony. In addition to the dataset, this work investigates the development of confidence-aware models for irony detection in a monolingual and multilingual setup. Results show that finetuning encoder-only models with confidence-aware labels improves the performance on binary irony detection and that finetuning on task-specific data in multiple languages also results in increased performance. Comparison to finetuned Llama3 indicates that generative decoder-only models perform better than confidence-aware models for English, but that encoder-only models perform best for less-resourced languages (both Dutch and Italian). Analysis of trigger words of both humans and automatic systems suggests that token-level importance differ significantly, but that n-gram based clustering can reveal deeper insights. In all three languages, automatic systems tend to rely more on hyperbolic positive sentiment and interjections, whereas humans more often identify topics that are relevant to understand irony.

Downloads

  • (...).pdf
    • full text (Accepted manuscript)
    • |
    • UGent only
    • |
    • PDF
    • |
    • 772.53 KB

Citation

Please use this url to cite or link to this publication:

MLA
Maladry, Aaron, et al. “TRIC : A Confidence-Aware Multilingual Dataset for Irony Detection and Rationales in English, Dutch and Italian.” NATURAL LANGUAGE PROCESSING, 2025.
APA
Maladry, A., Cignarella, A. T., Van Hee, C., Lefever, E., & Hoste, V. (2025). TRIC : a confidence-aware multilingual dataset for irony detection and rationales in English, Dutch and Italian. NATURAL LANGUAGE PROCESSING.
Chicago author-date
Maladry, Aaron, Alessandra Teresa Cignarella, Cynthia Van Hee, Els Lefever, and Veronique Hoste. 2025. “TRIC : A Confidence-Aware Multilingual Dataset for Irony Detection and Rationales in English, Dutch and Italian.” NATURAL LANGUAGE PROCESSING.
Chicago author-date (all authors)
Maladry, Aaron, Alessandra Teresa Cignarella, Cynthia Van Hee, Els Lefever, and Veronique Hoste. 2025. “TRIC : A Confidence-Aware Multilingual Dataset for Irony Detection and Rationales in English, Dutch and Italian.” NATURAL LANGUAGE PROCESSING.
Vancouver
1.
Maladry A, Cignarella AT, Van Hee C, Lefever E, Hoste V. TRIC : a confidence-aware multilingual dataset for irony detection and rationales in English, Dutch and Italian. NATURAL LANGUAGE PROCESSING. 2025;
IEEE
[1]
A. Maladry, A. T. Cignarella, C. Van Hee, E. Lefever, and V. Hoste, “TRIC : a confidence-aware multilingual dataset for irony detection and rationales in English, Dutch and Italian,” NATURAL LANGUAGE PROCESSING, 2025.
@article{01JX27SNS04D24WHFRJQ2YKKWC,
  abstract     = {{This article presents the novel manually annotated Trilingual Recognition of Irony with Confidence (TRIC) dataset for English, Dutch and Italian, publicly available on Hugging Face as Amala3/TRIC. The annotations for this dataset include irony likelihood labels, indicating how likely the annotators believe a text is ironic, as well as trigger words, indicating which words in a sentence are essential for understanding the irony. In addition to the dataset, this work investigates the development of confidence-aware models for irony detection in a monolingual and multilingual setup. Results show that finetuning encoder-only models with confidence-aware labels improves the performance on binary irony detection and that finetuning on task-specific data in multiple languages also results in increased performance. Comparison to finetuned Llama3 indicates that generative decoder-only models perform better than confidence-aware models for English, but that encoder-only models perform best for less-resourced languages (both Dutch and Italian). Analysis of trigger words of both humans and automatic systems suggests that token-level importance differ significantly, but that n-gram based clustering can reveal deeper insights. In all three languages,
automatic systems tend to rely more on hyperbolic positive sentiment and interjections, whereas humans more often identify topics that are relevant to understand irony.}},
  author       = {{Maladry, Aaron and Cignarella, Alessandra Teresa and Van Hee, Cynthia and Lefever, Els and Hoste, Veronique}},
  issn         = {{2977-0424}},
  journal      = {{NATURAL LANGUAGE PROCESSING}},
  language     = {{eng}},
  pages        = {{31}},
  title        = {{TRIC : a confidence-aware multilingual dataset for irony detection and rationales in English, Dutch and Italian}},
  year         = {{2025}},
}