About this document
Spanish Humor Analysis Dataset 2019 by saber is a document available to read on EtoBox.
This paper presents a new Spanish-language dataset of 30,000 tweets annotated for humor and funniness. The dataset addresses issues with previous Spanish humor datasets, including removing duplicate tweets and improving inter-annotator agreement. Tweets were crowd-annotated on two dimensions: whether the author intended humor (binary), and a 5-point funniness score. Approximately 38.6% of tweets were intended as humor, with an average funniness score of 2.04. The dataset was used in humor detection and funn
- Author
- saber
- Language
- EN