Skip to content

Opening book details…

About this document

Spanish Humor Analysis Dataset 2019 by saber is a document available to read on EtoBox.

This paper presents a new Spanish-language dataset of 30,000 tweets annotated for humor and funniness. The dataset addresses issues with previous Spanish humor datasets, including removing duplicate tweets and improving inter-annotator agreement. Tweets were crowd-annotated on two dimensions: whether the author intended humor (binary), and a 5-point funniness score. Approximately 38.6% of tweets were intended as humor, with an average funniness score of 2.04. The dataset was used in humor detection and funn

Author
saber
Language
EN