The dataset is composed of memes collected from public Facebook groups in Mexico and manually annotated according to the presence of hate speech, inappropriate content, or harmless content. Additionally, within the hate speech category, specific subcategories are included, such as classism, racism, sexism, and other forms of discrimination.
Language(s)
Spanish (Mexico)
Dataset description link
Year
2025
Domain
Social
Annotations
labels indicating the type of hate content
Format
json
Data access
Registration
Publication
Jarquín-Vásquez, H. et al. 2025. Overview of DIMEMEX at IberLEF 2025: Detection of Inappropriate Memes from Mexico. Procesamiento del Lenguaje Natural, 75, pp. 401-412.
NLP Topic
Number of units
3000
Size
3000.00MB

