DIMEMEX-2025

The dataset is composed of memes collected from public Facebook groups in Mexico and manually annotated according to the presence of hate speech, inappropriate content, or harmless content. Additionally, within the hate speech category, specific subcategories are included, such as classism, racism, sexism, and other forms of discrimination.

Language(s)
Spanish (Mexico)
Year
2025
Domain
Social
Annotations
labels indicating the type of hate content
Format
json
Data access
Registration

Publication
Jarquín-Vásquez, H. et al. 2025. Overview of DIMEMEX at IberLEF 2025: Detection of Inappropriate Memes from Mexico. Procesamiento del Lenguaje Natural, 75, pp. 401-412.
NLP Topic
Number of units
3000
Size
3000.00MB

If you have published a result better than those on the list, send a message to odesia-comunicacion@lsi.uned.es indicating the result and the DOI of the article, along with a copy of it if it is not published openly.