Understanding and Detecting Hateful Content using Contrastive Learning

González-Pizarro, Felipe; Zannettou, Savvas

Lokale TagsFreigabegeschichteDetailsÜbersicht

Understanding and Detecting Hateful Content using Contrastive Learning

González-Pizarro, F., & Zannettou, S. (2022). Understanding and Detecting Hateful Content using Contrastive Learning. Retrieved from https://arxiv.org/abs/2201.08387.

Item is Freigegeben

einblenden: alle ausblenden: alle

Basisdaten

einblenden: ausblenden:

Datensatz-Permalink: https://hdl.handle.net/21.11116/0000-000A-27F9-2 Versions-Permalink: https://hdl.handle.net/21.11116/0000-000A-27FA-1

Genre: Forschungspapier

Dateien

einblenden: Dateien

ausblenden: Dateien

:

arXiv:2201.08387.pdf (Preprint), 5MB

Öffnen Speichern

Datei-Permalink:
https://hdl.handle.net/21.11116/0000-000A-27FB-0

Name:
arXiv:2201.08387.pdf

Beschreibung:
File downloaded from arXiv at 2022-03-29 10:09

OA-Status:

Sichtbarkeit:
Öffentlich

MIME-Typ / Prüfsumme:
application/pdf / [MD5]

Technische Metadaten:

Öffnen

Copyright Datum:
-

Copyright Info:
-

Lizenz:
http://creativecommons.org/licenses/by/4.0/

Externe Referenzen

einblenden:

Urheber

einblenden:

ausblenden:

Urheber:
González-Pizarro, Felipe¹, Autor
Zannettou, Savvas², Autor

Affiliations:
1External Organizations, ou_persistent22
2Internet Architecture, MPI for Informatics, Max Planck Society, ou_2489697

Inhalt

einblenden:

ausblenden:

Schlagwörter: cs.SI,Computer Science, Computers and Society, cs.CY

Zusammenfassung: The spread of hate speech and hateful imagery on the Web is a significant
problem that needs to be mitigated to improve our Web experience. This work
contributes to research efforts to detect and understand hateful content on the
Web by undertaking a multimodal analysis of Antisemitism and Islamophobia on
4chan's /pol/ using OpenAI's CLIP. This large pre-trained model uses the
Contrastive Learning paradigm. We devise a methodology to identify a set of
Antisemitic and Islamophobic hateful textual phrases using Google's Perspective
API and manual annotations. Then, we use OpenAI's CLIP to identify images that
are highly similar to our Antisemitic/Islamophobic textual phrases. By running
our methodology on a dataset that includes 66M posts and 5.8M images shared on
4chan's /pol/ for 18 months, we detect 573,513 posts containing 92K
Antisemitic/Islamophobic images and 246K posts that include 420 hateful
phrases. Among other things, we find that we can use OpenAI's CLIP model to
detect hateful content with an accuracy score of 0.84 (F1 score = 0.58). Also,
we find that Antisemitic/Islamophobic imagery is shared in 2x more posts on
4chan's /pol/ compared to Antisemitic/Islamophobic textual phrases,
highlighting the need to design more tools for detecting hateful imagery.
Finally, we make publicly available a dataset of 420 Antisemitic/Islamophobic
phrases and 92K images that can assist researchers in further understanding
Antisemitism/Islamophobia and developing more accurate hate speech detection
models.

Details

einblenden:

ausblenden:

Sprache(n): eng - English

Datum: Erstellt: 2022-01-21Online veröffentlicht: 2022

Publikationsstatus: Online veröffentlicht

Seiten: 11 p.

Ort, Verlag, Ausgabe: -

Inhaltsverzeichnis: -

Art der Begutachtung: -

Identifikatoren: arXiv: 2201.08387
URI: https://arxiv.org/abs/2201.08387
BibTex Citekey: GonzalesPizarro22

Art des Abschluß: -

Datensatz

Basisdaten

Dateien

Externe Referenzen

Urheber

Inhalt

Details

Veranstaltung

Entscheidung

Projektinformation

Quelle