about
The Naughtyformer: A Transformer Understands Offensive Humor (arxiv.org)
7 points by leonardtang on Dec 1, 2022 | hide | past | pdf | discuss on HN

In plain words: A new collection of jokes pulled from Reddit is used to train a text-reading AI to sort jokes by type, especially spotting the offensive ones. It detects offensiveness in jokes more accurately than the best current systems.

Abstract

Jokes are intentionally written to be funny, but not all jokes are created the same. Some jokes may be fit for a classroom of kindergarteners, but others are best reserved for a more mature audience. While recent work has shown impressive results on humor detection in text, here we instead investigate the more nuanced task of detecting humor subtypes, especially of the less innocent variety. To that end, we introduce a novel jokes dataset filtered from Reddit and solve the subtype classification task using a finetuned Transformer dubbed the Naughtyformer. Moreover, we show that our model is significantly better at detecting offensiveness in jokes compared to state-of-the-art methods.

Leonard Tang, Alexander Cai, Steve Li, Jason Wang
arXiv:2211.14369 · cs.CL · submitted Nov 25, 2022
abstract · pdf · html · AAAI-23 Student Abstract

add comment on HN