Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Hannu Toivonen

Cards Against LLMs: Benchmarking Humor Alignment in Large Language Models

Apr 09, 2026

Yousra Fettach, Guillaume Bied, Hannu Toivonen, Tijl De Bie

Abstract:Humor is one of the most culturally embedded and socially significant dimensions of human communication, yet it remains largely unexplored as a dimension of Large Language Model (LLM) alignment. In this study, five frontier language models play the same Cards Against Humanity games (CAH) as human players. The models select the funniest response from a slate of ten candidate cards across 9,894 rounds. While all models exceed the random baseline, alignment with human preference remains modest. More striking is that models agree with each other substantially more often than they agree with humans. We show that this preference is partly explained by systematic position biases and content preferences, raising the question whether LLM humor judgment reflects genuine preference or structural artifacts of inference and alignment.

Via

Access Paper or Ask Questions

DopeLearning: A Computational Approach to Rap Lyrics Generation

Jun 09, 2016

Eric Malmi, Pyry Takala, Hannu Toivonen, Tapani Raiko, Aristides Gionis

Figure 1 for DopeLearning: A Computational Approach to Rap Lyrics Generation

Figure 2 for DopeLearning: A Computational Approach to Rap Lyrics Generation

Figure 3 for DopeLearning: A Computational Approach to Rap Lyrics Generation

Figure 4 for DopeLearning: A Computational Approach to Rap Lyrics Generation

Abstract:Writing rap lyrics requires both creativity to construct a meaningful, interesting story and lyrical skills to produce complex rhyme patterns, which form the cornerstone of good flow. We present a rap lyrics generation method that captures both of these aspects. First, we develop a prediction model to identify the next line of existing lyrics from a set of candidate next lines. This model is based on two machine-learning techniques: the RankSVM algorithm and a deep neural network model with a novel structure. Results show that the prediction model can identify the true next line among 299 randomly selected lines with an accuracy of 17%, i.e., over 50 times more likely than by random. Second, we employ the prediction model to combine lines from existing songs, producing lyrics with rhyme and a meaning. An evaluation of the produced lyrics shows that in terms of quantitative rhyme density, the method outperforms the best human rappers by 21%. The rap lyrics generator has been deployed as an online tool called DeepBeat, and the performance of the tool has been assessed by analyzing its usage logs. This analysis shows that machine-learned rankings correlate with user preferences.

* This is a pre-print of an article appearing at KDD'16

Via

Access Paper or Ask Questions