Cracking the Code of Juxtaposition: Can AI Models Understand the Humorous Contradictions

Published in Advances in Neural Information Processing Systems, Vol. 37 (NeurIPS 2024 Oral), 2024

We introduce the YesBut benchmark for evaluating how large vision-language models understand humorous contradictions in juxtaposed comic panels.

Recommended citation: Zhe Hu, Tuo Liang, Jing Li, Yiren Lu, Yunlai Zhou, Yiran Qiao, Jing Ma, and Yu Yin. "Cracking the Code of Juxtaposition: Can AI Models Understand the Humorous Contradictions." Advances in Neural Information Processing Systems 37 (2024): 47166-47188.
Download Paper