Multimodal Humor Β· Multilingual

A multimodal humor
benchmark across languages.

AfriHumor is an annotator-driven workbench for building a humor detection and generation dataset β€” African languages plus English and French. We capture word-level speech, prosody, visual cues, and linguistic structure from comedy β€” with per-language tonal and diacritic input support where needed. We start with the languages below and expand coverage as new annotators come on board.

β€” languages Annotating now Β· scroll to explore

Coverage grows with the team β€” new languages open up as annotators join.

What we capture

Multimodal humor, annotated at the clip level

Each comedy clip is annotated across four modalities so the dataset supports both humor detection and generation β€” without relying on laughter markers.

Speech & ASR
Word-level transcripts with AI-assisted diacritization and soft keyboards for tonal languages.
Prosody & delivery
Pitch, pauses, emphasis, and delivery tokens anchored to source-video timestamps.
Visual humor
Frame capture with VLM-assisted descriptions of gestures, facial expressions, and props.
Structured humor
Per-language taxonomy of humor types, linguistic style, and delivery β€” reviewed for agreement.
🌐
Built for the languages we annotate. Tonal and diacritic support (Keyman + in-app soft keyboards) is configured per language β€” including English and French β€” so new annotators can onboard with the right input tools from day one, and the language registry grows whenever the team does.
β€” Videos
β€” Clips
β€” Annotations

Live Β· updated β€”

Annotate with us

Sign in or register to annotate.

Open the Portal β†’