About Humansplain

Upload is the dare. Judge is the job. Points are how the job pays.

What is Humansplain?

Humansplain is an independent, crowdsourced benchmark that tests how well vision-language AI models can explain why something is funny. You upload a meme; several models each give a one-sentence take; humans decide who got it — or write a better explanation. Those judgments update public model rankings and the Hall of Shame.

The loop

  1. Upload — a dare to the models.
  2. Judge three other images while the models think — if you have already judged enough, you skip ahead. That is what settles consensus, not voting on your own meme.
  3. See your results — who got it, or humansplain it yourself. That peek doesn't score the Hall.

Most days you won't have a funny image handy. You can always judge. Signed-in judging earns Shame Points, which unlock more of the live Hall.

Shame Points

Signed-in judging earns Shame Points. The last 30 days of SP unlock more of the live Hall: a free account opens this month's #10, Regular (250 SP) opens ranks 6–9, Legend (1,000 SP) opens ranks 1–5 and the all-time archive. Closed months show every image to everyone; analysis still follows those tiers. Lifetime SP is prestige — it does not unlock access on its own.

Pro skips the treadmill. It isn't for sale, and is granted directly by us.

Your profile, handle, and avatar are yours to customize on My Page — Google is just how you sign in.

The Hall of Shame

A monthly competition for images that actually stump the models — humans agree there is a joke, and no AI answer was good enough. Live-month top 10 stay locked while the ranking is fresh; closed months show the images to everyone, while analysis stays tier-gated. Anonymous uploads never enter until you sign in. Each upload asks for three judgments on other people's images before yours can be considered.

On the 1st of each month the previous month is closed: the standings are snapshotted and placement points are paid out to the uploaders behind the top 50. Only the model explanations are graded — a run where a human wrote the winning answer still counts as a baffle, because the AIs still missed it.

Why does it exist?

Explaining humor is hard for AI: context, culture, irony, tone. Most benchmarks don't measure whether an explanation sounds like something a person would say. Crowdsourced judgments on real memes give a human-grounded signal — useful for researchers, labs, and anyone curious how "human" today's vision models really are.

Who runs it?

An independent project, not affiliated with any AI company. Methodology is public; the leaderboard is open. Details live in the FAQ.

Upload a meme · Start judging · Hall of Shame · FAQ