What the data shows

- Classic jokes: 0.78, close agreement. Jev knows which old chestnuts land.
- Edited headlines: 0.41, loose agreement, even though five judges per headline make the crowd itself noisy.
- Cartoon captions: 0.02, no relation at all (see "Jev can't tell which New Yorker captions are funny").
What it means, and what it doesn't
Jev's sense of humor tracks people's where the jokes are old and famous, and fades as the humor gets newer and more situational. That pattern is also what you'd expect if Jev were partly remembering reputations: the Jester jokes circulate widely online, while most contest captions are one-off entries few people ever saw.
It doesn't settle which explanation is right. A test with new jokes in the classic format, written after Jev was trained, would separate knowing a joke's reputation from finding it funny.