What the data shows
a rude postthe annotators: offensiveJev: hate speech- Up far more than down. Jev moves 33% of posts one or two rungs up and only 5% down.
- Offensive becomes hate. Of the posts the annotators unanimously called offensive, Jev calls 39% hate speech.
- Normal becomes a problem. Of the posts they called normal, Jev calls half (50%) offensive or hateful.
- Same lean elsewhere. Of tweets Davidson's annotators called neither offensive nor hateful, Jev flags 21%; of DynaHate statements labeled veiled hostility, Jev calls 49% something more explicit.
What it means, and what it doesn't
Used as a moderator, Jev would escalate. Posts people read as crude or as harmless would more often be treated as hate speech, a costly mistake to make against a user who was merely rude. The lean shows up in three datasets with three different labeling setups, so it isn't a quirk of one.
It doesn't mean Jev misses hate: on posts all three annotators called hate speech, it mostly agrees. And because slur-heavy posts were removed, this is a picture of the subtle middle of the scale, not of obvious abuse.