What the data shows
whichever joke is listed secondJevthe joke with more upvotesJev barely beats a coin: 53% on jokes, 56% on captions. What drives its pick is position.
- It picks the second item 61% of the time on jokes and 70% on captions.
- So its accuracy depends on where the answer sits. On captions it's right 76% of the time when the winner comes second and 36% when it comes first. On jokes, 64% and 41%.
- The answer buttons aren't the cause: reordering them changes Jev's pick in about 3% of pairs. The lean follows the order of the texts in the question.
What it means, and what it doesn't
If you ask Jev "which of these two is better?", the order you list them in can matter more than what they say, at least when Jev can't tell them apart. This kind of position bias isn't on TypeSafe's list of known weak spots. Position steers Jev elsewhere too, though not always the same way: on east-west questions about US cities it leans toward the city named first (see "West of what? Jev picks the city named first").
It doesn't mean Jev picks the second item whenever it's asked to compare. Here it had no real basis for a choice, and position filled the gap; on questions it can actually answer, knowledge should do more of the work. How much the lean survives there is worth testing directly, by asking the same pairs with the two texts swapped.