What the data shows
this is where I'd put "nothing here"if I picked it on more than half the chemistry questions- "Nothing" comes hard. When "none" is right, Jev picks it 49% of the time for chemical-protein relations, 61% for drug interactions, 65% for unfair clauses and 72% for questions a passage can't answer.
- It errs toward finding something. Wrongly picking "none" when there is something is rare: 3%, 1%, 2% and 1% for the same four tasks.
- Personal data is the exception. When a text has no personal data, Jev says so 91% of the time, at the cost of a slightly higher false "none" rate (6%).
What it means, and what it doesn't
In extraction, Jev leans toward answering. When a document doesn't contain what you asked for, there's a one-in-four to one-in-two chance it will produce something anyway, most of all in biomedical relations. Pipelines that extract facts into a database should check its non-"none" answers on documents where nothing is expected, or ask a separate yes/no question first ("does this sentence state any relation?").
It doesn't mean Jev hallucinates freely: when something is there, it rarely says "none", and some of its "something" answers on strictly labeled data are defensible readings.