Y
Hacker News
new
|
ask
|
show
|
jobs
by
kostaj
22 days ago
Agree. Human experts also struggle agreeing on this type of claims. The inter-annotator agreement on the verdicts on the AVeriTeC corpus across 50 organizations is κ=0.619 - substantial but well short of perfect.