| HN Mirror

Y	Hacker News new \| ask \| show \| jobs


	by kostaj 22 days ago
	Agree. Human experts also struggle agreeing on this type of claims. The inter-annotator agreement on the verdicts on the AVeriTeC corpus across 50 organizations is κ=0.619 - substantial but well short of perfect.