Y
Hacker News
new
|
ask
|
show
|
jobs
by
teleforce
6 days ago
>On this benchmark, a pure LLM generated an accuracy score of zero. Adding RAG, prompt engineering, and agentic AI raised accuracy to the 10+% range.
Oouch!