Y
Hacker News
new
|
ask
|
show
|
jobs
by
roselan
438 days ago
I'm surprised there was no human tested for a base reference point. I'm pretty sure some of us would not pass the test held by another human.
1 comments
Ukv
438 days ago
Human win rate would be 1 minus the model win rate, to my understanding. So 77% against ELIZA, 27% against GPT-4.5 with a human persona.
link