Hacker News new | ask | show | jobs
by bnfcl 6 days ago
In an attempt to uncover biases and default choices AI makes, I asked 100 AI models, 100 simple questions, 3 times each.

I am not sure what this tells, but I found it interesting that most of the models were very much in agreement. With some outliers like Mistral and Llama on some questions.

I even made a benchmark, ConsensusBench, to measure how aligned they were: https://www.modelbias.ai/consensus-bench

1 comments

The explore by prompt dropdown doesn’t seem to work.
How strange, works fine here. What OS/browser do you use?