Rendered at 17:51:15 GMT+0000 (Coordinated Universal Time) with Cloudflare Workers.
bnfcl 4 hours ago [-]
In an attempt to uncover biases and default choices AI makes, I asked 100 AI models, 100 simple questions, 3 times each.
I am not sure what this tells, but I found it interesting that most of the models were very much in agreement. With some outliers like Mistral and Llama on some questions.
I am not sure what this tells, but I found it interesting that most of the models were very much in agreement. With some outliers like Mistral and Llama on some questions.
I even made a benchmark, ConsensusBench, to measure how aligned they were: https://www.modelbias.ai/consensus-bench