In an attempt to uncover biases and default choices AI makes, I asked 100 AI models, 100 simple questions, 3 times each.
I am not sure what this tells, but I found it interesting that most of the models were very much in agreement. With some outliers like Mistral and Llama on some questions.
1 comment of 3
[ 2.0 ms ] story [ 3.1 ms ] threadI am not sure what this tells, but I found it interesting that most of the models were very much in agreement. With some outliers like Mistral and Llama on some questions.
I even made a benchmark, ConsensusBench, to measure how aligned they were: https://www.modelbias.ai/consensus-bench