Advertisement
Advertisement
β‘ Community Insights
Discussion Sentiment
50% Positive
Analyzed from 68 words in the discussion.
Trending Topics
#models#questions#attempt#uncover#biases#default#choices#makes#asked#simple
Discussion Sentiment
Analyzed from 68 words in the discussion.
Trending Topics
Discussion (3 Comments)Read Original on HackerNews
I am not sure what this tells, but I found it interesting that most of the models were very much in agreement. With some outliers like Mistral and Llama on some questions.
I even made a benchmark, ConsensusBench, to measure how aligned they were: https://www.modelbias.ai/consensus-bench