8 pointsby bnfcl6 hours ago1 comment
  • bnfcl6 hours ago
    In an attempt to uncover biases and default choices AI makes, I asked 100 AI models, 100 simple questions, 3 times each.

    I am not sure what this tells, but I found it interesting that most of the models were very much in agreement. With some outliers like Mistral and Llama on some questions.

    I even made a benchmark, ConsensusBench, to measure how aligned they were: https://www.modelbias.ai/consensus-bench

    • dd8601fn5 hours ago
      The explore by prompt dropdown doesn’t seem to work.
      • bnfcl4 hours ago
        How strange, works fine here. What OS/browser do you use?