Screenshot of this question was making the rounds last week. But this article covers testing against all the well-known models out there.

Also includes outtakes on the ‘reasoning’ models.

  • Iconoclast@feddit.uk
    link
    fedilink
    English
    arrow-up
    1
    arrow-down
    5
    ·
    4 hours ago

    Which is exactly why LLMs are useless.

    800 million weekly ChatGPT users disagree with that.

    • RichardDegenne@lemmy.zip
      link
      fedilink
      English
      arrow-up
      7
      ·
      3 hours ago

      And there are 1.3 billion smokers in the world according to the WHO.

      Does that make cigarettes useful?

      • Iconoclast@feddit.uk
        link
        fedilink
        English
        arrow-up
        3
        arrow-down
        3
        ·
        edit-2
        3 hours ago

        Something being useful doesn’t imply it’s good or beneficial. Those terms are not synonymous. Usefulness describes whether a thing achieves a particular goal or serves a specific purpose effectively.

        A torture device is useful for extracting information. A landmine is useful for denying an area to enemy troops.

        • Urist@leminal.space
          link
          fedilink
          English
          arrow-up
          2
          ·
          2 hours ago

          A torture device is useful for extracting information.

          No it fucking isn’t! This is a great analogy, actually, thank you for bringing it up. A person being tortured will tell you literally anything that they believe will stop you from torturing them. They will confess to crimes that never happened, tell you about all their accomplices who don’t exist, and all their daily schedules that were made up on the spot. Torture is useless but morons think it is useful. Just like AI.

    • Urist@leminal.space
      link
      fedilink
      English
      arrow-up
      1
      ·
      3 hours ago

      Those users are being harmed by it, not benefited. That isn’t useful, it’s a social disease.