Screenshot of this question was making the rounds last week. But this article covers testing against all the well-known models out there.

Also includes outtakes on the ‘reasoning’ models.

  • [deleted]@piefed.world
    link
    fedilink
    English
    arrow-up
    5
    ·
    9 hours ago

    If you can instantly verify it then you don’t need the torture.

    Getting the person to volunteer the information is proven to be far, far more successful and being able to instantly verofy means you know when you have the answers.