‹ BackHN Continuity

Thread

AI chatbots give wrong answers to financial queries 'most of the time'

157 points · 89 comments · 1vuio0pswjnm7

  1. simianwords · · focus · HN ↗
    These models do pretty well in benchmarks and real world so I'm highly suspicious of this article. Further more, in the original report, the examples of bad answers are from Haiku - at least 7 out of 10. Anyone who knows anything about LLMs know that haiku shouldn't be used for anything pretty much.

    There's no reproducible set either. I'm not gonna trust this report.

    1. stymaar · · focus · HN ↗
      Most people[1] interacting with chatbots don't have a paid subscription and they do interact with the free-tier LLMs that are Luna and Haiku, so I still think it's relevant.

      [1]: not on HN obviously, but IRL, and probably among FT's readership as well.

      1. cillian64 · · focus · HN ↗
        A free claude account with no subscription gets you access to sonnet and I believe uses it by default over haiku
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.