How Much Do LLMs Hallucinate in Document Q&A Scenarios? A 172-Billion-Token Study Across Temperatures, Context Lengths, and Hardware Platforms [TLDR: 25%]

RandAlThor@lemmy.ca · edit-2 21 days ago

How Much Do LLMs Hallucinate in Document Q&A Scenarios? A 172-Billion-Token Study Across Temperatures, Context Lengths, and Hardware Platforms [TLDR: 25%]

how_we_burned@lemmy.zip · 3 days ago

I refuse to call it AI

It’s a LM… Pure and simple. Anyway none of the LMs can come up with theory of relatively (if you gave them all of the known physics up to 1915).

Nor can they play paper scissors rock (they don’t realise it’s pointless).

As far as I can tell they’re wrong more times then they’re right and the only use I have for them is as a glorified search engine (and even then they’re still fricking wrong.

They’re only useful if you already know the answer because if you don’t know the answer you don’t know if they’ve given you the wrong answer.