Discussion about this post

User's avatar
David Holmer's avatar

The comparison to Wikipedia I think is interesting.

The usual objection is that Wikipedia is not a “primary” source and is instead a summary of primary sources by design. This applies to LLMs as well as they are inherently a kind of “summary” of their training data.

One thing a citation is SUPPOSED to provide is a reference which can be checked and followed up by the reader. Wikipedia cite does provide this if you include both the page and date of citation (even if it gets changed later history is preserved/recorded). The LLMs fail this because no one includes the full prompt in the citation or even if they did, LLMs are non deterministic by nature and may not say the same information consistently.

Mark F Radcliffe's avatar

Completely agree. The use of LLMs as a source is laughable. You don’t have to work with them very much to realize that they only partially reliable even for simple tasks. Just had both ChatGPT and Claude fail to identify a brand change on a medical device due to a merger which took place on 2021.

1 more comment...

No posts

Ready for more?