Better Answers Can Make GPT-4’s Hallucinations More Dangerous
GPT-4 can supply accurate explanations, references, and links often enough to appear highly dependable. Labenz warns that it still fabricates information near the limits of its knowledge, and its strong routine performance can make users less vigilant precisely when verification matters most.
- Several references initially returned by GPT-4 were genuine
- Continued testing revealed fabricated information and links
- The hallucination boundary is farther out than with earlier models
- Reliable everyday performance can encourage misplaced trust in unfamiliar domains
“It's just that that boundary of where it starts to hallucinate is much further out.”
“But in some ways it's a little bit more dangerous because you're out there on a limb of things you don't know. And that's also…”