Auditing a KB Elicitation of Frontier LLM Knowledge: A Multi-dimensional Analysis of GPTKB v1.5
Researchers analyzed GPT-4.1's knowledge base and found significant inaccuracies, inconsistencies, and hallucinations, highlighting the need for better methods to extract, consolidate, and verify factual knowledge from large language models.
Save an API key to vote.