כתבה
arXiv cs.CL ·
Constructing Disambiguated Knowledge Bases from Large Language Models at Scale
תקציר מקורי באנגליתarXiv:2608.03729v4 Announce Type: replace Abstract: Automated Knowledge Base Construction (AKBC) is a core NLP task, and recent work proposes generating knowledge bases directly from large language models (LLMs), treating the model itself as the knowledge source. However, LLMs natively possess no representation of entities, leading to duplicate entries as well as conflations. We propose GPTKB 2.0, a methodology for constructing disambiguated KBs directly from LLMs. GPTKB 2.0 incorporates on-the-fly disambiguation of entities, relations and classes, and is meticulously designed to satisfy both scalability and disambiguation accuracy. We analyze the central design decisions and characterize the trade-offs between accuracy, scale, and cost. We execute GPTKB 2.0 at scale, obtaining a materiali
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית