כתבה
arXiv cs.AI ·
Learning a Fact Is Not Learning How to Retrieve It
תקציר מקורי באנגליתarXiv:2610.03251v1 Announce Type: new Abstract: A model trained on "The capital of X is Y" may produce "Y" after "The capital of X is" but fail after "The capital of X:". We call these different ways of eliciting the same fact request forms. To separate learning a fact from retrieving it, we train two models in two stages. In the first stage (request-form training), one model sees each fact in five forms and the other sees the same facts only as statements. In the second stage (target-fact training), both receive identical training on new facts, all as statements. Both then retrieve the new facts almost equally well from statements, but differ sharply on other request forms. Thus, a model can learn how to retrieve through a request form before it learns the facts. To understand this differ
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית