all AI news
Locating Factual Knowledge in Large Language Models: Exploring the Residual Stream and Analyzing Subvalues in Vocabulary Space. (arXiv:2312.12141v2 [cs.CL] UPDATED)
cs.CL updates on arXiv.org arxiv.org
We find the location of factual knowledge in large language models by
exploring the residual stream and analyzing subvalues in vocabulary space. We
find the reason why subvalues have human-interpretable concepts when projecting
into vocabulary space. The before-softmax values of subvalues are added by an
addition function, thus the probability of top tokens in vocabulary space will
increase. Based on this, we find using log probability increase to compute the
significance of layers and subvalues is better than probability increase, …
arxiv concepts cs.cl human knowledge language language models large language large language models location reason residual softmax space values