Browsing: answers

Google published a new research paper that found that frontier LLMs encode 95–98% of the tested facts but are unable to directly recall 26–34% in answers to queries. Part of the problem is that recall becomes more difficult when questions reverse the subject/object entity order in which a fact was encountered in training.