about
The Factorization Curse: Which Tokens You Predict Underlie the Reversal Curse (arxiv.org)
1 point by ijk on Jul 1, 2024 | hide | past | pdf | discuss on HN

In plain words: Next-word training leaves models unable to recall facts when asked in a different order than they saw them — the factorization curse. Bigger models, reversed text, and letting words see both directions failed, but objectives that ignore word order cut errors sharply across five tasks.

Abstract · The Factorization Curse: Which Tokens You Predict Underlie the Reversal Curse and More

Today's best language models still struggle with hallucinations: factually incorrect generations, which impede their ability to reliably retrieve information seen during training. The reversal curse, where models cannot recall information when probed in a different order than was encountered during training, exemplifies this in information retrieval. We reframe the reversal curse as a factorization curse - a failure of models to learn the same joint distribution under different factorizations. Through a series of controlled experiments with increasing levels of realism including WikiReversal, a setting we introduce to closely simulate a knowledge intensive finetuning task, we find that the factorization curse is an inherent failure of the next-token prediction objective used in popular large language models. Moreover, we demonstrate reliable information retrieval cannot be solved with scale, reversed tokens, or even naive bidirectional-attention training. Consequently, various approaches to finetuning on specialized data would necessarily provide mixed results on downstream tasks, unless the model has already seen the right sequence of tokens. Across five tasks of varying levels of complexity, our results uncover a promising path forward: factorization-agnostic objectives can significantly mitigate the reversal curse and hint at improved knowledge storage and planning capabilities.

Ouail Kitouni, Niklas Nolte, Diane Bouchacourt, Adina Williams, Mike Rabbat, Mark Ibrahim
arXiv:2406.05183 · cs.LG, cs.AI, cs.CL · submitted Jun 7, 2024
abstract · pdf · html · 18 pages, 7 figures

add comment on HN
Also discussed: Jun 2024 (1 point, 0 comments)