Time-Reversal Provides Unsupervised Feedback to LLMs
Yerram Varun, Rahul Madhavan, Sravanti Addepalli, Arun Sai Suggala et autres
Large Language Models (LLMs) are typically trained to predict in the forward direction of time. However, recent works have shown that prompting these models to look back and critique their own generations can produce useful feedback. Motivated by this, we explore the …