DoLa: Decoding by Contrasting Layers Improves Factuality in Large Language Models
Yung-Sung Chuang, Yujia Xie, Hongyin Luo +3 authors
A decoding strategy that contrasts logits from different transformer layers reduces hallucinations in large language models without external knowledge or fine-tuning.