Writing in the Margins: Better Inference Pattern for Long Context Retrieval
Melisa Russak, Umar Jamil, Christopher Bryant +4 authors
Writing in the Margins (WiM) enhances Large Language Models' performance on long input sequences and retrieval tasks by using chunked prefill and marginal information, improving accuracy and F1-score without fine-tuning.