TensorX
返回文献探索

Paper · arXiv 2311.09277

Contrastive Chain-of-Thought Prompting

Yew Ken Chia, Guizhen Chen, Luu Anh Tuan, Soujanya Poria, Lidong Bing

35 upvotesNovember 15, 2023arXiv 预印本
AI 摘要

Contrastive chain of thought, utilizing both valid and invalid reasoning examples, improves language model reasoning and generalization compared to conventional methods.

chain of thoughtreasoningcontrastive chain of thoughtdemonstrationsreasoning mistakesgeneralization

Abstract

Despite the success of chain of thought in enhancing language model reasoning, the underlying process remains less well understood. Although logically sound reasoning appears inherently crucial for chain of thought, prior studies surprisingly reveal minimal impact when using invalid demonstrations instead. Furthermore, the conventional chain of thought does not inform language models on what mistakes to avoid, which potentially leads to more errors. Hence, inspired by how humans can learn from both positive and negative examples, we propose contrastive chain of thought to enhance language model reasoning. Compared to the conventional chain of thought, our approach provides both valid and invalid reasoning demonstrations, to guide the model to reason step-by-step while reducing reasoning mistakes. To improve generalization, we introduce an automatic method to construct contrastive demonstrations. Our experiments on reasoning benchmarks demonstrate that contrastive chain of thought can serve as a general enhancement of chain-of-thought prompting.

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号