Conceptual

Self-Refinement in Large Language Models

Post-hoc self-correction methods (e.g. Self-Refine) in which a language model iteratively generates output, produces feedback on its own output, and revises it — and the Degeneration-of-Thought limitation when a single agent grades itself.