Self-RAG
Self-RAG trains a model to decide when to retrieve, and to emit special reflection tokens that critique whether retrieved passages are relevant and whether its own output is supported by them.
Description
Retrieving on demand rather than for every query avoids adding irrelevant context and lets the model's critique steer generation.
Sources
- Asai et al. (2023). Self-RAG: Learning to Retrieve, Generate, and Critique through Self-Reflection.
Cite this entry
Protologue. (2026). Self-RAG. In Protologue: A Taxonomy of Prompting and LLM Techniques (v1.0.0, PTL-0057). https://protologue.com/t/self-rag/
BibTeX
@misc{protologue_self_rag,
title = {Self-RAG},
author = {{Protologue}},
year = {2026},
howpublished = {Protologue: A Taxonomy of Prompting and LLM Techniques, v1.0.0},
note = {Entry PTL-0057},
url = {https://protologue.com/t/self-rag/}
}