Retrieving the most relevant chunks of a knowledge base for a given query and injecting them into the prompt, so the model answers from that context instead of from memory alone.
More terms