RAG (retrieval-augmented generation)

Retrieving the most relevant chunks of a knowledge base for a given query and injecting them into the prompt, so the model answers from that context instead of from memory alone.

See it explained in full