---
title: "RAG-patronen"
source: "https://systems-analysis.info/int/RAG-patronen"
wiki: "systems-analysis.info/int"
article: "RAG-patronen"
language: "nl"
categories:
  - "Category:Dutch"
  - "Category:Large language models"
  - "Category:Prompt engineering"
revision_id: 6104
wiki_created_at: 2026-09-06T23:58:23Z
wiki_modified_at: 2026-09-06T23:58:23Z
downloaded_at: 2026-09-07T23:12:02Z
---

# RAG-patronen

**RAG-patronen** (Engels: *RAG Patterns*) — een verzameling architecturale en methodologische benaderingen voor het bouwen van **Retrieval-Augmented Generation** (RAG)-systemen. Deze patronen zijn bedoeld om fundamentele problemen van grote taalmodellen (LLM) op te lossen, zoals hallucinaties, verouderde kennis en gebrek aan domeinspecificiteit, door LLM's te integreren met externe, dynamisch toegankelijke gegevensbronnen<sup>[\[1\]](https://systems-analysis.info/int/RAG-patronen#cite_note-lewis2020-1)</sup>. De evolutie van RAG heeft een weg afgelegd van eenvoudige lineaire pipelines naar complexe modulaire en agentgebaseerde systemen<sup>[\[2\]](https://systems-analysis.info/int/RAG-patronen#cite_note-survey2024-2)</sup>.

## Belangrijkste RAG-patronen

Met de ontwikkeling van de technologie zijn er talloze RAG-patronen ontstaan, elk gericht op specifieke uitdagingen met eigen afwegingen tussen kwaliteit, snelheid en kosten.

- **Classic RAG (Klassiek RAG)** — de basisaanpak waarbij de gebruikersvraag wordt gevectoriseerd om relevante fragmenten (chunks) te zoeken in een vectordatabase; de gevonden chunks worden samen met de vraag aan het LLM aangeboden voor het genereren van een antwoord<sup>[\[1\]](https://systems-analysis.info/int/RAG-patronen#cite_note-lewis2020-1)</sup>.

<!-- -->

- **Multi‑Query RAG (Meervoudige zoekopdrachten)** — het LLM genereert meerdere herformuleerde/gepreciseerde varianten van de oorspronkelijke zoekopdracht; de zoekopdracht wordt uitgevoerd over alle varianten en de resultaten worden samengevoegd, wat de volledigheid (*recall*) verhoogt<sup>[\[3\]](https://systems-analysis.info/int/RAG-patronen#cite_note-langchain-multiquery-3)</sup>.

<!-- -->

- **HyDE (Hypothetical Document Expansion)** — bedoeld om de «semantische kloof» tussen een korte zoekopdracht en lange documenten te overbruggen. Het LLM genereert eerst een «hypothetisch» antwoorddocument, waarna de embedding daarvan wordt gebruikt voor de zoekopdracht, wat vaak de kwaliteit van het ophalen verbetert<sup>[\[4\]](https://systems-analysis.info/int/RAG-patronen#cite_note-hyde-4)</sup>.

<!-- -->

- **Hybrid Retrieval (Hybride zoeken)** — een combinatie van semantisch (vectorgebaseerd) en lexicaal (BM25) zoeken. Hybride schema's zijn de standaard geworden voor productiesystemen: vectorzoeken dekt semantische overeenkomsten, terwijl BM25 exacte termen/ID's/acroniemen vindt; de resultaten worden samengevoegd via fusion<sup>[\[5\]](https://systems-analysis.info/int/RAG-patronen#cite_note-weaviate-hybrid-5)[\[6\]](https://systems-analysis.info/int/RAG-patronen#cite_note-qdrant-hybrid-6)[\[7\]](https://systems-analysis.info/int/RAG-patronen#cite_note-milvus-fulltext-7)</sup>.

<!-- -->

- **Re‑ranking (Herrangschikking)** — een tweefasig proces: een snelle retriever levert een set kandidaten op (bijvoorbeeld top‑100), waarna een cross-encoder (of een andere reranker) de relevantie herberekent en de beste selecteert (bijvoorbeeld top‑5) voor het LLM<sup>[\[8\]](https://systems-analysis.info/int/RAG-patronen#cite_note-nogueira2019-8)[\[9\]](https://systems-analysis.info/int/RAG-patronen#cite_note-cohere-rerank-9)</sup>.

<!-- -->

- **Query Routing (Routering van zoekopdrachten)** — in systemen met meerdere heterogene gegevensbronnen (verschillende indexen/databases/API's) wordt de zoekopdracht via een router (LLM-selector of classifier) naar de beste bron geleid; bevat fallback-strategieën<sup>[\[10\]](https://systems-analysis.info/int/RAG-patronen#cite_note-llama-router-10)</sup>.

<!-- -->

- **Agentic/Web RAG (Agentgebaseerde RAG)** — het LLM fungeert als agent: het decomponeert complexe vragen, plant iteraties en gebruikt tools (vectorzoeken, webzoeken) met terugkoppeling. Een typische implementatie is het ReAct-paradigma<sup>[\[11\]](https://systems-analysis.info/int/RAG-patronen#cite_note-react-11)</sup>; voor webgerichte verzameling en verplichte bronvermelding zie WebGPT<sup>[\[12\]](https://systems-analysis.info/int/RAG-patronen#cite_note-webgpt-12)</sup>.

### Verwante en opkomende paradigma's

- **GraphRAG (Grafische RAG)** — maakt gebruik van een kennisgraaf als bron en mechanisme voor contextselectie; zoeken verloopt via de structuur van relaties tussen entiteiten en via tekst, wat de interpreteerbaarheid en kwaliteit bij multi‑hop vragen verhoogt<sup>[\[13\]](https://systems-analysis.info/int/RAG-patronen#cite_note-graphrag-13)[\[14\]](https://systems-analysis.info/int/RAG-patronen#cite_note-graphrag-project-14)</sup>.
- **MM‑RAG (Multimodale RAG)** — werkt met tekst en visuele bronnen (scans/schema's/tabellen). Voorbeeld: VisRAG demonstreert VLM-georiënteerd ophalen en genereren op multimodale documenten<sup>[\[15\]](https://systems-analysis.info/int/RAG-patronen#cite_note-visrag-15)</sup>.
- **Packaging & Context Handling (Contextinpakking)** — manieren om gevonden chunks in de prompt te integreren: *Stuff*, *Map‑Reduce*, *Refine*, *Tree‑of‑Chunks (RAPTOR)*<sup>[\[16\]](https://systems-analysis.info/int/RAG-patronen#cite_note-raptor-16)</sup>.

## Vergelijkingstabel van patronen

| Patroon              | Wanneer toepassen                                            | Invloed op kwaliteit                                                                                                                                                                                                                                                                                        | Kosten / Latentie | Risico's en beperkingen                                      |
|----------------------|--------------------------------------------------------------|-------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------|-------------------|--------------------------------------------------------------|
| **Classic RAG**      | PoC en eenvoudige Q&A op een homogene kennisbasis            | Basisniveau; sterk afhankelijk van embeddings<sup>[\[1\]](https://systems-analysis.info/int/RAG-patronen#cite_note-lewis2020-1)</sup>                                                                                                                                                                       | Laag              | Gevoeligheid voor formulering; risico op irrelevante context |
| **Hybrid Retrieval** | In de meeste productiescenario's; veel codes/acroniemen/ID's | Verhoogt volledigheid; dekt exacte termen<sup>[\[5\]](https://systems-analysis.info/int/RAG-patronen#cite_note-weaviate-hybrid-5)[\[6\]](https://systems-analysis.info/int/RAG-patronen#cite_note-qdrant-hybrid-6)[\[7\]](https://systems-analysis.info/int/RAG-patronen#cite_note-milvus-fulltext-7)</sup> | Laag/Gemiddeld    | Afstemming van fusion-gewichten; twee indexen                |
| **Re‑ranking**       | Kritisch wanneer hoge precisie belangrijk is                 | Significante verbetering van precision op top‑k<sup>[\[8\]](https://systems-analysis.info/int/RAG-patronen#cite_note-nogueira2019-8)[\[9\]](https://systems-analysis.info/int/RAG-patronen#cite_note-cohere-rerank-9)</sup>                                                                                 | Gemiddeld/Hoog    | Extra latentie/kosten                                        |
| **Multi‑Query**      | Korte/meervoudige zoekopdrachten                             | Verhoogt recall<sup>[\[3\]](https://systems-analysis.info/int/RAG-patronen#cite_note-langchain-multiquery-3)</sup>                                                                                                                                                                                          | Gemiddeld         | Overbodige/ruizige herformuleringen                          |
| **HyDE**             | Korte/ambigue zoekopdrachten met grote «semantische kloof»   | Verbetert kwaliteit van ophalen *zero‑shot*<sup>[\[4\]](https://systems-analysis.info/int/RAG-patronen#cite_note-hyde-4)</sup>                                                                                                                                                                              | Gemiddeld         | Afhankelijk van de kwaliteit van de «hypothetische» tekst    |
| **Query Routing**    | Meerdere bronnen (documentbasis, SQL, API, web)              | Verhoogt relevantie via de juiste bron<sup>[\[10\]](https://systems-analysis.info/int/RAG-patronen#cite_note-llama-router-10)</sup>                                                                                                                                                                         | Gemiddeld         | Routeringsfout = mislukte zoekopdracht                       |
| **Agentic/Web RAG**  | Complexe, onderzoeksmatige, meerstapsige zoekopdrachten      | Lost taken op buiten de lineaire pipeline<sup>[\[11\]](https://systems-analysis.info/int/RAG-patronen#cite_note-react-11)[\[12\]](https://systems-analysis.info/int/RAG-patronen#cite_note-webgpt-12)</sup>                                                                                                 | Hoog              | Complexiteit, risico op lussen; guardrails vereist           |

Vergelijking van belangrijke RAG-patronen

## Praktische implementatie en architectuur

### Implementatiefasen

1.  **Proof of Concept (PoC):** Begin met **Classic RAG** op een beperkte maar representatieve dataset om de kwaliteit van embeddings en het basisophalen te valideren<sup>[\[1\]](https://systems-analysis.info/int/RAG-patronen#cite_note-lewis2020-1)</sup>.
2.  **Minimum Viable Product (MVP):** Implementeer **Hybrid Retrieval** en **Re‑ranking** als de beste verhouding tussen inspanning en effect<sup>[\[5\]](https://systems-analysis.info/int/RAG-patronen#cite_note-weaviate-hybrid-5)[\[8\]](https://systems-analysis.info/int/RAG-patronen#cite_note-nogueira2019-8)</sup>.
3.  **Productie:** Voeg querytransformaties toe (**HyDE**, **Multi‑Query**) en indien nodig **Query Routing**; stel observability in (logging van ophalen/herrangschikking/antwoorden) en A/B-testen<sup>[\[3\]](https://systems-analysis.info/int/RAG-patronen#cite_note-langchain-multiquery-3)[\[10\]](https://systems-analysis.info/int/RAG-patronen#cite_note-llama-router-10)</sup>.

### Belangrijkste componenten

- **Chunking:** Een van de meest kritische kwaliteitsfactoren. Naïeve vaste groottes verbreken vaak semantische eenheden. Structuurgeoriënteerde (op opmaak gebaseerde) of recursieve splitters worden aanbevolen (alinea → zin → woord)<sup>[\[17\]](https://systems-analysis.info/int/RAG-patronen#cite_note-rcsplit-17)[\[18\]](https://systems-analysis.info/int/RAG-patronen#cite_note-llama-hier-18)</sup>.
- **Embeddings en metadata:** Sla bij elke chunk document_id, pagina/sectie, titel en datums op; dit is noodzakelijk voor filtering en correcte bronvermelding.
- **Hybride ophalen en herrangschikking:** Gebruik BM25+vector met fusion (of RRF), gevolgd door een cross-encoder voor herrangschikking op een kleine pool van kandidaten<sup>[\[5\]](https://systems-analysis.info/int/RAG-patronen#cite_note-weaviate-hybrid-5)[\[6\]](https://systems-analysis.info/int/RAG-patronen#cite_note-qdrant-hybrid-6)[\[8\]](https://systems-analysis.info/int/RAG-patronen#cite_note-nogueira2019-8)</sup>.
- **Contextinpakking:** Kies *Map‑Reduce*, *Refine* of *Tree‑of‑Chunks* voor lange corpora<sup>[\[16\]](https://systems-analysis.info/int/RAG-patronen#cite_note-raptor-16)[\[18\]](https://systems-analysis.info/int/RAG-patronen#cite_note-llama-hier-18)</sup>.

### Veelvoorkomende fouten (antipatronen)

- **Alleen vectorzoeken** zonder BM25 → mislukkingen bij codes/ID's/acroniemen<sup>[\[5\]](https://systems-analysis.info/int/RAG-patronen#cite_note-weaviate-hybrid-5)[\[7\]](https://systems-analysis.info/int/RAG-patronen#cite_note-milvus-fulltext-7)</sup>.
- **Te grote/kleine chunks** → verlies van context of «verdunning» van de embedding<sup>[\[17\]](https://systems-analysis.info/int/RAG-patronen#cite_note-rcsplit-17)</sup>.
- **Geen herrangschikking in productie** → het LLM ontvangt ruizige context<sup>[\[8\]](https://systems-analysis.info/int/RAG-patronen#cite_note-nogueira2019-8)</sup>.
- **Geen observability** en brontracing → het is onmogelijk om de oorzaken van fouten te analyseren (zie evaluatie van RAG).

## Kwaliteitsevaluatie en metrieken

Evaluatie wordt uitgevoerd op het niveau van ophalen (offline) en end‑to‑end (generatie).

### Metrieken van de retriever

- **Hit Rate, Recall@k, MRR** — dekking en positie van relevante documenten.
- **Context Precision & Recall** — in hoeverre de opgehaalde context vrij is van «ruis» en alles dekt wat nodig is (geïmplementeerd in RAGAS)<sup>[\[19\]](https://systems-analysis.info/int/RAG-patronen#cite_note-ragas-19)</sup>.

### Metrieken van de generator (end‑to‑end)

- **Faithfulness / Groundedness** — de mate waarin het antwoord overeenkomt met de aangeboden context.
- **Answer Relevancy (Relevantie van het antwoord)** — de mate waarin het antwoord overeenkomt met de oorspronkelijke vraag.

Voor de automatisering van metrieken worden open‑source frameworks gebruikt: **RAGAS**, **TruLens** (*RAG triad*: context relevance, groundedness, answer relevance), **DeepEval**<sup>[\[20\]](https://systems-analysis.info/int/RAG-patronen#cite_note-trulens-20)[\[21\]](https://systems-analysis.info/int/RAG-patronen#cite_note-deepeval-21)</sup>.

## Zie ook

- Retrieval-Augmented Generation (RAG)
- Vectordatabases
- Embedding
- AI-agent
- GraphRAG
- MM-RAG
- Evaluatie en benchmarks van LLM

## Literatuur

- Lewis, P., Perez, E., et al. (2020). *Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks*. NeurIPS. arXiv:2005.11401.
- Fan, W., Ding, Y., et al. (2024). *A Survey on RAG Meeting LLMs: Towards Retrieval-Augmented Large Language Models*. KDD. DOI:10.1145/3637528.3671470; arXiv:2405.06211.
- Gao, L., Ma, X., Lin, J., Callan, J. (2023). *Precise Zero‑Shot Dense Retrieval without Relevance Labels (HyDE)*. ACL 2023. ACL Anthology; arXiv:2212.10496.
- Nogueira, R., Cho, K. (2019). *Passage Re‑ranking with BERT*. arXiv:1901.04085.
- Weaviate Docs. *Hybrid search (BM25+Vector)*. <a href="https://docs.weaviate.io/weaviate/concepts/search/hybrid-search" class="external autonumber" rel="nofollow">[1]</a>.
- Qdrant Docs. *Hybrid Queries*. <a href="https://qdrant.tech/documentation/concepts/hybrid-queries/" class="external autonumber" rel="nofollow">[2]</a>.
- Milvus Docs. *Full‑Text Search* / *Hybrid Search*. <a href="https://milvus.io/docs/full-text-search.md" class="external autonumber" rel="nofollow">[3]</a> / <a href="https://milvus.io/docs/hybrid_search_with_milvus.md" class="external autonumber" rel="nofollow">[4]</a>.
- LangChain Docs. *MultiQueryRetriever*. <a href="https://python.langchain.com/docs/how_to/MultiQueryRetriever/" class="external autonumber" rel="nofollow">[5]</a>.
- Cohere Docs. *Rerank — best practices*. <a href="https://docs.cohere.com/docs/reranking-best-practices" class="external autonumber" rel="nofollow">[6]</a>.
- LlamaIndex Docs. *Routing (query routers/selectors)*. <a href="https://docs.llamaindex.ai/en/stable/module_guides/querying/router/" class="external autonumber" rel="nofollow">[7]</a>.
- Yao, S., et al. (2023). *ReAct: Synergizing Reasoning and Acting in Language Models*. ICLR. arXiv:2210.03629.
- Nakano, R., et al. (2021). *WebGPT: Browser‑assisted question‑answering with human feedback*. arXiv:2112.09332.
- Microsoft Research Blog. *GraphRAG: Unlocking LLM discovery on narrative private data*. (2024). <a href="https://www.microsoft.com/en-us/research/blog/graphrag-unlocking-llm-discovery-on-narrative-private-data/" class="external autonumber" rel="nofollow">[8]</a>.
- Microsoft Research. *Project GraphRAG*. (2024). <a href="https://www.microsoft.com/en-us/research/project/graphrag/" class="external autonumber" rel="nofollow">[9]</a>.
- Yu, S., et al. (2024). *VisRAG: Vision‑based Retrieval‑augmented Generation on Multi‑modality Documents*. arXiv:2410.10594.
- Sarthi, P., et al. (2024). *RAPTOR: Recursive Abstractive Processing for Tree‑Organized Retrieval*. arXiv:2401.18059.
- Es, S., et al. (2024). *RAGAs: Automated Evaluation of Retrieval Augmented Generation*. EACL (Demo). <a href="https://aclanthology.org/2024.eacl-demo.16/" class="external autonumber" rel="nofollow">[10]</a>.
- TruLens Docs. *RAG Triad*. <a href="https://www.trulens.org/getting_started/core_concepts/rag_triad/" class="external autonumber" rel="nofollow">[11]</a>.
- DeepEval (GitHub). *The LLM Evaluation Framework*. <a href="https://github.com/confident-ai/deepeval" class="external autonumber" rel="nofollow">[12]</a>.
- LangChain Docs. *RecursiveCharacterTextSplitter*. <a href="https://python.langchain.com/docs/how_to/recursive_text_splitter/" class="external autonumber" rel="nofollow">[13]</a>.
- LlamaIndex Docs. *HierarchicalNodeParser*; *Response Synthesis (Tree/Refine)*. <a href="https://docs.llamaindex.ai/en/stable/api/llama_index.core.node_parser.HierarchicalNodeParser.html" class="external autonumber" rel="nofollow">[14]</a>; <a href="https://docs.llamaindex.ai/en/stable/examples/low_level/response_synthesis/" class="external autonumber" rel="nofollow">[15]</a>.

## Noten

1.  <span id="cite_note-lewis2020-1">↑ <sup>[1.0](https://systems-analysis.info/int/RAG-patronen#cite_ref-lewis2020_1-0)</sup> <sup>[1.1](https://systems-analysis.info/int/RAG-patronen#cite_ref-lewis2020_1-1)</sup> <sup>[1.2](https://systems-analysis.info/int/RAG-patronen#cite_ref-lewis2020_1-2)</sup> <sup>[1.3](https://systems-analysis.info/int/RAG-patronen#cite_ref-lewis2020_1-3)</sup> Lewis, P., Perez, E., et al. (2020). *Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks*. NeurIPS. arXiv:2005.11401.</span>
2.  <span id="cite_note-survey2024-2">[↑](https://systems-analysis.info/int/RAG-patronen#cite_ref-survey2024_2-0) Fan, W., Ding, Y., et al. (2024). *A Survey on RAG Meeting LLMs: Towards Retrieval-Augmented Large Language Models*. KDD. DOI:10.1145/3637528.3671470; arXiv:2405.06211.</span>
3.  <span id="cite_note-langchain-multiquery-3">↑ <sup>[3.0](https://systems-analysis.info/int/RAG-patronen#cite_ref-langchain-multiquery_3-0)</sup> <sup>[3.1](https://systems-analysis.info/int/RAG-patronen#cite_ref-langchain-multiquery_3-1)</sup> <sup>[3.2](https://systems-analysis.info/int/RAG-patronen#cite_ref-langchain-multiquery_3-2)</sup> LangChain Docs. *MultiQueryRetriever*. <a href="https://python.langchain.com/docs/how_to/MultiQueryRetriever/" class="external free" rel="nofollow">https://python.langchain.com/docs/how_to/MultiQueryRetriever/</a></span>
4.  <span id="cite_note-hyde-4">↑ <sup>[4.0](https://systems-analysis.info/int/RAG-patronen#cite_ref-hyde_4-0)</sup> <sup>[4.1](https://systems-analysis.info/int/RAG-patronen#cite_ref-hyde_4-1)</sup> Gao, L., Ma, X., Lin, J., Callan, J. (2023). *Precise Zero‑Shot Dense Retrieval without Relevance Labels*. ACL 2023. arXiv:2212.10496; ACL Anthology: 2023.acl‑long.99.</span>
5.  <span id="cite_note-weaviate-hybrid-5">↑ <sup>[5.0](https://systems-analysis.info/int/RAG-patronen#cite_ref-weaviate-hybrid_5-0)</sup> <sup>[5.1](https://systems-analysis.info/int/RAG-patronen#cite_ref-weaviate-hybrid_5-1)</sup> <sup>[5.2](https://systems-analysis.info/int/RAG-patronen#cite_ref-weaviate-hybrid_5-2)</sup> <sup>[5.3](https://systems-analysis.info/int/RAG-patronen#cite_ref-weaviate-hybrid_5-3)</sup> <sup>[5.4](https://systems-analysis.info/int/RAG-patronen#cite_ref-weaviate-hybrid_5-4)</sup> Weaviate Docs. *Hybrid search (BM25+Vector)*. <a href="https://docs.weaviate.io/weaviate/concepts/search/hybrid-search" class="external free" rel="nofollow">https://docs.weaviate.io/weaviate/concepts/search/hybrid-search</a></span>
6.  <span id="cite_note-qdrant-hybrid-6">↑ <sup>[6.0](https://systems-analysis.info/int/RAG-patronen#cite_ref-qdrant-hybrid_6-0)</sup> <sup>[6.1](https://systems-analysis.info/int/RAG-patronen#cite_ref-qdrant-hybrid_6-1)</sup> <sup>[6.2](https://systems-analysis.info/int/RAG-patronen#cite_ref-qdrant-hybrid_6-2)</sup> Qdrant Docs. *Hybrid Queries*. <a href="https://qdrant.tech/documentation/concepts/hybrid-queries/" class="external free" rel="nofollow">https://qdrant.tech/documentation/concepts/hybrid-queries/</a></span>
7.  <span id="cite_note-milvus-fulltext-7">↑ <sup>[7.0](https://systems-analysis.info/int/RAG-patronen#cite_ref-milvus-fulltext_7-0)</sup> <sup>[7.1](https://systems-analysis.info/int/RAG-patronen#cite_ref-milvus-fulltext_7-1)</sup> <sup>[7.2](https://systems-analysis.info/int/RAG-patronen#cite_ref-milvus-fulltext_7-2)</sup> Milvus Docs. *Full‑Text Search* и *Hybrid Search*. <a href="https://milvus.io/docs/full-text-search.md" class="external free" rel="nofollow">https://milvus.io/docs/full-text-search.md</a>; <a href="https://milvus.io/docs/hybrid_search_with_milvus.md" class="external free" rel="nofollow">https://milvus.io/docs/hybrid_search_with_milvus.md</a></span>
8.  <span id="cite_note-nogueira2019-8">↑ <sup>[8.0](https://systems-analysis.info/int/RAG-patronen#cite_ref-nogueira2019_8-0)</sup> <sup>[8.1](https://systems-analysis.info/int/RAG-patronen#cite_ref-nogueira2019_8-1)</sup> <sup>[8.2](https://systems-analysis.info/int/RAG-patronen#cite_ref-nogueira2019_8-2)</sup> <sup>[8.3](https://systems-analysis.info/int/RAG-patronen#cite_ref-nogueira2019_8-3)</sup> <sup>[8.4](https://systems-analysis.info/int/RAG-patronen#cite_ref-nogueira2019_8-4)</sup> Nogueira, R., Cho, K. (2019). *Passage Re‑ranking with BERT*. arXiv:1901.04085.</span>
9.  <span id="cite_note-cohere-rerank-9">↑ <sup>[9.0](https://systems-analysis.info/int/RAG-patronen#cite_ref-cohere-rerank_9-0)</sup> <sup>[9.1](https://systems-analysis.info/int/RAG-patronen#cite_ref-cohere-rerank_9-1)</sup> Cohere Docs. *Rerank — best practices*. <a href="https://docs.cohere.com/docs/reranking-best-practices" class="external free" rel="nofollow">https://docs.cohere.com/docs/reranking-best-practices</a></span>
10. <span id="cite_note-llama-router-10">↑ <sup>[10.0](https://systems-analysis.info/int/RAG-patronen#cite_ref-llama-router_10-0)</sup> <sup>[10.1](https://systems-analysis.info/int/RAG-patronen#cite_ref-llama-router_10-1)</sup> <sup>[10.2](https://systems-analysis.info/int/RAG-patronen#cite_ref-llama-router_10-2)</sup> LlamaIndex Docs. *Routing (query routers/selectors)*. <a href="https://docs.llamaindex.ai/en/stable/module_guides/querying/router/" class="external free" rel="nofollow">https://docs.llamaindex.ai/en/stable/module_guides/querying/router/</a></span>
11. <span id="cite_note-react-11">↑ <sup>[11.0](https://systems-analysis.info/int/RAG-patronen#cite_ref-react_11-0)</sup> <sup>[11.1](https://systems-analysis.info/int/RAG-patronen#cite_ref-react_11-1)</sup> Yao, S., et al. (2023). *ReAct: Synergizing Reasoning and Acting in Language Models*. ICLR 2023. arXiv:2210.03629.</span>
12. <span id="cite_note-webgpt-12">↑ <sup>[12.0](https://systems-analysis.info/int/RAG-patronen#cite_ref-webgpt_12-0)</sup> <sup>[12.1](https://systems-analysis.info/int/RAG-patronen#cite_ref-webgpt_12-1)</sup> Nakano, R., et al. (2021). *WebGPT: Browser‑assisted question‑answering with human feedback*. arXiv:2112.09332.</span>
13. <span id="cite_note-graphrag-13">[↑](https://systems-analysis.info/int/RAG-patronen#cite_ref-graphrag_13-0) Microsoft Research Blog. *GraphRAG: Unlocking LLM discovery on narrative private data*. 2024. <a href="https://www.microsoft.com/en-us/research/blog/graphrag-unlocking-llm-discovery-on-narrative-private-data/" class="external free" rel="nofollow">https://www.microsoft.com/en-us/research/blog/graphrag-unlocking-llm-discovery-on-narrative-private-data/</a></span>
14. <span id="cite_note-graphrag-project-14">[↑](https://systems-analysis.info/int/RAG-patronen#cite_ref-graphrag-project_14-0) Microsoft Research. *Project GraphRAG*. <a href="https://www.microsoft.com/en-us/research/project/graphrag/" class="external free" rel="nofollow">https://www.microsoft.com/en-us/research/project/graphrag/</a></span>
15. <span id="cite_note-visrag-15">[↑](https://systems-analysis.info/int/RAG-patronen#cite_ref-visrag_15-0) Yu, S., et al. (2024). *VisRAG: Vision‑based Retrieval‑augmented Generation on Multi‑modality Documents*. arXiv:2410.10594; OpenReview: zG459X3Xge.</span>
16. <span id="cite_note-raptor-16">↑ <sup>[16.0](https://systems-analysis.info/int/RAG-patronen#cite_ref-raptor_16-0)</sup> <sup>[16.1](https://systems-analysis.info/int/RAG-patronen#cite_ref-raptor_16-1)</sup> Sarthi, P., et al. (2024). *RAPTOR: Recursive Abstractive Processing for Tree‑Organized Retrieval*. arXiv:2401.18059.</span>
17. <span id="cite_note-rcsplit-17">↑ <sup>[17.0](https://systems-analysis.info/int/RAG-patronen#cite_ref-rcsplit_17-0)</sup> <sup>[17.1](https://systems-analysis.info/int/RAG-patronen#cite_ref-rcsplit_17-1)</sup> LangChain Docs. *RecursiveCharacterTextSplitter*. <a href="https://python.langchain.com/docs/how_to/recursive_text_splitter/" class="external free" rel="nofollow">https://python.langchain.com/docs/how_to/recursive_text_splitter/</a></span>
18. <span id="cite_note-llama-hier-18">↑ <sup>[18.0](https://systems-analysis.info/int/RAG-patronen#cite_ref-llama-hier_18-0)</sup> <sup>[18.1](https://systems-analysis.info/int/RAG-patronen#cite_ref-llama-hier_18-1)</sup> LlamaIndex Docs. *HierarchicalNodeParser* и *Tree Summarization*. <a href="https://docs.llamaindex.ai/en/stable/api/llama_index.core.node_parser.HierarchicalNodeParser.html" class="external free" rel="nofollow">https://docs.llamaindex.ai/en/stable/api/llama_index.core.node_parser.HierarchicalNodeParser.html</a>; <a href="https://docs.llamaindex.ai/en/stable/examples/low_level/response_synthesis/" class="external free" rel="nofollow">https://docs.llamaindex.ai/en/stable/examples/low_level/response_synthesis/</a></span>
19. <span id="cite_note-ragas-19">[↑](https://systems-analysis.info/int/RAG-patronen#cite_ref-ragas_19-0) Es, S., et al. (2024). *RAGAs: Automated Evaluation of Retrieval Augmented Generation*. EACL (Demo). <a href="https://aclanthology.org/2024.eacl-demo.16/" class="external free" rel="nofollow">https://aclanthology.org/2024.eacl-demo.16/</a></span>
20. <span id="cite_note-trulens-20">[↑](https://systems-analysis.info/int/RAG-patronen#cite_ref-trulens_20-0) TruLens Docs. *RAG Triad*. <a href="https://www.trulens.org/getting_started/core_concepts/rag_triad/" class="external free" rel="nofollow">https://www.trulens.org/getting_started/core_concepts/rag_triad/</a></span>
21. <span id="cite_note-deepeval-21">[↑](https://systems-analysis.info/int/RAG-patronen#cite_ref-deepeval_21-0) DeepEval (GitHub). <a href="https://github.com/confident-ai/deepeval" class="external free" rel="nofollow">https://github.com/confident-ai/deepeval</a></span>
