
Engineering 9 min read· 14 Aug 2026· By Engineering
Chunking strategies for RAG
Fixed-size, semantic, and recursive chunking — with the tradeoffs.
The single biggest RAG quality lever isn't the model — it's the chunker.
Fixed-size (baseline)#
512-1024 tokens with 10% overlap. Cheap, easy, works for most content.
Recursive#
Split on paragraphs, then sentences, then words — collapsing until each chunk fits the size cap. Preserves structure.
Semantic#
Embed each sentence, cluster on cosine distance, chunk cluster boundaries. Best quality, 3× slower to index.