All posts
Chunking strategies for RAG
Engineering 9 min read· 14 Aug 2026· By Engineering

Chunking strategies for RAG

Fixed-size, semantic, and recursive chunking — with the tradeoffs.

The single biggest RAG quality lever isn't the model — it's the chunker.

Fixed-size (baseline)#

512-1024 tokens with 10% overlap. Cheap, easy, works for most content.

Recursive#

Split on paragraphs, then sentences, then words — collapsing until each chunk fits the size cap. Preserves structure.

Semantic#

Embed each sentence, cluster on cosine distance, chunk cluster boundaries. Best quality, 3× slower to index.

Ready to build?

Grab a key, keep your OpenAI SDK, and pay in INR.

Start free