What is late chunking or contextual retrieval?
Simple meaning
Late chunking embeds a long passage with full context then pools token vectors into chunk embeddings so boundaries lose less meaning.
Open the full page for Why, Steps, Example and Key takeaway.