
Better Vector Search for Long Documents: Chunking Strategies Inside Manticore Search
An embedding model reads only the first few hundred tokens of a document and silently drops the rest. Manticore Search now splits long documents for you at INSERT time: add chunk_strategy to the vector column and pick one of five chunking strategies. No ingest pipeline, no text splitter library. On our own manual, recall@5 for deep content went from 55% to ...









