MDKeyChunker: What Does One LLM Call per Chunk Buy for Markdown Retrieval?
October 6, 2026
MDKeyChunker optimizes Markdown retrieval by utilizing structural headers for chunking and performing a single LLM call per chunk to generate a rolling dictionary of metadata. Evaluations using Qwen2.5-7B across Qasper and FreshStack datasets show that these structural chunks outperform standard 512-character windows under hybrid retrieval.