●builderYou can use this benchmark to evaluate how well your RAG pipelines extract paired visual and textual evidence from complex documents.
●researcherThis provides a new evaluation framework for joint image-text extraction tasks in long-form document processing.