ExtractBench is a benchmark for schema-guided document extraction that evaluates value accuracy, record completeness, grounding, and cost. The dataset comprises 4,869 pages across 370 enterprise documents to test the ability of agents to follow user-defined schemas accurately.
HOW THIS AFFECTS YOU
●
builderYou can use this to measure how well your extraction agents perform against real-world business constraints like cost and completeness.