TACT Benchmark for Intent-Conditioned Full-Duplex Dialogue Turn-Taking
September 24, 2026
The TACT benchmark introduces a continuous ranked probability score to evaluate full-duplex dialogue models based on speaker intent rather than binary timing windows. It utilizes 73.2 hours of dyadic corpora to penalize inappropriate silences or overlaps by weighting them against human-derived timing kernels.
HOW THIS AFFECTS YOU
●
builderThis provides a more realistic evaluation framework for developing low-latency, conversational voice agents.
●
researcherYou can now evaluate spoken dialogue models using intent-aware timing distributions instead of rigid windows.