Hybrid LLM System Achieves High Accuracy in L2 English Assessment
August 28, 2026
An interpretable feature-plus-LLM hybrid model for L2 English speaking assessment achieved a Spearman rho of 0.818 against human consensus. The system outperforms the median human rater by combining deterministic speech-timing with LLM fluency judgments.
HOW THIS AFFECTS YOU
●
builderYou can build more trustworthy automated scoring tools by combining deterministic signal processing with LLMs.
●
designerThis enables new, low-anxiety interfaces for language learners to practice speaking.