Uncertainty-Aware Trust Estimation for Heterogeneous Multi-LLM Ensembles
July 24, 2026
This method uses Cooke-style log weighting and context-aware calibration to weight individual LLM contributions in an ensemble. It penalizes overconfident incorrect predictions to improve reliability in heterogeneous multi-model systems, tested on MMLU and MMLU-Pro.
HOW THIS AFFECTS YOU
●
builderYou can implement this to build more reliable ensembles that avoid being misled by unreliable experts.
●
researcherThe use of structured expert judgment for calibration offers a new way to handle model heterogeneity.