When models disagree on predictions, it often signals tricky or risky inputs...
https://wiki-canyon.win/index.php/How_to_Set_a_Disagreement_Threshold_for_Human_Review
When models disagree on predictions, it often signals tricky or risky inputs worth a closer look. Tracking metrics like ensemble variance or prediction margin helps spot these uncertain cases