Artificial Intelligence #llm#artificial intelligence
LLM Confidence Is Epistemically Vacuous: New Method Detects Blind Spots in Clinical Data
A new study reveals that large language models (LLMs) fail to recognize their own knowledge limits on structured clinical data, outputting near-constant confidence scores regardless of accuracy. Researchers propose a cross-model calibrator using attribution divergence between LLM and XGBoost, reducing calibration error from 0.254 to 0.080 and improving accuracy from 49% to 75.3% without training.
Jun 20, 2026 1 source