# Introduction A model that says it is 90% confident should be right 90% of the time. When that relationship breaks down, you get a miscalibration problem. The model’s scores stop telling you anything useful about reliability. For large language models (LLMs), miscalibration is widespread. A 2024 NAACL survey found that confidence scores diverge […]
We Should Train AI to Betray Its Users
The dilemma employee at an engineering company but have uncovered a deadly secret. Your company is performing ill-advised engineering activities that have already killed six contractors in a landslide. Despite this the company is pressing ahead, creating risks of further landslides, a catastrophic dam breach and/ or groundwater contamination. Instead of dealing with the problem, […]

