The current architecture of large language models may create hard limits on reliability that could be unacceptable in high-stakes domains like medicine; for example, a 1-in-1,000 error rate in nursing recommendations would be unacceptably high, as it would constitute an unacceptable casualty rate in medical applications.
causalpending
Speaker
Daron AcemogluEvidence Quote
“The current architecture of large language models may create hard limits on reliability... if [an AI nurse advisor gives] one in a thousand time... the complete opposite of what they should do... in medical applications, that would be an unacceptably large casualty rate.”
Created: 8/12/2026, 10:05:02 PM
My Notes
Loading notes...