safety
fact
bearish
Autoregressive large language models routinely generate factually incorrect outputs with high decoding confidence, limiting their deployment in high-stakes workflows
Autoregressive large language models (LLMs) routinely generate factually incorrect outputs with high decoding confidence, limiting their deployment in high-stakes workflows.
Computation and Language30 Aug 2026