Do LLMs Estimate Uncertainty Well in Instruction-Following?
Large language models (LLMs) could be valuable personal AI agents across various domains, provided they can precisely follow user instructions. However, recent studies have shown significant limitations in LLMs’ instruction-following capabilities, raising concerns about their reliability in high-stakes applications. Accurately estimating LLMs’ uncertainty in adhering to instructions is critical to mitigating deployment risks. We present, …
Read more “Do LLMs Estimate Uncertainty Well in Instruction-Following?”