Preprint2024
Uncertainty-aware reward model: Teaching reward models to know what is unknown
Unknown
This paper introduces an uncertainty-aware reward model that teaches reward models to recognize and quantify their own uncertainty, improving alignment of LLMs with human expectations.
0Oct 1, 2024Large Language Models
arXiv