Article
Improved Value Alignment in Large Language Models Using Variational Best-of-N Techniques
2024-07-25
Abstract excerpt
<title>Abstract</title> <p>Large language models have shown high capabilities in generating human-like text and performing complex language-related tasks, yet they face significant challenges regarding value alignment to prevent the generation of harmful or biased content. The novel integration of the Variational Best-of-N technique within the Llama model enhances the ability to generate ethically aligned content...
Topics
Open a Topic to create a Post that cites this publication.
Identifiers and source
- Literature Corpus work
- 0d9ac18d-42e6-5254-aeca-bf69c8b23c54
- DOI
- 10.21203/rs.3.rs-4794797/v1
