Back to search

Article

Improved Value Alignment in Large Language Models Using Variational Best-of-N Techniques

2024-07-25

Abstract excerpt

<title>Abstract</title> <p>Large language models have shown high capabilities in generating human-like text and performing complex language-related tasks, yet they face significant challenges regarding value alignment to prevent the generation of harmful or biased content. The novel integration of the Variational Best-of-N technique within the Llama model enhances the ability to generate ethically aligned content...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
0d9ac18d-42e6-5254-aeca-bf69c8b23c54
DOI
10.21203/rs.3.rs-4794797/v1
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
Improved Value Alignment in Large Language Models Using Variational Best-of-N TechniquesDOI 10.21203/rs.3.rs-4794797/v1
Select a neighboring publication to make it the new centre.