Skip to content

Opening book details…

About this document

Reinforcement Learning Enhanced LLMs A Survey - 2412.10400v3 by Aemy Jaffary is a document available to read on EtoBox.

This document is a survey on reinforcement learning (RL) enhanced large language models (LLMs), highlighting their performance improvements and the complexities involved in their implementation. It reviews the basics of RL, popular RL-enhanced LLMs, and various training techniques such as Reinforcement Learning from Human Feedback (RLHF) and Direct Preference Optimization (DPO). The survey aims to consolidate existing research, identify challenges, and suggest avenues for further advancements in the field o

Author
Aemy Jaffary
Language
EN