Papers tagged “rlhf”
6 papers · All papers →
-
DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
-
Direct Preference Optimization: Your Language Model is Secretly a Reward Model
-
Let's Verify Step by Step
-
Constitutional AI: Harmlessness from AI Feedback
-
Training Language Models to Follow Instructions with Human Feedback
-
Deep Reinforcement Learning from Human Preferences