Fine-Tuning Language Models from Human Preferences
TM
Trevor McFedries
@trevvyboi
Reward learning enables the application of reinforcement learning (RL) to tasks where reward is defined by human judgment, building a model of reward by asking humans questions. Most work on reward learning has used simulated environments, but complex infor...
- Uploaded
- Uploaded Jul 10, 2026
- Queried
- Queried 0 times
No preview text is available for this document yet.
Want to learn more?
Ask a question