RLHF
Quick Definition
Aligning LLMs with human preferences using reinforcement learning
Full Definition
Reinforcement Learning from Human Feedback aligning language models with human preferences.
Examples
ChatGPT alignment, preference learning, safety tuning
Related Terms
reinforcement-learning
reward-model
alignment
More Large Language Models Terms
Sliding Window Attention
Attention limiting tokens to nearby window only
RAG
Enhancing LLMs by retrieving relevant external knowledge
Tool Use
Capability of LLMs to call external tools and APIs
Llama
Meta AI's open-source large language models
Synthetic Data Generation
Using AI to create artificial training data replicating patterns
Fine-Tuning
Further training pre-trained models on specific domain data