1 min read
Constitutional AI: Self-Supervised Alignment
Understanding Constitutional AI and how it enables scalable alignment through self-critique and revision.
3 articles
Understanding Constitutional AI and how it enables scalable alignment through self-critique and revision.
Understanding Direct Preference Optimization as a simplified approach to aligning LLMs with human preferences.
Understanding Reinforcement Learning from Human Feedback and its role in aligning LLMs with human preferences.