Arena · AI · ML & GenAI Fundamentals · advanced
RLHF and preference optimization risks
What can go wrong with RLHF or preference optimization?
Sign in to see what a strong answer covers and to get AI feedback on your own.
Practice this question on PrepGraph