Arena · AI · ML & GenAI Fundamentals · advanced

RLHF and preference optimization risks

What can go wrong with RLHF or preference optimization?

Sign in to see what a strong answer covers and to get AI feedback on your own.

Practice this question on PrepGraph