What is RLHF?

A training method where humans rate an AI's responses and the AI is retrained on those ratings, the key technique behind making chatbots more natural and useful.

Briefings mentioning this
2026-09-10

Terms seen alongside this one