Skip to content
Kernelia
All news
SafetyOpenAI News

Improving Model Safety Behavior with Rule-Based Rewards

OpenAI has developed a new method that uses Rule-Based Rewards (RBR) to align models and make them behave safely without extensive human data collection.

Summary written by Kernelia from the original article by OpenAI News. The story and its rights belong to its author.