SafetyOpenAI News
Improving Model Safety Behavior with Rule-Based Rewards
OpenAI has developed a new method that uses Rule-Based Rewards (RBR) to align models and make them behave safely without extensive human data collection.
Summary written by Kernelia from the original article by OpenAI News. The story and its rights belong to its author.

