SafetyOpenAI News
Improving instruction hierarchy in frontier LLMs
The IH-Challenge trains models to prioritize trusted instructions, improving instruction hierarchy, safety steerability, and resistance to prompt injection attacks.
Summary written by Kernelia from the original article by OpenAI News. The story and its rights belong to its author.

