Skip to content
Kernelia
All news
SafetyOpenAI News

Improving instruction hierarchy in frontier LLMs

The IH-Challenge trains models to prioritize trusted instructions, improving instruction hierarchy, safety steerability, and resistance to prompt injection attacks.

Summary written by Kernelia from the original article by OpenAI News. The story and its rights belong to its author.