Skip to content
AI Lehel Briefing
← Back to latest
Models & research OpenAI

Improving instruction hierarchy in frontier LLMs

IH-Challenge trains models to prioritize trusted instructions, improving instruction hierarchy, safety steerability, and resistance to prompt injection attacks.

Excerpt supplied by the publisher’s feed

Read the original article at OpenAI Opens in a new tab