All News
Neural Networks

The Instruction Hierarchy: Training LLMs to Prioritize Privileged Instructions

April 19, 20242 views1 min read

Today's LLMs are susceptible to prompt injections, jailbreaks, and other attacks that allow adversaries to overwrite a model's original instructions with their own malicious prompts.

Want to try AI?

Compare the best neural networks in one place — for free

Go to neural networks
Share:
Источник: OpenAI Blog

Related News