overfeed.news

The Instruction Hierarchy: Training LLMs to Prioritize Privileged Instructions

OpenAI Newsen

2a

Idade

Publicado
Coletado

Today's LLMs are susceptible to prompt injections, jailbreaks, and other attacks that allow adversaries to overwrite a model's original instructions with their own malicious prompts.

Leia o artigo completo em openai.com

O overfeed.news indexa e aponta. Publicamos um trecho curto — o artigo completo fica em OpenAI News.

Mais de OpenAI News

Entre para seguir esta fonte
The Instruction Hierarchy: Training LLMs to Prioritize…