Prompt injection attacks against GPT-3
Riley Goodside, yesterday: Exploiting GPT-3 prompts with malicious inputs that order the model to ignore its previous directions. pic.twitter.com/I0NVr9LOJq- Riley Goodside (@goodside) September 12, 2022 Riley provided several examples. Here’s …
Prompt injection attacks against GPT-3 Simon Willison’s Weblog Subscribe Sponsored by: Atlassian - Give your agents a plan. Not a prompt. New Jira capabilities unlock full-context for AI-native software development. Assign tasks to Claude, Cursor, or GitHub Copilot, now directly from Jira. Learn more Prompt injection attacks against GPT-3 12th September 2022 Riley Goodside, yesterday : Exploiting GPT-3 prompts with malicious inputs that order the model to ignore its previous directions. pic.twitter.com/I0NVr9LOJq - Riley Goodside (@goodside) September 12, 2022 Riley provided several examples.
related reading
- GitHub - brexhq/prompt-engineering: Tips and tricks for working with Large Language Models like OpenAI's GPT-4.github.com
- llm-security/README.md at main · greshake/llm-securitygithub.com
- CaMeL offers a promising new direction for mitigating prompt injection attackssimonwillison.net
- You can’t solve AI security problems with more AIsimonwillison.net
- Prompt injection and jailbreaking are not the same thingsimonwillison.net
- A Mechanistic Explanation of Prompt Injection (and why you should study roles) — LessWronglesswrong.com
- Prompt Injection as Role Confusionrole-confusion.github.io
- GPT-4openai.com
- gpt-4.pdfcdn.openai.com
- HackAPromptpaper.hackaprompt.com
- The lethal trifecta for AI agents: private data, untrusted content, and external communicationsimonwillison.net
- Prompt Engineering Tips and Tricks with GPT-3 · andrew makes thingsblog.andrewcantino.com