Prompt injection attacks against GPT-3
Riley Goodside, yesterday: Exploiting GPT-3 prompts with malicious inputs that order the model to ignore its previous directions. pic.twitter.com/I0NVr9LOJq- Riley Goodside (@goodside) September 12, 2022 Riley provided several examples. Here’s …
Prompt injection attacks against GPT-3 Simon Willison’s Weblog Subscribe Sponsored by: Atlassian - Give your agents a plan. Not a prompt. New Jira capabilities unlock full-context for AI-native software development. Assign tasks to Claude, Cursor, or GitHub Copilot, now directly from Jira. Learn more Prompt injection attacks against GPT-3 12th September 2022 Riley Goodside, yesterday : Exploiting GPT-3 prompts with malicious inputs that order the model to ignore its previous directions. pic.twitter.com/I0NVr9LOJq - Riley Goodside (@goodside) September 12, 2022 Riley provided several examples.
Explore this link on the map →related reading
- You can’t solve AI security problems with more AIsimonwillison.net
- Prompt injection and jailbreaking are not the same thingsimonwillison.net
- GPT-4openai.com
- CaMeL offers a promising new direction for mitigating prompt injection attackssimonwillison.net
- gpt-4.pdfcdn.openai.com
- HackAPromptpaper.hackaprompt.com
- Prompt Engineering Tips and Tricks with GPT-3 · andrew makes thingsblog.andrewcantino.com
- llm-security/README.md at main · greshake/llm-security · GitHubgithub.com
- The lethal trifecta for AI agents: private data, untrusted content, and external communicationsimonwillison.net
- The Dual LLM pattern for building AI assistants that can resist prompt injectionsimonwillison.net
- ASCII art elicits harmful responses from 5 major AI chatbots - Ars Technicaarstechnica.com
- Prompt Engineering | Kagglekaggle.com