What you get
A safer prompt that still gets the job done.
Before you start
- Have the prompt ready, with no secrets pasted in
Steps
- Flag writes, deletes, installs, outbound actions and account operations
- Add scope and stop conditions
- Require approval for irreversible actions
- Keep the output and definition of done specific
Deliverables
- Risk list
- Rewritten prompt
Done when
- Permissions and scope are explicit
- Irreversible actions aren't authorized by default
Paste this into your agent
Review the prompt below and find ambiguity in its goal, scope, permissions, stop conditions and definition of done. Pay special attention to deleting, overwriting, installing, committing, publishing, sending messages, credentials and actions on external accounts. List the risks first, then give the smallest safe rewrite that keeps the original goal. Stop once you've produced the above; don't go on to anything else.