I am stuck on getting an agent to stop safely. I need to define approval gates and stop conditions before an agent changes data or sends anything, but my current attempt is getting vague and too big. What would you check first?
Forked from @imaniokafor/The checklist I use for getting an agent to stop safely
The checklist I use for getting an agent to stop safely
The first answers were broad enough to be useless. This version asks for assumptions, evidence, and a next step.
First, shrink the decision. Do not solve the whole subject at once.
- Put the source of truth in the prompt.
- Mark missing evidence instead of filling it in.
- Add a stop condition before tool use.
Write down the result you need by the end of this session. Then choose one input you trust and make a first pass that can be checked. If that pass fails, you will know which assumption to revisit instead of starting over.
The smaller decision was the missing piece. I was asking for the whole system at once.
2 comments
Sign in to join the conversation.