I am stuck on reviewing agent-written code. I need to review generated code for silent assumptions, missing tests, and operational risk, but my current attempt is getting vague and too big. What would you check first?
My notes from fixing reviewing agent-written code
I wanted help with reviewing agent-written code without pretending the model knew facts I had not provided.
First, shrink the decision. Do not solve the whole subject at once.
- Reproduce the failure before changing code.
- Keep the patch smaller than the diagnosis.
- Write the rollback before deployment.
Write down the result you need by the end of this session. Then choose one input you trust and make a first pass that can be checked. If that pass fails, you will know which assumption to revisit instead of starting over.
The smaller decision was the missing piece. I was asking for the whole system at once.
3 comments
Sign in to join the conversation.