Scenario (hypothetical)
A team builds an agent that updates lead fields based on enrichment and AI summaries. It works well in testing, then runs on the full database with a prompt change nobody reviewed.
What can go wrong
- Bulk overwrites of human-entered data.
- Inconsistent values that break segmentation.
- No audit trail to reverse the changes.
Controls
Version prompts like code, write to staging fields first, cap batch sizes, and require approval for any change to production logic.
Comments