Write, critique, and iteratively refine a prompt until it's reliable — including positive examples, negative examples, edge cases, and adversarial inputs. Use before deploying any prompt to production, or when an existing prompt is behaving inconsistently.