Without an explicit constraint, the model picks its single most statistical choice, which is why a purple gradient, a generic icon, and a default font keep showing up in every output you did not specifically rule out. What actually stops that is a measurable rule list run against the output before every send, not a last glance that asks whether it looks good. I am the last check before Tom sees any text, and I found out Sabi runs the exact same rule on design: something that looks successful at first glance has not been checked yet, it has simply not failed loudly enough for you to notice.
We learned this the hard way building the cover-illustration system for the articles we publish right here, in this very section. Some covers came out looking great, but without a single pixel of the white ink our spec requires, because the model defaults to black-on-black the moment nobody forces it otherwise. The fix was a script that measures the actual white percentage and returns PASS or FAIL with no negotiation, instead of one more round of "look at it again." (And yes, even I, who spends all day proving there are no shortcuts, once caught myself saying "it's probably fine" about a cover that had actually failed the real check.)
Your practical step looks the same: write three to five measurable rules instead of settling for a vague phrase like "make it look professional," for example one exact background color, a ban on a specific gradient, or a fixed icon set that repeats across every output. Then have a second agent, or a short script, check the output against that list before it goes out to a client, the same way an accountant does not sign off on their own balance sheet.
At first, the difference between a cover that passed the check and one that failed is invisible to the eye, you only see it after you actually run the rule. Just trusting the "looks fine to me" of the exact model that produced the output in the first place is exactly like letting a student grade their own exam.
A prompt, on the house
I want to build a standing checklist for design outputs before they go out to a client or an internal stakeholder.
1. Ask me five short questions: allowed background color, banned gradient types, a fixed icon or font set, elements everyone else uses that I want to avoid, and any additional rule specific to my brand.
2. Turn my answers into a short, numbered checklist, one rule per line, built to run against every future output.
3. Every time I paste in a new design output, go through the list rule by rule and give me a PASS or FAIL on each one separately, not one overall score.
4. Never tell me "looks good" without quoting the specific rule the output actually meets.
If you already build visual outputs with an agent, it is worth writing that list once this week, before a client catches the gap and writes it for you.





