The AI design crit: how to run reviews like elite creative teams
Playbook
Satej Sirur
|
Co-founder & CEO
What a good crit actually checks
Regardless of the industry, channel, or campaign objective, we have found that the best crits check these 7 areas.
Brief adherence
Compare the asset against the brief it came from: objective, audience, funnel stage, channel, placement, primary message, mandatory elements, call to action. If any of those are missing or wrong, stop and rework the asset. No other checks are needed if this fails.
Brand
Logo usage, colour, typography, claims, legal language, accessibility. Failure here should also lead to immediate rework.
Squint test
Shrink the asset to 25%, blur it, and check the focal point, a readable eye path, visible product, and a headline that draws the eye before anyone reads a word.
Craft
Layout, alignment, grid, spacing rhythm, negative space, balance, cropping, type hierarchy, colour hierarchy, lighting, motion.
Taste
Originality, contemporary feel, emotional impact, premium quality, distinctiveness, whether the work leads the category or follows it. This is subjective but it is perhaps the most important in today's age of volume and churn.
AI fingerprints
Patterns that make work feel AI-generated rather than human-crafted. Examples are everything centred, spacing that is uniform instead of rhythmic, products floating with no ground shadow, default typography, lifestyle imagery that could belong to any brand.
Quality
Resolution, compression artifacts, cutout edges, shadows, colour profile, gradient banding, text rendering, safe zones, accessibility, platform specifications.
How to give feedback a designer can act on
The first thing that the crit should share is a verdict: Ship, Tweak, Redo. Mincing words here will waste time.
Next, for each of the 7 areas, share any misses. For each miss, share:
What the miss is in one sentence
A tag: Wrong, Weak, or Taste
How confident you are
What would fix it
Evidence to back up your claim
The tag is perhaps the most overlooked part of this. Wrong means an objective failure such as a brief miss or a broken brand rule. Weak means a craft problem measured against a standard you can point to. Taste means personal preference.
Building an AI Crit for your own brand
Each of the individual checks of the crit can be built into their own prompts. What matters is grounding the prompt with context about your brand. Here are 6 things to add as context to your AI Crit.
Brand book containing colour, logo claims, legal language, accessibility rules, and more
Approved assets with annotations about their tier, channel, and audience
Industry best practices drawn from competitors, adjacent categories, and winning campaigns
Anti-references, meaning work you explicitly rejected or dislike
Past feedback pulled from where you collaborate on designs
Craft standards such as grid system, type scale, spacing scale, and photography treatment
Testing your AI Crit
Calibrate your AI Crit so you can trust it for every campaign. Here are a few tests you can run.
Run it on the last 20 assets that were approved the first time and the last 20 that came back with comments. The approved set has to score higher.
Run it on 20 competitor assets. It should flag them as off-brand rather than passing them, because passing a competitor's work means it is scoring general design quality and not your brand.
Run it on 20 AI-generated assets. It should catch the fingerprints.
AI Crit out of the box with AI Studio
AI Crit is one of the many tools built into AI Studio that allows us to ship tens of thousands of on-brand assets on time for global brands. We fastidiously measure the first-pass rate, which tells us how many assets were approved by customers without any reviews. Tools like AI Crit help us keep our first-pass rate to above 96%.
Where human expertise comes in
No AI Crit can replace your and your team's taste. A model can tell you an asset looks like every other asset in the category, and it can tell you which of your own approved work it resembles. Whether that sameness is a problem for this campaign, in this quarter, against this competitor, is a judgement that belongs to a person with a name, a job title, and accountability.
The 7 checks buy you the right to spend your judgement on that question instead of on whether the logo is in the safe zone. Brand guardians do not need a machine with opinions. They need everything below the opinion handled, at volume, before the work reaches them.
If your reviews keep uncovering problems that were really brief problems, here is how to craft a brief that works.
Frequently asked questions
What is a design crit?
A design crit is a structured review where designers, copywriters, and Creative Directors evaluate work in progress before it goes out. It differs from an approval in that its purpose is to improve the work rather than to authorise it. The rules are that feedback attacks the work rather than the person, and that every criticism comes with a reason.
Who should give feedback on creative work?
Whoever owns the business outcome makes the final call, usually the brand or channel manager. Everyone else contributes within their area: a creative director on craft and taste, a legal or compliance reviewer on claims, an e-commerce manager on platform specification. The common failure is treating all opinions as equally weighted, which turns a review into a negotiation. Settle who decides what before the first round, not during the third.
How do you review AI-generated designs differently?
The 7 checks stay the same, with one addition. AI-generated assets fail in a way human work rarely does: they pass every objective test and still read as machine-made. Centred layouts, uniform spacing, weak scale contrast, floating products, and interchangeable lifestyle imagery are the usual signals.
What is the squint test?
The squint test means resizing an asset to about 25% and blurring it until no text is legible, then judging what survives. It shows you the focal point, the eye path, product visibility, and message clarity as a scrolling shopper would experience them. It costs little, takes seconds, and catches hierarchy problems that a full-size review hides because the reviewer can read.
How many revision rounds are normal, and what causes them?
Two rounds is a healthy target for a standard adaptation, three for new creative. Most teams run more, and the extra rounds usually trace back to one of two causes: an incomplete brief, or a review that started at craft and only reached brief adherence much later. Both are process problems.
Can AI review creative work?
AI can reliably check the objective layers such as brief coverage, brand rules, platform specifications, asset resolutions, safe zones, and accessibility. It is reasonable at craft when it has been trained on the brand's own approved and rejected work. It is unreliable at taste, and any tool that claims otherwise is scoring general design convention rather than your brand. Treat the output as a first pass that clears the mechanical failures before a person spends attention on the interesting ones.
How is a crit different from an approval?
An approval answers one question: can this ship. A crit answers a longer one: what is wrong with this, how badly, and what would fix it. Teams that only run approvals accumulate the same errors across quarters, because nothing in an approval creates a written standard. A crit produces one every time somebody has to explain why a preference is a preference.


