Critique without a maker
The most common advice for design teams using AI tools in 2026: treat the output like a junior designer's work. Review it, critique it, give it feedback. The advice sounds practical. It misreads what made critique work.
Design critique works because you can ask the maker why. Why did you put the action here, and what user need drove this layout? The junior designer answers, sometimes badly, but the answer reveals intent. Critique operates on intent. You interrogate the choice, push on the reasoning, and the work improves because the maker adjusts their thinking, not just their pixels.
AI output has no reasoning to push on. It has a prompt, a model, and a probability distribution. When you sit in front of an AI-generated screen and ask "why is the call to action below the fold," nobody answers. The AI in Design Report 2026 found that AI has become excellent at generating possibilities but still struggles with intent. Intent is not a feature on a roadmap.
So what happens when teams try to critique intentionless output? They evaluate. They look at the thing and ask whether it is good. thecrit.co draws the line: critique makes designs better, review makes decisions about them. Most teams running "critique" on AI output are actually running review, but they have not updated the label.
The agent extension test exposes this
Here is the test: can you describe your critique process clearly enough that an agent, human or software, could apply it to a new case and produce a result you would endorse? If yes, you have a process. If no, you have a habit that depends on the people in the room.
Try it: write down how your team critiques a screen. Most teams produce something like: "We look at it together and give feedback based on experience." That describes a social ritual where the maker's presence does most of the structural work. The maker explains intent and the critics react to it. Remove the maker and the ritual collapses.
The teams that will survive this transition are the ones writing evaluation criteria before the output exists. Adam Elman at NN/g argues that good judging criteria must be as objective as possible without becoming arbitrary. Itamar Medeiros at designative.info makes the standard concrete: a good criterion is a testable statement like "for this task type, in this user context, the agent must do this behavior to this standard."
That level of specificity is uncommon in design teams. 91% of surveyed designers now use AI weekly, according to the AI in Design Report 2026, but only 28% of leaders say their companies have made formal updates to evaluation, comp, or hiring. The gap is not subtle.
This is where the agent extension test bites hardest. If you cannot write your criteria down, you cannot delegate judgment to an AI or a new hire. You were relying on the maker to bring the structure, then calling your reaction to that structure a "process."
NN/g names the shift: the output of research and design is moving from documents written for humans to curated context that guides AI. The work is writing the criteria that make feedback possible before any output exists.
Some designers will resist this because it feels bureaucratic. An ACM DIS 2026 study on cognitive outcomes in generative AI work found that some participants experienced questioning as adversarial but acknowledged its cognitive value. Writing criteria before output exists forces you to articulate what good looks like when you cannot point at a screen. That is uncomfortable, and it is the work.
The junior designer analogy is comfortable because it preserves the existing workflow: you still sit in a room and give feedback. The valuable part was always the interrogation of intent, and that interrogation required a mind on the other side with reasons it could defend. Without that mind, you need criteria. Without criteria, you are voting on aesthetics with extra steps.
The question for every design team using AI: if you removed every human maker from the room, could your critique process still function on its own terms? If the answer is no, what you have is a dependency you have not named.