Copilotly

5 copilots · Design & Creative

An AI designer,
available right now.

UX review, brand identity, layout and colour.

5 specialist copilots for design & creative, included in one subscription with 126 more across 19 other domains.

Free plan, no card. Pro from $4.99/week for everything.

Design & Creative Copilot5 copilots
Contrast audit
  • Body on background12.1:1
  • Muted label3.2:1
  • Button text6.8:1

One row fails AA. It is the one on every card.

What the Design & Creative Copilot actually does

  • Critique a layout against what it is asking the user to do
  • Build a type scale and a spacing system that hold together
  • Check colour contrast and fix it without abandoning the palette
  • Write the design rationale that makes a decision defensible
  • Audit a flow for the steps that lose people
  • Prepare a critique so feedback is about the work, not taste

What the human equivalent costs

$75-200/hour

Freelance product and brand design in the US. A full identity package from a studio starts around $5,000 and rises steeply.

Indicative range, not a surveyed figure.

Indicative range, not a surveyed figure. Verified July 2026.

Copilotly Pro is $4.99/week for every copilot across all 20 domains - and the free plan needs no card.

Critique against a goal, not against taste

The reason design reviews go badly is that without a stated goal, every comment is a preference and the most senior preference wins. Nobody learns anything and the work does not improve.

The Design Copilot starts from what the screen is trying to achieve, which converts "I do not like the button" into "the primary action is competing with three others" - a claim that can be evaluated.

Systems are what stop drift

Almost every product that looks inconsistent looks that way because decisions were made one screen at a time. Twelve type sizes, five greys, spacing values that nobody chose.

A type scale and a spacing scale are an afternoon's work and they prevent the slow accumulation that takes a quarter to unwind. This is the least glamorous and highest-leverage thing in product design.

Accessibility is a constraint, not a review stage

Contrast ratios, focus states, target sizes and text alternatives are cheap at design time and expensive as a retrofit - particularly contrast, which usually means renegotiating a palette that has already been approved.

WCAG AA is the working standard: 4.5:1 for body text, 3:1 for large text and interface components. Checking a palette against it before anything is built is a five-minute step that saves a fortnight.

What people actually bring to it

Not hypotheticals. These are the situations this copilot sees most.

  • A layout that feels wrong with no vocabulary for why
  • A palette that fails contrast and a brand that will not budge
  • Twelve font sizes across one product and no scale
  • A stakeholder review with no rationale prepared
  • A signup flow with a drop-off nobody has located
  • A design system that started well and drifted

A worked example, start to finish

A signup form looks clean, tests fine with the team, and loses 60% of users between step one and step two.

  1. 01

    Establish what step two asks for

    Drop-off at a specific step is almost always about what that step demands rather than how it looks. Phone number, payment details or company size at step two are all known drop-off causes, and none of them are visual problems.

  2. 02

    Check whether the ask is justified in place

    Users abandon when a request feels unearned. The same field converts differently depending on whether the reason is visible next to it, and that is a copy decision as much as a design one.

  3. 03

    Look at the progress signal

    A flow with no visible end is a flow people leave. Knowing there are three steps rather than an unknown number materially changes completion, which is a cheap fix.

  4. 04

    Test the actual mobile rendering

    Team testing happens on a large screen with autofill. Real users are on a phone with a keyboard covering the submit button. This is a genuinely common and genuinely invisible failure.

  5. 05

    Change one thing

    Moving the phone field to after signup, or justifying it in a line of copy. One change, measured. Redesigning the whole flow gives a number and no understanding of what caused it.

It was the phone number field, and it did not need to be there yet. Most drop-offs are a demand problem wearing a design problem's clothes.

What helps before you start

Design critique needs the goal. Without it, feedback collapses into preference, which is where most reviews go wrong.

  • What this screen is asking the user to do
  • Who the user is and what they know
  • The constraints - brand, platform, existing system
  • What has already been tried and rejected
  • The screenshot or the description, in detail

What goes wrong most often

Designing for the portfolio rather than the task

A screen can be beautiful and fail. The question is always whether the user does the thing, and that is measurable.

Ignoring contrast until accessibility review

Retrofitting contrast into a finished palette is painful. Checking at palette stage costs nothing.

Presenting work without rationale

Without a stated reason, feedback defaults to preference and the loudest opinion wins. Rationale turns a review into a discussion about goals.

A type scale that grew by accident

Twelve sizes means no scale. It is the most common source of the feeling that something is off without an obvious cause.

When to use this, and when to hire a designer

Including the rows that send you elsewhere. A tool that never does that is not being honest with you.

  • Structured critique of a layoutThis copilotAgainst principles, not taste.
  • Building type and spacing scalesThis copilotSystematic and quickly checked.
  • Contrast and accessibility checksThis copilotRatios are arithmetic.
  • Brand identityA professionalJudgement and craft. Hire someone.
  • Illustration and visual craftA professionalNot a description problem.
  • Research with actual usersA professionalNothing substitutes for watching people.

Visual hierarchy, and why layouts feel wrong

The complaint that something feels off is almost always a hierarchy problem, and hierarchy is one of the few areas of design where the rules are concrete enough to check.

A layout has a hierarchy whether or not one was designed. The eye goes somewhere first, and if that place is not what the screen is for, the screen is working against itself. Two elements competing for primary attention means neither has it.

The tools that create hierarchy are size, weight, colour, contrast, and space - and space is the one that gets used least and works best. Elements grouped tightly read as related; elements separated read as distinct. Most cluttered interfaces are not carrying too much content, they are carrying content with undifferentiated spacing, so nothing groups.

Contrast is the second underused tool. A page where everything is bold has no emphasis. Making one thing prominent usually means making the other things less so, which feels like a loss and is how emphasis works.

The practical test is to squint at the screen or blur the screenshot. What remains legible is the hierarchy as a user experiences it in the first second. If the primary action is not what survives, that is the problem, and it is a specific fixable one rather than a matter of taste.

Building a system that survives contact with a team

A design system fails in a predictable way: it is built thoroughly, adopted partially, and then bypassed under deadline until it describes a product that no longer exists.

The systems that survive are the ones that are easier to follow than to ignore. That means fewer decisions, not more documentation. A type scale with five sizes gets used; one with twelve is a menu, and a menu recreates the inconsistency it was meant to solve.

The same applies to spacing. A scale based on a consistent step - four or eight pixels, multiplied - removes the per-decision judgement entirely. Any value not on the scale becomes visibly a choice, which is exactly the friction you want.

Colour needs roles rather than names. A palette of hex values invites arbitrary use. A set of semantic tokens - surface, ink, accent, border, and their states - encodes intent, which means a theme change is a token change rather than a search across the codebase.

And components need to cover the awkward states. A button component with default, hover, active, focus, disabled and loading is a component. One with default and hover is a suggestion, and the missing states get invented separately by everyone who needs them.

The cheapest useful research you are not doing

Formal user research is expensive and often skipped entirely, which leaves teams designing on assumption. There is a substantial middle ground that costs almost nothing.

Watching five people attempt a task is the classic example, and the finding that a handful of participants surfaces the majority of usability problems has held up well in practice. It does not need a lab, a script, or a research function - it needs five people, one task, and the discipline not to help them.

Not helping is the hard part. The instinct when someone gets stuck is to explain, and the moment of getting stuck is the entire data. Watching in silence is uncomfortable and it is where the value is.

Support tickets are a research corpus that already exists. Recurring confusion in support is a design problem that has already been reported, categorised and paid for, and almost nobody in design reads them.

Analytics tells you where people stop and never why, which makes it a good instrument for choosing what to investigate and a poor one for deciding what to change. Pairing a drop-off number with five people attempting that step is the combination that actually answers a question.

The states that get designed last and break first

Designs are made in the ideal case - real content, sensible length, everything loaded, nothing wrong. Products spend a meaningful share of their time in none of those conditions.

Empty states are the first gap. A dashboard with no data, a list with no items, a search with no results. These are the screens a new user sees first, which makes them the highest-stakes screens in the product and the ones designed last. An empty state that explains what goes here and offers the action to create it is onboarding; a blank panel is confusion.

Loading states are the second. Something has to occupy the gap, and a layout that shifts when content arrives is worse than one that reserves the space. Skeletons work because they set the expectation of shape.

Error states are the third and most neglected. An error message that names what went wrong and what to do next is a design artefact. One that says something went wrong is an admission that nobody thought about it.

Then the extremes of real content. A name three times longer than the placeholder, a list with four hundred items, a title with no spaces. Every one of those exists in production and none of them exist in the mockup.

Designing the awkward states is unglamorous and it is most of the difference between a product that demos well and one that holds up.

What it will not do

Stated before the pitch rather than after it. On a page titled “AI designer” this is the part that matters most.

  • It does not generate images or visual mockups
  • It cannot replace judgement on brand and craft
  • Contrast checking is not a full accessibility audit
  • It has not seen your users or your product analytics
  • For identity work, hire a designer

AI designer: common questions

Can it critique a design I show it?

Yes - layout, hierarchy, contrast, spacing and whether the primary action is actually primary. It critiques against principles rather than taste, which is what makes it useful in a review.

Tell it what the screen is for. Critique without a goal is just opinion with more words.

Will it generate images or mockups?

No. It works in words - structure, systems, critique and rationale. For visual generation you want a dedicated image tool, and for actual craft you want a designer.

Is it useful for someone who is not a designer?

Very. Founders and engineers making interface decisions without a designer are the clearest case - it supplies the vocabulary and the principles that turn a vague feeling into a specific fix.

It will not make you a designer, and it will meaningfully raise the floor.

Can it check accessibility?

It checks contrast ratios against WCAG, flags common structural issues, and reviews for focus states and target sizes.

That is not a full audit. Real accessibility testing involves assistive technology and actual users, and neither is something software can simulate.

How do I give design feedback without being vague?

State the goal, then the observation, then the effect. Something like: the goal is signup, and the two buttons have the same weight, so the primary action does not stand out.

That is a claim about the work rather than a preference, and it can be discussed. It will help you translate a feeling that something is off into that form, which is most of what makes a review productive.

Should we build a design system?

A type scale, a spacing scale and semantic colour tokens are worth it almost immediately, and they take an afternoon. A full component library with documentation is worth it once more than a couple of people are building interface.

The failure mode is building the large version first. It gets bypassed under deadline and then describes a product that no longer exists.

Can an AI designer replace a real one?

It replaces the hour you would have spent working it out alone, not a professional engagement. The Design & Creative Copilot gives you a structured starting point, drafts you can use, and the specific questions worth asking - so you move faster and arrive better prepared.

How is this different from asking ChatGPT about design & creative?

A general-purpose assistant has to stay safe across every subject at once, so on design & creative questions it hedges. The Design & Creative Copilot is configured for this field alone - its own system prompt, model and parameters - which is the difference between "you may want to check your local rules" and a named rule, a deadline and a draft you can send.

OpenAI has also been narrowing what ChatGPT will say about professional matters, which is precisely the gap these copilots exist to fill.

What can the Design & Creative Copilot actually do?

UX review, brand identity, layout and colour.

There are 5 specialist copilots inside this domain, each tuned to a narrower job, so you are not asking one generalist to cover everything.

What does it cost?

The free plan gives you three copilots of your choice, 50 messages a day and the browser extension, with no card required. Pro starts at $4.99/week and unlocks all 131 copilots across all 20 domains, with unlimited messages, document upload and the mobile apps. Annual works out at $24.17/month.

There is a 3-day free trial and a 7-day money-back guarantee.

Is what I share private?

Conversations are encrypted in transit and at rest. We do not use your data to train models and we do not share it with third parties. Given how much of what people bring to a designer is sensitive, that is a requirement rather than a feature.

What if it gets something wrong?

It can. Treat any answer as a well-informed starting point rather than a verified conclusion, particularly where money, health or a deadline is involved. You can rate any response, which feeds back into how copilots are tuned.

For consequential decisions, use it to understand the situation and prepare your questions, then confirm with a qualified professional.

Do I only get the Design & Creative copilots?

No. Pro includes every copilot in every domain, with no per-domain upsell - which is the whole point. Problems rarely stay in one lane: a design & creative question usually has a financial consequence, and that is one click away rather than another subscription.

Need a different expert?

Try it on your own case

Get help with this from the Design & Creative Copilot

Describe your situation and get specific, actionable guidance - not the generic hedging a general-purpose chatbot gives you on design & creative questions.

Free plan, no card. Pro from $4.99/week for every copilot across all 20 domains - about what one hour with any single professional costs per year.