Skip to content
Web Development27 September 2026 · 11 min read

The Popular Design Concept Is Often the Wrong One

Ninety percent of voters chose the concept we did not ship. Why a design poll measures the wrong thing accurately, and six criteria that should outrank it when you pick a direction.

The Popular Design Concept Is Often the Wrong One

Ninety percent of voters preferred the concept we did not ship. That is not an argument against asking people what they think. It is an argument for knowing, before you ask, what their answer is allowed to decide.

The question of how to choose between website design concepts usually gets answered badly in one of two directions. One habit treats audience preference as noise for a designer's taste to override. The other treats a vote as a verdict and hands a branding decision to whoever happened to be scrolling that afternoon. Both skip the same step: naming what evidence you are collecting and what that evidence can settle.

A design poll is better at measuring immediate drama than long-term brand fit

Two small concrete tiles standing on a bone-white sheet, the tall chipped one throwing a long hard shadow while the low broad one sits in soft light with almost none

A design poll is better at measuring immediate drama than long-term brand fit. Voters look at two frames for a few seconds each and pick the one that produces a reaction. Brand fit is a different property entirely. It asks whether a returning customer still recognises the company, whether the identity the business already owns survives contact with the new layout, and whether someone can put the site in front of a cautious buyer without a preamble.

None of that is visible in a feed. The voter has no relationship with the brand, no memory of the previous site, and no stake in what the company still has to look like in three years.

So the votes are not wrong. They are answers to a narrower question than the one being decided, and the narrower question happens to be the one where drama wins. Strong contrast and unexpected composition both read as quality in a four-second glance. That glance is real, and it matters on a first visit. It is simply not the same measurement as brand fit, and treating the two as interchangeable is where the mistake enters.

What does a design poll actually measure?

A chrome dial gauge with a completely blank unmarked face, its probe resting against a folded bone-white card on a concrete ledge

A poll measures attitude: how a design is perceived at a glance, not whether it helps a buyer trust the company.

Nielsen Norman Group draws the line the same way, splitting visual-design research into preference testing and behavioural testing and reserving preference work for brand alignment and first impressions. The same article names a failure mode worth internalising: the query effect, where the two options are not different enough for a non-designer to distinguish, so people manufacture a preference rather than report none. Ask a yes-or-no question and you will get a yes or a no whether or not one exists.

There is a second distortion underneath it. The classic 1995 study by Kurosu and Kashimura, covering 252 participants across 26 ATM interface variations and summarised in NN/g's write-up of the aesthetic-usability effect, found that beauty tracked perceived ease of use more closely than actual ease of use. Attractive work hides problems. That is useful in production and actively unhelpful in evaluation.

Was the sample too small to trust?

Seen from above, a dense cluster of concrete pebbles on one side of a bone-white sheet and two isolated pebbles on the other, one of them circled in cyan chalk

No: a 90 percent share of twenty-one votes is about nineteen to two, and the 95 percent Wilson interval runs from roughly 71 to 97 percent. Sauro's guidance on testing preference data with the one-sample binomial is the method; run it and the result holds.

Dismissing the poll as underpowered would have been the comfortable move. It also would have been wrong.

Who answered matters more than how many

The defect in a public design poll is construct validity. Wrong population, wrong question, right arithmetic.

A poll posted to a professional network reaches designers and developers who look at interfaces for a living. They are a fine audience for craft. They are not the relocation manager or the investor comparing three agencies before sending one enquiry. When the voting audience and the buying audience differ, a clean statistical result on the wrong population tells you what your peers admire, which is worth knowing and is not the thing you were trying to decide.

Here is how that played out on Elpida Solutions.

For Elpida Solutions, the more dramatic concept won 90 percent of 21 public votes. I still chose the quieter direction because it matched the existing dove-and-house identity and the trust the agency needed to communicate. The poll was useful evidence, but treating it as the decision would have produced the wrong site.

The Elpida Solutions real-estate website case study records what the quieter direction had to carry: company, service, contact and legal pages, a contact flow, and a five-language architecture generating Serbian, English, German, Greek and Russian routes from one shared structure. Five locales is the detail that settles a lot of visual arguments on its own. A layout that depends on a particular line length or a tight headline rhythm has to survive the same headline in five languages, and German or Serbian will not cooperate with an English-tuned grid.

Six criteria that outrank the vote

Once a poll stops being the verdict, it needs something to be evidence for. These are the criteria I weigh, and the honest version of this table admits that the louder concept wins some of them.

CriterionThe question it answersFavours the dramatic conceptFavours the quieter concept
Brand continuityDoes a returning customer still recognise the company?Rarely, unless the brand is being deliberately resetUsually, because it inherits existing marks and palette
Customer trustDoes the intended buyer read this as credible?In creative and consumer categoriesIn legal, financial, medical and property categories
Content legibilityCan long copy and service detail sit here comfortably?When the page is short and visualWhen the page carries real explanatory text
Motion toleranceCan effects be reduced without the design collapsing?Seldom, since motion is often the conceptUsually, because motion is decoration rather than structure
Production costWhat does it cost to build and to extend?Higher, especially for bespoke interactionLower, and cheaper to hand over
MaintainabilityCan the client's next contractor keep it consistent?Only with a documented systemYes, with conventional components

Two notes on reading it. Brand continuity and maintainability both improve when the concept is expressed as reusable components with named tokens rather than as one beautiful page; this is the practical argument for a real design system rather than a style guide, and it applies at four pages as much as at forty. Content legibility is the criterion clients underrate most, because a concept is presented with placeholder copy and lives with the real thing — the same trap that shows up when a marketing site has to convert rather than just impress.

Novelty has a ceiling, and it is measurable

Preference rises with novelty and then falls, so the most unusual concept in a set is frequently past the point where unfamiliarity starts costing more than it earns.

Hung and Chen put numbers on this in the International Journal of Design. In their 2012 study of novelty and aesthetic preference, 60 participants rated 88 chair designs and the relationship came out as an inverted U: moderately novel chairs were rated most beautiful, while both the most conventional and the most radical scored lower. It is the empirical shape of Raymond Loewy's older rule that a design should be as advanced as possible while staying acceptable.

That gives the poll a legitimate job. A vote is decent evidence about where a concept sits on the novelty axis. It is poor evidence about where the ceiling is for a specific audience, because the ceiling moves with category. A property agency's buyers sit closer to the typical end than a studio's do.

Motion is no longer only a question of taste

For any service offered into the EU, a concept built on heavy motion now carries a compliance obligation as well as an aesthetic one.

Article 31 of the European Accessibility Act set 28 June 2025 as the date from which member states apply its measures, and the European Commission's scope list includes e-commerce among the covered services. The web part of meeting it is unremarkable engineering. WCAG's Pause, Stop, Hide criterion is Level A and requires a mechanism to pause, stop or hide any moving content that starts automatically, runs longer than five seconds and sits alongside other content; honouring prefers-reduced-motion is how that usually gets built. The design consequence is what interests me here. If a concept's identity is the scroll-driven reveal, then the reduced-motion version is a different and worse design that nobody reviewed, and it is what a real share of visitors will see.

Ask the question during selection instead of after. What does this concept look like with motion disabled? If the answer is "flat and unfinished", the concept has a dependency rather than a feature. Tooling choice interacts with this too, which is part of why Webflow and Framer pull in different directions on motion and control.

What it costs when the popular direction wins anyway

Tropicana is the canonical case. In January 2009 the brand replaced its long-running straw-in-orange packaging with a cleaner, more contemporary design, and sales fell about 20 percent in two months — roughly 30 million dollars before the old packaging was reinstated in February.

The replacement was not ugly. It was better looking by most conventional measures, and it would very likely have won a poll against the packaging it replaced. Let me back up, because "better looking" is carrying too much weight in that sentence. It was cleaner. Cleanliness turned out to be a different property from findability, and findability was the one doing the work. What the new design lost was recognition: shoppers scanning a chiller could no longer find the brand, and some assumed they were looking at a house label.

Websites fail more quietly. There is no shelf, no sales figure moving inside eight weeks, and often no second measurement at all. The equivalent damage shows up as enquiries that do not arrive and a client who cannot say why the new site feels less like their company.

How do you present two directions without hiding the trade-offs?

Show both concepts against the same criteria, state your recommendation with its reasoning, and make the decision a choice between named trade-offs.

  1. Send the criteria before the concepts. Agree what the decision is about (continuity, trust, legibility, motion, cost, maintenance) while nothing is on screen to react to.
  2. Present the recommendation first, then the alternative. Presentation order biases preference, so being explicit about your own position is more honest than pretending the running order is neutral.
  3. Show each concept with real content. Real headlines in every language the site ships, a real service description, the actual legal text. Placeholder copy is how a concept passes review and then fails in build.
  4. Show the reduced-motion state of each. One screenshot with animation disabled, next to the full version.
  5. Give each concept a cost and a maintenance note. One sentence each, in plain numbers or plain hours.
  6. Record the decision and who made it. A short written rationale in the repository or the project doc, dated. It costs ten minutes and settles the question that arrives four months later about why the site looks like this.
  7. Include the poll as a labelled input. Put the result in the deck with its population named: twenty-one peers, not twenty-one buyers.

That last item is what turns an uncomfortable number into a useful one. You are not hiding the vote. You are telling the client exactly what it measured, which is the only way a client can weigh it properly.

Your turn to run the test: take the concept you are about to recommend, disable motion, replace every placeholder line with the client's real copy in their longest language, and look at what is left. Would it still win the poll? Should it have to?

Free resource

Free SaaS MVP Scope Template

A Notion document with the full feature checklist, MVP vs. nice-to-have table, pre-build questions, and cost signals — so you walk into any developer call knowing exactly what to ask for.

Get the template →
DL

Dusko Licanin

Full-Stack Developer · Banja Luka, Bosnia

Full-stack developer shipping SaaS MVPs, web apps, and mobile apps using AI-augmented workflows — without agency coordination overhead. Live portfolio: BookBed, Callidus, Pizzeria Bestek.

Frequently Asked Questions

Should a design poll decide which website concept wins?

No: a poll is evidence about first impressions, while the decision also needs evidence about brand continuity, content legibility, motion tolerance and long-term maintenance. Treat the result as one labelled input rather than a verdict, and record which population voted. A public poll usually reaches peers rather than buyers, so a clean result can be a precise measurement of an audience that will never purchase anything. The useful move is to publish the vote inside your recommendation with its population named, then show why the criteria point where they point. Clients accept a contrary recommendation far more easily when the uncomfortable number is on the slide rather than missing from it.

How do you test website design concepts properly?

Test them against tasks and content, not against each other in isolation, and keep attitudinal questions separate from behavioural ones. Nielsen Norman Group splits this into preference testing and behavioural testing, and recommends running behavioural tasks first so aesthetic opinions do not colour what people report about usability. In practice that means three things worth doing before any vote: load each concept with real copy in the longest language the site ships, view each with motion disabled, and ask a small number of actual target customers to find a specific piece of information. A concept that survives all three is the one to recommend.

What is brand fit, and how does it differ from a popular design?

Brand fit is whether a design keeps the identity a business already owns legible to the people who already know it, which is a different question from whether a stranger finds it attractive. Popularity is measured in a glance; fit is measured across returning visits, print collateral, sales conversations and whatever the company looked like last year. The clearest cautionary case is Tropicana's 2009 packaging: a cleaner, more contemporary design that cost roughly 20 percent of sales in two months because shoppers could no longer recognise the brand on the shelf. Attractive and recognisable are not the same property, and only one of them is what a returning customer uses.

How do you agree a visual direction with a client without endless rounds?

Agree the decision criteria in writing before any concept is shown, then present a recommendation against those criteria rather than a gallery of options. Rounds multiply when the conversation has no shared standard, because every review becomes a fresh referendum on taste. Send a short list first: brand continuity, customer trust, content legibility, motion tolerance, production cost, maintainability. Present two directions at most, state which one you recommend and why, and give each a cost and maintenance note. Then write the decision and its date into the repository or project document. That record is what prevents the same argument reopening four months later when someone new joins the client's team.

Does heavy animation in a concept create legal risk?

For services offered into the EU it can, because accessibility obligations now cover moving content rather than leaving it to taste. WCAG's Pause, Stop, Hide criterion sits at Level A and requires a mechanism to pause, stop or hide anything that moves automatically, runs longer than five seconds and appears alongside other content, and the European Accessibility Act's measures have applied since 28 June 2025. The practical consequence during concept selection is simple: if a concept collapses when animation is disabled, its reduced-motion state is an unreviewed second design that some visitors will get by default. Review that state before choosing, not after the build.