Back to the Models tab

Metrics by Value

Not every impressive-sounding result actually helps you get hired. On a product design portfolio, some metrics make a hiring manager lean in — and some make them quietly discount the case study. This scale ranks the evidence you might put in a case study by how much it moves a hiring manager, from low value (looks nice, proves little) to high value(concrete outcomes that are hard to fake). Select any point to see what it is, how to get the number, what a strong version looks like, and why it does — or doesn't — land in an interview.

Behavior changeHigh value

Awards & recognition

Low value
What it is
Industry awards, “best of” lists, and design-community recognition for the work.
How to get it
Just count them — awards, nominations, publication features. There's nothing to instrument.
Why it impresses hiring managers — or doesn't
Awards catch a hiring manager's eye for a second, but most have learned to discount them — they signal taste and effort, not business impact, and everyone suspects the politics behind them. Worth a small mention; never the spine of a case study.
2awards won

2 industry awards + a “Site of the Day” feature.

WeakAward-winning redesign.

StrongRecognized externally — but I led the case study with the 13-pt conversion lift, not the trophy.

Press & social buzz

Low value
What it is
Media coverage, launch chatter, and social shares around a release.
How to get it
Media-monitoring / PR tools (Meltwater, Google Alerts) plus each platform's share counts.
Why it impresses hiring managers — or doesn't
A launch that got press feels exciting, but hiring managers read it as marketing, not design outcome. It says the company promoted the thing, not that your design worked. Use it as color, never as evidence.
6press mentions

Featured in 6 outlets; ~40k launch-week impressions.

WeakOur launch got a ton of press.

StrongPress drove awareness — and I tracked whether it converted to sign-ups (+8%).

Output shipped

Low value
What it is
The volume of design delivered — features, screens, flows, components.
How to get it
Count what shipped from your design files or tickets. It's an inventory, not a measurement.
Why it impresses hiring managers — or doesn't
Leading with how much you shipped reads as activity, and activity is the weakest thing a portfolio can open with. Hiring managers assume you did work; they want to know what it changed. Volume without outcome signals a designer who mistakes motion for impact.
40screens shipped

40 screens, 12 flows, a 60-component library.

WeakI designed 40 screens for the app.

StrongShipped the flow that lifted activation 9 pts — the volume was the cost, not the result.

Testimonials & praise

Low value
What it is
Positive quotes and anecdotes from users, customers, or stakeholders.
How to get it
Pull quotes from interviews, reviews, support tickets, and sales calls.
Why it impresses hiring managers — or doesn't
A warm quote adds humanity and can seal a strong story — but on its own it reads as cherry-picked, because it is. Hiring managers weight it far below a number. Pair a great quote with the metric it illustrates and it lands; alone, it's garnish.
“½ the time”— a customer

“This cut our onboarding in half.” — Customer PM

WeakUsers loved it — see the quotes.

Strong“Cut onboarding in half” — and the data backed it: setup time −47%.

Stakeholder buy-in

Low value
What it is
How confident and satisfied leadership, partners, and the team felt about the work.
How to get it
No tool — read the room: design reviews, sponsor feedback, and what gets greenlit.
Why it impresses hiring managers — or doesn't
“Leadership loved it” shows you can navigate an org, which senior roles do value — but it measures internal politics, not user or business outcomes. A hiring manager hears influence, not impact. Useful supporting detail for lead and principal stories; weak as your headline.
Phase 2greenlit

Exec sponsor greenlit phase 2 and doubled the team.

WeakLeadership loved the work.

StrongEarned exec buy-in for phase 2 — on the back of the phase-1 retention lift.

Brand lift & perception

Situational
What it is
Shifts in how people perceive the brand — trust, quality, preference — that design contributed to.
How to get it
Brand-tracking surveys (Qualtrics, YouGov) and pre/post preference studies.
Why it impresses hiring managers — or doesn't
Moving brand perception sounds impressive and is increasingly design-driven — but it's slow, fuzzy, and obviously shared with marketing and PR. A sharp interviewer will ask which part was actually yours. Credible inside a brand or 0-to-1 story; shaky as a solo claim.

Brand-trust score 6.1 → 6.8 / 10 over two quarters.

WeakThe rebrand made us look more trustworthy.

StrongTrust score +0.7 in two quarters — with the honest caveat that marketing moved in parallel.

Qualitative feedback themes

Situational
What it is
Recurring patterns in what users say — frustrations, requests, mental models.
How to get it
Affinity-map interview notes, tickets, and reviews — in Dovetail or a spreadsheet.
Why it impresses hiring managers — or doesn't
Turning messy user input into clear themes signals research rigor, which hiring managers genuinely respect — especially for mid-level and research-leaning roles. It proves how you think. Its ceiling: it shows process, not outcome, so it strengthens a case study rather than closing it.
3themes · 24 interviews

3 themes across 24 interviews; the top one hit 71% of users.

WeakResearch showed users were confused.

Strong71% of interviewees hit the same blocker — which is exactly what I redesigned around.

Usability findings

Situational
What it is
What testing reveals about whether people can actually use the design — task observations, issue severity, a SUS score.
How to get it
Moderated or unmoderated tests (UserTesting, Maze), scored with SUS (0–100).
Why it impresses hiring managers — or doesn't
Task success rates and severity ratings show you validate instead of guess, which reassures a hiring manager about your process. Strong supporting evidence — but it reads as “competent,” not “exceptional,” unless you connect it to a real-world result you shipped.

SUS 62 → 81; task success 68% → 94%.

WeakUsability testing showed improvement.

StrongSUS climbed 62 → 81 (below-average → excellent) across two test-and-fix rounds.

Satisfaction & sentiment

Situational
What it is
How users feel about the experience — NPS, CSAT, star ratings.
How to get it
In-product surveys (Delighted, Pendo) — an NPS or 1–5 CSAT prompt.
Why it impresses hiring managers — or doesn't
A satisfaction lift is easy to state and easy to grasp, so it reads clean in a case study — but hiring managers know it's attitudinal and gameable, and they've seen it inflated. Fine as one data point among several; unconvincing as your single proof.

CSAT 3.4 → 4.3 / 5; NPS +12 → +34.

WeakUsers were more satisfied.

StrongCSAT 3.4 → 4.3 — I frame it as a companion to the task-success data, not the headline.

Accessibility & compliance

Situational
What it is
How well the design works for people with disabilities, and whether it meets standards like WCAG.
How to get it
Automated audits (axe, Lighthouse) plus manual screen-reader testing; report WCAG conformance.
Why it impresses hiring managers — or doesn't
Accessibility wins are increasingly a differentiator — they signal craft, empathy, and awareness of risk, and more teams now screen for it. Framed as more users reached or legal exposure reduced, it lands well. Framed as “we passed the audit,” it reads as a checkbox.

Critical a11y issues 47 → 3; reached WCAG 2.1 AA.

WeakMade the product accessible / AA compliant.

StrongCut critical issues 47 → 3 to hit AA — and screen-reader task success rose to 90%.

Feature adoption

Situational
What it is
The share of users who took up a new capability the design introduced.
How to get it
Product analytics (Amplitude, Mixpanel) — feature-usage events over a set window.
Why it impresses hiring managers — or doesn't
“X% of users adopted it” is a concrete, believable outcome hiring managers like — it ties your design to real usage. Strong. It gets stronger the moment you show adopters stuck around or succeeded, rather than just tried it once.

42% of active users tried the new export in 2 weeks.

WeakThe new feature saw good adoption.

Strong42% adopted in two weeks — and 60% of them were still using it a month on.

Engagement & usage depth

Situational
What it is
How deeply and how often people use the product because of the design.
How to get it
Product analytics — actions per active user per week, tied to a value action.
Why it impresses hiring managers — or doesn't
An engagement lift sounds impressive and charts beautifully, which makes it a portfolio favorite — but savvy hiring managers probe it, because more usage can mean value or friction. It lands when you've tied it to a value action and can explain why more was the goal.

Value actions per active week 3.1 → 5.4.

WeakEngagement went up.

StrongValue actions per week 3.1 → 5.4 — and I can show it's genuine use, not confusion.

Behavior change

High value
What it is
A measurable shift in what people actually do — the outcome a design set out to create.
How to get it
An A/B test or pre/post with a baseline (ideally a control), tracking the target action.
Why it impresses hiring managers — or doesn't
A measured shift in what users do is the outcome hiring managers most want to see, because it's the actual job of design and it's hard to fake. Show a clear before/after and you've demonstrated cause and effect — this is what separates a portfolio that gets callbacks from one that lists screens.

Repeat-booking rate 22% → 31% vs. a holdout.

WeakThe redesign changed user behavior.

StrongRepeat bookings 22% → 31% against a control group — a clean causal read.

Task efficiency & time saved

High value
What it is
How much faster, or with how much less effort, people reach their goal.
How to get it
Time-on-task from analytics or timed usability sessions; count steps or clicks.
Why it impresses hiring managers — or doesn't
“Cut task time 40%” is concrete, easy to picture, and trivially converted to money — hiring managers love it, especially for enterprise and internal-tool work. It reads as a designer who understands efficiency as value, and it travels well between industries.

Avg. task time 90s → 54s (−40%).

WeakMade the workflow faster.

StrongCut task time 90s → 54s — roughly 2,400 hours a year saved across the team.

Error & risk reduction

High value
What it is
A fall in mistakes, failures, or unsafe actions the design helped prevent.
How to get it
Error and exception logs, failed-transaction counts, incident tracking.
Why it impresses hiring managers — or doesn't
Fewer errors or failures is a highly credible, business-legible win — and in regulated domains like healthcare or fintech it can be the whole point. Squarely a design lever, through clarity, constraints, and defaults. Among the most defensible impact claims you can make.

Payment error rate 18% → 5%.

WeakReduced user errors.

StrongPayment errors 18% → 5% — about $300k a year in recovered failed orders.

Retention & loyalty

High value
What it is
Whether the people the design served keep coming back over time.
How to get it
Cohort analysis in analytics (Amplitude, Mixpanel) or SQL — read by signup cohort.
Why it impresses hiring managers — or doesn't
A retention or churn improvement is among the most impressive things on a portfolio, because everyone knows it's hard to move and impossible to fake with a launch. It signals durable value. Show a cohort curve bending and you've made your strongest possible case.

Day-30 retention 18% → 27%, held across 3 cohorts.

WeakImproved retention.

StrongD30 retention 18% → 27% across three cohorts — hard to fake, easy to defend.

Conversion & activation

High value
What it is
Whether people cross the moments that matter — reaching first value, or committing (sign up, buy, upgrade).
How to get it
Funnel analytics (GA4, Amplitude), validated with an A/B test.
Why it impresses hiring managers — or doesn't
“Increased conversion 25%” is the line hiring managers scan for — concrete, tied to revenue, and unmistakably an outcome. One of the most persuasive results in a portfolio. Just be ready to explain the funnel context, because good interviewers test whether the win was real or borrowed.

Checkout completion 61% → 74% (+13 pts).

WeakImproved the checkout.

StrongCheckout completion 61% → 74% in an A/B test — ~$240k a year incremental.

Financial impact

High value
What it is
The design's effect in the language of the business — revenue, cost savings, ROI, lifetime value.
How to get it
Attribute revenue or cost to the change via an A/B test or pre/post, worked out with finance.
Why it impresses hiring managers — or doesn't
Impact in dollars is the most impressive evidence a portfolio can carry, because it's the language executives and hiring managers both fund. Tie your design to money credibly and you stand out instantly. The catch: strong interviewers will probe attribution, so claim only what you can defend.
+$1.2Mannual revenue

+$1.2M annual revenue; −$180k support cost.

WeakDrove significant business value.

Strong+$1.2M annual revenue in a controlled test — attribution I could walk a CFO through.