NPS, SUS, and CSAT all produce numbers, but they answer different questions. NPS concerns the overall relationship and willingness to recommend, SUS evaluates perceived usability of a system, and CSAT typically measures one interaction. Ranking the three together creates a complete-looking dashboard and more confused decisions.
NPS, SUS, and CSAT all produce numbers, but they answer different questions. NPS concerns the overall relationship and willingness to recommend, SUS evaluates perceived usability of a system, and CSAT typically measures one interaction. Ranking the three together creates a complete-looking dashboard and more confused decisions.
01 Start With a Comparison
| Metric | Primary Construct | Typical Timing | Output | Better Question |
|---|---|---|---|---|
| NPS | Overall relationship and willingness to recommend | After a period of use or in a relationship survey | -100 to 100 | Will users recommend us, and is the relationship improving? |
| SUS | Perceived usability of a system or product | After representative tasks | 0 to 100 | Does the system feel learnable, consistent, and confidence-building? |
| CSAT | Satisfaction with a service, feature, or touchpoint | After support, payment, delivery, or another interaction | Percentage or mean | Was the user satisfied with what just happened? |
The three can be used together, but cannot replace one another or be compared directly merely because they look like percentages or 100-point scores.
02 NPS Measures a Relationship, Not a Specific Interface
NPS normally asks users, on a 0–10 scale, how likely they are to recommend a brand, product, or service. Scores 9–10 are promoters, 7–8 passives, and 0–6 detractors.
The calculation is:
The range is therefore -100 to 100, not “72% satisfied.” NPS supports longitudinal tracking, segmentation, and relationship changes, but does not explain why a button is difficult. Price, brand reputation, service outcomes, and many other factors affect it.
Pair NPS with an open question—“What is the main reason for your score?”—and segment answers by customer type, tenure, plan, and churn status.

03 SUS Measures System Usability, Not an Exam Percentage
SUS contains ten statements rated on a five-point agreement scale, alternating positive and negative phrasing. A calculation transforms answers to a 0–100 score.
The score is not “the percentage answered correctly” and should not be described as “80% satisfied.” It is a standardized scale whose fixed questions and short administration support version comparisons and longitudinal tracking after representative tasks.
SUS indicates whether overall perceived usability improved but has limited diagnostic power. To locate problems, combine it with task completion, errors, observation, and interviews.
Three cautions:
- Do not change wording or order casually, or comparability will decline;
- Let users operate the product before answering instead of rating an imagined experience;
- Small samples fluctuate, so show sample size and uncertainty rather than treating a two- or three-point difference as certain.
04 CSAT Measures What Just Happened
CSAT usually asks, after one interaction, “How satisfied are you with this service?” A common implementation uses a 1–5 scale and counts 4 and 5 as satisfied:
Some teams use a mean, a 1–7 scale, or another definition. The choice matters less than keeping the scale, calculation, and trigger stable so months and channels remain comparable.
CSAT fits support, order, delivery, and feature-completion touchpoints. Its weakness is response bias: very happy and very unhappy users may answer more often, and satisfaction with one interaction does not equal long-term loyalty.

05 Choose the Metric According to the Decision
Assessing Customer-Relationship Stability
Prioritize NPS alongside renewal, retention, expansion, and churn reasons. Do not chase a higher score alone; identify which promoters and detractors changed and why.
Comparing Whether Two Versions Are Easier to Use
Prioritize SUS after users complete the same or comparable tasks. Also record success, time, errors, and assistance; otherwise the team sees a changed score without knowing what to improve.
Evaluating One Support, Delivery, or Payment Experience
Prioritize CSAT immediately after the interaction, keep the question short, and allow users to explain why.
Measuring the Entire Product Experience
Do not choose one “overall score.” Build a set across the journey: NPS for the relationship, SUS for system usability, CSAT at key touchpoints, plus behavioral and business results.
06 Example Metric Set for a SaaS Product
Suppose enterprise software is optimizing the path from customer activation to first configuration.
| Stage | Suggested Metrics | Reason |
|---|---|---|
| First configuration completed | Task success, critical errors, SUS | Assess whether the product itself is usable |
| After implementation consulting | CSAT plus open feedback | Assess whether the service solved the problem |
| After 90 days | NPS, renewal intention, depth of product use | Assess the overall relationship, not one interaction |
| Before and after a release | SUS and behavioral metrics for the same tasks | Maintain comparability and assess redesign impact |
This set does not combine every number into one index. Each metric supports a different decision, making diagnosis easier.

07 Five Common Interpretation Mistakes
Looking Only at the Mean, Not the Distribution
The same average may represent clustered opinions or polarization between new and long-term users. Segment at least by role, tenure, region, plan, or critical scenario.
Comparing Different Products Directly
Industry, user task, and survey timing all influence scores. External benchmarks provide context but do not replace your own history and segments.
Changing the Trigger Frequently
Asking CSAT immediately after resolution and emailing one week later reach different respondents and memories. A metric definition must include when it is asked.
Treating Correlation as Causation
NPS rising with renewals does not prove the former caused the latter. Analyze product use, pricing, customer mix, and service quality.
Optimizing the Survey for the Score
Hiding low-score options, sending only to active users, or rewarding positive ratings creates prettier data, not a better experience.
08 Create a Metric Definition Card
Record for every metric:
- What question it answers and what it does not;
- Exact question wording and scale;
- Trigger event and channel;
- Formula, missing values, and valid-sample definition;
- User attributes required for segmentation;
- Behavioral and business metrics used for interpretation;
- Data owner and review frequency.
This card is more useful than another dashboard number because it tells the team when comparison is valid.
Frequently Asked Questions
Does high NPS mean the product is easy to use?
Not necessarily. Recommendation may come from brand, price, service, or outcomes. Evaluate interface and workflow with usability tests, task data, or SUS.
Can SUS use only some questions?
Do not call a reduced questionnaire standard SUS. If a short scale fits the scenario, state its name, validation basis, and formula, and do not compare it directly with full SUS.
Should CSAT use an average or satisfaction rate?
Either can work if defined and kept consistent. Reports must state the scale, satisfaction threshold, sample size, and trigger timing.
Can all three metrics become one experience score?
A weighted score is technically possible but can hide different problems. Without a defined model, stable data, and clear business purpose, retain each meaning and view them along the same journey.
09 Put Research and Design Into Product Decisions
| Service | View |
|---|---|
| UI/UX design services | View service details |
| Project inquiry | Contact JVDS |
| Design and website articles | Read more related articles |