HomeToolsProduct managementCSAT & CES calculator
CSAT & CES calculator with period comparison
Enter the distribution of satisfaction ratings (CSAT, 1–5) or effort ratings (CES, 1–7) and the calculator returns the share of satisfied customers, the mean rating and their confidence intervals. Add the previous period and it tests whether the metrics really changed or it is just noise.
—
Formulas. CSAT = (4s + 5s) / all responses; Wilson interval. Period difference: two-proportion z-test; difference in means: Welch’s t-test.
—
Formulas. CES = mean rating on a 1–7 scale ± t · s / √n; “easy” = share of 5–7. Period difference: Welch’s t-test for means and a z-test for shares.
Everything is calculated in your browser — nothing you enter is sent anywhere.
How to use it
- Pick a metricCSAT is satisfaction with a specific interaction on a 1–5 scale. CES is how easy it was to get a task done, on a 1–7 agreement scale.
- Enter the distributionHow many times each rating was given — from your survey tool’s report. Or paste a list of ratings and let the calculator count them.
- Add the previous periodThe same question last month, before a release or in a control group. The periods must be independent samples.
- Read the verdictThe calculator compares both the satisfied share and the mean rating, and shows intervals for the difference and p-values. If an interval includes zero, the change is not proven.
How to calculate CSAT: top-2-box and mean rating
CSAT (Customer Satisfaction Score) measures satisfaction with a specific moment: a support conversation, a purchase, finishing setup. The most common version is the share of satisfied customers — ratings of 4 and 5 on a five-point scale, known as “top-2-box”:
The mean rating (“4.27 out of 5”) is useful too, but very different distributions can share the same mean. That is why the calculator shows both numbers and the full distribution. The share uses a Wilson interval: at shares around 80–90% the textbook “share ± margin” formula produces noticeably skewed bounds.
If you need an overall loyalty index rather than a rating of a single interaction, calculate NPS with the NPS calculator, which also gives the margin of error and a period comparison.
How to calculate Customer Effort Score
CES (Customer Effort Score) shows how easy it is for a customer to get something done. The metric followed Dixon, Freeman and Toman’s 2010 Harvard Business Review article, which argued that loyalty depends more on reducing effort than on trying to delight. The current version asks for agreement with “The company made it easy for me to handle my issue” on a 1–7 scale:
Beyond the mean, look at the share of 1–3 ratings: these people found it hard, and their journeys are the first to investigate. CES is especially useful after onboarding, data import or checkout — wherever the interface gets in the way most.
Did CSAT or CES change significantly?
When CSAT goes from 78% to 85%, the first question is whether it is chance. The calculator tests two hypotheses independently.
Difference in shares: z-test
Difference in means: Welch’s t-test
Welch’s test does not assume equal variances and works with samples of different sizes. Ratings are an ordinal scale, so strictly speaking a mean is a convention; with hundreds of responses the t-test is reliable, but a 0.05-point difference is not worth interpreting even at p < 0.05.
Limitations
- The periods must be independent: different respondents or different conversations. If the same people answered twice, you need a paired test.
- The test covers random error only. If the new survey was shown in a different place or to a different audience, the difference may come from the method rather than the product.
- Do not re-check significance every day until it “works”: repeated peeking guarantees false positives. The survey sample size calculator tells you how many responses to plan for.
Sources
- Dixon M., Freeman K., Toman N. Stop Trying to Delight Your Customers. Harvard Business Review, July–August 2010 — Customer Effort Score.
- Dixon M., Toman N., DeLisi R. The Effortless Experience. Portfolio/Penguin, 2013 — CES on a 1–7 scale.
- Welch B. L. The Generalization of “Student’s” Problem when Several Different Population Variances are Involved. Biometrika, 34(1/2), 1947 — t-test for unequal variances.
- Wilson E. B. Probable Inference, the Law of Succession, and Statistical Inference. Journal of the American Statistical Association, 22(158), 1927.
- ISO 10004:2018. Quality management — Customer satisfaction — Guidelines for monitoring and measuring.
FAQ
How do you calculate CSAT?
Divide the number of 4 and 5 ratings by all responses on a 1–5 scale and multiply by 100%. For example, 213 satisfied out of 250 responses is a CSAT of 85.2%.
What is top-2-box?
The share of responses in the two highest options of a scale: for CSAT on a 1–5 scale, the 4s and 5s. It is easier to interpret than a mean and more robust to different distributions.
How do you calculate CES?
In the 1–7 version, CES is the mean agreement with “It was easy to get my issue resolved”. It also helps to track the share of 5–7 ratings (“easy”) and 1–3 ratings (“difficult”).
How do I know if CSAT changed significantly?
Compare the two periods’ shares with a z-test: if the p-value is below 0.05 (at 95% confidence) and the interval for the difference excludes zero, the change is significant. The calculator does this automatically.
Which test should I use for a difference in mean ratings?
Welch’s t-test: it does not assume equal variances and suits samples of different sizes. For ordinal ratings with few responses, check the result against the distribution as well.
How are CSAT, CES and NPS different?
CSAT is satisfaction with a specific interaction, CES is how easy a task was, NPS is willingness to recommend the product overall. CSAT and CES are collected right after an event, NPS less often, typically quarterly.
How many responses do I need to compare periods?
It depends on the difference you want to detect. With 80% power at a 5% significance level, detecting a rise in CSAT from 80% to 85% needs about 900 responses per period; from 80% to 90%, about 200.
More in Product management
- RICE & ICE prioritization calculator
Which backlog items should go first by RICE or ICE score?
- Eisenhower matrix
What do I do now, what gets a date, what gets dropped — and do the first ones even fit this week?
- SWOT analysis with a TOWS matrix
What are the strengths, weaknesses, opportunities and threats — and which strategies follow from crossing them?
- Kano model survey analyzer
Which features are must-haves, performance features or delighters for users?
- Activation revenue calculator
How much revenue does a lift in activation rate bring?
- Feature adoption calculator
How many active users actually adopt a feature, and how fast?
- Time to value calculator
How long do new users take to reach the first value, and where do they drop off?
- Retention calculator
How many users stay after day 1, 7 and 30, and what does the curve say?
- A/B test calculator
Is the difference between A and B significant, and how many users does the test need?
- NPS, CSAT and CES calculator
What is the Net Promoter Score of these answers, with a confidence interval?
- SUS calculator
What is the System Usability Scale score, and is the usability good?
- NASA-TLX calculator
How heavy did this task feel, and which kind of demand made it heavy?