Likert scales: 5 or 7 points and what to do with the middle
A Likert item is a statement plus a symmetric, labelled agreement scale. The three decisions that matter are how many points it has, whether the middle exists, and whether every point is labelled — and each one changes how you are allowed to analyze the results.
What it is, precisely
A Likert item presents a statement — "The checkout process was straightforward" — and asks how strongly the respondent agrees, on a scale with the same number of positive and negative options around a middle. Strictly speaking, a Likert scale is several such items about one underlying attitude, summed into a single score; the individual row is a Likert item. Most people use "Likert scale" for both, which is fine in conversation and matters in analysis.
It matters because a single item is a noisy, ordinal measurement of an attitude, while a set of items about the same thing averages out some of that noise. If you intend to combine items into a score, they need to be about the same construct. Averaging "the checkout was easy" with "delivery was on time" produces a number that describes nothing.
One more distinction worth keeping straight: a Likert item asks for agreement with a statement, while a rating scale asks for a judgement on a dimension. Where you can, prefer the dimension — "How easy or difficult was checkout?" rather than "Do you agree checkout was easy?" Agreement wording invites acquiescence, the well-documented tendency for people to agree by default when they are unsure or in a hurry.
The five decisions
How many points
Five and seven are the common choices. More points only help if respondents can genuinely discriminate at that resolution — a customer rating their broadband is not resolving seven levels of agreement, while a trained assessor rating performance might. On a phone, seven labelled options either wrap awkwardly or shrink to unreadable.
Whether the middle exists
An odd number of points gives a defined middle; an even number forces respondents to lean one way. Removing the middle does not remove ambivalence — it records it as a mild opinion, which moves the noise into your data instead of labelling it.
Which points get labels
Fully labelled scales make each point mean roughly the same thing to everyone. Endpoint-only labels are faster to read and work well with numbers, but the middle points then mean whatever the respondent decides they mean.
Which direction it runs
Negative on the left and positive on the right, or the reverse — either is defensible. What is not defensible is changing direction partway through a survey, which reliably produces wrong answers from people who had learned the pattern.
Where "does not apply" lives
If an item might not apply to some respondents, give them an explicit option outside the scale. A "not applicable" hidden inside the neutral midpoint is indistinguishable from genuine ambivalence once it is in your results, and the two mean opposite things.
The defaults we would pick
None of these is a law. They are the choices that go wrong least often, and consistency matters more than any single one of them.
| Decision | The trade-off | A sensible default |
|---|---|---|
| 5 points or 7 | 7 gives finer resolution to respondents who can use it; 5 is easier to read and fits a phone | 5 for general audiences, 7 for expert or repeat raters |
| Neutral midpoint | Keeping it invites satisficers to park; removing it forces ambivalent people to pick a side at random | Keep it, and label it "Neither agree nor disagree" |
| Label every point | Full labels standardise meaning; endpoint labels are cleaner and read faster | Label every point on agreement scales, endpoints only on 0–10 numeric scales |
| Scale direction | Either direction works; only inconsistency does damage | Negative on the left, unchanged for the whole survey |
| Not applicable | On-scale is tidier; off-scale keeps the data interpretable | A separate option outside the scale, placed last |
| Changing scale between waves | A better scale next quarter breaks the comparison with last quarter | Keep the old scale unless it was actively broken |
The midpoint argument, settled as far as it can be
The case against the midpoint is that it is a parking spot. Someone who has not thought about the question can select the middle and move on, so removing it forces a real judgement and increases the spread of your data.
The case for it is stronger: ambivalence and mild opinion are genuinely different states, and a forced choice makes them look the same. If you take the middle away, the person with no real view still has to answer, and they answer somewhere — you have not gained information, you have added noise you can no longer identify.
The practical resolution is to keep the midpoint and be precise about what it says. "Neither agree nor disagree" describes a balanced view. "Neutral" is vaguer and gets used as a shrug. "No opinion" and "Don't know" are different again, and both belong off the scale next to "not applicable" if your item might not apply. Distinguishing balanced view, no view, and no basis for a view costs you one extra option and saves you from reporting three different things as one.
Analysing it without kidding yourself
Likert data is ordinal. You know that "strongly agree" is more than "agree," but not that the gap between them is the same size as the gap between "disagree" and "strongly disagree." Taking a mean assumes those gaps are equal, which is an assumption, not a fact.
In practice most teams report means anyway, and for tracking the same question over time that is defensible — the assumption is wrong in the same way every quarter, so the direction of change is still informative. What is not defensible is showing the mean on its own. A mean of 3 out of 5 can be everyone choosing the middle or half the sample at each extreme, and those two situations call for completely different responses.
The reporting summary that survives scrutiny is top-two-box: the percentage of respondents choosing the two most positive points. It does not assume equal spacing, it is easy for a non-specialist to interpret, and it is stable enough to track. Pick one summary, show the full distribution beside it, and use the same summary every time.
Sample size deserves the same caution here as anywhere else. A five-point item from 30 respondents has enough sampling variation that a shift of a tenth of a point means nothing. Report the response count next to the score, always.
Likert scale examples by number of points
Every common size, with the agreement labels that go with it and the honest reason to pick it. The labels are the standard ones; you can copy them as they stand.
| Points | Labels (agreement) | When it fits |
|---|---|---|
| 2-point | Disagree · Agree | A yes/no in disguise. Use when you need a decision, not a degree — screening questions, eligibility. |
| 3-point | Disagree · Neither · Agree | Very fast surveys, children, or low-literacy audiences. Too coarse to track change over time. |
| 4-point | Strongly disagree · Disagree · Agree · Strongly agree | A forced choice with no midpoint. Use only when ambivalence genuinely isn't a valid answer — and expect some random leaning. |
| 5-point | Strongly disagree · Disagree · Neither agree nor disagree · Agree · Strongly agree | The default. Readable on a phone, every point labelled, a defined middle. If in doubt, this one. |
| 6-point | Strongly disagree · Disagree · Slightly disagree · Slightly agree · Agree · Strongly agree | Forced choice with finer grain. Popular in employee research where 'neutral' was being used as a hiding place. |
| 7-point | Strongly disagree · Disagree · Somewhat disagree · Neither · Somewhat agree · Agree · Strongly agree | Expert or repeat raters who can genuinely tell 'somewhat' from 'agree'. Slightly better statistical properties than 5. |
| 9-point | Endpoints labelled, numbers between | Rare outside academic instruments. Label the ends and the middle; nobody can name nine degrees of agreement. |
| 10-point | 1 = strongly disagree … 10 = strongly agree | A numeric rating, not strictly a Likert scale. No midpoint (5.5), which people find awkward for agreement — better for satisfaction or likelihood. |
| 11-point (0–10) | 0 = not at all likely … 10 = extremely likely | The NPS scale. Endpoints only, a true middle at 5. Use it when you will report the result as a score rather than as a distribution. |
Label sets for the other things you measure
Agreement is only one dimension. These five-point sets cover the rest; drop the middle for a four-point version and add 'somewhat' on each side for seven.
Satisfaction
Very dissatisfied · Dissatisfied · Neither satisfied nor dissatisfied · Satisfied · Very satisfied. This is the CSAT scale; the top two boxes are what count as satisfied.
Frequency
Never · Rarely · Sometimes · Often · Always. Define the words if it matters — 'often' means once a week to some people and every day to others. A concrete version (Never · Less than monthly · Monthly · Weekly · Daily) removes the guessing.
Likelihood
Very unlikely · Unlikely · Neither likely nor unlikely · Likely · Very likely. For a numeric alternative use 0–10 from 'not at all likely' to 'extremely likely', the NPS wording.
Importance
Not at all important · Slightly important · Moderately important · Very important · Extremely important. Importance scales skew high — almost everything is 'very important' — so rank questions often work better.
Quality and extent
Poor · Fair · Good · Very good · Excellent. Or for extent: Not at all · A little · Somewhat · Quite a bit · Very much. Three-point variants (Low · Medium · High) are fine for quick internal ratings.
Ease (Customer Effort)
Very difficult · Difficult · Neither · Easy · Very easy. Ask it about one specific task. Reported as the share choosing easy or very easy, or as a mean on a 1–7 version.
Scoring a Likert scale
Give each point a number, low to high — 1 to 5 on a five-point scale, with the most negative label as 1. If a statement is worded negatively ("the checkout was confusing"), reverse it before you do anything else, so that 5 always means the good end. Skipping this step is the most common reason a summed score comes out meaningless.
For one item, the summaries worth reporting are the distribution (how many chose each point), the top-two-box percentage, and, if your team insists, the mean. Weighted mean = Σ(point × count) ÷ responses. On a four-point scale the neutral line is 2.5, on five it is 3, on seven it is 4 — a mean just above that line leans positive and just below leans negative.
For a set of items about the same attitude, sum or average the item scores per respondent after reverse-coding, then summarise that total the same way. Do not combine items about different things; a score built from "delivery was on time" and "the app is easy to use" describes neither.
The Likert scale calculator does the arithmetic from your counts per point — mean, median, mode, standard deviation and top-two-box — and shows the ordinal and interval readings side by side so you can see when they disagree.
Building one in Zunoform
For a single item with numbers, use a rating question. It runs from 5 to 10 points, optionally starting at zero, with labels on the two ends — which makes it a good fit for numeric scales and a poor fit if you need each point named.
When every point needs a word, use a choice question with the five or seven labels as options. You lose the compact scale look and gain unambiguous meaning, which is usually the better trade on an agreement scale.
When several statements share one scale, use a matrix question: each row is a statement, and the columns carry the scale labels, which you set yourself. Keep matrices short. A grid of fifteen statements is where people abandon a survey on a phone — split by topic and put the most important block first.
All three question types are on the free plan, along with 500 responses a month and conditional logic, so an entire attitude survey can be built and run without paying anything.
Questions people ask
Should a Likert scale have a neutral option?
Usually yes. Removing it does not remove ambivalence, it just relabels ambivalent respondents as mildly positive or negative and makes that noise invisible. Keep the midpoint, label it "Neither agree nor disagree," and add a separate "not applicable" option if the item might not apply.
Is a 5-point or 7-point Likert scale better?
For general audiences on phones, 5. For expert or trained raters who can genuinely distinguish finer degrees, 7. If you are continuing an existing survey, keep whatever the previous wave used — a changed scale breaks your comparison more than a slightly suboptimal scale hurts your data.
Can I calculate an average from Likert responses?
You can, and most teams do, but it assumes the gaps between points are equal, which is not guaranteed. It is reasonable for tracking the same question over time, provided you show the distribution alongside the mean. Top-two-box percentages avoid the assumption entirely.
What's the difference between a Likert scale and a rating scale?
A Likert item asks how strongly you agree with a statement. A rating scale asks you to judge something directly on a dimension, like easy to difficult. Where both would work, the rating version is usually better, because agreement questions invite people to agree by default.
How many statements should a matrix question have?
Enough to fit on one phone screen without side-scrolling — roughly five to eight rows. Longer grids are a common abandonment point, and splitting them by topic costs you nothing but a page break.
What are the five options on a standard Likert scale?
Strongly disagree, Disagree, Neither agree nor disagree, Agree, Strongly agree. Some instruments write the middle as "Neutral" or "Undecided"; "Neither agree nor disagree" is the most precise, because it describes a balanced view rather than an absence of one.
Is a 10-point or 0–10 scale a Likert scale?
Strictly, no — a Likert item has labelled, symmetric agreement options. A 0–10 numeric scale is a rating scale, and it is the right tool when you plan to report a score (NPS, likelihood, satisfaction). It has no labelled midpoint, which is why it works less well for agreement questions.
How do you score a 4-point Likert scale?
Number the points 1 to 4 with strongly disagree as 1, reverse-code any negatively worded statements, and take the weighted mean or the top-two-box share (those choosing 3 or 4). The neutral line is 2.5, not 3, because there is no middle point.
Should the scale run from strongly disagree to strongly agree, or the reverse?
Either. Negative on the left is the more common convention and reads naturally left-to-right for most audiences. What matters is never changing direction within a survey — a flipped scale halfway through reliably produces wrong answers from people who had learned the pattern.
Keep reading
Better forms.
Better data.
Build your first form in under 60 seconds. Free forever for personal use, no credit card required.