1-5 rating scale wording: the labels for each point
The words that go on points 1 to 5, taken from 644 rating questions in a published template library, and the one job that gets a 0 to 10 instead.
Point 1 carries the least of whatever the question measures, point 5 carries the most, and the three between them run in order. For satisfaction: Very dissatisfied, Dissatisfied, Neither satisfied nor dissatisfied, Satisfied, Very satisfied. Change the pair at the ends and you change what the question measures: Very difficult to Very easy, Poor to Excellent, Not at all confident to Completely confident.
A higher number means more of the attribute, which is usually, but not always, the better answer. On Much too little to Much too much the best answer is the 3 in the middle. The end words carry the meaning: all 644 rating questions in our template library label both ends, 34 different end pairs are in use, and 637 of the 644 run 1 to 5.
SuperSurvey Question Corpus, 109 live templates and 1,422 questions, read 11 September 2026.
1-5 rating scale wording that people actually use
Twelve label sets, counted from a published library rather than invented for a blog post.
A rating question can be built two ways, and the choice decides how much wording you need. The first way is a numbered row with a word under each end, which is how all 644 rating questions in our template library are built. The second is a list of five options that each carry their own words, which the same library uses 141 times. Only the second kind needs all five labels written out, and those are the ones below.
Agreement wording dominates. It is the Likert family's home ground, compared with plain rating questions further down this page, so it takes two rows here and then gets out of the way. The other constructs are the ones you reach for when the thing you are measuring is not agreement. Where a construct has more than one published wording, the alternatives sit together; pick one and keep it for the whole survey.
| Measuring | 1 | 2 | 3 | 4 | 5 | Questions |
|---|---|---|---|---|---|---|
| Agreement, standard middle | Strongly disagree | Disagree | Neither agree nor disagree | Agree | Strongly agree | 71 |
| Agreement, short middle | Strongly disagree | Disagree | Neutral | Agree | Strongly agree | 4 |
| How often, most used | Never | Rarely | Sometimes | Often | Always | 14 |
| How often, formal register | Never | Seldom | Sometimes | Often | Always | 11 |
| How often, softer top | Never | Rarely | Sometimes | Usually | Always | 9 |
| Satisfaction | Very dissatisfied | Dissatisfied | Neither satisfied nor dissatisfied | Satisfied | Very satisfied | 8 |
| Level | Very low | Low | Neither high nor low | High | Very high | 4 |
| How well | Very poorly | Poorly | Adequately | Well | Very well | 3 |
| Amount | Much too little | Slightly too little | Just about right | Slightly too much | Much too much | 3 |
| Quality | Very poor | Poor | Acceptable | Good | Very good | 2 |
| Ease | Very hard | Hard | Neither easy nor hard | Easy | Very easy | 2 |
| How often, fixed period, published, flagged needs revision | Never | Rarely | Monthly | Weekly | Daily | 2 |
| How often, fixed period, revised, editorial | Not at all | On 1 to 3 days | On 4 to 9 days | On 10 to 19 days | On 20 or more days | editorial |
Agreement
The onboarding materials prepared me for my first week.
Standard middle, 71 questions
- Strongly disagree
- Disagree
- Neither agree nor disagree
- Agree
- Strongly agree
Short middle, 4 questions
- Strongly disagree
- Disagree
- Neutral
- Agree
- Strongly agree
How often
How often does your manager give you feedback you can act on?
Most used, 14 questions
- Never
- Rarely
- Sometimes
- Often
- Always
Formal register, 11 questions
- Never
- Seldom
- Sometimes
- Often
- Always
Softer top, 9 questions
- Never
- Rarely
- Sometimes
- Usually
- Always
Satisfaction
Overall, how satisfied are you with the support you received?
Labels, 8 questions
- Very dissatisfied
- Dissatisfied
- Neither satisfied nor dissatisfied
- Satisfied
- Very satisfied
Level
How would you describe your workload over the last month?
Labels, 4 questions
- Very low
- Low
- Neither high nor low
- High
- Very high
How well
How well does the team communicate changes that affect your work?
Labels, 3 questions
- Very poorly
- Poorly
- Adequately
- Well
- Very well
Amount
How much guidance did you get during your first month?
Labels, 3 questions
- Much too little
- Slightly too little
- Just about right
- Slightly too much
- Much too much
The best answer here is 3. A higher number means more of the attribute, not a better result.
Quality
How would you rate the quality of the training materials?
Labels, 2 questions
- Very poor
- Poor
- Acceptable
- Good
- Very good
Ease
How easy was it to find what you needed on the site?
Labels, 2 questions
- Very hard
- Hard
- Neither easy nor hard
- Easy
- Very easy
How often, over a fixed period
In the last 4 weeks, on how many days did you use the app?
Published, flagged, 2 questions
- Never
- Rarely
- Monthly
- Weekly
- Daily
Published twice in the library, kept here as an example that needs revision: there is no recall period, Rarely overlaps Monthly, and Weekly and Daily are not exclusive of each other. Use the revised set below.
Revised, editorial
- Not at all
- On 1 to 3 days
- On 4 to 9 days
- On 10 to 19 days
- On 20 or more days
Editorial replacement, not from the library: a 28-day recall period and five categories that cannot overlap.
Every set with a count is published in the library, written here low to high. The twelve published sets cover 133 of the 141 fully labelled five-option questions; the other eight are one-offs, and a set published in both directions is counted once at its combined total. The example prompts and the set marked editorial are ours, not the library's. Source: SuperSurvey Question Corpus, 109 live templates and 1,422 questions, read 11 September 2026.
Copy the 1 to 5 rating scale wording
Agreement, standard middle: Strongly disagree / Disagree / Neither agree nor disagree / Agree / Strongly agree
Agreement, short middle: Strongly disagree / Disagree / Neutral / Agree / Strongly agree
How often, most used: Never / Rarely / Sometimes / Often / Always
How often, formal register: Never / Seldom / Sometimes / Often / Always
How often, softer top: Never / Rarely / Sometimes / Usually / Always
Satisfaction: Very dissatisfied / Dissatisfied / Neither satisfied nor dissatisfied / Satisfied / Very satisfied
Level: Very low / Low / Neither high nor low / High / Very high
How well: Very poorly / Poorly / Adequately / Well / Very well
Amount: Much too little / Slightly too little / Just about right / Slightly too much / Much too much
Quality: Very poor / Poor / Acceptable / Good / Very good
Ease: Very hard / Hard / Neither easy nor hard / Easy / Very easy
How often, over a fixed period, revised, editorial (last 4 weeks): Not at all / On 1 to 3 days / On 4 to 9 days / On 10 to 19 days / On 20 or more days
Eleven published sets plus the revised fixed-period set. The flagged calendar wording is left out on purpose.
What the number means
The code behind each button counts how much of the attribute the respondent reports: 1 is the least, 5 is the most. That is not the same as worst to best. On eight of the nine constructs above the two coincide, so 5 is the answer you hope for. On the Amount set they part company: Much too little is 1, Much too much is 5, and the answer you hope for is the 3 in the middle, Just about right.
Averaging a scale that peaks in the middle rewards nobody, because a room split between too little and too much scores the same as a room that is content. Report the share at 3 instead, or split the scale into "too little" and "too much" and report both.
Direction is the other decision nobody writes down. Every numbered row in the library runs low on the left and high on the right, but 102 of the 141 labelled lists are published with the positive option first, because a list is read top down and the good news goes at the top. Both are defensible. Running both inside one survey is not, and it is the fastest way to collect answers about your layout.
Frequency wording has its own trap, which is why one published set is flagged above. Never, Rarely, Monthly, Weekly, Daily mixes two kinds of word: the first two describe an impression, the last three describe a calendar, and they overlap, because someone who uses a thing every Monday is both Weekly and, over a year, Monthly. Fix it by naming the recall period in the question, the last 4 weeks, and offering categories that cannot both be true: Not at all, On 1 to 3 days, On 4 to 9 days, On 10 to 19 days, On 20 or more days. Nobody has to guess what Rarely means, and every respondent falls into exactly one box.
Agreement wording carries 443 of the 644 rating questions
The concentration is the useful part. Five pairs cover 553 of the 644 questions, so a reader who has answered any survey in the last year has seen your end words before and does not have to stop and work them out. That is worth more than a bespoke phrase. It also has a cost: the agreement pair invites people to agree. To avoid biased answer options where the thing you want is an attribute, how clear, how easy, how fast, ask for the attribute instead and the pull disappears.
Every live template in the library was parsed question by question on 10 September 2026: 109 templates, 1,422 questions, of which 644 use the rating widget and 141 more are fully labelled five-option lists. Scale ranges were read from the rendered buttons, not from the question text. Drafts and redirects were excluded.
The limit matters. This is one survey vendor showing what it publishes as good practice, not a census of the world's surveys. It tells you what wording is conventional enough to be understood. It does not tell you that five points are better than seven.
1 to 5 or 0 to 10: which rating scale to use
Of the 644 rating questions in the library, 637 run 1 to 5 and seven run 0 to 10. There is no third option and no middle ground: nothing runs 1 to 7 or 1 to 10, and the ranges were read from the rendered buttons rather than from the question text.
What makes the seven interesting is that they are the same question. Every 0 to 10 scale in the library asks how likely somebody is to recommend something, anchored Not at all likely to Extremely likely, one per template. That is the honest answer to when a 0 to 10 earns its extra six buttons: when one number is going to be tracked on its own, over time, against an outside benchmark, and needs room to move.
| Template | Question |
|---|---|
| NPS-Survey | How likely are you to recommend us to a friend or colleague? |
| enps-survey | How likely are you to recommend this organization as a place to work? |
| Employee-Engagement | How likely are you to recommend this organization as a place to work to a friend? |
| App-Feedback-Survey | How likely are you to recommend this app to a friend or colleague? |
| Event-Feedback-Survey | How likely are you to recommend this event to a colleague? |
| LPI-interview-experience | How likely are you to recommend applying here to a friend or colleague? |
| LPD-construction-customer-satisfaction-survey | How likely are you to recommend us to someone planning similar work? |
All seven carry the same end anchors, Not at all likely to Extremely likely, and each of these templates also asks 1 to 5 questions. Source: SuperSurvey Question Corpus, 109 live templates and 1,422 questions, read 11 September 2026.
The evidence behind the range is older than any of this. Preston and Colman had 149 people rate a shop or restaurant they had recently visited, on scales that differed only in how many categories they offered, from two up to eleven.2 Two, three and four categories did poorly on reliability, validity and discriminating power, and the indices were significantly higher for scales with more categories, up to about seven; test-retest reliability tended to fall again past ten. Respondents themselves liked the 10-point scale best, then the seven and the nine. So the scale people enjoy most and the scale that measures best are not quite the same scale.
A longer scale is not a repair, either. Dell-Kuster and colleagues had the same hospital patients answer five questions twice, once on a short labelled scale and once on an 11-point end-anchored numeric one. Ceiling effects were high on both, and the longer scale did not substantially reduce them.4 Their conclusion was that the questions needed the work, not the scale. They also found three response choices too few, which is the floor worth remembering.
The working rule: 1 to 5 with words when the answer has to mean something specific, 0 to 10 when one headline number is being tracked and compared. Whether that headline number should start at 0 or at 1 is a separate decision, and a smaller one than it looks.
1-10 versus 0-10: ten choices or eleven
A 1 to 10 scale offers ten choices and has no exact middle; 0 to 10 offers eleven, with 5 in the centre. That is the whole mechanical difference. The one that matters in practice is compatibility: the Net Promoter Score is defined on 0 to 10, with 9 and 10 counted as promoters, 7 and 8 as passives and 0 to 6 as detractors, so a recommendation question that wants to be compared against any NPS benchmark has to use the eleven-button version. That is why all seven of the library's long scales run 0 to 10 and none runs 1 to 10, and why the NPS survey template ships the question, the anchors and the follow-up already wired.
Use 1 to 10 when the number is yours alone and nobody will benchmark it: an internal quality score, a rating that stakeholders already read out of ten, a form that has always used it. It is a rating question, not an NPS question, and its results cannot be compared with NPS results even if the end words match.
Overall, how would you rate the training you completed this quarter? 1 Very poor to 10 Excellent. Ten buttons, only the ends labelled, reported as the distribution and the share choosing 8 or higher.
Rating scale formats beyond 1 to 5
Numbers and words are two of five formats a rating question can take. The library ships no stars, no faces and no sliders, so the counts on this page say nothing about them; what follows is what each control does to the answer, which is a matter of mechanism rather than measurement.
| Format | What the respondent sees | Where it fits | What it costs |
|---|---|---|---|
| Numbered row, ends labelled | Five or eleven buttons in a row, with a word under each end | Anything you will repeat or track. It is what all 644 rating questions here use. | The buttons between the ends mean whatever the respondent decides they mean. |
| Fully labelled list | Five options stacked down the page, each carrying its own words | Short surveys where every point has to mean one thing. | Four lines of vertical space per question, and long labels wrap on a phone. |
| Stars | Five stars filled from the left | The moment straight after a purchase, where the reader already knows the gesture. | It reads as rating the product rather than answering your question, and there is no honest neutral. |
| Faces | Three to five faces, unhappy through happy | Kiosks, children, and audiences reading in a second language. | The middle face means fine to some people and unimpressed to others. |
| Slider | A handle dragged along a line | Desktop panels where the extra precision is genuinely used. | Whatever position the handle starts in is an answer somebody will submit untouched. |
Five ways to put the same question on screen. The 644 figure in the first row is the count of rating questions in the library; the last three rows carry no counts because the library publishes none of those formats. Source: SuperSurvey Question Corpus, 109 live templates and 1,422 questions, read 11 September 2026.
A fully labelled list is not really a rating widget at all: it is a single-select question whose options happen to be in order, which is why option order and option count matter to it in the way answer option design describes. That is also why the wording table above exists. Nothing forces you to write five labels when you use a numbered row, and almost nobody does.
Whether you label every point or only the ends is a real choice with real consequences. Weijters, Cabooter and Schillewaert tested both the number of response categories and whether those categories carried labels, and found that scale format moves response style: how often people take an end, how often they sit in the middle.3 de Rezende and de Medeiros examined the same four outcomes: reliability, use of the extreme points, use of the middle point, and which scale respondents said they preferred.1 The practical reading is not that one format wins. It is that changing format mid-programme changes your numbers while nothing has changed outside.
Likert scale vs rating scale
A Likert scale is one kind of rating scale. Rating scale is the parent term: any question that asks for a judgement on an ordered set of options, whether those options are numbers, words, stars or faces.
What makes something Likert is not the word "agree". It is that several statements share one answer set and are scored together as a single index. Likert-type items are written for satisfaction, frequency and likelihood as readily as for agreement; the answer set just has to stay identical down the block so the items can be summed. A one-off question with its own bespoke labels is a rating question and nothing more, however much it looks like the others. If you need to score multiple items together, that is the battery to build, and the wording on this page is only its first row.
So the difference is scope, not vocabulary. Every Likert item is a rating question. Most rating questions are not Likert items, because they stand alone and nothing is summed. The test is whether removing one item from the set would change the score you report.
Survey templates that use a 1 to 5 rating scale
Eighty-eight of the 109 templates in the library ask at least one rating question, and 81 of those never leave the 1 to 5 scale. The split by category is uneven in a way that matches what the questions are for: 11 of 11 technology templates use one, 27 of 33 employee templates, 17 of 20 business, 15 of 22 satisfaction, 10 of 13 marketing and 8 of 10 training.
Seven templates in 109 reach past 1 to 5
If a template is the fastest route, the survey template library is sorted by those six categories and every one of them is editable before you send it. The nine questions below are lifted from it word for word, one per end-anchor family, so they can be pasted straight into a draft.
Rating scale questions from the template library
Overall, how satisfied were you with your most recent experience with us?
1 to 5, Very dissatisfied to Very satisfied
How easy was it to get what you came for?
1 to 5, Very difficult to Very easy
How would you rate this visit overall?
1 to 5, Poor to Excellent
How confident are you that we would get it right next time?
1 to 5, Not at all confident to Completely confident
Did you get an appointment as soon as you needed one?
1 to 5, No, not at all to Yes, definitely
How likely are you to use us again?
1 to 5, Very unlikely to Very likely
A disagreement inside this team ends in a decision rather than in silence.
1 to 5, Almost never to Almost always
I could explain to a new joiner how my account list was decided.
1 to 5, Not at all true to Completely true
How likely are you to recommend us to a friend or colleague?
0 to 10, Not at all likely to Extremely likely
How to report a 1 to 5 rating scale
Lead with the distribution: the percentage who chose each of the five points, with the number of people it is based on. Then add the top-two-box, the share choosing 4 or 5. Those two together answer the question a mean cannot, which is whether a middling result is a room full of people who feel lukewarm or two camps who disagree. The steps that follow, coding, cleaning and cross-breaks, are set out where we analyse rating-scale answers as quantitative data.
Whether you are entitled to average the numbers at all has been argued for twenty years, since Jamieson set out the case in Medical Education that answers of this kind are ordinal, so the gaps between points cannot be assumed equal.7 Uher goes further and asks whether a rating scale encodes the property it claims to measure closely enough for the arithmetic to mean anything.6
You do not have to settle that argument to publish a defensible number. A share needs no assumption about spacing, so top-two-box is safe in a way a mean is not, and a room full of people understands it without a footnote. Publish the mean beside it if your organisation expects one, never instead of it, and never for a scale whose best answer sits in the middle.
Track the shape, not just the headline. A top-two-box that holds steady while the bottom box doubles is a result moving in two directions at once, and a mean shows you neither half. Put one text box straight after the rating and you get the reason as well as the number, which is what open-ended questions are for.
Once a scale is in the field, leave it alone. Relabelling the points or switching the range breaks the series, and afterwards you cannot separate the change in wording from the change in opinion. The neighbouring decisions, who to ask and how many replies you need before the number holds, are indexed in the survey learning centre.
Frequently asked questions
Is 1 the best or the worst on a 1 to 5 rating scale?
Usually the worst, never automatically. The number is the code the button contributes, and by the layout of every numbered row in this library it counts upward: 1 is the least of the thing being measured and 5 the most. Whether the most is the best depends on the labels. Very dissatisfied to Very satisfied, yes. Much too little to Much too much, no: there the good answer is Just about right, coded 3. Read the end words before you read the number, and if you invert the usual direction, expect every reader who glanced at the question to have assumed the usual thing.
What is a 1-5 scoring scale?
The same control used as a number instead of a label: each button carries a value, and the values are added or averaged into a score. It works when all five points mean the same thing to every person filling it in, which is why marking rubrics spell out what a 3 requires rather than leaving it at "average". If you are scoring work or people rather than opinions, write the criterion for every point, not only the two ends.
Where does Not applicable go on a rating scale?
Outside the row, never at one end and never hidden in the middle. Somebody who did not use the thing has no honest button, so they abandon the question or take the midpoint, and a midpoint padded with non-users looks exactly like mild approval. In this library 44 questions carry a choice outside the ordered row. "Cannot rate" does that work 30 times, "Not sure" four and "Not applicable" once; the other nine are "Prefer not to say", the same trick borrowed for a privacy problem.
What goes on point 3 of a 1 to 5 rating scale?
Whatever sits genuinely between your two ends, spelled out. "Neither agree nor disagree" and "Neutral" are offered as the same thing and are not: the first describes a position, the second invites a shrug. "Just about right" is a third case altogether, because too little and too much are both faults and the middle is the good answer. If your point 3 would really mean "I have no idea", that belongs in its own option.
How many rating scale questions in one survey?
Of the 88 templates that use them, the median asks seven rating questions and the longest asks 22. Past some length people stop reading each row and start answering the column, which Krosnick named satisficing: when answering properly costs more than it is worth, respondents settle for an answer that will do.5 Twenty-two is survivable if the statements are short and grouped by subject. Break the block with a different question type before it becomes a wall.
References
- de Rezende, N. A., and de Medeiros, D. D. (2022). How rating scales influence responses' reliability, extreme points, middle point and respondent's preferences. Journal of Business Research, 138, 266-274. doi.org/10.1016/j.jbusres.2021.09.031
- Preston, C. C., and Colman, A. M. (2000). Optimal number of response categories in rating scales: reliability, validity, discriminating power, and respondent preferences. Acta Psychologica, 104(1), 1-15. doi.org/10.1016/S0001-6918(99)00050-5
- Weijters, B., Cabooter, E., and Schillewaert, N. (2010). The effect of rating scale format on response styles: the number of response categories and response category labels. International Journal of Research in Marketing, 27(3), 236-247. doi.org/10.1016/j.ijresmar.2010.02.004
- Dell-Kuster, S., Sanjuan, E., Todorov, A., Weber, H., Heberer, M., and Rosenthal, R. (2014). Designing questionnaires: healthcare survey to compare two different response scales. BMC Medical Research Methodology, 14, 96. doi.org/10.1186/1471-2288-14-96
- Krosnick, J. A. (1991). Response strategies for coping with the cognitive demands of attitude measures in surveys. Applied Cognitive Psychology, 5(3), 213-236. doi.org/10.1002/acp.2350050305
- Uher, J. (2018). Quantitative data from rating scales: an epistemological and methodological enquiry. Frontiers in Psychology, 9, 2599. doi.org/10.3389/fpsyg.2018.02599
- Jamieson, S. (2004). Likert scales: how to (ab)use them. Medical Education, 38(12), 1217-1218. doi.org/10.1111/j.1365-2929.2004.02012.x
- SuperSurvey Question Corpus, 109 live templates and 1,422 questions, read 11 September 2026.
Correction, 10 September 2026: an earlier version of this page credited reference 1 to four authors. The paper has two. Every reference above was checked against the publisher's deposited record before publication.
Michael Hodge: Survey methodologist and editor, SuperSurvey. Bachelor of Science (Psychology), University of Wollongong, with coursework in psychometrics and research methods. Designing surveys since 2003. About the author and how these guides are reviewed
Further reading
Build a survey with a 1 to 5 scale
Every wording set on this page comes from a template you can open, edit and send.
Create a rating-scale surveyFree to send. No card.