What is Thurstone Scale?
A Thurstone scale is an attitude-measurement method that uses a set of pre-rated statements and numeric weights to produce an interval-level score for a respondent’s opinion. It relies on a panel of judges to assign values to statements so individual responses can be averaged into a single attitude score.
The Thurstone scale (often called the Thurstone method or the method of equal-appearing intervals) measures attitudes by presenting respondents with a list of statements about a topic and asking them whether they agree with each statement. Before fielding the survey, a group of expert judges rates each statement for how favorable or unfavorable it is toward the attitude in question. Each statement is assigned a numeric weight (usually the median or mean judge rating), and a respondent’s score is calculated from the weights of the statements they endorse. Because statements carry calibrated weights, the resulting scores approximate interval-level measurement rather than just ordinal ranks. The method is most useful when you need rigorously scaled attitude measures, but it requires time and a reliable judging panel to create the weighted statement set.
Usage example
A research team studying trust in local services compiled 40 statements about the council’s responsiveness. A panel of 15 judges rated each statement from strongly negative to strongly positive; the research team assigned each statement its median weight and created a Thurstone scale. Respondents checked which statements matched their views, and the researchers calculated each person’s attitude score by averaging the weights of endorsed statements.
Practical application
Why it matters: Thurstone scaling produces more precise, comparable attitude scores than simple agree/disagree items because statements are pre-calibrated by experts. Practical uses include academic research, psychometrics and any context where measured differences in intensity of opinion matter (for example, comparing community attitudes across neighbourhoods). Limitations: developing a Thurstone scale takes time, requires a representative panel of judges, and depends on well‑crafted statements—factors that can make it a poor fit for quick operational surveys. In multilingual or community engagement work, the Thurstone approach can be valuable when measuring nuanced attitudes across groups, but it requires careful translation and re-calibration of statement weights in each language to ensure equivalence. For most everyday multilingual surveys, simpler approaches (like well-tested Likert items with good translation and review workflows) are faster and easier to maintain, while Thurstone-style calibration is worth the investment when precision and defensibility are essential.
FAQ
How is a Thurstone scale different from a Likert scale?
A Likert scale asks respondents to rate their level of agreement with each statement using the same numeric scale (e.g., 1–5) and treats each item equally when scoring. A Thurstone scale first has experts assign weights to statements based on how favorable they are; respondents indicate which statements they agree with, and scores are calculated from the weighted statements. Thurstone aims for interval-level measurement through pre-calibration, while Likert is simpler and more commonly used.
Is the Thurstone method still used today?
Yes, but less often for routine surveys. Modern practice favors Likert scales and item-response theory because they are easier to implement and analyse. Thurstone scales remain useful when a high degree of measurement precision is required or when researchers want defensible, calibrated statement weights.
Can I use a Thurstone scale in a multilingual survey?
You can, but it requires extra work: statements must be translated carefully and judges who understand each language (or bilingual judges) must re-rate statements to ensure weights are equivalent across languages. Without re-calibration, weights derived in one language may not reflect how statements are perceived in another, undermining comparability.