REVIEW 3 major objections 4 minor 60 references
Don't Let Your Likert Scales Grow Up To Be Visual Analog Scales: Understanding the Relationship Between Number of Response Categories and Measurement Error
T0 review · 3 major / 4 minor · reviewed 2026-08-09 · deepseek-v4-flash
Pith's one-line read If measurement error rises with the number of response options, the optimal Likert scale has 4-7 categories, and converting to a 100-point visual analog scale can sharply reduce reliability.
desk verdict A useful formal framework for a classic scale-design question, but the headline VAS warning rests on assumed error-growth slopes not estimated from data. read the letter →
The pith
A machine-rendered reading of the paper's core claim, the machinery that carries it, and where it could break.
The reading
What carries the argument
The item variable construction of the Graded Response Model (GRM) is the mechanism that carries the argument. Each item has a continuous latent response variable $\gamma_{ij} \sim N(\theta_i, \sigma_j)$, where $\sigma_j$ is the item's measurement error on the latent scale; observed category choices are obtained by thresholding $\gamma_{ij}$. This reparameterization of Samejima's GRM lets the authors define measurement error independently of the response format, then impose different linear dependencies between $\sigma_j$ and the number of categories and trace how the optimal category count shifts.
What would settle it
Estimate item-level measurement error directly by administering the same item text with 2, 5, 7, 11, and 101 response options to the same respondents, then check whether item discrimination or test-retest reliability declines as categories increase; if reliability stays flat or keeps improving through 100 options, the predicted reliability collapse and the VAS caution are falsified for that construct.
Extended reading notes
Core claim
The paper's central claim is that the optimal number of response options is governed by how item measurement error depends on category count. Using the item variable construction of the graded response model, the authors simulate responses across 2 to 100 categories. When error is independent of the number of categories, recovery of the true score improves and then plateaus after roughly 5 to 10 categories, so there is no single optimum. When error increases linearly with the number of categories, a clear optimum emerges at 7, 5, and 4 response options for small, medium, and large dependency structures, with reliability declining past that point; the standard error of a regression coefficient follows the same pattern. The paper therefore argues that a VAS, considered as a 101-point Likert scale, will typically have lower reliability than a shorter Likert version of the same item, and that a format change should trigger re-validation.
Load-bearing premise
The argument rests on the assumption that an item's measurement error grows linearly with the number of response categories at the specific rates used in the simulation; if the true relationship is flat, nonlinear, or varies by construct, the predicted optima and the warning against VAS conversion do not follow.
Editorial extensions
If this is right
- If measurement error grows with category count, reliability peaks at 7, 5, or 4 response options depending on the growth rate, then declines beyond that point.
- Converting an existing Likert item to a 100-point visual analog scale will decrease reliability when such error growth is present.
- When measurement error is independent of category count, there is no true optimum; reliability just increases and then plateaus.
- Changing the response format of a validated measure requires re-validation because the measurement error of the scale is likely to change.
- Adding more items (three instead of one) improves true-score recovery but does not move the optimal number of response categories.
Reading between the lines
- The specific optima of 7, 5, and 4 are direct consequences of the chosen linear error-growth slopes; real-world error could grow nonlinearly, so the robust qualitative prediction is that an optimum exists well below VAS-like counts.
- A direct empirical test would estimate item discrimination or test-retest reliability for the same item text under 2, 5, 7, 11, and 101 response options, which would map the true dependency structure and identify construct-specific optima.
- For ecological momentary assessment, where single-item measures dominate, the reliability penalty of VAS conversion may be largest, but momentary-state constructs might exhibit flatter error growth than trait measures, making construct type a moderator worth testing.
Signed reviews
Editorial analysis
A structured set of objections, weighed in public.
Referee Report
Summary. The paper simulates Likert- and VAS-style response data using the graded response model with an item-variable parameterization, and examines how the number of response categories affects recovery of the latent trait (Spearman correlation) and the standard error of a regression coefficient. In the independent-error condition, reliability increases and then plateaus. In the dependent-error condition, measurement error is assumed to increase linearly with the number of categories at three slopes (small, medium, large), yielding optima at 7, 5, and 4 categories, respectively. The authors conclude that converting Likert items to VAS (0-100) will drastically reduce reliability and that any such conversion requires re-validation.
Significance. If taken as a conditional simulation study, the paper is a useful proof of concept: it demonstrates that under a monotonically increasing error-category relationship, a finite optimum exists and the location of that optimum shifts with the error-growth rate. The item-variable construction of the GRM is a clean way to separate categorization error from item-level measurement error. However, the practical conclusions are not empirically supported because the linear error-growth slopes in Section 2.2 are chosen by fiat, not estimated from data, and the reported optima are direct consequences of those slopes. The paper is also transparent about this dependence in Section 5, but the abstract and conclusion overstate the generalizability. With appropriate reframing and additional simulation detail, the paper could be a worthwhile contribution to the response-format literature.
major comments (3)
- [§2.2, §3.2, §5] The optima of 7, 5, and 4 categories are direct outputs of the three linear error-dependency sequences specified in Section 2.2 (σ_j from 0.05-0.5, 0.1-1.0, and 0.2-2.0 across k=2..20). These slopes are not estimated from any empirical data, and the independent-error condition in Section 3.1 shows that no optimum arises without them. The abstract's claim that 'conversion of any Likert scale item to VAS will result in a drastic decrease in reliability' is therefore not supported by the simulations alone. The authors should reframe the contribution as a conditional demonstration, explicitly state that the optima are functions of the assumed error-growth function, and ideally include a sensitivity analysis across a range of functional forms (e.g., sublinear, concave, or stepwise) to show how the conclusion depends on the shape of σ(k).
- [§2.2, §3.2, Figures 5-6] The dependent-error simulations only vary the number of response categories from 2 to 20, yet the abstract and conclusion extrapolate the 'drastic decrease' to VAS scales with 0-100 response points. This extrapolation assumes that linear error growth continues unchanged from k=20 to k=100, which is an untested assumption. The authors should either simulate the full 2-100 range under their dependency rules, or explicitly restrict the practical claim to the range of categories actually simulated and discuss what additional evidence would be needed to extend it to VAS.
- [§2.2, all figures] The number of Monte Carlo replications is never reported, and none of the figures include error bars or confidence bands. For a study whose central claim is the location of an optimum and the existence of a decline beyond it, the reader cannot distinguish a real turning point from Monte Carlo noise. The authors should report the number of replications and add uncertainty estimates (e.g., standard errors or confidence bands across replications) to Figures 1-6, or at least report the Monte Carlo standard error for the reported optima.
minor comments (4)
- [§1.1] The text states there is a one-to-one relationship between σ_j and the GRM discrimination parameter but does not provide the analytic mapping; giving the formula would improve reproducibility and help readers interpret the error sequences.
- [§2.2] The independent-error condition sets σ_j to values from 0.1 to 1.0, but Figure 1's legend shows only 0.25, 0.50, 0.75, 1.00; clarify whether these are the full set of levels or a selected subset, and if the latter, why the subset was chosen.
- [§3.2, Figure 6] The text says the SE for large dependency 'rises again after approximately five response options,' but Figure 6 appears to show the increase beginning around four categories; please align the description with the plotted results.
- [General] The manuscript references Supplementary Materials for regression-coefficient bias figures but does not state how to access them; please include a data/code availability statement.
Circularity Check
No significant circularity: the reported optima are conditional simulation outcomes, and the paper explicitly disclaims design-dependence.
full rationale
The paper's central 'optimum' results are obtained by Monte Carlo simulation under an explicitly stated data-generating process. Section 2.2 fixes sigma_j as a linear function of the number of categories (small 0.05-0.5, medium 0.1-1.0, large 0.2-2.0). That is an input assumption, not a fitted parameter, and the reliability optima (7, 5, and 4) are legitimate derived outcomes of that model rather than a renaming of the input. The conclusion that VAS conversion would decrease reliability is explicitly conditional: the abstract says 'If measurement error increases with the number of response categories,' and Section 5 states 'these optima are highly dependent on our simulation design.' The question of whether the assumed error-growth relationship is realistic or empirically calibrated is a validity/external-evidence concern, not a circularity concern under the rubric. The only self-citation (Schmidt, 2010) is used as supporting literature for the general point that reducing response categories can improve measurement; it is not load-bearing for the simulation or the central derivation. No self-definitional equivalence, fitted-input-as-prediction, imported uniqueness, or ansatz-via-citation was found.
Assumptions & free parameters
free parameters (4)
- Error-dependency slopes (small/medium/large) =
0.025, 0.05, and 0.1 per additional category over categories 2 to 20
- Item measurement error sigma_j grid =
0.1 to 1.0 in increments of 0.1
- Threshold span and spacing =
Evenly spaced thresholds in [-2, 2]; binary threshold 0
- Predictor relation constants =
Regression slope 0.5, residual SD 0.2
assumptions (6)
- domain assumption Latent trait theta_i follows N(0,1) and item variable gamma_ij follows N(theta_i, sigma_j)
- standard math Response category is determined by thresholds on gamma_ij with probability 1, producing an ogive graded response model
- domain assumption Measurement error is independent of the number of response categories in the first simulation set
- ad hoc to paper Measurement error increases linearly with the number of response categories in the second simulation set
- domain assumption Unidimensionality and a stable trait within a measurement occasion
- domain assumption Composite or mean scoring of items is used rather than model-based scores
Cite this review
Pith. "Pith review of Don't Let Your Likert Scales Grow Up To Be Visual Analog Scales: Understanding the Relationship Between Number of Response Categories and Measurement Error." pith.science (2026). https://pith.science/paper/KPYTVCUN
@misc{pith2026250202846,
author = {Pith},
title = {Pith review of: Don't Let Your Likert Scales Grow Up To Be Visual Analog Scales: Understanding the Relationship Between Number of Response Categories and Measurement Error},
year = {2026},
howpublished = {\url{https://pith.science/paper/KPYTVCUN}},
note = {Machine review of arXiv:2502.02846}
}
read the original abstract
The use of Visual Analog Scales (VAS), which can be broadly conceptualized as items where the response scale is 0-100, has surged recently due to the convenience of digital assessments. However, there is no consensus as to whether the use of VAS scales is optimal in a measurement sense. Put differently, in the 90+ years since Likert introduced his eponymous scale, the field does not know how to determine the optimal number of response options for a given item. In the current work, we investigate the optimal number of response categories using a series of simulations. We find that when the measurement error of an item is not dependent on the number of response categories, there is no true optimum; rather, reliability increases with number of response options and then plateaus. However, under the more realistic assumption that the measurement error of an item increases with the number of response categories, we find a clear optimum that depends on the rate of that increase. If measurement error increases with the number of response categories, then conversion of any Likert scale item to VAS will result in a drastic decrease in reliability. Finally, if researchers do want to change the response scale of a validated measure, they must re-validate the new measure as the measurement error of the scale is likely to change.
Figures
Figures from the paper (3 more)
Reference graph
Works this paper leans on
-
[1]
abend_reliability_2014 APACrefauthors Abend, R. , Dan, O. , Maoz, K. , Raz, S. \ Bar-Haim, Y. APACrefauthors \ 2014 . Reliability, validity and sensitivity of a computerized visual analog scale measuring state anxiety Reliability, validity and sensitivity of a computerized visual analog scale measuring state anxiety . Journal of behavior therapy and exper...
work page 2014
-
[2]
adelson_measuring_2010 APACrefauthors Adelson, J L. \ McCoach, D B. APACrefauthors \ 2010 . Measuring the Mathematical Attitudes of Elementary Students : The Effects of a 4- Point or 5- Point Likert - Type Scale Measuring the Mathematical Attitudes of Elementary Students : The Effects of a 4- Point or 5- Point Likert - Type Scale . Educational and Psychol...
-
[3]
alan_effect_2020 APACrefauthors Alan, U. \ Atalay Kabasakal, K. APACrefauthors \ 2020 . Effect of number of response options on the psychometric properties of Likert -type scales used with children Effect of number of response options on the psychometric properties of Likert -type scales used with children . Studies in Educational Evaluation 66 100895 . A...
-
[4]
allen_single_2022 APACrefauthors Allen, M S. , Iliescu, D. \ Greiff, S. APACrefauthors \ 2022 . Single Item Measures in Psychological Science Single Item Measures in Psychological Science . European Journal of Psychological Assessment 38 1 1--5 . APACrefDOI doi:10.1027/1015-5759/a000699 APACrefDOI
-
[5]
alwin_information_1992 APACrefauthors Alwin, D F. APACrefauthors \ 1992 . Information transmission in the survey interview: Number of response categories and the reliability of attitude measurement Information transmission in the survey interview: Number of response categories and the reliability of attitude measurement . Sociological methodology 83--118
work page 1992
-
[6]
alwin_feeling_1997 APACrefauthors Alwin, D F. APACrefauthors \ 1997 . Feeling Thermometers Versus 7- Point Scales : Which are Better ? Feeling Thermometers Versus 7- Point Scales : Which are Better ? Sociological Methods & Research 25 3 318--340 . APACrefDOI doi:10.1177/0049124197025003003 APACrefDOI
-
[7]
alwin_reliability_1991 APACrefauthors Alwin, D F. \ Krosnick, J A. APACrefauthors \ 1991 . The Reliability of Survey Attitude Measurement : The Influence of Question and Respondent Attributes The Reliability of Survey Attitude Measurement : The Influence of Question and Respondent Attributes . Sociological Methods & Research 20 1 139--181 . APACrefDOI doi...
-
[8]
averbuch_assessment_2004 APACrefauthors Averbuch, M. \ Katzper, M. APACrefauthors \ 2004 . Assessment of Visual Analog versus Categorical Scale for Measurement of Osteoarthritis Pain Assessment of Visual Analog versus Categorical Scale for Measurement of Osteoarthritis Pain . The Journal of Clinical Pharmacology 44 4 368--372 . APACrefDOI doi:10.1177/0091...
Show all 60 references
-
[9]
APACrefauthors \ 1994
baddeley_magical_1994 APACrefauthors Baddeley, A. APACrefauthors \ 1994 . The magical number seven: Still magic after all these years? The magical number seven: Still magic after all these years? Psychological Review 101 2 535--356
1994
-
[10]
, Silver, W
bijur_reliability_2001 APACrefauthors Bijur, P E. , Silver, W. \ Gallagher, E J. APACrefauthors \ 2001 . Reliability of the Visual Analog Scale for Measurement of Acute Pain Reliability of the Visual Analog Scale for Measurement of Acute Pain . Academic Emergency Medicine 8 12...
2001
-
[11]
, Revilla, M
bosch_measurement_2019 APACrefauthors Bosch, O J. , Revilla, M. , DeCastellarnau, A. \ Weber, W. APACrefauthors \ 2019 . Measurement reliability, validity, and quality of slider versus radio button scales in an online probability-based panel in Norway Measurement reliability, ...
2019
-
[12]
, Eilers, P H
camarda_modelling_2008 APACrefauthors Camarda, C G. , Eilers, P H. \ Gampe, J. APACrefauthors \ 2008 . Modelling general patterns of digit preference Modelling general patterns of digit preference . Statistical Modelling 8 4 385--401 . APACrefDOI doi:10.1177/1471082X0800800404...
2008 doi
-
[13]
\ Marshall, H
champney_optimal_1939 APACrefauthors Champney, H. \ Marshall, H. APACrefauthors \ 1939 . Optimal refinement of the rating scale. Optimal refinement of the rating scale. Journal of Applied Psychology 23 3 323
1939
-
[14]
, Barratt, A L
davey_one-item_2007 APACrefauthors Davey, H M. , Barratt, A L. , Butow, P N. \ Deeks, J J. APACrefauthors \ 2007 . A one-item question with a Likert or Visual Analog Scale adequately measured current anxiety A one-item question with a Likert or Visual Analog Scale adequately m...
2007
-
[15]
Krieg, J
edward_f__krieg_biases_1999 APACrefauthors Edward F. Krieg, J. APACrefauthors \ 1999 . Biases Induced by Coarse Measurement Scales Biases Induced by Coarse Measurement Scales . Educational and Psychological Measurement . APACrefDOI doi:10.1177/00131649921970125 APACrefDOI
1999 doi
-
[16]
\ Reise, S P
embretson_item_2000 APACrefauthors Embretson, S E. \ Reise, S P. APACrefauthors \ 2000 . Item response theory for psychologists Item response theory for psychologists . Mahwah, NJ, US Lawrence Erlbaum Associates Publishers
2000
-
[17]
APACrefauthors \ 2003
ferrando_kernel_2003 APACrefauthors Ferrando, P J. APACrefauthors \ 2003 . A Kernel Density Analysis of Continuous Typical - Response Scales A Kernel Density Analysis of Continuous Typical - Response Scales . Educational and Psychological Measurement 63 5 809--824 . APACrefDOI...
2003 doi
-
[18]
, Van Schaik, P
flynn_comparison_2004 APACrefauthors Flynn, D. , Van Schaik, P. \ Van Wersch, A. APACrefauthors \ 2004 . A comparison of multi-item likert and visual analogue scales for the assessment of transactionally defined coping function1 A comparison of multi-item likert and visual ana...
2004
-
[19]
\ Reips, U D
funke_why_2012 APACrefauthors Funke, F. \ Reips, U D. APACrefauthors \ 2012 . Why Semantic Differentials in Web - Based Research Should Be Made from Visual Analogue Scales and Not from 5- Point Scales Why Semantic Differentials in Web - Based Research Should Be Made from Visua...
2012 doi
-
[20]
, Townsend, M
guyatt_comparison_1987 APACrefauthors Guyatt, G H. , Townsend, M. , Berman, L B. \ Keller, J L. APACrefauthors \ 1987 . A comparison of Likert and visual analogue scales for measuring change in function A comparison of Likert and visual analogue scales for measuring change in ...
1987
-
[21]
, Martínez, A J
haslbeck_comparing_2024 APACrefauthors Haslbeck, J. , Martínez, A J. , Roefs, A. , Fried, E I. , Lemmens, L H J M. , Groot, E. \ Edelsbrunner, P. APACrefauthors \ 2024 . Comparing Likert and Visual Analogue Scales in Ecological Momentary Assessment . Comparing Likert and Visua...
2024 doi
-
[22]
\ Arnetz, B B
hasson_validation_2005 APACrefauthors Hasson, D. \ Arnetz, B B. APACrefauthors \ 2005 . Validation and findings comparing VAS vs. Likert scales for psychosocial measurements. Validation and findings comparing VAS vs. Likert scales for psychosocial measurements. International E...
2005
-
[23]
APACrefauthors \ 1921
hayes_experimental_1921 APACrefauthors Hayes, M H. APACrefauthors \ 1921 . Experimental development of the graphic rating method Experimental development of the graphic rating method . Psychological Bulletin 18 98--99
1921
-
[24]
APACrefauthors \ 2015
hilbert_influence_2015 APACrefauthors Hilbert, S. APACrefauthors \ 2015 . The influence of the response format in a personality questionnaire: An analysis of a dichotomous, a Likert -type, and a visual analogue scale The influence of the response format in a personality questi...
2015 doi
-
[25]
, Thomas, S
jacob_over_2024 APACrefauthors Jacob, B M. , Thomas, S. \ Joseph, J. APACrefauthors \ 2024 . Over two decades of research on choice overload: An overview and research agenda Over two decades of research on choice overload: An overview and research agenda . International Journa...
2024 doi
-
[26]
, Singer, J
jaeschke_comparison_1990 APACrefauthors Jaeschke, R. , Singer, J. \ Guyatt, G H. APACrefauthors \ 1990 . A comparison of seven-point and visual analogue scales: data from a randomized trial A comparison of seven-point and visual analogue scales: data from a randomized trial . ...
1990
-
[27]
\ Srivastava, S
john_big-five_1999 APACrefauthors John, O P. \ Srivastava, S. APACrefauthors \ 1999 . The Big Five Trait taxonomy: History , measurement, and theoretical perspectives The Big Five Trait taxonomy: History , measurement, and theoretical perspectives . Handbook of personality: Th...
1999
-
[28]
, Dantlgraber, M
kuhlmann_investigating_2017 APACrefauthors Kuhlmann, T. , Dantlgraber, M. \ Reips, U D. APACrefauthors \ 2017 . Investigating measurement equivalence of visual analogue scales and Likert -type scales in Internet -based personality questionnaires Investigating measurement equiv...
2017 doi
-
[29]
\ Paek, I
lee_search_2014 APACrefauthors Lee, J. \ Paek, I. APACrefauthors \ 2014 . In Search of the Optimal Number of Response Categories in a Rating Scale In Search of the Optimal Number of Response Categories in a Rating Scale . Journal of Psychoeducational Assessment 32 7 663--673 ....
2014 doi
-
[30]
APACrefauthors \ 1994
lei_chang_psychometric_1994 APACrefauthors Lei Chang . APACrefauthors \ 1994 . A Psychometric Evaluation of 4- Point and 6- Point Likert - Type Scales in Relation to Reliability and Validity A Psychometric Evaluation of 4- Point and 6- Point Likert - Type Scales in Relation to...
1994 doi
-
[31]
, Berjot, S
lesage_clinical_2012 APACrefauthors Lesage, F X. , Berjot, S. \ Deschamps, F. APACrefauthors \ 2012 . Clinical stress assessment using a visual analogue scale Clinical stress assessment using a visual analogue scale . Occupational medicine 62 8 600--605
2012
-
[32]
APACrefauthors \ 1932
likert_technique_1932 APACrefauthors Likert, R. APACrefauthors \ 1932 . A technique for the measurement of attitudes A technique for the measurement of attitudes . Archives of Psychology
1932
-
[33]
APACrefauthors \ 1954
loevinger_attenuation_1954 APACrefauthors Loevinger, J. APACrefauthors \ 1954 . The attenuation paradox in test theory The attenuation paradox in test theory . Psychological Bulletin 51 5 493--504 . APACrefDOI doi:10.1037/h0058543 APACrefDOI
1954 doi
-
[34]
\ Novick, M R
lord_statistical_1968 APACrefauthors Lord, F M. \ Novick, M R. APACrefauthors \ 1968 . Statistical theories of mental test scores Statistical theories of mental test scores . IAP
1968
-
[35]
, García-Cueto, E
lozano_effect_2008 APACrefauthors Lozano, L M. , García-Cueto, E. \ Muñiz, J. APACrefauthors \ 2008 . Effect of the Number of Response Categories on the Reliability and Validity of Rating Scales Effect of the Number of Response Categories on the Reliability and Validity of Rat...
2008 doi
-
[36]
, Lemmens, L H J M
martinez_validation_2024 APACrefauthors Mart\'inez, A J. , Lemmens, L H J M. , Fried, E I. , Gu\"omundsdóttir, G R. \ Roefs, A. APACrefauthors \ 2024 . Validation of a transdiagnostic psychopathology EMA protocol in a university students sample. Validation of a transdiagnostic...
2024 doi
-
[37]
, Lemmens, L
martinez_developing_2023 APACrefauthors Martínez, A J. , Lemmens, L. , Fried, E I. \ Roefs, A. APACrefauthors \ 2023 . Developing a Transdiagnostic Ecological Momentary Assessment Protocol for Psychopathology . Developing a Transdiagnostic Ecological Momentary Assessment Proto...
2023 doi
-
[38]
, Kramp, U
maydeu-olivares_effect_2009 APACrefauthors Maydeu-Olivares, A. , Kramp, U. , García-Forero, C. , Gallardo-Pujol, D. \ Coffman, D. APACrefauthors \ 2009 . The effect of varying the number of response alternatives in rating scales: Experimental evidence from intra-individual eff...
2009 doi
-
[39]
APACrefauthors \ 2009
mcardle_latent_2009 APACrefauthors McArdle, J J. APACrefauthors \ 2009 . Latent Variable Modeling of Differences and Changes with Longitudinal Data Latent Variable Modeling of Differences and Changes with Longitudinal Data . Annual Review of Psychology 60 1 577--605 . APACrefD...
2009
-
[40]
, Mitchell, M
mcgrath_evidence_2010 APACrefauthors McGrath, R E. , Mitchell, M. , Kim, B H. \ Hough, L. APACrefauthors \ 2010 . Evidence for response bias as a source of error variance in applied assessment. Evidence for response bias as a source of error variance in applied assessment. Psy...
2010
-
[41]
APACrefauthors \ 1978
mckelvie_graphic_1978 APACrefauthors McKelvie, S J. APACrefauthors \ 1978 . Graphic rating scales — How many categories? Graphic rating scales — How many categories? British Journal of Psychology 69 2 185--202 . APACrefDOI doi:10.1111/j.2044-8295.1978.tb01647.x APACrefDOI
1978
-
[42]
APACrefauthors \ 1956
miller_magical_1956 APACrefauthors Miller, G A. APACrefauthors \ 1956 . The magical number seven, plus or minus two: Some limits on our capacity for processing information The magical number seven, plus or minus two: Some limits on our capacity for processing information . Psy...
1956 doi
-
[43]
, Garc\'ia-Cueto, E
muniz_item_2005 APACrefauthors Mu\ niz, J. , Garc\'ia-Cueto, E. \ Lozano, L M. APACrefauthors \ 2005 . Item format and the psychometric properties of the Eysenck Personality Questionnaire Item format and the psychometric properties of the Eysenck Personality Questionnaire . Pe...
2005 doi
-
[44]
, Drasgow, F
olsson_polyserial_1982 APACrefauthors Olsson, U. , Drasgow, F. \ Dorans, N J. APACrefauthors \ 1982 . The polyserial correlation coefficient The polyserial correlation coefficient . Psychometrika 47 3 337--347 . APACrefDOI doi:10.1007/BF02294164 APACrefDOI
1982 doi
-
[45]
APACrefauthors \ 1991
paulhus_measurement_1991 APACrefauthors Paulhus, D L. APACrefauthors \ 1991 . Measurement and Control of Response Bias Measurement and Control of Response Bias . J P. Robinson, P R. Shaver \ L S. Wrightsman\ ( ), Measures of Personality and Social Psychological Attitudes Measu...
1991 doi
-
[46]
, MacKenzie, S B
podsakoff_common_2003 APACrefauthors Podsakoff, P M. , MacKenzie, S B. , Lee, J Y. \ Podsakoff, N P. APACrefauthors \ 2003 . Common method biases in behavioral research: a critical review of the literature and recommended remedies. Common method biases in behavioral research: ...
2003
-
[47]
\ Colman, A M
preston_optimal_2000 APACrefauthors Preston, C C. \ Colman, A M. APACrefauthors \ 2000 . Optimal number of response categories in rating scales: reliability, validity, discriminating power, and respondent preferences Optimal number of response categories in rating scales: reli...
2000 doi
-
[48]
APACrefauthors \ 1969
samejima_estimation_1969 APACrefauthors Samejima, F. APACrefauthors \ 1969 . Estimation of latent ability using a response pattern of graded scores Estimation of latent ability using a response pattern of graded scores . Psychometrika Society
1969
-
[49]
\ Saris, W E
scherpenzeel_validity_1997 APACrefauthors Scherpenzeel, A C. \ Saris, W E. APACrefauthors \ 1997 . The Validity and Reliability of Survey Questions : A Meta - Analysis of MTMM Studies The Validity and Reliability of Survey Questions : A Meta - Analysis of MTMM Studies . Sociol...
1997 doi
-
[50]
APACrefauthors \ 2010
schmidt_more_2010 APACrefauthors Schmidt, K M. APACrefauthors \ 2010 . More is Not Better : Rescoring Combinations for Length Rating Scales More is Not Better : Rescoring Combinations for Length Rating Scales . International Conference on Measurement ( ICOM -2010). Internation...
2010
-
[51]
, Zelazny, K
simms_does_2019 APACrefauthors Simms, L J. , Zelazny, K. , Williams, T F. \ Bernstein, L. APACrefauthors \ 2019 . Does the number of response options matter? Psychometric perspectives using personality questionnaire data. Does the number of response options matter? Psychometri...
2019
-
[52]
\ Wu, J S
sung_visual_2018 APACrefauthors Sung, Y T. \ Wu, J S. APACrefauthors \ 2018 . The Visual Analogue Scale for Rating , Ranking and Paired - Comparison ( VAS - RRP ): A new technique for psychological measurement The Visual Analogue Scale for Rating , Ranking and Paired - Compari...
2018 doi
-
[53]
, Uldall, B
thomas_how_2004 APACrefauthors Thomas, R K. , Uldall, B. \ Krosnick, J A. APACrefauthors \ 2004 . How many are too many? Number of response categories and validity How many are too many? Number of response categories and validity . 59th Annual Conference of the American Associ...
2004
-
[54]
, Van Der Zaag‐Loonen, H
van_laerhoven_comparison_2004 APACrefauthors Van Laerhoven, H. , Van Der Zaag‐Loonen, H. \ Derkx, B. APACrefauthors \ 2004 . A comparison of Likert scale and visual analogue scales as response options in children's questionnaires A comparison of Likert scale and visual analogu...
2004
-
[55]
, Bergen, M
viswanathan_does_1996 APACrefauthors Viswanathan, M. , Bergen, M. , Dutta, S. \ Childers, T. APACrefauthors \ 1996 . Does a single response category in a scale completely capture a response? Does a single response category in a scale completely capture a response? Psychology &...
1996 doi
-
[56]
, Zhang, Z
wang_investigating_2008 APACrefauthors Wang, L. , Zhang, Z. , McArdle, J J. \ Salthouse, T A. APACrefauthors \ 2008 . Investigating Ceiling Effects in Longitudinal Data Analysis Investigating Ceiling Effects in Longitudinal Data Analysis . Multivariate Behavioral Research 43 3...
2008 doi
-
[57]
APACrefauthors \ 1998
weng_scale_1998 APACrefauthors Weng, L J. APACrefauthors \ 1998 . Scale values of anchor labels in Chinese rating scales: Responses on frequency and agreement Scale values of anchor labels in Chinese rating scales: Responses on frequency and agreement . Chinese Journal of Psyc...
1998
-
[58]
, Morlock, R J
williams_psychometric_2010 APACrefauthors Williams, V S. , Morlock, R J. \ Feltner, D. APACrefauthors \ 2010 . Psychometric evaluation of a visual analog scale for the assessment of anxiety Psychometric evaluation of a visual analog scale for the assessment of anxiety . Health...
2010 doi
-
[59]
\ Leung, S O
wu_can_2017 APACrefauthors Wu, H. \ Leung, S O. APACrefauthors \ 2017 . Can Likert Scales be Treated as Interval Scales ?— A Simulation Study Can Likert Scales be Treated as Interval Scales ?— A Simulation Study . Journal of Social Service Research 43 4 527--532 . APACrefDOI d...
2017
-
[60]
\ Aitken, R C B
zealley_growing_1969 APACrefauthors Zealley, A K. \ Aitken, R C B. APACrefauthors \ 1969 . A Growing Edge of Measurement of Feelings [ Abridged ]: Measurement of Mood A Growing Edge of Measurement of Feelings [ Abridged ]: Measurement of Mood . Proceedings of the Royal Society...
1969 doi
Reviewed August 9, 2026 · model on record in the stance chip above.
Discussion (0). Continue with ORCID to comment.