Curation methodology

Not a database.
A curated resource.

Every scale in this collection has been individually assessed, verified, and entered by experienced researchers in marketing and consumer psychology. There is no automated ingestion pipeline, no crowdsourced submission queue, no algorithm deciding what belongs here. Each entry reflects a deliberate decision.

Scales indexed
.80 Min. reliability
18 Domains covered

A database grows by accumulation.
PsychScales grows by careful, expert review.

The validated scale literature is vast, fragmented, and uneven in quality. Finding the right instrument takes time — and too often that search leads to measures that are poorly validated, single-item proxies, verbatim reproductions of prior work, or inadequately tested for the context at hand. PsychScales exists to reduce that friction. Not by being exhaustive, but by being selective — so researchers can spend less time vetting instruments and more time advancing their work.

Including a poorly validated scale alongside a rigorous one creates false equivalence and undermines the usefulness of any list. The approach here is intentional: a smaller collection of instruments you can trust is more valuable than a larger one where quality varies.

Inclusion criteria

Scales are included where they represent a meaningful contribution to the literature and have demonstrated strong reliability and validity evidence in their primary validation. The following standards define what that means in practice.

01

Reliability

A minimum Cronbach's alpha or composite reliability of .80 is required across the primary validation sample. The widely cited .70 threshold, originating from Nunnally's early psychometric work, was a pragmatic heuristic rather than a theoretically derived standard — and one Nunnally later revised upward.

At .70, approximately 30% of observed score variance is attributable to measurement error rather than the construct of interest. For researchers building structural models or testing fine-grained hypotheses, that level of measurement imprecision matters. The .80 threshold reduces that error proportion meaningfully, producing more precise parameter estimates, stronger statistical power in hypothesis testing, and greater confidence that relationships observed between constructs reflect theoretical reality rather than shared measurement noise.

In practical terms, scales meeting the .80 standard are more likely to detect true effects, less likely to produce spurious ones, and more defensible in peer review. A small number of seminal legacy instruments with exceptional citation impact and established disciplinary use are assessed on a case-by-case basis.

02

Multi-item measurement

Single-item scales are excluded. Without multiple indicators, it is impossible to separate true score variance from random measurement error, and reliability in the classical sense cannot be estimated at all — a scale of one item has no internal consistency to assess.

Single items are also acutely sensitive to question wording effects, momentary response biases, and context-order artefacts that multi-item scales partially cancel out through aggregation. Beyond reliability, single items preclude meaningful tests of dimensionality, convergent validity, or discriminant validity — the core evidence base that justifies treating a measure as capturing a distinct psychological construct.

Researchers selecting instruments from this collection are typically building structural models where measurement quality propagates directly into the validity of latent variable estimates. A single item introduced at that stage does not just underperform — it contaminates.

03

Original instrumentation

Scales that reproduce items verbatim from a prior instrument without substantive modification are generally excluded. The collection prioritises genuinely new instruments and bona fide adaptations — where the focal paper has made meaningful changes and established independent validation evidence.

However, the boundary is not purely syntactic. Where a prior instrument is deployed in a substantively novel context — a new population, domain, or theoretical framework — and the focal paper provides the first rigorous demonstration of predictive validity in that context, inclusion is considered on its merits. The relevant question is not whether the items are new, but whether the paper makes an original psychometric contribution.

Verbatim reproductions serving purely as manipulation checks, covariates, or convenience measures without independent validation evidence remain excluded.

PsychScales draws from peer-reviewed work across three broad disciplinary areas — Marketing & Consumer Psychology, Tourism & Hospitality, and Psychology — spanning organisational behaviour, health, sport, and adjacent social sciences. The collection is deepest in marketing and consumer behaviour, where the curation process draws on close familiarity with the literature. Adjacent fields are represented and growing, but if your work sits outside these areas, coverage may be less comprehensive. The psychometric criteria — not journal prestige or author profile — determine what belongs here. A well-validated scale from a specialist outlet is treated the same as one from a flagship journal.

Edge cases and grey areas

Most inclusion decisions are straightforward — a scale either meets the psychometric criteria or it doesn't. A small number of cases require more careful consideration, and it is worth being transparent about how those are handled.

Legacy instruments. A small number of seminal scales have shaped entire research programmes and remain in active use despite reliability coefficients that fall just below the .80 threshold. These are assessed individually, weighing citation impact, disciplinary entrenchment, and whether more recent validation evidence brings the instrument closer to the standard. Where a legacy instrument is included under this discretion, the entry notes it explicitly.

Adaptations and redeployments. Where an existing instrument has been deployed in a genuinely novel context — a new population, domain, or theoretical framework — and the paper provides rigorous independent validation evidence, that contribution is considered on its merits. The relevant question is not whether the items are new, but whether the paper makes an original psychometric contribution.

When in doubt, the entry is deferred. A scale not yet in the collection may simply not have been reviewed yet, or may be sitting in a queue pending a closer read. If you believe an instrument belongs here, the Suggest a Scale button on the main page is the right place to flag it.

The entry standard

Each record in PsychScales goes beyond a simple citation and item list. Every entry includes original descriptive prose summarising the scale's development and validation history, its theoretical grounding, and its appropriate use context. Notes fields document known limitations, cross-cultural caveats, scoring guidance, and relevant comparisons to related instruments. This annotation layer is what separates a curated resource from a glorified spreadsheet.

Full item wording is included wherever the original publication makes items publicly available, enabling researchers to evaluate fit before sourcing the full paper.

Construct Precise theoretical construct being measured
Reliability statistics Alpha, composite reliability, and AVE with sample details
Factor structure Number of dimensions and dimension labels
Item wording Complete item wording organised by dimension
Validation history Scale development, theoretical context, and appropriate use
Limitations & scoring Known caveats, cross-cultural notes, and scoring guidance
Related instruments Comparable measures for triangulation or alternatives
Known adaptations Derivative versions and modifications in the literature

Ongoing maintenance

PsychScales is a living resource updated on a continuous basis as new instruments are validated and published. The same inclusion criteria are applied to new entries as to the existing collection — there is no relaxation of standards over time, and no backlog of unreviewed material waiting for attention.

When an entry is updated to reflect a correction, replication concern, or material change in the evidence base, that change is documented. The collection does not quietly revise history.

What this is for

PsychScales is free, openly accessible, and will remain so. It exists because finding the right validated instrument should not require access to expensive database subscriptions, hours of citation chasing, or insider knowledge of what the field actually uses.

Rigorous measurement is foundational to good research. This resource is built on the conviction that making it easier to measure things well is worth doing.

Browse the database →