Survey Data Quality Screen

Verified

Psychometrics (legacy hub)

Independently verified. Every statistic this test reports has been re-derived against an independent reference — never the library the pipeline itself calls — the rendered output was read back in a browser, and the result is locked with a committed regression suite.

Run this test straight away on a free built-in teaching dataset — no data of your own needed — or bring your own. Either opens the guided workspace: variable setup, assumption diagnostics, results with effect sizes and confidence intervals, figures, and APA-ready reporting.

Loading teaching datasets…

Or use your own dataset

Loading your datasets…

Pre-analysis integrity check for survey-response data.

Identifies responses that are likely INVALID and should be excluded or flagged before any substantive analysis. Six diagnostics: (1) missing-data pattern (overall + per-respondent + per-item); (2) Little's MCAR test; (3) straightlining (respondent checks same option for most items); (4) pattern responding (zigzag / repeating sequences); (5) speeders (completion time < 1/3 median); (6) careless-response flags summary. Recommended exclusion list + per-flag counts.

Worked example

Is the dataset clean enough to analyse?

A data-quality scan reviews missing values, duplicates, out-of-range and disguised-missing codes before any test is run.

Result

The data were clean: 0.4% missing (consistent with MCAR), no duplicates, and no out-of-range values.

How you'd report it (APA)

Data-quality screening found the dataset analysis-ready (0.4% missing, no duplicates or out-of-range values).

When to use it

  • Online-survey pre-analysis integrity screen
    Online personality survey on n=500 Prolific respondents.
  • Longitudinal-survey attrition + missing-pattern check
    4-wave panel (n=800 at baseline) with decreasing n across waves.

When NOT to — use instead

Assumptions (and what to do if they fail)

Items are numeric and use a structured response setmedium

Check: See the assumption diagnostics in the workspace.

If violated: The workspace flags this and suggests a robust or nonparametric alternative.

Each row = one respondentmedium

Check: See the assumption diagnostics in the workspace.

If violated: The workspace flags this and suggests a robust or nonparametric alternative.

Response times (if supplied) are total seconds per respondentmedium

Check: See the assumption diagnostics in the workspace.

If violated: The workspace flags this and suggests a robust or nonparametric alternative.

Ready to run a Survey Data Quality Screen on your own data?

Guided setup, automatic assumption checks, effect sizes, figures and an APA write-up.

Run this test →