Manuel B. Garcia is a professor of information technology and the founding director of the Educational Innovation and Technology Hub (EdITH) at FEU Institute of Technology, Manila, Philippines. Read More

Contact Info

1607, FEU Tech Building,
P. Paredes St, Sampaloc,
Manila, Philippines
mbgarcia@feutech.edu.ph

Follow Me

Statistical Test Selector

Answer a few questions about your research objective, variables, and study design to identify statistical tests that may fit your analysis.

Find an appropriate statistical test

The recommendation updates as you describe your study.

No data uploaded
1

Research objective

What are you primarily trying to do?
2

Outcome variable

What type of outcome are you analyzing?
3

Predictor or grouping variable

What type of variable defines the comparison or predictor?
4

Number of groups or measurements

How many groups, conditions, or measurements are involved?
5

Study design

Are the observations independent or related?
6

Distribution and assumptions

For a continuous outcome, which description fits best?
Assumptions should be evaluated using the design, residuals, plots, sample size, and subject-matter context.
3

Agreement design

What kind of ratings or measurements are being compared?

Your recommendation will appear here

Start by choosing your research objective.

Educational guidance onlyThis tool cannot evaluate sampling, power, missing data, complex designs, measurement validity, or causal assumptions.

How the statistical test selector works

A statistical test should follow the research question and data structure—not the other way around.

The selector first identifies whether your goal is comparison, association, prediction, agreement, survival analysis, or a one-sample distributional question. It then considers the measurement level of the outcome, the predictor or grouping variable, the number of groups, and whether observations are independent or related.

The resulting recommendation includes a commonly used primary method, possible alternatives, and issues to verify before analysis.

Research objective

Different questions require different families of statistical procedures.

Variable type

Continuous, ordinal, categorical, count, and survival outcomes are modeled differently.

Data dependence

Independent groups require different tests from matched or repeated observations.

Assumptions

Model form, residuals, outliers, sample size, and variance structure matter.

A practical statistical test guide

Common tests organized by question and data structure.

Research questionTypical designCommon test or model
Compare one continuous sample with a reference valueOne sampleOne-sample t testWilcoxon signed-rank
Compare two independent groupsContinuous outcomeWelch's t testMann–Whitney U
Compare two related measurementsPre–post or matched pairsPaired t testWilcoxon signed-rank
Compare three or more independent groupsContinuous or ordinal outcomeOne-way ANOVAWelch ANOVAKruskal–Wallis
Compare three or more related measurementsRepeated observationsRepeated-measures ANOVAFriedman testMixed model
Assess association between categorical variablesContingency tableChi-square testFisher's exact test
Assess association between numeric or ordinal variablesTwo variablesPearson correlationSpearman correlationKendall's tau
Predict a continuous outcomeOne or more predictorsLinear regressionRobust regressionMixed-effects model
Predict a binary or categorical outcomeOne or more predictorsBinary logisticOrdinal logisticMultinomial logistic
Analyze count or rate dataEvent counts or exposure-adjusted ratesPoisson regressionNegative binomial regression
Analyze time until an eventCensoring may occurKaplan–MeierLog-rank testCox regression
Measure reliability or agreementRaters, items, or methodsCohen's kappaFleiss' kappaICCBland–Altman

Before choosing a test

  • Define the estimand. State exactly what difference, association, probability, or effect you want to estimate.
  • Identify the unit of analysis. Participants, classrooms, schools, repeated visits, and observations are not interchangeable.
  • Respect dependence. Repeated, matched, clustered, or nested observations usually need methods that model their correlation.
  • Inspect the data. Check coding, missingness, impossible values, outliers, distributions, and group sizes.
  • Plan effect sizes. A p value alone does not show magnitude or practical relevance.

When a simple test is not enough

A basic test selector cannot fully handle factorial experiments, multilevel sampling, longitudinal trajectories, latent variables, survey weights, propensity scores, mediation, moderation, multiple imputation, compositional data, network data, Bayesian analysis, or machine-learning validation.

For clustered or repeatedly measured data, mixed-effects models or generalized estimating equations may be more appropriate than treating every observation as independent. For observational causal questions, statistical adjustment alone does not guarantee causal identification.

Good practice: choose the analysis during study planning, justify it from the design and research questions, and document any deviations from the preregistered plan.

Frequently asked questions

The appropriate test depends on your research objective, outcome variable, predictor or grouping variable, number of groups or measurements, whether observations are independent or paired, and whether important assumptions are reasonably satisfied.

No. It provides an educational starting point, not a definitive analysis plan. Complex sampling, multilevel data, repeated observations, missing data, small samples, multiple outcomes, and causal questions may require specialist advice.

Parametric tests model particular distributional features and usually make assumptions about errors or residuals. Nonparametric or rank-based tests use fewer distributional assumptions but still have design requirements and may test a different statistical hypothesis.

Normality is generally most relevant to model residuals or within-group distributions, not merely the raw pooled outcome. Graphs, sample size, outliers, design balance, and robustness should be considered alongside formal tests.

Use it to compare the mean of a continuous outcome between two independent groups when the observations are independent and the model assumptions are acceptable. Welch's t test is often preferred when group variances or sample sizes differ.

Use it for two related measurements, such as pretest and posttest scores from the same participants or matched pairs, when the pairwise differences are approximately normal and serious outliers are absent.

Use ANOVA to compare a continuous outcome across three or more groups or conditions. The exact form depends on whether groups are independent, measurements are repeated, covariates are included, or the design contains multiple factors.

A chi-square test of independence is commonly used to assess association between two categorical variables in independent observations when expected cell counts are adequate. Fisher's exact test is an alternative for sparse small tables.

Pearson correlation describes linear association between continuous variables. Spearman correlation describes monotonic rank association and may be more suitable for ordinal variables, strong outliers, or clearly non-linear monotonic relationships.

The model is largely determined by the outcome: linear regression for continuous outcomes, binary logistic regression for two-category outcomes, ordinal logistic regression for ordered categories, multinomial logistic regression for unordered categories, and count models such as Poisson or negative binomial regression for counts.

Report descriptive statistics, the test statistic, degrees of freedom when applicable, exact p value, confidence interval, an appropriate effect size, assumption checks, missing-data handling, and enough methodological detail for reproducibility.

No. Statistical significance does not establish practical importance, causality, measurement quality, or replicability. Interpret p values together with effect sizes, confidence intervals, study design, and substantive context.

Disclaimer: This selector provides general educational guidance and does not constitute statistical consulting. The final analysis should reflect the study design, sampling process, measurement quality, assumptions, missing data, multiplicity, and substantive research context.