Conceptual and Statistical Knowledge

 

Statistical Reporting Errors and Collaboration on Statistical Analyses in Psychological Science

Statistical analysis is error prone. A best practice for researchers using statistics would therefore be to share data among co-authors, allowing double-checking of executed tasks just as co-pilots do in aviation. To document the extent to which this …

The (mis)reporting of statistical results in psychology journals

In order to study the prevalence, nature (direction), and causes of reporting errors in psychology, we checked the consistency of reported test statistics, degrees of freedom, and p values in a random sample of high- and low-impact psychology …

The garden of forking paths: Why multiple comparisons can be a problem, even when there is no “fishing expedition” or “p-hacking” and the research hypothesis was posited ahead of time

Data-dependent analysis—a “garden of forking paths”— explains why many statistically significant comparisons don't hold up.

An exploratory test for an excess of significant findings

Background The published clinical research literature may be distorted by the pursuit of statistically significant results. Purpose: We aimed to develop a test to explore biases stemming from the pursuit of nominal statistical significance. Methods …

Bayes Factor

Bayes factors are somewhat essential to Bayesian statistics. Tony O’Hagan explains their basics

Intro to the special issue

A paper about a special issue on cognitive modelling

Power failure: why small sample size undermines the reliability of neuroscience

A study with low statistical power has a reduced chance of detecting a true effect, but it is less well appreciated that low power also reduces the likelihood that a statistically significant result reflects a true effect. Here, we show that the …

Psychological testing and psychological assessment: A review of evidence and issues.

This article summarizes evidence and issues associated with psychological assessment. Data from more than 125 meta-analyses on test validity and 800 samples examining multimethod assessment suggest 4 general conclusions: (a) Psychological test …

Statistical power analysis

A paper about statistical power

The earth is round (p < .05).

After 4 decades of severe criticism, the ritual of null hypothesis significance testing (mechanical dichotomous decisions around a sacred .05 criterion) still persists. This article reviews the problems with this practice, including near universal …
JUST-OS