Conceptual and Statistical Knowledge

 

On the origins of the .05 level of statistical significance

Examination of the literature in statistics and probability that predates Fisher's Statistical Methods for Research Workers indicates that although Fisher is responsible for the first formal statement of the .05 criterion for statistical …

Statistical errors: P values, the ‘gold standard’ of statistical validity, are not as reliable as many scientists assume

P values, the 'gold standard' of statistical validity, are not as reliable as many scientists assume.

Statistical significance in psychological research.

MOST THEORIES IN THE AREAS OF PERSONALITY, CLINICAL, AND SOCIAL PSYCHOLOGY PREDICT ONLY THE DIRECTION OF A CORRELATION, GROUP DIFFERENCE, OR TREATMENT EFFECT. SINCE THE NULL HYPOTHESIS IS NEVER STRICTLY TRUE, SUCH PREDICTIONS HAVE ABOUT A 50-50 …

Surrogate Science: The Idol of a Universal Method for Scientific Inference

The application of statistics to science is not a neutral act. Statistical tools have shaped and were also shaped by its objects. In the social sciences, statistical methods fundamentally changed research practice, making statistical inference its …

The appropriate use of null hypothesis testing.

The many criticisms of null hypothesis testing suggest when it is not useful and what is should not be used for. This article explores when and why its use is appropriate. Null hypothesis testing is insufficient when size of effect is important, but …

The ASA Statement on p-Values: Context, Process, and Purpose

An editorial about p value

The case against statistical significance testing

In recent years the use of traditional statistical methods in educational research has increasingly come under attack. In this article, Ronald P. Carver exposes the fantasies often entertained by researchers about the meaning of statistical …

The harm done by tests of significance

Three historical episodes in which the application of null hypothesis significance testing (NHST) led to the mis-interpretation of data are described. It is argued that the pervasive use of this statistical ritual impedes the accumulation of …

Randomization Does Not Help Much, Comparability Does

According to R.A. Fisher, randomization “relieves the experimenter from the anxiety of considering innumerable causes by which the data may be disturbed.” Since, in particular, it is said to control for known and unknown nuisance factors that may …

Research Practices That Can Prevent an Inflation of False-Positive Rates

Recent studies have indicated that research practices in psychology may be susceptible to factors that increase false-positive rates, raising concerns about the possible prevalence of false-positive findings. The present article discusses several …
JUST-OS