
p-value
Sign in to saveIn null-hypothesis significance testing, the '''p-value' is the probability of obtaining test results at least as extreme as the result actually observed, under the assumption that the null hypothesis is correct. A very small p-value means that such an extreme observed outcome would be very unlikely under the null hypothesis. Even though reporting p''-values of statistical tests is common practice in academic publications of many quantitative fields, misinterpretation and misuse of p-values is widespread and has been a major topic in mathematics and metascience.
Wikidata facts
Show 2 more facts
- Stack Exchange tag
- stackoverflow.com/tags/p-value
- Commons category
- P-value
Sources (3)
via Wikidata · CC0
~26 min read
Article
19 sectionsContents
- Basic concepts
- Definition and interpretation
- Definition
- Interpretations
- Distribution
- Distribution for composite hypothesis
- Usage
- Misuse
- Calculation
- Example
- Testing the fairness of a coin
- Optional stopping
- History
- Related indices
- See also
- Notes
- References
- Further reading
- External links
In null-hypothesis significance testing, the '''p-value' is the probability of obtaining test results at least as extreme as the result actually observed, under the assumption that the null hypothesis is correct. A very small p-value means that such an extreme observed outcome would be very unlikely under the null hypothesis. Even though reporting p-values of statistical tests is common practice in academic publications of many quantitative fields, misinterpretation and misuse of p-values is widespread and has been a major topic in mathematics and metascience.
In 2016, the American Statistical Association (ASA) made a formal statement that "p-values do not measure the probability that the studied hypothesis is true, or the probability that the data were produced by random chance alone" and that "a p-value, or statistical significance, does not measure the size of an effect or the importance of a result", and "does not provide a good measure of evidence regarding a model or hypothesis" without "context or other evidence". That said, a 2019 task force by ASA has issued a statement on statistical significance and replicability, concluding with: "p-values and significance tests, when properly applied and interpreted, increase the rigor of the conclusions drawn from data".