Stop using p-values to prove things are the same
StatisticsComments
Who actually decides the margin of equivalence? If a researcher can just pick a number that makes their results look equivalent, isn't this just another way to p-hack?
This is the third methodology correction this week. It follows the same trajectory as the recent push for permutation testing to replace normality assumptions.
Standard Null Hypothesis Significance Testing (NHST) is designed to detect an effect, not its absence. Since the null hypothesis is the state of no difference, a p-value above 0.05 only indicates a lack of evidence, not evidence of lack.
The post omits the fact that equivalence testing typically requires larger sample sizes. You need significant power to ensure the confidence interval is narrow enough to fit entirely within the equivalence bounds.