%0 Report %A Jaeger, David A. %T Robustness? Range Tests for Equality and Equivalence Across Specifications %D 2026 %8 2026 Aug %I Institute of Labor Economics (IZA) %C Bonn %7 IZA Discussion Paper %N 18851 %U https://www.iza.org/publications/dp18851 %X Applied economists routinely compare estimates across specifications, observe that they are "similar," and conclude that their results are "robust.'" This common procedure makes an implicit inferential claim about the range of estimates, but usually does not account for their joint sampling distribution. I formalize informal practice with two bootstrap statistics. The minimum equivalence bound, $R^*_{1-\alpha}$, is the smallest tolerance within which the estimates can be judged equivalent. The range-based equality $p$-value, $p_R$, tests whether the estimates are statistically distinguishable. Together they distinguish failure to detect differences from affirmative evidence of agreement. Simulations show approximately correct size and coverage. Applications to five prominent papers validate some robustness claims while revealing cases in which apparent agreement reflects imprecision rather than stability. A survey of CEPR and NBER affiliates shows that expert judgments aligns with the framework in obvious cases but diverges in intermediate cases. I suggest that $R^*_{.95}$ and $p_R$ be reported whenever multiple specifications are presented as evidence of robustness. %K robustness %K specification sensitivity %K equivalence testing %K bootstrap inference %K joint inference %K model uncertainty