Limiting the Shrinkage for the Exceptional by Objective Robust Bayesian Analysis: The “Clemente Problem”
Pub. online: 8 July 2026
Type: Methodology Article
Open Access
Area: Statistical Methodology
Accepted
22 May 2026
22 May 2026
Published
8 July 2026
8 July 2026
Notes
The authors want to dedicate this work to Roberto Clemente Walker, first Latin American to reach the Baseball Hall of Fame.
Abstract
Modern Statistics is made of the sensible combination of direct evidence (the data directly relevant or the “individual data”) and indirect evidence (the data and knowledge indirectly relevant or the “group data”). The admissible procedures combine the two sources of information, and technology advances make indirect evidence more substantial and ubiquitous. It has been pointed out, however, that an important problem of Statistics when “borrowing strength” is to treat in a fundamentally different way exceptional cases that do not adapt to the central “aurea mediocritas”. This is what has been coined as “the Clemente problem” in honor of R. Clemente, an exceptional batter [6]. In this article, we argue that the problem is caused by the simultaneous use of square loss function and conjugate (light-tailed) priors, which is the usual procedure. We propose in their place to use robust penalties, in the form of losses that penalize more severely huge errors, or (equivalently) priors of heavy tails which make the exceptional more probable. Using heavy-tailed priors, we can reproduce, in a Bayesian structured way, Efron and Morris’ “limited translated estimators” (with Double Exponential Priors) and “discarding priors estimators” (with Cauchy-like priors), which discard the prior in the presence of prior-likelihood conflict. We show that both Empirical Bayes and Full Bayes approaches can alleviate the Clemente problem and beat the James-Stein estimator in terms of smaller square errors, for sensible Robust Bayes priors. We model in parallel Empirical Bayes and Fully Bayesian hierarchical models, illustrating that the differences among sensible versions of both are relatively small, as compared with the effect due to the robust assumptions. We follow [16] in using a heavy-tailed (scaled) Beta2 distribution for (squared) scales that arise naturally as an alternative to the usual Inverted-Gamma distribution. The combination of a Cauchy Prior for location and Scaled Beta2 for square scales yields a novel closed-form prior for location, extremely suitable for Objective Robust Bayesian Analysis. Finally, we calculate the predictive intervals and the Robust models covers the Clemente average at $80\% $ of probability, which the conjugate do not. Our approach has connections with [25], which employs a completely different methodology. This is an instance of the convergence of the best frequentist and Bayesian analyses [see 3].
References
Andrade, J. A. A. and O’Hagan, A. (2006). Bayesian Robustness Modeling Using Regularly Varying Distribution. Bayesian Analysis 1 169–188. https://doi.org/10.1214/06-BA106. MR2227369
Berger, J. O. (1985) Statistical Decision Theory and Bayesian Analysis. Springer-Verlag, New York. https://doi.org/10.1007/978-1-4757-4286-2. MR0804611
Berger, J. O. (2023). Four Types of Frequentism and Their Interplay with Bayesianism. New England Journal of Statistics in Data Science 1(2) 126–137. https://doi.org/10.51387/22-NEJSDS4.
Brown, L. D. (2008). In-Season Prediction of Batting Averages: a Field Test of Empirical Bayes and Bayes Methodologies. The Annals of Applied Statistics 2 113–214. https://doi.org/10.1214/07-AOAS138. MR2415597
Carvalho, C. M., Polson, N. G. and Scott, J. G. (2010). The Horseshoe Estimator for Sparse Signals. Biometrika 97(2) 465–480. https://doi.org/10.1093/biomet/asq017. MR2650751
Efron, B. (2010). The Future of Indirect Evidence. Statistical Science 25 145–157. https://doi.org/10.1214/09-STS308. MR2789983
Efron, B. and Morris, C. (1972). Limiting the Risk of Bayes and Empirical Bayes Estimators-Part II: The Empirical Bayes Case. Journal of the American Statistical Association 67 130–139. MR0323015
Ferguson, T. (1967) Mathematical Statistics. A Decision Theoretic Approach. Academic Press. MR0215390
Fúquene, J., Pérez, M. E. and Pericchi, L. (2014). An alternative to the Inverted Gamma for the variances to modelling outliers and structural breaks in dynamic models. Brazilian Journal of Probability and Statistics 28(2). https://doi.org/10.1214/12-BJPS207. https://doi.org/10.1214/12-BJPS207. MR3189499
Fúquene, J. A., Cook, J. D. and Pericchi, L. R. (2009). A Case for Robust Bayesian Priors with Applications to Clinical Trials. Bayesian Analysis 4 817–846. https://doi.org/10.1214/09-BA431. MR2570090
Gelman, A. (2006). Prior Distributions for Variance Parameters in Hierarchical Models. Bayesian Analysis 1(3) 515–533. https://doi.org/10.1214/06-BA117A. MR2221284
Johnson, N. L., Kotz, S. and Balakrishnan, N. (1995) Continuous Univariate Distributions 2. Wiley. MR1326603
Normand, S. T. and Shahian, D. M. (2007). Statistical and Clinical Aspects of Hospital Outcomes Profiling. Statistical Science 22(2) 206–226. https://doi.org/10.1214/088342307000000096. MR2408959
O’Hagan, A. and Pericchi, L. (2012). Bayesian heavy-tailed models and conflict resolution: A review. Brazilian Journal of Probability and Statistics 26(4) 372–401. https://doi.org/10.1214/11-BJPS164. https://doi.org/10.1214/11-BJPS164. MR2949085
Perez, M. E., Pericchi, L. R. and Ramirez, I. C. (2017). The Scaled Beta2 Distribution as a Robust Prior for Scales. Bayesian Analysis 12(3) 615–637. https://doi.org/10.1214/16-BA1015. MR3655869
Pericchi, L. R. and Smith, A. F. M. (1992). Exact and approximate posterior moments for a Normal location parameter. Journal of the Royal Statistical Society B 54(3) 793–804. MR1185223
Polson, N. G. and Scott, J. G. (2011). Shrink Globally, Act Locally: Sparse Bayesian Regularization and Prediction. In Bayesian Statistics 9 501–538 Oxford University Press. https://doi.org/10.1093/acprof:oso/9780199694587.003.0017. MR3204017
Stan Development Team (2023). Stan Modeling Language Users Guide and Reference Manual, version 2.32. https://mc-stan.org.
Stan Development Team (2025). RStan: the R interface to Stan. R package version 2.32.7. https://mc-stan.org/.
Vehtari, A., Gelman, A. and Gabry, J. (2017). Practical Bayesian model evaluation using leave-one-out cross-validation and WAIC. Statistics and Computing 27 1413–1432. https://doi.org/10.1007/s11222-016-9696-4. https://doi.org/10.1007/s11222-016-9696-4. MR3647105
Vehtari, A., Gabry, J., Magnusson, M., Yao, Y., Bürkner, P. -C., Paananen, T. and Gelman, A. (2025). loo: Efficient leave-one-out cross-validation and WAIC for Bayesian models. R package version 2.9.0. https://mc-stan.org/loo/.
Watanabe, S. (2010). Asymptotic Equivalence of Bayes Cross Validation and Widely Applicable Information Criterion in Singular Learning Theory. J. Mach. Learn. Res. 11 3571–3594. MR2756194
Yu, C. and Hoff, P. D. (2018). Adaptive multigroup confidence intervals with constant coverage. Biometrika 105 (2) 319–335. https://doi.org/10.1093/biomet/asy009. https://doi.org/10.1093/biomet/asy009. MR3804405