Articles

Regional analysis of specific costs in agricultural production: using the Wasserstein distance

Abstract

In the context of the forthcoming reform of the Common Agricultural Policy, analysing agricultural production costs  is a crucial step in developing fair pricing policies, particularly at regional level, to avoid potential distortions of competition. This article presents the application of Wasserstein distance-based statistical techniques to the regional analysis of specific costs  in agricultural production. Taking into account the heterogeneity and asymmetry of the distributions, it combines the conditional quantile estimation methodology with those based on the Wasserstein distance (factorial analysis of the distance matrix, unsupervised divisive classification, quadratic test). The procedure is applied to the comparative regional analysis of the distributions of specific fertiliser costs  for crop production, based on the European Agricultural Accounting Network, at the level of the European basic regions (NUTS2). The results obtained for specific input costs and associated margins appear consistent with those of previous work carried out under the European Union’s 7th Framework Programme for Research.

References

  • Benzécri, J.-P. et al. (1973). L’analyse des données. Tome 1 : La taxinomie. Tome 2 : L’analyse des correspondances. Dunod.
  • Billard, L., & Diday, E. (2007). Symbolic Data Analysis: Conceptual Statistics and Data Mining (Wiley Series in Computational Statistics). Dans John Wiley & Sons, Inc. eBooks. http://dl.acm.org/citation.cfm?id=1206598
  • Birnbaum, Z. W., & McCarty, R. C. (1958). A Distribution-Free Upper Confidence Bound for $ \Pr \{Y. The Annals Of Mathematical Statistics, 29(2), 558 562. https://doi.org/10.1214/aoms/1177706631
  • Bofinger, E. (1975). Non-Parametric Estimation of Density for Regularly Varying Distributions. Australian Journal Of Statistics, 17(3), 192 195. https://doi.org/10.1111/j.1467-842x.1975.tb00957.x
  • Bureau, J.-C., & Cyncynatus, M. (1991). Estimation de coûts de production et de coefficients input-output à partir de données comptables : Méthodes et applications aux produits agricoles sur la base du RICA (Notes et Documents, n° 38). INRA, Département Économie et Sociologie Rurales.
  • Cameron, A. C., & Trivedi, P. K. (2005). Microeconometrics : Methods and Applications. Cambridge University Press.
  • Chavent, M., Lechevallier, Y., & Briant, O. (2007). DIVCLUS-T : A monothetic divisive hierarchical clustering method. Computational Statistics & Data Analysis, 52(2), 687 701. https://doi.org/10.1016/j.csda.2007.03.013
  • Clement, P., & Desch, W. (2007). An elementary proof of the triangle inequality for the Wasserstein metric. Proceedings Of The American Mathematical Society, 136(1), 333 339. https://doi.org/10.1090/s0002-9939-07-09020-x
  • Desbois, D. (2023). Coûts spécifiques et marges brutes du blé en Europe : Pratiques innovantes d’estimations pour l’analyse des changements d’échelles régionaux ou structurels. NOV’AE, (8), 1–23. https://doi.org/10.17180/novae-2023-NO-art08
  • Desbois, D., Butault, J., & Surry, Y. (2013). Estimation des coûts de production en phytosanitaires pour les grandes cultures. Une approche par la régression quantile. Économie Rurale, 333, 27 49. https://doi.org/10.4000/economierurale.3857
  • Desbois, D., Butault, J., & Surry, Y. (2017). Distribution des coûts spécifiques de production dans l’agriculture de l’Union européenne : une approche reposant sur la régression quantile. Économie Rurale, 361, 3 22. https://doi.org/10.4000/economierurale.5320
  • Desgraupes, B. (2017). Clustering indices [Vignette R]. CRAN. https://cran.r-project.org/web/packages/clusterCrit/vignettes/clusterCrit.pdf
  • Diday, E. (1971). Une nouvelle méthode en classification automatique et reconnaissance des formes la méthode des nuées dynamiques. https://www.numdam.org/item/RSA_1971__19_2_19_0/
  • Divay, J.-F., & Meunier, F. (1980). Deux méthodes de confection du tableau entrées-sorties. Annales de l’Insee, 37, 59–108.
  • Dvoretzky, A., Kiefer, J., & Wolfowitz, J. (1956). Asymptotic Minimax Character of the Sample Distribution Function and of the Classical Multinomial Estimator. The Annals Of Mathematical Statistics, 27(3), 642 669. https://doi.org/10.1214/aoms/1177728174
  • Gabriel, K. R. (1971). The Biplot Graphic Display of Matrices with Application to Principal Component Analysis. Biometrika, 58(3), 453. https://doi.org/10.2307/2334381
  • Hall, P., & Sheather, S. J. (1988). On the Distribution of a Studentized Quantile. Journal Of The Royal Statistical Society Series B (Statistical Methodology), 50(3), 381-391. https://doi.org/10.1111/j.2517-6161.1988.tb01735.x
  • Haultfoeuille, X. d’, & Givord, P. (2014). La régression quantile en pratique. Economie et Statistique / Economics And Statistics, 471(1), 85 111. https://doi.org/10.3406/estat.2014.10484
  • He, X., & Hu, F. (2002). Markov Chain Marginal Bootstrap. Journal Of The American Statistical Association, 97(459), 783 795. https://doi.org/10.1198/016214502388618591
  • Irpino, A., Verde, R., & Lechevallier, Y. (2006). Dynamic clustering of histograms using Wasserstein metric. In A. Rizzi & M. Vichi (Eds.), COMPSTAT 2006: Proceedings in computational statistics (pp. 869–876). Physica-Verlag.
  • Khmaladze, E. V. (1982). Martingale Approach in the Theory of Goodness-of-Fit Tests. Theory Of Probability And Its Applications, 26(2), 240 257. https://doi.org/10.1137/1126027
  • Kocherginsky, M., He, X., & Mu, Y. (2005). Practical Confidence Intervals for Regression Quantiles. Journal Of Computational And Graphical Statistics, 14(1), 41 55. https://doi.org/10.1198/106186005x27563
  • Koenker, R., & Bassett, G. (1978). Regression quantiles. Econometrica, 46(1), 33-50. https://doi.org/10.2307/1913643
  • Koenker, R., & Bassett, G. (1982). Robust Tests for Heteroscedasticity Based on Regression Quantiles. Econometrica, 50(1), 43-61. https://doi.org/10.2307/1912528
  • Koenker, R., & Machado, J. A. F. (1999). Goodness of Fit and Related Inference Processes for Quantile Regression. Journal Of The American Statistical Association, 94(448), 1296 1310. https://doi.org/10.1080/01621459.1999.10473882
  • Koenker, R., & Xiao, Z. (2002). Inference on the Quantile Regression Process. Econometrica, 70(4), 1583 1612. https://doi.org/10.1111/1468-0262.00342
  • Koenker, R., & Zhao, Q. (1994). L -estimation for linear heteroscedastic models. Journal Of Nonparametric Statistics, 3(3 4), 223 235. https://doi.org/10.1080/10485259408832584$
  • Massart, P. (1990). The Tight Constant in the Dvoretzky-Kiefer-Wolfowitz Inequality. The Annals Of Probability, 18(3), 1269-1283. https://doi.org/10.1214/aop/1176990746
  • SAS Institute Inc. (2009). The QUANTREG procedure. In SAS/STAT 9.2 User’s Guide (2nd ed., Chapter 72, pp. 5352–5425). SAS Institute Inc.
  • Torgerson, W. S. (1958). Theory and methods of scaling. Wiley.
  • Verde, R., & Irpino, A. (2008). Comparing histogram data using a Mahalanobis–Wasserstein distance. In COMPSTAT 2008: Proceedings in computational statistics (pp. 77–89). Physica-Verlag.
  • Ward, J. H. (1963). Hierarchical Grouping to Optimize an Objective Function. Journal Of The American Statistical Association, 58(301), 236 244. https://doi.org/10.1080/01621459.1963.10500845

Authors


Dominique Desbois

dominique.desbois@inrae.fr

Affiliation : INRAE, AgroParisTech, UMR PSAE, 91123, Palaiseau, France

Country : France

Attachments

No supporting information for this article

Article statistics

Views: 9