Treatment of sample under-representation and skewed heavy-tailed distributions in survey-based microsimulation: An analy

  • PDF / 887,914 Bytes
  • 38 Pages / 439.37 x 666.142 pts Page_size
  • 3 Downloads / 178 Views

DOWNLOAD

REPORT


Treatment of sample under-representation and skewed heavy-tailed distributions in survey-based microsimulation: An analysis of redistribution effects in compulsory health care insurance in Switzerland Tobias Schoch

· André Müller

Received: 30 December 2019 / Accepted: 5 August 2020 © The Author(s) 2020

Abstract The credibility of microsimulation modeling with the research community and policymakers depends on high-quality baseline surveys. Quality problems with the baseline survey tend to impair the quality of microsimulation built on top of the survey data. We address two potential issues that both relate to skewed and heavytailed distributions. First, we find that ultra-high-income households are under-represented in the baseline household survey. Moreover, the sample estimate of average income underestimates the known population average. Although the Deville–Särndal calibration method corrects the under-representation, it cannot achieve alignment of estimated average income in the right tail of the distribution with known population values without distorting the empirical income distribution. To overcome the problem, we introduce a Pareto tail model. With the help of the tail model, we can adjust the sample income distribution in the tail to meet the alignment targets. Our method can be a useful tool for microsimulation modelers working with survey income data. The second contribution refers to the treatment of an outlier-prone variable that has been added to the survey by record linkage (our empirical example is health care cost). The nature of the baseline survey is not affected by record linkage, that is, the baseline survey still covers only a small part of the population. Hence, the sampling weights are relatively large. An outlying observation together with a high sampling weight can heavily influence or even ruin an estimate of a population characteristic. Thus, we argue that it is beneficial—in terms of mean square error—to use robust T. Schoch () School of Business, Institute ICC, University of Applied Sciences Northwestern Switzerland, Riggenbachstrasse 16, 4600 Olten, Switzerland E-Mail: [email protected] A. Müller Ecoplan AG – Research in Economics and Policy Consultancy, Monbijoustrasse 14, 3011 Bern, Switzerland

K

T. Schoch, A. Müller

estimation and alignment methods, because robust methods are less affected by the presence of outliers. Keywords Simulation · Pareto distribution · Representative outliers · Nonresponse · Calibration · Imputation JEL classification C15 · C54 · C63 · C83 · I13 · I18

Methoden zur Behandlung von Unterrepräsentation bei schiefen Verteilungen für die stichprobenbasierte Mikrosimulation: Eine Analyse zu den Umverteilungseffekten in der obligatorischen Krankenversicherung der Schweiz Zusammenfassung Eine qualitativ hochstehende Stichprobenerhebungen ist eine wesentliche Voraussetzung, um gültige und zuverlässige Aussagen mit stichprobenbasierten Mikrosimulationsstudien zu tätigen. Sind die stichprobenbasierten Schätzer massiv verzerrt (z. B. infolge von Antwortau