2000RePEc: Research Papers in EconomicsOpen access

On dispersion preserving estimation of the mean of a binary variable from small areas

Li‐Chun Zhang

Open full text 32 citations

Abstract

Abstract:\nOver-shrinkage is a common problem in small area (or domain) estimation. It happens when the estimated small-area parameters have less between-area variation than their true values. To deal with this problem, Louis (1984), Ghosh (1992) and Spjøtvoll and Thomsen (1987) have proposed various constrained empirical and hierarchical Bayes methods. In this paper we study two non-Bayesian methods based on, respectively, the synthetic estimator and a variance-component model. We show first that the synthetic estimator entails loss of dispersion in general, from which it follows that the coverage level of the confidence intervals could be far below the nominal level of confidence, when these are derived from the sampling error alone. A bivariate variance-component model at the area-level, as well as its simplification, can greatly improve the efficiency of the confidence intervals. However, super-population approaches as such are unable to capture the distribution of the true area-parameters. We develop a finite-population approach based on an empirical finite-population distribution function of the area-parameters, which provides the necessary adjustment. The various methods will be illustrated using the data of the Census 1990. Finally, we notice that several European countries will base the upcoming Census on their administrative register systems, instead of collecting the information in the field. Improved small area estimation methods may prove to be valuable for assessing the quality of such Register Counting. \nKeywords: Over-shrinkage, synthetic estimator, variance-component model

Open-access reader

About this research paper

What this paper is about

Abstract:\nOver-shrinkage is a common problem in small area (or domain) estimation. It happens when the estimated small-area parameters have less between-area variation than their true values. To deal with this problem, Louis (1984), Ghosh (1992) and Spjøtvoll and Thomsen (1987) have proposed various constrained empirical and hierarchical Bayes methods. In this paper we study two non-Bayesian methods based on, respectively, the synthetic estimator and a variance-component model. We show first that the synthetic estimator entails loss of dispersion in general, from which it follows that the coverage level of the confidence intervals could be far below the nominal level of confidence, when these are derived from the sampling error alone. A bivariate variance-component model at the area-level, as well as its simplification, can greatly improve the efficiency of the confidence intervals. However, super-population approaches as such are unable to capture the distribution of the true area-parameters. We develop a finite-population approach based on an empirical finite-population distribution function of the area-parameters, which provides the necessary adjustment. The various methods will be illustrated using the data of the Census 1990. Finally, we notice that several European countries will base the upcoming Census on their administrative register systems, instead of collecting the information in the field. Improved small area estimation methods may prove to be valuable for assessing the quality of such Register Counting. \nKeywords: Over-shrinkage, synthetic estimator, variance-component model

Why it matters

OpenAlex reports 32 citations for this work. Citation counts describe recorded attention and do not establish research quality.

Key contribution

A contribution statement is not available in the OpenAlex record.

Method / approach

Method details are not available in the OpenAlex metadata.

Main findings

Findings are not separately available in the OpenAlex metadata.

Limitations

Limitations are not available in the OpenAlex metadata.

Applications

Application details are not available in the OpenAlex metadata.

Available abstract

Abstract:\nOver-shrinkage is a common problem in small area (or domain) estimation. It happens when the estimated small-area parameters have less between-area variation than their true values. To deal with this problem, Louis (1984), Ghosh (1992) and Spjøtvoll and Thomsen (1987) have proposed various constrained empirical and hierarchical Bayes methods. In this paper we study two non-Bayesian methods based on, respectively, the synthetic estimator and a variance-component model. We show first that the synthetic estimator entails loss of dispersion in general, from which it follows that the coverage level of the confidence intervals could be far below the nominal level of confidence, when these are derived from the sampling error alone. A bivariate variance-component model at the area-level, as well as its simplification, can greatly improve the efficiency of the confidence intervals. However, super-population approaches as such are unable to capture the distribution of the true area-parameters. We develop a finite-population approach based on an empirical finite-population distribution function of the area-parameters, which provides the necessary adjustment. The various methods will be illustrated using the data of the Census 1990. Finally, we notice that several European countries will base the upcoming Census on their administrative register systems, instead of collecting the information in the field. Improved small area estimation methods may prove to be valuable for assessing the quality of such Register Counting. \nKeywords: Over-shrinkage, synthetic estimator, variance-component model

Key concepts: Small area estimation, Statistics, Estimator, Population, Mathematics, Econometrics, Variance (accounting), Computer science

Related papers

Back to paper searchBrowse research topicsOriginal source
On dispersion preserving estimation of the mean of a binary variable from small areas — Research Paper | ScholarLens