Weighting and estimation
Filter results by
Search HelpKeyword(s)
Type
Survey or statistical program
- Survey of Labour and Income Dynamics (5)
- Census of Population (5)
- Survey of Household Spending (2)
- Longitudinal and International Study of Adults (2)
- Survey of Employment, Payrolls and Hours (1)
- Canadian Cancer Registry (1)
- Canadian Community Health Survey - Annual Component (1)
- Uniform Crime Reporting Survey (1)
- Quarterly Demographic Estimates (1)
- Annual Demographic Estimates: Canada, Provinces and Territories (1)
- Estimates of the number of census families for July 1st, Canada, provinces and territories (1)
- Annual Demographic Estimates : Subprovincial Areas (1)
- Labour Force Survey (1)
- Longitudinal Administrative Databank (1)
- General Social Survey - Social Identity (1)
- Canadian Community Health Survey - Nutrition (1)
- Canadian Income Survey (1)
- Residential Property Values (1)
- Canadian Survey on Business Conditions (1)
Results
All (638)
All (638) (600 to 610 of 638 results)
- Articles and reports: 12-001-X198400214357Description:
A finite population of size N is supposed to contain M (unknown) units of a specified category A (say) constituting a domain with mean \mu. A procedure which involves drawing units using simple random sampling without replacement till a preassigned number of members of the domain is reached is proposed. An unbiased estimator of \mu is also derived. This is seen to be superior to the corresponding possibly biased estimator based on a comparable SRSWOR scheme with a fixed number of draws. The proposed scheme is also shown to admit unbiased estimators of M and the domain total T.
Release date: 1984-12-14 - Articles and reports: 12-001-X198300214342Description:
This study considers the suitability of composite estimation techniques for the Canadian Labour Force Survey. The performance of a class of AK composite estimators introduced initially by Gurney and Daly is investigated for several characteristics. While the ordinary composite estimate has a large bias, the AK composite estimate is capable of reducing the bias. Composite estimates having minimum variance and minimum mean square error are compared.
Release date: 1983-12-15 - Articles and reports: 12-001-X198300214344Description:
In order to improve the timeliness, accuracy and consistency of population estimates for different geographic areas, Statistics Canada has developed new methods of estimation for sub-provincial areas (census divisions and census metropolitan areas). Beginning with 1982, two sets of population estimates (regression and component based) will be published yearly, appearing 3-4 months and 12-15 months, respectively, from the reference date.
The regression technique uses family allowance recipients as the main symptomatic indicator and where available, additional indicators - reference population from provincial health insurance files and hydro accounts - to derive population change for the current year. The first set is obtained by adding this change to the second set for the previous year produced by the component method, with births and deaths from vital registers, and estimated migration from Revenue Canada taxation files. The two sets were found to be statistically similar with respect to accuracy, though the first set is more timely, and the second provides more details on the components of population change.
Release date: 1983-12-15 - 604. The methodology of the Canadian Air Scheduled International Passenger Origin and Destination estimation system ArchivedArticles and reports: 12-001-X198300114333Description:
The Air Scheduled International Passenger Origin and Destination (ASIPOD) estimation system uses the data from two air traffic surveys to produce origin-destination estimates of international passengers. The “assignment technique” is the solution to the problem caused by the non-coverage of non-interlining traffic. The assumptions of the technique are sufficiently questionable to warrant an evaluation of the bias of the estimates. However, major improvements will be made in the new system which will decrease the bias in the estimates. Also, estimates of reliability will be produced. And as a result, knowledge of the strength of the inferences made with respect to air traffic markets from these estimates will be improved in international bilateral air negotiations.
Release date: 1983-06-15 - Articles and reports: 12-001-X198300114335Description:
The Canadian Labour Force Survey is a household survey conducted each month for the purpose of producing point-in-time estimates of the number of persons employed, unemployed and not in the labor force. The survey has a rotating panel design in which all individuals in a sampled household location are interviewed each month, for six consecutive months. In the past, little use has been made of this longitudinal structure, although considerable interest has been expressed in the month-to-month gross flows (transitions) amongst the labour force status categories. In this paper we discuss methods being considered by Statistics Canada for the production of gross flow estimates, but from a model-based perspective.
Release date: 1983-06-15 - Articles and reports: 12-001-X198300114336Description:
The peach, sour cherry and the grape objective yield surveys have been carried out annually in the Niagara Peninsula since 1964 in order to forecast the magnitude of change in marketable fruit production from the previous year. Timeliness of the estimates is essential in order to enable the Ontario Tender Fruit Growers Marketing Board (OTFGMB) and the Ontario Grape Growers Marketing Board (OGGMB) to establish the marketing strategies well ahead of the harvest. This paper summarizes the major changes due to the second redesign initiated in 1982. In particular, the sample design, data collection operation and modifications of the estimation procedures are elaborated upon.
Release date: 1983-06-15 - 607. A timely and accurate potato acreage estimate from Landsat: Results of a demonstration ArchivedArticles and reports: 12-001-X198300114337Description:
This paper describes the procedures used and results of a joint Canada Centre for Remote Sensing (CCRS) and Statistics Canada project to provide a timely potato acreage estimate for New Brunswick, a major potato producing province in Canada. The project has demonstrated that satellite imagery combined with more traditional potato area estimation procedures can lower respondent burden, produce timely crop distribution maps and produce reliable estimates for subregions.
Release date: 1983-06-15 - 608. Sampling on two occasions with probabilities proportional to size without replacement (PPSWOR) ArchivedArticles and reports: 12-001-X198300114340Description:
A theory of sampling on two occasions with unequal probabilities and without replacement is presented. Fellegi’s (1963) method, which yields the same selection probabilities for a given unit on each occasion, is used to select the units for the rotation sample. The variances of composite estimators of the population total on the second occasion are developed. Numerical results are presented for small sample sizes and efficiency comparisons are made with a competing strategy.
Release date: 1983-06-15 - Articles and reports: 12-001-X198200114328Description:
Estimates from sample surveys are sometimes required for domains whose boundaries do not coincide with those of design strata. Taking the Canadian Labour Force Survey as an example of a survey utilizing a clustered sample design, some alternative small area estimation techniques available in the literature are evaluated empirically including synthetic, domain (simple and post-stratified) and composite estimators which are linear combinations of synthetic and post-stratified domain estimators. A sample dependent estimator which attaches weight to the post-stratified domain estimate depending on the amount of sample in the domain is proposed and its performance is also evaluated.
Release date: 1982-06-15 - 610. Computerization of complex survey estimates ArchivedArticles and reports: 12-001-X198200114331Description:
Survey data collected by statistical agencies is most likely to be processed through to the tabulation stage by these agencies. The computer programs associated with this processing are also most likely tailored to the particular design and variables used. The statistics computed from such surveys typically range from simple descriptive totals and means to these required for analytic studies such as comparison of domains, regression analysis and contingency tables analysis. This paper describes a computer program which computes these statistics and their associated sampling errors for commonly used sampling designs.
Release date: 1982-06-15
- Previous Go to previous page of All results
- 1 Go to page 1 of All results
- ...
- 58 Go to page 58 of All results
- 59 Go to page 59 of All results
- 60 Go to page 60 of All results
- 61 (current) Go to page 61 of All results
- 62 Go to page 62 of All results
- 63 Go to page 63 of All results
- 64 Go to page 64 of All results
- Next Go to next page of All results
Data (0)
Data (0) (0 results)
No content available at this time.
Analysis (610)
Analysis (610) (50 to 60 of 610 results)
- Articles and reports: 11-522-X202200100015Description: We present design-based Horvitz-Thompson and multiplicity estimators of the population size, as well as of the total and mean of a response variable associated with the elements of a hidden population to be used with the link-tracing sampling variant proposed by Félix-Medina and Thompson (2004). Since the computation of the estimators requires to know the inclusion probabilities of the sampled people, but they are unknown, we propose a Bayesian model which allows us to estimate them, and consequently to compute the estimators of the population parameters. The results of a small numeric study indicate that the performance of the proposed estimators is acceptable.Release date: 2024-03-25
- Articles and reports: 11-522-X202200100018Description: The Longitudinal Social Data Development Program (LSDDP) is a social data integration approach aimed at providing longitudinal analytical opportunities without imposing additional burden on respondents. The LSDDP uses a multitude of signals from different data sources for the same individual, which helps to better understand their interactions and track changes over time. This article looks at how the ethnicity status of people in Canada can be estimated at the most detailed disaggregated level possible using the results from a variety of business rules applied to linked data and to the LSDDP denominator. It will then show how improvements were obtained using machine learning methods, such as decision trees and random forest techniques.Release date: 2024-03-25
- Articles and reports: 12-001-X202300200002Description: Being able to quantify the accuracy (bias, variance) of published output is crucial in official statistics. Output in official statistics is nearly always divided into subpopulations according to some classification variable, such as mean income by categories of educational level. Such output is also referred to as domain statistics. In the current paper, we limit ourselves to binary classification variables. In practice, misclassifications occur and these contribute to the bias and variance of domain statistics. Existing analytical and numerical methods to estimate this effect have two disadvantages. The first disadvantage is that they require that the misclassification probabilities are known beforehand and the second is that the bias and variance estimates are biased themselves. In the current paper we present a new method, a Gaussian mixture model estimated by an Expectation-Maximisation (EM) algorithm combined with a bootstrap, referred to as the EM bootstrap method. This new method does not require that the misclassification probabilities are known beforehand, although it is more efficient when a small audit sample is used that yields a starting value for the misclassification probabilities in the EM algorithm. We compared the performance of the new method with currently available numerical methods: the bootstrap method and the SIMEX method. Previous research has shown that for non-linear parameters the bootstrap outperforms the analytical expressions. For nearly all conditions tested, the bias and variance estimates that are obtained by the EM bootstrap method are closer to their true values than those obtained by the bootstrap and SIMEX methods. We end this paper by discussing the results and possible future extensions of the method.Release date: 2024-01-03
- Articles and reports: 12-001-X202300200003Description: We investigate small area prediction of general parameters based on two models for unit-level counts. We construct predictors of parameters, such as quartiles, that may be nonlinear functions of the model response variable. We first develop a procedure to construct empirical best predictors and mean square error estimators of general parameters under a unit-level gamma-Poisson model. We then use a sampling importance resampling algorithm to develop predictors for a generalized linear mixed model (GLMM) with a Poisson response distribution. We compare the two models through simulation and an analysis of data from the Iowa Seat-Belt Use Survey.Release date: 2024-01-03
- Articles and reports: 12-001-X202300200004Description: We present a novel methodology to benchmark county-level estimates of crop area totals to a preset state total subject to inequality constraints and random variances in the Fay-Herriot model. For planted area of the National Agricultural Statistics Service (NASS), an agency of the United States Department of Agriculture (USDA), it is necessary to incorporate the constraint that the estimated totals, derived from survey and other auxiliary data, are no smaller than administrative planted area totals prerecorded by other USDA agencies except NASS. These administrative totals are treated as fixed and known, and this additional coherence requirement adds to the complexity of benchmarking the county-level estimates. A fully Bayesian analysis of the Fay-Herriot model offers an appealing way to incorporate the inequality and benchmarking constraints, and to quantify the resulting uncertainties, but sampling from the posterior densities involves difficult integration, and reasonable approximations must be made. First, we describe a single-shrinkage model, shrinking the means while the variances are assumed known. Second, we extend this model to accommodate double shrinkage, borrowing strength across means and variances. This extended model has two sources of extra variation, but because we are shrinking both means and variances, it is expected that this second model should perform better in terms of goodness of fit (reliability) and possibly precision. The computations are challenging for both models, which are applied to simulated data sets with properties resembling the Illinois corn crop.Release date: 2024-01-03
- Articles and reports: 12-001-X202300200012Description: In recent decades, many different uses of auxiliary information have enriched survey sampling theory and practice. Jean-Claude Deville contributed significantly to this progress. My comments trace some of the steps on the way to one important theory for the use of auxiliary information: Estimation by calibration.Release date: 2024-01-03
- Articles and reports: 12-001-X202300200013Description: Jean-Claude Deville is one of the most prominent researcher in survey sampling theory and practice. His research on balanced sampling, indirect sampling and calibration in particular is internationally recognized and widely used in official statistics. He was also a pioneer in the field of functional data analysis. This discussion gives us the opportunity to recognize the immense work he has accomplished, and to pay tribute to him. In the first part of this article, we recall briefly his contribution to the functional principal analysis. We also detail some recent extension of his work at the intersection of the fields of functional data analysis and survey sampling. In the second part of this paper, we present some extension of Jean-Claude’s work in indirect sampling. These extensions are motivated by concrete applications and illustrate Jean-Claude’s influence on our work as researchers.Release date: 2024-01-03
- Articles and reports: 12-001-X202300200014Description: Many things have been written about Jean-Claude Deville in tributes from the statistical community (see Tillé, 2022a; Tillé, 2022b; Christine, 2022; Ardilly, 2022; and Matei, 2022) and from the École nationale de la statistique et de l’administration économique (ENSAE) and the Société française de statistique. Pascal Ardilly, David Haziza, Pierre Lavallée and Yves Tillé provide an in-depth look at Jean-Claude Deville’s contributions to survey theory. To pay tribute to him, I would like to discuss Jean-Claude Deville’s contribution to the more day-to-day application of methodology for all the statisticians at the Institut national de la statistique et des études économiques (INSEE) and at the public statistics service. To do this, I will use my work experience, and particularly the four years (1992 to 1996) I spent working with him in the Statistical Methods Unit and the discussions we had thereafter, especially in the 2000s on the rolling census.Release date: 2024-01-03
- Articles and reports: 12-001-X202300200015Description: This article discusses and provides comments on the Ardilly, Haziza, Lavallée and Tillé’s summary presentation of Jean-Claude Deville’s work on survey theory. It sheds light on the context, applications and uses of his findings, and shows how these have become engrained in the role of statisticians, in which Jean-Claude was a trailblazer. It also discusses other aspects of his career and his creative inventions.Release date: 2024-01-03
- Articles and reports: 12-001-X202300200016Description: In this discussion, I will present some additional aspects of three major areas of survey theory developed or studied by Jean-Claude Deville: calibration, balanced sampling and the generalized weight-share method.Release date: 2024-01-03
- Previous Go to previous page of Analysis results
- 1 Go to page 1 of Analysis results
- 2 Go to page 2 of Analysis results
- 3 Go to page 3 of Analysis results
- 4 Go to page 4 of Analysis results
- 5 Go to page 5 of Analysis results
- 6 (current) Go to page 6 of Analysis results
- 7 Go to page 7 of Analysis results
- ...
- 61 Go to page 61 of Analysis results
- Next Go to next page of Analysis results
Reference (28)
Reference (28) (0 to 10 of 28 results)
- Surveys and statistical programs – Documentation: 11-633-X2026002Description: Recent changes in Canada’s immigration levels have heightened interest in understanding how immigration affects housing demand. This article develops a methodological framework for projecting housing use associated with permanent residents (PRs) and non-permanent residents (NPRs) under alternative immigration scenarios. The framework applies observed per capita housing use rates from the Census of Population to estimate incremental housing use by tenure over time.Release date: 2026-04-24
- Surveys and statistical programs – Documentation: 91-528-XDescription: The Technical Guide on Demographic Estimates at Statistics Canada provides detailed descriptions of the most current data sources and methods used by the Centre for demography at Statistics Canada to produce demographic estimates as part of the Demographic estimates program. They comprise postcensal and intercensal population estimates; base population; births and deaths; immigrants; emigrants; returning emigrants; non-permanent residents; interprovincial migration; subprovincial estimates of population and intraprovincial migration; population estimates by age and gender; and census family estimates. A glossary of commonly used terms is available at the end of the guide.Release date: 2025-12-17
- Surveys and statistical programs – Documentation: 98-306-XDescription:
This report describes sampling, weighting and estimation procedures used in the Census of Population. It provides operational and theoretical justifications for them, and presents the results of the evaluations of these procedures.
Release date: 2023-10-04 - Notices and consultations: 75F0002M2019006Description:
In 2018, Statistics Canada released two new data tables with estimates of effective tax and transfer rates for individual tax filers and census families. These estimates are derived from the Longitudinal Administrative Databank. This publication provides a detailed description of the methods used to derive the estimates of effective tax and transfer rates.
Release date: 2019-04-16 - 5. Revisions to 2006 to 2011 income data ArchivedSurveys and statistical programs – Documentation: 75F0002M2015003Description:
This note discusses revised income estimates from the Survey of Labour and Income Dynamics (SLID). These revisions to the SLID estimates make it possible to compare results from the Canadian Income Survey (CIS) to earlier years. The revisions address the issue of methodology differences between SLID and CIS.
Release date: 2015-12-17 - Surveys and statistical programs – Documentation: 13-605-X201500414166Description:
Estimates of the underground economy by province and territory for the period 2007 to 2012 are now available for the first time. The objective of this technical note is to explain how the methodology employed to derive upper-bound estimates of the underground economy for the provinces and territories differs from that used to derive national estimates.
Release date: 2015-04-29 - Surveys and statistical programs – Documentation: 99-002-X2011001Description:
This report describes sampling and weighting procedures used in the 2011 National Household Survey. It provides operational and theoretical justifications for them, and presents the results of the evaluation studies of these procedures.
Release date: 2015-01-28 - Surveys and statistical programs – Documentation: 99-002-XDescription: This report describes sampling and weighting procedures used in the 2011 National Household Survey. It provides operational and theoretical justifications for them, and presents the results of the evaluation studies of these procedures.Release date: 2015-01-28
- Surveys and statistical programs – Documentation: 92-568-XDescription:
This report describes sampling and weighting procedures used in the 2006 Census. It reviews the history of these procedures in Canadian censuses, provides operational and theoretical justifications for them, and presents the results of the evaluation studies of these procedures.
Release date: 2009-08-11 - Surveys and statistical programs – Documentation: 71F0031X2006003Description:
This paper introduces and explains modifications made to the Labour Force Survey estimates in January 2006. Some of these modifications include changes to the population estimates, improvements to the public and private sector estimates and historical updates to several small Census Agglomerations (CA).
Release date: 2006-01-25