Survey design
Filter results by
Search HelpKeyword(s)
Type
Survey or statistical program
- Survey of Labour and Income Dynamics (6)
- Census of Population (4)
- Canadian Community Health Survey - Annual Component (2)
- Labour Force Survey (2)
- General Social Survey - Social Identity (2)
- Canadian Income Survey (2)
- National Gross Domestic Product by Income and by Expenditure Accounts (1)
- National Population Health Survey: Household Component, Longitudinal (1)
- Uniform Crime Reporting Survey (1)
- Adult Education and Training Survey (1)
- Annual Income Estimates for Census Families and Individuals (T1 Family File) (1)
- Biotechnology Use and Development Survey (1)
- General Social Survey - Family (1)
- General Social Survey - Caregiving and Care Receiving (1)
- Longitudinal and International Study of Adults (1)
Results
All (334)
All (334) (0 to 10 of 334 results)
- Articles and reports: 12-001-X202600100004Description: We test the notion that a quasi-probabilistic method of selecting individuals within households (last birthday, LB) draws in a different sample compared to a non-probabilistic approach that selects respondents according to known parameters on age and gender (frequency matching, FM). With data from an original field experiment, we evaluate fieldwork efficiency (time and completed cases), economy (cost), success in recruiting a representative sample, and differences across a set of attitudinal and behavioral measures. We find that the FM approach performs better on efficiency and cost and achieves a comparable sample; importantly, this comparability extends across measures of personality traits and public opinion. With appropriate caveats, we conclude that researchers’ choice of selection methods should be guided by both theoretical benefits and practical tradeoffs.Release date: 2026-06-29
- Articles and reports: 12-001-X202600100006Description: We introduce a general framework for constructing master samples that preserve desirable design properties across panels. The core procedure is to order an initial probability sample. Since the final sequence must be robust to a uniform random rotation, we define and minimize an objective that aggregates panel-level performance across all possible circular panels. A final random rotation is applied to ensure design validity. The framework is flexible with respect to the choice of design criteria, such as spatial balance or marginal balance, and can be implemented efficiently using simulated annealing to obtain high-quality approximate solutions. By construction, the approach supports both positive and negative sample coordination for spatially balanced, marginally balanced, and doubly balanced samples. The method’s versatility is demonstrated through three applications: constructing a master sample with spatially balanced panels, marginally balanced panels, and doubly balanced panels.Release date: 2026-06-29
- Articles and reports: 12-001-X202600100007Description: National statistical institutes operate sample coordination systems to spread the response burden in business surveys. Despite the applied sample coordination and monitoring the response burden, some businesses might still be heavily sampled within a short period. This may lead to a peaking response burden for individual businesses, which could affect response rates and response quality. This paper proposes a new sample coordination method based on Adapted Spatially Correlated Poisson (ASCP) sampling that focuses on businesses with a high response burden. The effects on the response burden will be evaluated in two simulation studies and compared with a stratified approach, a pragmatic method in which sampling fractions are manually adjusted and with the baseline method of ignoring the response burden. For the simulations, real-world scenarios and data from Statistics Netherlands are used. The first simulation study considers a practical situation in which a given sample is adjusted with the aim to avoid the occurrence of businesses with a peaking response burden. The second simulation study analyzes the longer-term effects of the different sample coordination methods and focuses both on the reduction and spread of the response burden. The advantages and disadvantages of the different methods will be explained and discussed in detail, and recommendations for applying these methods at national statistical institutes and other survey agencies will be given.Release date: 2026-06-29
- Articles and reports: 12-001-X202600100008Description: This paper introduces an innovative and intuitive finite population sampling method that has been developed using a unique graphical framework. In this approach, first-order inclusion probabilities are represented as bars on a two-dimensional graph. By manipulating the positions of these bars, researchers can create a wide range of different sampling designs. This graphical visualization of sampling designs facilitates the exploration of alternative designs and may simplify certain aspects of the implementation compared to traditional mathematical algorithms. This novel approach holds significant promise for tackling complex challenges in sampling, such as achieving an optimal design. By applying a version of the greedy best-first search algorithm to this graphical approach, the potential for integrating intelligent algorithms into finite population sampling is demonstrated.Release date: 2026-06-29
- Journals and periodicals: 75F0002MDescription: This series provides detailed documentation on income developments, including survey design issues, data quality evaluation and exploratory research.Release date: 2026-05-20
- Articles and reports: 12-001-X202500200013Description: This article examines the methodological complexities associated with the design of business surveys, with particular emphasis on sampling strategies implemented by National Statistical Offices (NSOs). It addresses the inherent challenges posed by the dynamic nature of the business population, which necessitates continual updates to the sampling frame to ensure representativeness and relevance. Critical design considerations include the determination of optimal sample sizes, stratification across key dimensions such as industry, geographic region, and enterprise size, as well as the treatment of business births and the exclusion of inactive (or “dead”) units. The article applies Bankier’s (1988) power allocation method to a two-way stratification scheme defined by industry and geography, evaluating its performance by comparing the resulting coefficients of variation with those obtained via a raking algorithm applied to the marginal coefficients. Furthermore, the approach is extended to a multivariate context to accommodate multiple estimation domains. The discussion also encompasses practical issues related to sample rotation and coordination, which are critical for maintaining data quality and minimizing respondent burden over time.Release date: 2025-12-23
- Articles and reports: 75-005-M2025001Description: Since 2010, engaging Canadians to participate in the LFS has become more challenging due to a variety of social and technological changes. The decline in the LFS response rate accelerated in 2020, exacerbated by public health measures during the COVID-19 pandemic. This technical paper presents preliminary results of two collection initiatives implemented using an online first strategy to improve the LFS response rates by confirming respondent contact information and expanding the availability of online response. Through these and other planned initiatives, Statistics Canada is working to ensure that the LFS estimates continue to provide an accurate and representative portrait of the Canadian labour market.Release date: 2025-10-21
- Articles and reports: 11-522-X202500100004Description: The Survey of Household Spending (SHS) conducted by Statistics Canada collects paper diaries and shopping receipts as a source of household expenditure data. An auto-capturing algorithm was created for SHS 2023 to reduce statistical clerks' manual work of extracting important information from scanned receipts of common store brands. The algorithm used Tesseract optical character recognition (OCR) to extract text characters from images of receipts, and it identified store and product entities using regular expressions, also known as regex. The goal of this study was to enhance the current auto-capture algorithm by experimenting with more advanced OCR and machine learning methods. As a result, PaddleOCR, an open-source OCR toolkit, was selected as the new default OCR engine due to its overall performance in recognizing texts, especially digits, accurately across receipts of various qualities. Additionally, entity classifiers based on support vector machines were trained on historical SHS records and existing regex patterns. By using classifiers to categorize different elements present on receipts instead of relying solely on regex patterns, product and store recognition improved. It is expected that this new algorithm will be used for SHS 2025 to improve the auto-capture quality and reduce the manual burden associated with capturing receipt variables.Release date: 2025-09-08
- 9. Data-driven Imputation Strategies and their Associated Quality Indicators in Economic Surveys ArchivedArticles and reports: 11-522-X202500100011Description: The use of modern "data"-driven imputation methods to treat non-response in the context of surveys processed in the Integrated Business Statistics Program at Statistics Canada has previously been explored. It was observed that these methods can lead to high quality imputation and further have the potential to result in broad efficiencies when setting up a particular survey's edit and imputation strategy. However, estimation of the associated total variance, more specifically the component due to imputation, remains a challenge. In this article, two methods for estimation of total variance are proposed and show preliminary results that have motivated us to pursue further research in this area.Release date: 2025-09-08
- Articles and reports: 11-522-X202500100029Description: J.N.K. Rao has contributed to almost every subdiscipline of survey research, including unequal-probability and two-phase sampling, variance estimation, regression and categorical data analysis, small area estimation, and data integration. For each of these topics, Rao's work anticipated and led future research directions. His contributions will be discussed in the context of broader research trends as seen in the articles of Survey Methodology over the journal's 50-year history.Release date: 2025-09-08
- Previous Go to previous page of All results
- 1 (current) Go to page 1 of All results
- 2 Go to page 2 of All results
- 3 Go to page 3 of All results
- 4 Go to page 4 of All results
- 5 Go to page 5 of All results
- 6 Go to page 6 of All results
- 7 Go to page 7 of All results
- ...
- 34 Go to page 34 of All results
- Next Go to next page of All results
Data (0)
Data (0) (0 results)
No content available at this time.
Analysis (305)
Analysis (305) (10 to 20 of 305 results)
- 11. Contributions of J.N.K. Rao to Complex Survey Multilevel Models and Composite Likelihood ArchivedArticles and reports: 11-522-X202500100030Description: In the setting of multilevel models to be estimated using data from surveys with complex sampling designs, this paper outlines some contributions of the landmark paper by Rao, Verret and Hidiroglou (Survey Methodology, 2013) and subsequent related work.Release date: 2025-09-08
- Articles and reports: 11-522-X202500100032Description: Although non-probability data sources are not new to official statistics, a revived interest in the topic has emerged from pressures due to falling survey response rates, increasing data collection costs and a desire to take advantage of new data source opportunities from the ongoing societal digitalisation. Due to the exclusion of certain segments of the target population, inference derived solely from a non-probability data source is likely to result in bias. This work approaches the challenge of addressing the bias by integrating non-probability data with reference probability samples. The focus will be on methods to model the propensity of inclusion in the non-probability dataset with the help of the accompanying reference sample, with the modelled propensities then applied in an inverse probability weighting approach to produce population estimates. The reference sample is sometimes assumed as given. In this presentation however, an objective of finding an optimal strategy will be pursued that is, the combination of a data integration-based estimator and sample design for the reference probability sample. Recent work is discussed in which advantage is taken of the good unit identification possibilities in business surveys to study an estimator based on propensities and derive optimal (unequal) selection probabilities for the reference sample.Release date: 2025-09-08
- 13. Including Non-binary Gender in the Calibration Strategy for the Canadian Long-Form Sample Survey Weights ArchivedArticles and reports: 11-522-X202500100033Description: Aligning with recent needs for increased disaggregated data, in 2021 Canada became the first country to collect and disseminate data on gender diversity in a national census giving Canadians the option to select male, female, or non-binary. Due to their small size, non-binary population counts were not used in the 2021 Census long-form sample calibration procedure due to the risk of increasing the variance of estimates. This paper presents an alternative long-form calibration strategy which allows for small populations, such as the non-binary group, to be incorporated while mitigating methodological concerns. The strategy put forward can incorporate multiple small populations simultaneously while also being flexible enough to fit the calibration systems of other National Statistical Offices (NSOs). The results of a Monte Carlo (MC) simulation are presented showing improved data quality for the non-binary population under the alternative calibration strategy.Release date: 2025-09-08
- Articles and reports: 12-001-X202500100010Description: The discussants highlight promising research topics for improving the quality and granularity of estimates from surveys. We agree that continued research is needed to evaluate models used for inference, and suggest development of measures of model dependence.Release date: 2025-06-30
- Articles and reports: 12-001-X202500100011Description: This discussion examines some advancements in survey design and estimation, inspired by the comprehensive appraisal of Professors Jon Rao and Sharon Lohr on current trends in the field. It delves into three specific areas: balanced sampling, calibration, and small area estimation. Probabilistic balanced sampling methods, such as the cube method and penalized balanced sampling, are explored, with an emphasis on addressing emerging challenges, including extensions to linear mixed models, nonparametric regression models, and spatially balanced designs. Calibration is discussed using a modular framework that incorporates modern regression techniques, and highlights innovative uses of model calibration for data editing and causal inference. Small area estimation is considered in the context of latent variable modeling and data integration, emphasizing its role when the variable(s) of interest cannot be measured either directly or without error. Applications in integrating probability and non-probability data and conducting causal analysis at local level are also discussed.Release date: 2025-06-30
- Articles and reports: 12-001-X202500100012Description: In this discussion, we complement the excellent overview by Profs. Lohr and Rao with some additional topics. The first topic is a call for more recognition of the central role of modeling in survey estimation. The second is a brief discussion of the use of partial frame information in survey design. Finally, we draw the attention to recent increases of synthetic methods, in particular, multilevel regression and poststratification (MRP) in small area estimation applications.Release date: 2025-06-30
- Articles and reports: 12-001-X202400200003Description: The optimum sample allocation in stratified sampling is one of the basic issues of survey methodology. It is a procedure of dividing the overall sample size into strata sample sizes in such a way that for given sampling designs in strata the variance of the stratified \pi estimator of the population total (or mean) for a given study variable assumes its minimum. In this work, we consider the optimum allocation of a sample, under lower and upper bounds imposed jointly on sample sizes in strata. We are concerned with the variance function of some generic form that, in particular, covers the case of the simple random sampling without replacement in strata. The goal of this paper is twofold. First, we establish (using the Karush-Kuhn-Tucker conditions) a generic form of the optimal solution, the so-called optimality conditions. Second, based on the established optimality conditions, we derive an efficient recursive algorithm, named RNABOX, which solves the allocation problem under study. The RNABOX can be viewed as a generalization of the classical recursive Neyman allocation algorithm, a popular tool for optimum allocation when only upper bounds are imposed on sample strata-sizes. We implement RNABOX in R as a part of our package stratallo which is available from the Comprehensive R Archive Network (CRAN) repository.Release date: 2024-12-20
- Articles and reports: 12-001-X202400200016Description: Joseph Waksberg was an important figure in survey statistics mainly through his applied work in the design of samples. He took a design-based approach to sample design by emphasizing uses of randomization with the goal of creating estimators with good design-based properties. Since his time on the scene, advances have been made in the use of models to construct designs and in software to implement elaborate designs. This paper reviews uses of models in balanced sampling, cutoff samples, stratification using models, multistage sampling, and mathematical programming for determining sample sizes and allocations.Release date: 2024-12-20
- Articles and reports: 75-005-M2024005Description: This article provides information about how wage data is collected in the Labour Force Survey (LFS). In particular, it examines aspects of the LFS methodology which may impact wage trends.Release date: 2024-12-13
- Articles and reports: 75F0002M2024005Description: The Canadian Income Survey (CIS) has introduced improvements to the methods and data sources used to produce income and poverty estimates with the release of its 2022 reference year estimates. Foremost among these improvements is a significant increase in the sample size for a large subset of the CIS content. The weighting methodology was also improved and the target population of the CIS was changed from persons aged 16 years and over to persons aged 15 years and over. This paper describes the changes made and presents the approximate net result of these changes on the income estimates and data quality of the CIS using 2021 data. The changes described in this paper highlight the ways in which data quality has been improved while having little impact on key CIS estimates and trends.Release date: 2024-04-26
- Previous Go to previous page of Analysis results
- 1 Go to page 1 of Analysis results
- 2 (current) Go to page 2 of Analysis results
- 3 Go to page 3 of Analysis results
- 4 Go to page 4 of Analysis results
- 5 Go to page 5 of Analysis results
- 6 Go to page 6 of Analysis results
- 7 Go to page 7 of Analysis results
- ...
- 31 Go to page 31 of Analysis results
- Next Go to next page of Analysis results
Reference (29)
Reference (29) (20 to 30 of 29 results)
- 21. Calculation of change for annual business surveys ArchivedSurveys and statistical programs – Documentation: 11-522-X19980015027Description:
The disseminated results of annual business surveys inevitably contain statistics that are changing. Since the economic sphere is increasingly dynamic, a simple difference of aggregates between n-l and n is no longer sufficient to provide an overall description of what has happened. The change calculation module in the new generation of annual business surveys divides overall change into various components (births, deaths, inter-industry migration) and calculates change on the basis of a constant field, assigning special importance to restructurings. The main difficulties lie in establishing subsamples, reweighting, calibrating according to calculable changes, and taking account of restructuring.
Release date: 1999-10-22 - Surveys and statistical programs – Documentation: 11-522-X19980015029Description:
In longitudinal surveys, sample subjects are observed over several time points. This feature typically leads to dependent observations on the same subject, in addition to the customary correlations across subjects induced by the sample design. Much research in the literature has focussed on modeling the marginal mean of a response as a function of covariates. Liang and Zeger (1986) used generalized estimating equations (GEE), requiring only correct specification of the marginal mean, and obtained standard errors of regression parameter estimates and associated Wald tests, assuming a "working" correlation structure for the repeated measurements on a sample subject. Rotnitzky and Jewell (1990) developed quasi-score tests and Rao-Scott adjustments to "working" quasi-score tests under marginal models. These methods are asymptotically robust to misspecification of the within-subject correlation structure, but assume independence of sample subjects which is not satisfied for complex longitudinal survey data based on stratified multi-stage sampling. We proposed asymptotically valid Wald and quasi-score tests for longitudinal survey data, using the Taylor Linearization and jackknife methods. Alternative tests, based on Rao-Scott adjustments to naive tests that ignore survey design features and on Bonferroni-t, are also developed. These tests are particularly useful when the effective degrees of freedom, usually taken as the total number of sample primary units (clusters) minus the number of strata, is small.
Release date: 1999-10-22 - 23. Estimating the incidence of dementia from longitudinal two-phase sampling with nonignorable missing data ArchivedSurveys and statistical programs – Documentation: 11-522-X19980015030Description:
Two-phase sampling designs have been conducted in waves to estimate the incidence of a rare disease such as dementia. Estimation of disease incidence from longitudinal dementia study has to appropriately adjust for data missing by death as well as the sampling design used at each study wave. In this paper we adopt a selection model approach to model the missing data by death and use a likelihood approach to derive incidence estimates. A modified EM algorithm is used to deal with data missing by sampling selection. The non-paramedic jackknife variance estimator is used to derive variance estimates for the model parameters and the incidence estimates. The proposed approaches are applied to data from the Indianapolis-Ibadan Dementia Study.
Release date: 1999-10-22 - 24. Estimation with partial overlap longitudinal samples ArchivedSurveys and statistical programs – Documentation: 11-522-X19980015035Description:
In a longitudinal survey conducted for k periods some units may be observed for less than k of the periods. Examples include, surveys designed with partially overlapping subsamples, a pure panel survey with nonresponse, and a panel survey supplemented with additional samples for some of the time periods. Estimators of the regression type are exhibited for such surveys. An application to special studies associated with the National Resources Inventory is discussed.
Release date: 1999-10-22 - Notices and consultations: 13F0026M1999001Description:
The main objectives of a new Canadian survey measuring asset and debt holding of families and individuals will be to update wealth information that is over one decade old; to improve the reliability of the wealth estimates; and, to provide a primary tool for analysing many important policy issues related to the distribution of assets and debts, future consumption possibilities, and savings behaviour that is of interest to governments, business and communities.
This paper is the document that launched the development of the new asset and debt survey, subsequently renamed the Survey of Financial Security. It looks at the conceptual framework for the survey, including the appropriate unit of measurement (family, household or person) and discusses measurement issues such as establishing an accounting framework for assets and debts. The variables proposed for inclusion are also identified. The paper poses several questions to readers and asks for comments and feedback.
Release date: 1999-03-23 - Notices and consultations: 13F0026M1999002Description:
This document summarizes the comments and feedback received on an earlier document: Towards a new Canadian asset and debt survey - A content discussion paper. The new asset and debt survey (now called the Survey of Financial Security) is to update the wealth information on Canadian families and unattached individuals. Since the last data collection was conducted in 1984, it was essential to include a consultative process in the development of the survey in order to obtain feedback on issues of concern and to define the conceptual framework for the survey.
Comments on the content discussion paper are summarized by major theme and sections indicate how the suggestions are being incorporated into the survey or why they could not be incorporated. This paper also mentions the main objectives of the survey and provides an overview of the survey content, revised according to the feedback from the discussion paper.
Release date: 1999-03-23 - 27. Proposal for an Asset and Debt Survey ArchivedSurveys and statistical programs – Documentation: 13F0026M1999003Description:
This paper presents a proposal for conducting a Canadian asset and debt survey. The first step in preparing this proposal was the release, in February 1997, of a document entitled Towards a new Canadian asset and debt survey whose intent was to elicit feedback on the initial thinking regarding the content of the survey.
This paper reviews the conceptual framework for a new asset and debt survey, data requirements, survey design, collection methodology and testing. It provides also an overview of the anticipated data processing system, describes the analysis and dissemination plan (analytical products and microdata files), and identifies the survey costs and major milestones. Finally, it presents the management/coordination approach used.
Release date: 1999-03-23 - Surveys and statistical programs – Documentation: 75F0002M1993019Description:
This paper examines the issues and the procedures designed to maintain a representative sample of the population for the Survey of Labour and Income Dynamics (SLID).
Release date: 1995-12-30 - Surveys and statistical programs – Documentation: 75F0002M1994001Description:
This paper describes the Survey of Labour and Income Dynamics (SLID) following rules, which govern who is traced and who is interviewed. It also outlines the conceptual basis for these procedures.
Release date: 1995-12-30