Other content related to Statistical methods

Filter results by

Search Help
Currently selected filters that can be removed

Keyword(s)

Geography

2 facets displayed. 0 facets selected.

Survey or statistical program

1 facets displayed. 0 facets selected.

Content

1 facets displayed. 0 facets selected.
Sort Help
entries

Results

All (162)

All (162) (40 to 50 of 162 results)

  • Stats in brief: 11-627-M2020051
    Description:

    This infographic provides an overview of national statistical standards, explaining what they are and where they are used, the advantages of using them, and the role they play in the collection and dissemination of disaggregated data.

    Release date: 2020-07-24

  • Articles and reports: 62F0014M2020011
    Description:

    A summary of methodological treatments as applied to the June 2020 CPI in response to the effects of the COVID-19 pandemic on price collection, price availability, and business closure.

    Release date: 2020-07-22

  • Articles and reports: 82-003-X202000300001
    Description:

    This study describes the characteristics of residential postal codes of the Canadian population using the 2016 Census and determines how frequently these postal codes are matched to one or more dissemination areas, a unit of census geography.

    Release date: 2020-06-17

  • Articles and reports: 62F0014M2020009
    Description:

    A summary of methodological treatments as applied to the May 2020 CPI in response to the effects of the COVID-19 pandemic on price collection, price availability, and business closures.

    Release date: 2020-06-17

  • 19-23-0005
    Description:

    The service aims to provide clients with various advice and customise training on statistical methodology.

    Release date: 2020-06-12

  • Articles and reports: 62F0014M2020008
    Description:

    This document describes the methodology and data source for the provincial monthly average retail prices table. This supplement also explains the difference between the Consumer Price Index and average retail prices in context of inflation.

    Release date: 2020-06-10

  • Surveys and statistical programs – Documentation: 89-26-0003
    Description:

    Statistics Canada Data Strategy (SCDS) provides a course of action for managing and leveraging the agency’s data assets to ensure their optimal use and value while maintaining public trust. As Statistics Canada is the nation’s trusted provider of high-quality data and information to support evidence-based policy and decision making, the SCDS also naturally includes the agency’s plan for providing support and data expertise to other government organizations (federal, provincial and territorial), non-governmental organizations, the private sector, academia, and other national and international communities).

    The SCDS provides a roadmap for how Statistics Canada will continue to govern and manage its valuable data assets as part of its modernization agenda and in alignment with and response to other federal government strategies and initiatives. These federal strategies include the Data Strategy for the Federal Public Service, Canada’s 2018-2020 National Action Plan on Open Government, and the Treasury Board Secretariat Digital Operations Strategic Plan: 2018-2022.

    Release date: 2020-04-30

  • Stats in brief: 11-631-X2020001
    Description:

    This booklet provides a snapshot of data offered by Statistics Canada.

    Release date: 2020-01-16

  • Journals and periodicals: 92F0138M
    Description:

    The Geography working paper series is intended to stimulate discussion on a variety of topics covering conceptual, methodological or technical work to support the development and dissemination of the division's data, products and services. Readers of the series are encouraged to contact the Geography Division with comments and suggestions.

    Release date: 2019-11-13

  • Surveys and statistical programs – Documentation: 99-011-X
    Description:

    This topic presents data on the Aboriginal peoples of Canada and their demographic characteristics. Depending on the application, estimates using any of the following concepts may be appropriate for the Aboriginal population: (1) Aboriginal identity, (2) Aboriginal ancestry, (3) Registered or Treaty Indian status and (4) Membership in a First Nation or Indian band. Data from the 2011 National Household Survey are available for the geographical locations where these populations reside, including 'on reserve' census subdivisions and Inuit communities of Inuit Nunangat as well as other geographic areas such as the national (Canada), provincial and territorial levels.

    Analytical products

    The analytical document provides analysis on the key findings and trends in the data, and is complimented with the short articles found in NHS in Brief and the NHS Focus on Geography Series.

    Data products

    The NHS Profile is one data product that provides a statistical overview of user selected geographic areas based on several detailed variables and/or groups of variables. Other data products include data tables which represent a series of cross tabulations ranging in complexity and are available for various levels of geography.

    Release date: 2019-10-29
Data (1)

Data (1) ((1 result))

  • Table: 82-567-X
    Description:

    The National Population Health Survey (NPHS) is designed to enhance the understanding of the processes affecting health. The survey collects cross-sectional as well as longitudinal data. In 1994/95 the survey interviewed a panel of 17,276 individuals, then returned to interview them a second time in 1996/97. The response rate for these individuals was 96% in 1996/97. Data collection from the panel will continue for up to two decades. For cross-sectional purposes, data were collected for a total of 81,000 household residents in all provinces (except people on Indian reserves or on Canadian Forces bases) in 1996/97.

    This overview illustrates the variety of information available by presenting data on perceived health, chronic conditions, injuries, repetitive strains, depression, smoking, alcohol consumption, physical activity, consultations with medical professionals, use of medications and use of alternative medicine.

    Release date: 1998-07-29
Analysis (102)

Analysis (102) (10 to 20 of 102 results)

  • Articles and reports: 11-633-X2021006
    Description:

    This paper describes the current thinking at Statistics Canada about future directions in social statistics. It describes how the system of statistics on social statistics (which would be renamed quality of life statistics) will look like in the next 5 to 10 years if Statistics Canada adopts the transformative methodologies and dissemination products that are needed to meet the growing demand for more disaggregated, timely, granular, accessible and more responsive statistics on quality of life.

    Release date: 2022-01-31

  • Articles and reports: 11-633-X2021007
    Description:

    Statistics Canada continues to use a variety of data sources to provide neighbourhood-level variables across an expanding set of domains, such as sociodemographic characteristics, income, services and amenities, crime, and the environment. Yet, despite these advances, information on the social aspects of neighbourhoods is still unavailable. In this paper, answers to the Canadian Community Health Survey on respondents’ sense of belonging to their local community were pooled over the four survey years from 2016 to 2019. Individual responses were aggregated up to the census tract (CT) level.

    Release date: 2021-11-16

  • Articles and reports: 75F0002M2021007
    Description:

    This discussion paper describes the proposed methodology for a Northern Market Basket Measure (MBM-N) for Yukon and the Northwest Territories, as well as identifies research which could be conducted in preparation for the 2023 review. The paper presents initial MBM-N thresholds and provides preliminary poverty estimates for reference years 2018 and 2019. A review period will follow the release of this paper, during which time Statistics Canada and Employment and Social Development Canada will welcome feedback from interested parties and work with experts, stakeholders, indigenous organizations, federal, provincial and territorial officials to validate the results.

    Release date: 2021-11-12

  • Articles and reports: 11-522-X202100100010
    Description:

    As part of processing for the 2021 Canadian Census, the write-in responses to 31 census questions must be coded. Up until, and including, 2016, this was a three stage process, including an “interactive (human) coding” step as the second stage. This human coding step is both lengthy and expensive, spanning many months and requiring the hiring and training of a large number of temporary employees. With this in mind, for 2021, this stage was either augmented with or replaced entirely by machine learning models using the "fastText" algorithm. This presentation will discuss the implementation of this algorithm and the challenges and decisions taken along the way.

    Key Words: Natural Language Processing, Machine Learning, fastText, Coding

    Release date: 2021-11-05

  • Articles and reports: 11-522-X202100100011
    Description: The ways in which AI may affect the world of official statistics are manifold and Statistics Netherlands (CBS) is actively exploring how it can use AI within its societal role. The paper describes a number of AI-related areas where CBS is currently active: use of AI for its own statistics production and statistical R&D, the development of a national AI monitor, the support of other government bodies with expertise on fair data and fair algorithms, data sharing under safe and secure conditions, and engaging in AI-related collaborations.

    Key Words: Artificial Intelligence; Official Statistics; Data Sharing; Fair Algorithms; AI monitoring; Collaboration.

    Release date: 2021-11-05

  • Articles and reports: 11-522-X202100100012
    Description: The modernization of price statistics by National Statistical Offices (NSO) such as Statistics Canada focuses on the adoption of alternative data sources that include the near-universe of all products sold in the country, a scale that requires machine learning classification of the data. The process of evaluating classifiers to select appropriate ones for production, as well as monitoring classifiers once in production, needs to be based on robust metrics to measure misclassification. As commonly utilized metrics, such as the Fß-score may not take into account key aspects applicable to prices statistics in all cases, such as unequal importance of categories, a careful consideration of the metric space is necessary to select appropriate methods to evaluate classifiers. This working paper provides insight on the metric space applicable to price statistics and proposes an operational framework to evaluate and monitor classifiers, focusing specifically on the needs of the Canadian Consumer Prices Index and demonstrating discussed metrics using a publicly available dataset.

    Key Words: Consumer price index; supervised classification; evaluation metrics; taxonomy

    Release date: 2021-11-05

  • Articles and reports: 11-522-X202100100013
    Description: Statistics Canada’s Labour Force Survey (LFS) plays a fundamental role in the mandate of Statistics Canada. The labour market information provided by the LFS is among the most timely and important measures of the Canadian economy’s overall performance. An integral part of the LFS monthly data processing is the coding of respondent’s industry according to the North American Industrial Classification System (NAICS), occupation according to the National Occupational Classification System (NOC) and the Primary Class of Workers (PCOW). Each month, up to 20,000 records are coded manually. In 2020, Statistics Canada worked on developing Machine Learning models using fastText to code responses to the LFS questionnaire according to the three classifications mentioned previously. This article will provide an overview on the methodology developed and results obtained from a potential application of the use of fastText into the LFS coding process. 

    Key Words: Machine Learning; Labour Force Survey; Text classification; fastText.

    Release date: 2021-11-05

  • Articles and reports: 11-522-X202100100028
    Description:

    Many Government of Canada groups are developing codes to process and visualize various kinds data, often duplicating each other’s efforts, with sub-optimal efficiency and limited level of code quality reviewing. This paper informally presents a working-level approach to addressing this technical problem. The idea is to collaboratively build a common repository of code and knowledgebase for use by anyone in the public sector to perform many common data science tasks, and, in doing that, help each other to master both the data science coding skills and the industry standard collaborative practices. The paper explains why R language is used as the language of choice for collaborative data science code development. It summaries R advantages and addresses its limitations, establishes the taxonomy of discussion topics of highest interested to the GC data scientists working with R, provides an overview of used collaborative platforms, and presents the results obtained to date. Even though the code knowledgebase is developed mainly in R, it is meant to be valuable also for data scientists coding in Python and other development environments. Key Words: Collaboration; Data science; Data Engineering; R; Open Government; Open Data; Open Science

    Release date: 2021-10-29

  • Articles and reports: 11-522-X202100100001
    Description:

    We consider regression analysis in the context of data integration. To combine partial information from external sources, we employ the idea of model calibration which introduces a “working” reduced model based on the observed covariates. The working reduced model is not necessarily correctly specified but can be a useful device to incorporate the partial information from the external data. The actual implementation is based on a novel application of the empirical likelihood method. The proposed method is particularly attractive for combining information from several sources with different missing patterns. The proposed method is applied to a real data example combining survey data from Korean National Health and Nutrition Examination Survey and big data from National Health Insurance Sharing Service in Korea.

    Key Words: Big data; Empirical likelihood; Measurement error models; Missing covariates.

    Release date: 2021-10-15

  • Articles and reports: 11-522-X202100100002
    Description:

    A framework for the responsible use of machine learning processes has been developed at Statistics Canada. The framework includes guidelines for the responsible use of machine learning and a checklist, which are organized into four themes: respect for people, respect for data, sound methods, and sound application. All four themes work together to ensure the ethical use of both the algorithms and results of machine learning. The framework is anchored in a vision that seeks to create a modern workplace and provide direction and support to those who use machine learning techniques. It applies to all statistical programs and projects conducted by Statistics Canada that use machine learning algorithms. This includes supervised and unsupervised learning algorithms. The framework and associated guidelines will be presented first. The process of reviewing projects that use machine learning, i.e., how the framework is applied to Statistics Canada projects, will then be explained. Finally, future work to improve the framework will be described.

    Keywords: Responsible machine learning, explainability, ethics

    Release date: 2021-10-15
Reference (54)

Reference (54) (50 to 60 of 54 results)

  • Surveys and statistical programs – Documentation: 92-353-X
    Description:

    This report deals with age, sex, marital status and common-law status. It is aimed at informing users about the complexity of the data and any difficulties that could affect their use. It explains the theoretical framework and definitions used to gather the data, and describes unusual circumstances that could affect data quality. Moreover, the report touches upon data capture, edit and imputation, and deals with the historical comparability of the data.

    Release date: 1999-04-16

  • Surveys and statistical programs – Documentation: 75F0002M1998005
    Description:

    This article gives an overview of the main goals of the Survey of Labour and Income Dynamics (SLID) and the methodology used.

    Release date: 1998-12-30

  • Surveys and statistical programs – Documentation: 5190
    Description: The Data Inventory Project is a government-wide stock-taking of federal data holdings within departments that are part of the Policy Research Data Group to determine the broad range of data holdings that could address the medium to longer-term priorities. The inventory is comprised of the metadata on datasets held within the various departments and will be linked, when possible, to specific key policy issues.

  • Surveys and statistical programs – Documentation: 8014
    Description: This study will be used to determine which method would be the most effective to select households in Canada for any given survey that is conducted by Statistics Canada.
Date modified: