Data analysis
Filter results by
Search HelpKeyword(s)
Type
Survey or statistical program
- Census of Population (12)
- Canadian Community Health Survey - Annual Component (7)
- Labour Force Survey (7)
- Survey of Household Spending (6)
- Canadian Income Survey (4)
- Survey of Labour and Income Dynamics (3)
- Longitudinal Immigration Database (3)
- Canadian Health Measures Survey (3)
- Gross Domestic Product by Industry - National (Monthly) (2)
- Monthly Oil and Other Liquid Petroleum Products Pipeline Survey (2)
- Uniform Crime Reporting Survey (2)
- Census of Agriculture (2)
- Households and the Environment Survey (2)
- Time Use Survey (2)
- Biennial Drinking Water Plants Survey (2)
- Longitudinal Employment Analysis Program (2)
- Canada's International Transactions in Services (1)
- Waste Management Industry Survey: Government Sector (1)
- National Balance Sheet Accounts (1)
- National Gross Domestic Product by Income and by Expenditure Accounts (1)
- National Tourism Indicators (1)
- Biennial Waste Management Survey (1)
- Monthly Electricity Supply and Disposition Survey (1)
- Annual Electricity Supply and Disposition Survey (1)
- Consumer Price Index (1)
- Monthly New Motor Vehicle Sales Survey (1)
- Survey of Employment, Payrolls and Hours (1)
- Survey of Financial Security (1)
- Monthly Passenger Bus and Urban Transit Survey (1)
- Stock and Consumption of Fixed Non-residential Capital (1)
- Tuition and Living Accommodation Costs (1)
- Vital Statistics - Death Database (1)
- Annual Demographic Estimates: Canada, Provinces and Territories (1)
- Homeowner Repair and Renovation Survey (1)
- Annual Income Estimates for Census Families and Individuals (T1 Family File) (1)
- Annual Survey of Research and Development in Canadian Industry (1)
- Research and Development of Canadian Private Non-Profit Organizations (1)
- General Social Survey - Victimization (1)
- General Social Survey - Social Identity (1)
- Culture Services Trade (1)
- Canadian Community Health Survey - Nutrition (1)
- Canadian System of Environmental-Economic Accounts - Physical Flow Accounts (1)
- Air Quality Indicators (1)
- Freshwater Quality Indicator (1)
- Longitudinal and International Study of Adults (1)
- Government Finance Statistics (1)
- National Household Survey (1)
- Gross Domestic Expenditures on Research and Development (1)
- Survey of Safety in Public and Private Spaces (1)
- Canadian Housing Statistics Program (1)
- Study on International Money Transfers (1)
- Canadian Housing Survey (1)
- Survey on Early Learning and Child Care Arrangements (SELCCA) (1)
- Canadian Perspectives Survey Series (CPSS) (1)
- Canada Mortgage and Housing Corporation (1)
Results
All (274)
All (274) (40 to 50 of 274 results)
- Articles and reports: 11-522-X202100100027Description:
Privacy concerns are a barrier to applying remote analytics, including machine learning, on sensitive data via the cloud. In this work, we use a leveled fully Homomorphic Encryption scheme to train an end-to-end supervised machine learning algorithm to classify texts while protecting the privacy of the input data points. We train our single-layer neural network on a large simulated dataset, providing a practical solution to a real-world multi-class text classification task. To improve both accuracy and training time, we train an ensemble of such classifiers in parallel using ciphertext packing.
Key Words: Privacy Preservation, Machine Learning, Encryption
Release date: 2021-10-29 - 42. Statistical Disclosure Control and Developments in Formal Privacy: In Memoriam to Chris Skinner ArchivedArticles and reports: 11-522-X202100100022Description:
I provide an overview of the evolution of Statistical Disclosure Control (SDC) research over the last decades and how it has evolved to handle the data revolution with more formal definitions of privacy. I emphasize the many contributions by Chris Skinner in the research areas of SDC. I will review his seminal research, starting in the 1990’s with his work on the release of UK Census sample microdata. This led to a wide-range of research on measuring the risk of re-identification in survey microdata through probabilistic models. I also focus on other aspects of Chris’ research in SDC. Chris was the recipient of the 2019 Waksberg Award and sadly never got a chance to present his Waksberg Lecture at the Statistics Canada International Methodology Symposium. This paper follows the outline that Chris had prepared in preparation for that lecture, and provided to me by his son, Tom Skinner. Keywords: Risk of Re-identification, Data Revolution, Privacy Models, Differential Privacy
Release date: 2021-10-22 - Articles and reports: 11-522-X202100100021Description: Istat has started a new project for the Short Term statistical processes, to satisfy the coming new EU Regulation to release estimates in a shorter time. The assessment and analysis of the current Short Term Survey on Turnover in Services (FAS) survey process, aims at identifying how the best features of the current methods and practices can be exploited to design a more “efficient” process. In particular, the project is expected to release methods that would allow important economies of scale, scope and knowledge to be applied in general to the STS productive context, usually working with a limited number of resources. The analysis of the AS-IS process revealed that the FAS survey incurs substantial E&I costs, especially due to intensive follow-up and interactive editing that is used for every type of detected errors. In this view, we tried to exploit the lessons learned by participating to the High-Level Group for the Modernisation of Official Statistics (HLG-MOS, UNECE) about the Use of Machine Learning in Official Statistics. In this work, we present a first experiment using Random Forest models to: (i) predict which units represent “suspicious” data, (ii) to assess the prediction potential use over new data and (iii) to explore data to identify hidden rules and patterns. In particular, we focus on the use of Random Forest modelling to compare some alternative methods in terms of error prediction efficiency and to address the major aspects for the new design of the E&I scheme.Release date: 2021-10-15
- Articles and reports: 12-001-X202100100003Description:
One effective way to conduct statistical disclosure control is to use scrambled responses. Scrambled responses can be generated by using a controlled random device. In this paper, we propose using the sample empirical likelihood approach to conduct statistical inference under complex survey design with scrambled responses. Specifically, we propose using a Wilk-type confidence interval for statistical inference. Our proposed method can be used as a general tool for inference with confidential public use survey data files. Asymptotic properties are derived, and the limited simulation study verifies the validity of theory. We further apply the proposed method to some real applications.
Release date: 2021-06-24 - 19-22-0005Description:
In this session, we will attempt to demystify the concept of confidence intervals as they relate to sample data. A practical approach is used, placing emphasis on the meaning and interpretation of results rather than the mathematics. The goal is to make sense of some common challenges faced by data users when interpreting confidence intervals. The session is intended for a beginner audience. Some familiarity with basic statistical concepts would be beneficial/advantageous but not required.
https://www.statcan.gc.ca/eng/wtc/information/19220005
Release date: 2021-05-28 - 46. Statistics 101: Correlation and Causality ArchivedStats in brief: 89-20-00062021002Description:
This video is intended for viewers who wish to gain a basic understanding of correlation and causality. As a prerequisite, before beginning this video, we highly recommend having already completed our videos titled “What is Data? An Introduction to Data Terminology and Concepts” and “Types of Data: Understanding and Exploring Data”.
Release date: 2021-05-03 - Articles and reports: 11-633-X2021003Description:
Canada continues to experience an opioid crisis. While there is solid information on the demographic and geographic characteristics of people experiencing fatal and non-fatal opioid overdoses in Canada, there is limited information on the social and economic conditions of those who experience these events. To fill this information gap, Statistics Canada collaborated with existing partnerships in British Columbia, including the BC Coroners Service, BC Stats, the BC Centre for Disease Control and the British Columbia Ministry of Health, to create the Statistics Canada British Columbia Opioid Overdose Analytical File (BC-OOAF).
Release date: 2021-02-17 - Articles and reports: 11-633-X2021001Description:
Using data from the Canadian Housing Survey, this project aimed to construct a measure of social inclusion, using indicators identified by the Canada Mortgage and Housing Corporation (CMHC), to report a social inclusion score for each geographic stratum separately for dwellings that are and are not in social and affordable housing. This project also sought to examine associations between social inclusion and a set of economic, social and health variables.
Release date: 2021-01-05 - Articles and reports: 12-001-X202000200004Description:
This article proposes a weight scaling method for Firth’s penalized likelihood for proportional hazards regression models. The method derives a relationship between the penalized likelihood that uses scaled weights and the penalized likelihood that uses unscaled weights, and it shows that the penalized likelihood that uses scaled weights have some desirable properties. A simulation study indicates that the penalized likelihood using scaled weights produces smaller biases in point estimates and standard errors than the biases produced by the penalized likelihood using unscaled weights. The weighted penalized likelihood is applied to estimate hazard rates for heart attacks by using a public-use data set from the National Health and Epidemiology Followup Study (NHEFS). SAS® statements to estimate hazard rates using data from complex surveys are given in the appendix.
Release date: 2020-12-15 - 50. Validation of the Food Security Module in the 2018 Longitudinal and International Study of AdultsArticles and reports: 89-648-X2020004Description:
This technical report is intended to validate the Longitudinal and International Study of Adults (LISA) Wave 4 (2018) Food Security (FSC) module and provide recommendations for analytical use. Section 2 of this report provides an overview of the LISA data. Section 3 provides some background information of food security measures in national surveys and why it is significant in today's literature. Section 4 analyzes FSC data by presenting key descriptive statistics and logic checks using LISA methodology as well as outside researcher information. In section 5, certification validation was done by comparing other Canadian national surveys that have used the FSC module to the one used by LISA. Finally in section 6, key findings and their implications with regard to LISA are outlined.
Release date: 2020-11-02
- Previous Go to previous page of All results
- 1 Go to page 1 of All results
- 2 Go to page 2 of All results
- 3 Go to page 3 of All results
- 4 Go to page 4 of All results
- 5 (current) Go to page 5 of All results
- 6 Go to page 6 of All results
- 7 Go to page 7 of All results
- ...
- 28 Go to page 28 of All results
- Next Go to next page of All results
Data (2)
Data (2) ((2 results))
- Data Visualization: 71-607-X2020010Description: The Canadian Statistical Geospatial Explorer empowers users to discover geo enabled data holdings of Statistics Canada at various levels of geography including at the neighbourhood level. Users are able to visualize, thematically map, spatially explore and analyze, export and consume data in various formats. Users can also view the data superimposed on satellite imagery, topographic and street layers.Release date: 2024-08-21
- 2. Housing Data Viewer ArchivedData Visualization: 71-607-X2019010Description: The Housing Data Viewer is a visualization tool that allows users to explore Statistics Canada data on a map. Users can use the tool to navigate, compare and export data.Release date: 2019-10-30
Analysis (246)
Analysis (246) (0 to 10 of 246 results)
- Journals and periodicals: 11-633-XDescription: Papers in this series provide background discussions of the methods used to develop data for economic, health, and social analytical studies at Statistics Canada. They are intended to provide readers with information on the statistical methods, standards and definitions used to develop databases for research purposes. All papers in this series have undergone peer and institutional review to ensure that they conform to Statistics Canada's mandate and adhere to generally accepted standards of good professional practice.Release date: 2024-09-11
- 2. Labour Force Survey initiatives under Statistics Canada’s Disaggregated Data Action Plan ArchivedArticles and reports: 11-522-X202200100004Description: In accordance with Statistics Canada’s long-term Disaggregated Data Action Plan (DDAP), several initiatives have been implemented into the Labour Force Survey (LFS). One of the more direct initiatives was a targeted increase in the size of the monthly LFS sample. Furthermore, a regular Supplement program was introduced, where an additional series of questions are asked to a subset of LFS respondents and analyzed in a monthly or quarterly production cycle. Finally, the production of modelled estimates based on Small Area Estimation (SAE) methodologies resumed for the LFS and will include a wider scope with more analytical value than what had existed in the past. This paper will give an overview of these three initiatives.Release date: 2024-03-25
- 3. ABS DataLab output checking tools ArchivedArticles and reports: 11-522-X202200100006Description: The Australian Bureau of Statistics (ABS) is committed to improving access to more microdata, while ensuring privacy and confidentiality is maintained, through its virtual DataLab which supports researchers to undertake complex research more efficiently. Currently, the DataLab research outputs need to follow strict rules to minimise disclosure risks for clearance. However, the clerical-review process is not cost effective and has potential to introduce errors. The increasing number of statistical outputs from different projects can potentially introduce differencing risks even though these outputs from different projects have met the strict output rules. The ABS has been exploring the possibility of providing automatic output checking using the ABS cellkey methodology to ensure that all outputs across different projects are protected consistently to minimise differencing risks and reduce costs associated with output checking.Release date: 2024-03-25
- Articles and reports: 11-522-X202200100009Description: Education and training is acknowledged as fundamental for the development of a society. It is a complex multidimensional phenomenon, which determinants are ascribable to several interrelated familiar and socio-economic conditions. To respond to the demand of supporting statistical information for policymaking and its monitoring and evaluation process, the Italian National Statistical Institute (Istat) is renewing the education and training statistical production system, implementing a new thematic statistical register. It will be part of the Istat Integrated System of Registers, thus allowing relating the education and training phenomenon to other relevant phenomena, e.g. transition to work.Release date: 2024-03-25
- Stats in brief: 11-001-X202402237898Description: Release published in The Daily – Statistics Canada’s official release bulletinRelease date: 2024-01-22
- Articles and reports: 11-633-X2024001Description: The Longitudinal Immigration Database (IMDB) is a comprehensive source of data that plays a key role in the understanding of the economic behaviour of immigrants. It is the only annual Canadian dataset that allows users to study the characteristics of immigrants to Canada at the time of admission and their economic outcomes and regional (inter-provincial) mobility over a time span of more than 35 years.Release date: 2024-01-22
- Articles and reports: 12-001-X202300200007Description: Conformal prediction is an assumption-lean approach to generating distribution-free prediction intervals or sets, for nearly arbitrary predictive models, with guaranteed finite-sample coverage. Conformal methods are an active research topic in statistics and machine learning, but only recently have they been extended to non-exchangeable data. In this paper, we invite survey methodologists to begin using and contributing to conformal methods. We introduce how conformal prediction can be applied to data from several common complex sample survey designs, under a framework of design-based inference for a finite population, and we point out gaps where survey methodologists could fruitfully apply their expertise. Our simulations empirically bear out the theoretical guarantees of finite-sample coverage, and our real-data example demonstrates how conformal prediction can be applied to complex sample survey data in practice.Release date: 2024-01-03
- Articles and reports: 45-20-00022023004Description: Gender-based Analysis Plus (GBA Plus) is an analytical tool developed by Women and Gender Equality Canada (WAGE) to support the development of responsive and inclusive initiatives, including policies, programs, and other initiatives. This information sheet presents the usefulness of GBA Plus for disaggregating and analyzing data to identify the groups most affected by certain issues, such as overqualification.Release date: 2023-11-27
- Stats in brief: 89-20-00062023001Description: This course is intended for Government of Canada employees who would like to learn about evaluating the quality of data for a particular use. Whether you are a new employee interested in learning the basics, or an experienced subject matter expert looking to refresh your skills, this course is here to help.Release date: 2023-07-17
- Articles and reports: 82-003-X202300200003Description: Utility scores are an important tool for evaluating health-related quality of life. Utility score norms have been published for Canadian adults, but no nationally representative utility score norms are available for non-adults. Using Health Utilities Index Mark 3 (HUI3) data from two recent cycles of the Canadian Health Measures Survey (i.e., 2016-2017 and 2018-2019), this is the first study to provide utility score norms for children aged 6 to 11 years and adolescents aged 12 to 17 years.Release date: 2023-02-15
- Previous Go to previous page of Analysis results
- 1 (current) Go to page 1 of Analysis results
- 2 Go to page 2 of Analysis results
- 3 Go to page 3 of Analysis results
- 4 Go to page 4 of Analysis results
- 5 Go to page 5 of Analysis results
- 6 Go to page 6 of Analysis results
- 7 Go to page 7 of Analysis results
- ...
- 25 Go to page 25 of Analysis results
- Next Go to next page of Analysis results
Reference (22)
Reference (22) (0 to 10 of 22 results)
- Surveys and statistical programs – Documentation: 32-26-0006Description: This report provides data quality information pertaining to the Agriculture–Population Linkage, such as sources of error, matching process, response rates, imputation rates, sampling, weighting, disclosure control methods and data quality indicators.Release date: 2023-08-25
- Surveys and statistical programs – Documentation: 98-20-00032021011Description: This video explains the key concepts of different levels of aggregation of income data such as household and family income; income concepts derived from key income variables such as adjusted income and equivalence scale; and statistics used for income data such as median and average income, quartiles, quintiles, deciles and percentiles.Release date: 2023-03-29
- Surveys and statistical programs – Documentation: 98-20-00032021012Description: This video builds on concepts introduced in the other videos on income. It explains key low-income concepts - Market Basket Measure (MBM), Low income measure (LIM) and Low-income cut-offs (LICO) and the indicators associated with these concepts such as the low-income gap and the low-income ratio. These concepts are used in analysis of the economic well-being of the population.Release date: 2023-03-29
- Notices and consultations: 98-26-0001Description:
This white paper presents Statistics Canada’s planned approach to the 2021 Census of Population and provides a clear explanation of the processes behind the census program, touching on historical, legal, operational and content aspects. Statistics Canada recognizes that it is important to not only successfully conduct the census, but also to be transparent and informative about the way in which those efforts are accomplished. Painting a Portrait of Canada: The 2021 Census of Population gives readers an exclusive, detailed look at how census data is collected, analyzed and given back to Canadians, in the form of high-quality statistical information, used to make evidence-based decisions in Canadian society.
Release date: 2020-07-20 - Surveys and statistical programs – Documentation: 91F0015M2016012Description:
This article provides information on using family-related variables from the microdata files of Canada’s Census of Population. These files exist internally at Statistics Canada, in the Research Data Centres (RDCs), and as public-use microdata files (PUMFs). This article explains certain technical aspects of all three versions, including the creation of multi-level variables for analytical purposes.
Release date: 2016-12-22 - 6. The Data Warehouse and analytical tools to facilitate the integration of the Canadian Macroeconomic Accounts ArchivedSurveys and statistical programs – Documentation: 11-522-X201700014710Description:
The Data Warehouse has modernized the way the Canadian System of Macroeconomic Accounts (MEA) are produced and analyzed today. Its continuing evolution facilitates the amounts and types of analytical work that is done within the MEA. It brings in the needed element of harmonization and confrontation as the macroeconomic accounts move toward full integration. The improvements in quality, transparency, and timeliness have strengthened the statistics that are being disseminated.
Release date: 2016-03-24 - Notices and consultations: 75-513-X2014001Description:
Starting with the 2012 reference year, annual individual and family income data is produced by the Canadian Income Survey (CIS). The CIS is a cross-sectional survey developed to provide information on the income and income sources of Canadians, along with their individual and household characteristics. The CIS reports on many of the same statistics as the Survey of Labour and Income Dynamics (SLID), which last reported on income for the 2011 reference year. This note describes the CIS methodology, as well as the main differences in survey objectives, methodology and questionnaires between CIS and SLID.
Release date: 2014-12-10 - 8. Using a Trend-cycle Approach to Estimate Changes in Southern Canada's Water Yield from 1971 to 2004 ArchivedSurveys and statistical programs – Documentation: 16-001-M2010014Description: Quantifying how Canada's water yield has changed over time is an important component of the water accounts maintained by Statistics Canada. This study evaluates the movement in the series of annual water yield estimates for Southern Canada from 1971 to 2004. We estimated the movement in the series using a trend-cycle approach and found that water yield for southern Canada has generally decreased over the period of observation.Release date: 2010-09-13
- 9. Finding and Using Statistics ArchivedSurveys and statistical programs – Documentation: 11-533-XDescription:
This guide has been created especially for users needing a step-by-step review on how to find, read and use data, with quick tips on locating information on the Statistics Canada website. Originally published in paper format in the 1980s, revised as part of the 1994 Statistics Canada Catalogue, and then transformed into an electronic version, this guide is continually being updated to maintain its currency and usefulness.
Release date: 2007-11-19 - Surveys and statistical programs – Documentation: 81-595-M2007056Geography: CanadaDescription:
This handbook discusses the collection and interpretation of statistical data on Canada's trade in culture services.
Release date: 2007-10-31
- Date modified: