首页    期刊浏览 2024年11月27日 星期三
登录注册

文章基本信息

  • 标题:Task-specific information outperforms surveillance-style big data in predictive analytics
  • 本地全文:下载
  • 作者:Andreas Bjerre-Nielsen ; Valentin Kassarnig ; David Dreyer Lassen
  • 期刊名称:Proceedings of the National Academy of Sciences
  • 印刷版ISSN:0027-8424
  • 电子版ISSN:1091-6490
  • 出版年度:2021
  • 卷号:118
  • 期号:14
  • 页码:1
  • DOI:10.1073/pnas.2020258118
  • 出版社:The National Academy of Sciences of the United States of America
  • 摘要:Increasingly, human behavior can be monitored through the collection of data from digital devices revealing information on behaviors and locations. In the context of higher education, a growing number of schools and universities collect data on their students with the purpose of assessing or predicting behaviors and academic performance, and the COVID-19–induced move to online education dramatically increases what can be accumulated in this way, raising concerns about students’ privacy. We focus on academic performance and ask whether predictive performance for a given dataset can be achieved with less privacy-invasive, but more task-specific, data. We draw on a unique dataset on a large student population containing both highly detailed measures of behavior and personality and high-quality third-party reported individual-level administrative data. We find that models estimated using the big behavioral data are indeed able to accurately predict academic performance out of sample. However, models using only low-dimensional and arguably less privacy-invasive administrative data perform considerably better and, importantly, do not improve when we add the high-resolution, privacy-invasive behavioral data. We argue that combining big behavioral data with “ground truth” administrative registry data can ideally allow the identification of privacy-preserving task-specific features that can be employed instead of current indiscriminate troves of behavioral data, with better privacy and better prediction resulting.
  • 关键词:academic performance ; prediction ; big data ; privacy
国家哲学社会科学文献中心版权所有