细分种族数据揭示临床风险评分表现的差异性 (Coarse race data conceals disparities in clinical risk score performance) - 专知论文

会员服务 ·

0

度量 · 变异性 · 医疗健康 · 医院 · 类别 ·

2023 年 4 月 18 日

Coarse race data conceals disparities in clinical risk score performance

翻译：细分种族数据揭示临床风险评分表现的差异性

Rajiv Movva,Divya Shanmugam,Kaihua Hou,Priya Pathak,John Guttag,Nikhil Garg,Emma Pierson

from arxiv, The first two authors contributed equally. Under review

Healthcare data in the United States often records only a patient's coarse race group: for example, both Indian and Chinese patients are typically coded as ``Asian.'' It is unknown, however, whether this coarse coding conceals meaningful disparities in the performance of clinical risk scores across granular race groups. Here we show that it does. Using data from 418K emergency department visits, we assess clinical risk score performance disparities across granular race groups for three outcomes, five risk scores, and four performance metrics. Across outcomes and metrics, we show that there are significant granular disparities in performance within coarse race categories. In fact, variation in performance metrics within coarse groups often exceeds the variation between coarse groups. We explore why these disparities arise, finding that outcome rates, feature distributions, and the relationships between features and outcomes all vary significantly across granular race categories. Our results suggest that healthcare providers, hospital systems, and machine learning researchers should strive to collect, release, and use granular race data in place of coarse race data, and that existing analyses may significantly underestimate racial disparities in performance.

翻译：美国的医疗健康数据通常只记录患者的粗略种族群体：例如，印度和中国患者通常被编码为“亚洲人”。然而，尚不清楚这种粗编码是否掩盖了种族群体之间临床风险评分表现差异的有意义差异。本文向大家展示了确实存在这样的差异。我们利用418千急诊患者数据，评估了三个结果、五个风险评分和四个表现度量之间精细种族群体之间的诊所风险评分表现差异。在不同的结果和度量标准之间，我们发现在粗略种族类别内部存在显著的精细差异。事实上，在粗略族群之间的变异之外，度量标准的变异性通常更大。我们探究了这些差异的原因，发现精细种族群体之间的结果率、特征分布以及特征和结果之间的关系都存在显著差异。我们的结果表明，医疗保健提供者，医院系统和机器学习研究人员应该努力收集，发布和使用精细种族数据代替粗略种族数据，并且现有的分析可能严重低估了表现的种族差异。

0

相关内容

Nat. Biotechnol. | 利用生成式深度学习模型发现Ⅱ型糖尿病药物-组学相关性

Nat. Biotechnol. | 利用生成式深度学习模型发现Ⅱ型糖尿病药物-组学相关性

专知会员服务

14+阅读 · 2023年1月9日

【牛津大学】电子医疗记录的生成式对抗网络:应用、评估措施和数据来源综述，A review of Generative Adversarial Networks for Electronic Health Records: applications, evaluation measures and data sources

【牛津大学】电子医疗记录的生成式对抗网络:应用、评估措施和数据来源综述，A review of Generative Adversarial Networks for Electronic Health Records: applications, evaluation measures and data sources

专知会员服务

24+阅读 · 2022年3月15日

【斯坦福博士论文】机器学习的模型解释和数据评估，206页pdf

专知会员服务

128+阅读 · 2021年8月3日

不可错过！斯坦福<人工智能疾病诊断与信息推荐>2021课程，附Slides下载

不可错过！斯坦福<人工智能疾病诊断与信息推荐>2021课程，附Slides下载

专知会员服务

47+阅读 · 2021年4月29日

最新《联邦学习Federated Learning》报告，Federated Learning

最新《联邦学习Federated Learning》报告，Federated Learning

专知会员服务

89+阅读 · 2020年12月2日

【MIT】反偏差对比学习，Debiased Contrastive Learning

【MIT】反偏差对比学习，Debiased Contrastive Learning

专知会员服务

91+阅读 · 2020年7月4日

回顾机器学习公平的数学框架，Review of Mathematical frameworks for Fairness in Machine Learning

回顾机器学习公平的数学框架，Review of Mathematical frameworks for Fairness in Machine Learning

专知会员服务

38+阅读 · 2020年5月30日

【清华大学】诊断和增强VAE模型，Diagnosing and Enhancing VAE Models

【清华大学】诊断和增强VAE模型，Diagnosing and Enhancing VAE Models

专知会员服务

37+阅读 · 2020年2月27日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

机器学习入门的经验与建议

机器学习入门的经验与建议

专知会员服务

94+阅读 · 2019年10月10日

直播 | Interpretable and Trustworthy Graph Geometric Deep Learning

直播 | Interpretable and Trustworthy Graph Geometric Deep Learning

图与推荐

2+阅读 · 2022年11月2日

GNN 新基准！Long Range Graph Benchmark

GNN 新基准！Long Range Graph Benchmark

图与推荐

0+阅读 · 2022年10月18日

KDD 2022最佳论文 | HyperSCI：在超图上学习因果效应

KDD 2022最佳论文 | HyperSCI：在超图上学习因果效应

PaperWeekly

2+阅读 · 2022年9月20日

NeurIPS'22上的GNN好文集合 (表示能力、架构设计、图对比/自监督学习、分布偏移、可解释、推荐系统等)

NeurIPS'22上的GNN好文集合 (表示能力、架构设计、图对比/自监督学习、分布偏移、可解释、推荐系统等)

图与推荐

3+阅读 · 2022年9月20日

计算机 | 入门级EI会议ICVRIS 2019诚邀稿件

计算机 | 入门级EI会议ICVRIS 2019诚邀稿件

Call4Papers

10+阅读 · 2019年6月24日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

Focal Loss for Dense Object Detection

Focal Loss for Dense Object Detection

统计学习与视觉计算组

12+阅读 · 2018年3月15日

疼痛敏感性个体差异的神经机制研究

国家自然科学基金

0+阅读 · 2015年12月31日

高血压患者Corin基因变异对其蛋白结构及酶功能影响的研究

国家自然科学基金

0+阅读 · 2015年12月31日

ATP13A2基因亚型Ala746Thr和Thr12met突变与新疆维吾尔族早发型和家族型帕金森病临床的相关研究

国家自然科学基金

0+阅读 · 2014年12月31日

S3AGA样本（Spitzer-SDSS Spectral Atlas of Galaxies and AGNs)及其AGN研究

国家自然科学基金

0+阅读 · 2014年12月31日

肾癌的磁共振扩散峰度成像研究

国家自然科学基金

0+阅读 · 2013年12月31日

UGT1A新等位基因功能的研究及UGT1A基因多态性与伊立替康疗效和毒性的关系

国家自然科学基金

0+阅读 · 2013年12月31日

血管紧张素II受体基因多态性与原发性醛固酮增多症发病风险、亚型及预后的相关性研究

国家自然科学基金

0+阅读 · 2012年12月31日

线粒体相关基因遗传变异与麻风易感性关联分析

国家自然科学基金

0+阅读 · 2012年12月31日

Notch家族四种蛋白质在胰腺癌细胞基因组的靶基因谱和调控机制的比较研究

国家自然科学基金

0+阅读 · 2012年12月31日

Reality-based Interaction用户界面模型和评估方法研究

国家自然科学基金

0+阅读 · 2011年12月31日

FilFL: Client Filtering for Optimized Client Participation in Federated Learning

Arxiv

0+阅读 · 2023年6月5日

Data Quality in Imitation Learning

Arxiv

0+阅读 · 2023年6月4日

Temporal-spatial Correlation Attention Network for Clinical Data Analysis in Intensive Care Unit

Arxiv

0+阅读 · 2023年6月3日

Subject Membership Inference Attacks in Federated Learning

Arxiv

0+阅读 · 2023年6月2日

Transformers in Medical Image Analysis: A Review

Transformers in Medical Image Analysis: A Review

Arxiv

40+阅读 · 2022年2月24日

Transformers in Medical Imaging: A Survey

Arxiv

15+阅读 · 2022年1月24日

GeomGCL: Geometric Graph Contrastive Learning for Molecular Property Prediction

Arxiv

11+阅读 · 2021年9月24日

Domain Generalization in Vision: A Survey

Arxiv

16+阅读 · 2021年7月18日

Curriculum Learning: A Survey

Arxiv

24+阅读 · 2021年1月25日

Train Large, Then Compress: Rethinking Model Size for Efficient Training and Inference of Transformers

Arxiv

12+阅读 · 2020年6月23日

VIP会员

文章信息

相关主题

相关VIP内容

Nat. Biotechnol. | 利用生成式深度学习模型发现Ⅱ型糖尿病药物-组学相关性

Nat. Biotechnol. | 利用生成式深度学习模型发现Ⅱ型糖尿病药物-组学相关性

专知会员服务

14+阅读 · 2023年1月9日

【牛津大学】电子医疗记录的生成式对抗网络:应用、评估措施和数据来源综述，A review of Generative Adversarial Networks for Electronic Health Records: applications, evaluation measures and data sources

【牛津大学】电子医疗记录的生成式对抗网络:应用、评估措施和数据来源综述，A review of Generative Adversarial Networks for Electronic Health Records: applications, evaluation measures and data sources

专知会员服务

24+阅读 · 2022年3月15日

【斯坦福博士论文】机器学习的模型解释和数据评估，206页pdf

专知会员服务

128+阅读 · 2021年8月3日

不可错过！斯坦福<人工智能疾病诊断与信息推荐>2021课程，附Slides下载

不可错过！斯坦福<人工智能疾病诊断与信息推荐>2021课程，附Slides下载

专知会员服务

47+阅读 · 2021年4月29日

最新《联邦学习Federated Learning》报告，Federated Learning

最新《联邦学习Federated Learning》报告，Federated Learning

专知会员服务

89+阅读 · 2020年12月2日

【MIT】反偏差对比学习，Debiased Contrastive Learning

【MIT】反偏差对比学习，Debiased Contrastive Learning

专知会员服务

91+阅读 · 2020年7月4日

回顾机器学习公平的数学框架，Review of Mathematical frameworks for Fairness in Machine Learning

回顾机器学习公平的数学框架，Review of Mathematical frameworks for Fairness in Machine Learning

专知会员服务

38+阅读 · 2020年5月30日

【清华大学】诊断和增强VAE模型，Diagnosing and Enhancing VAE Models

【清华大学】诊断和增强VAE模型，Diagnosing and Enhancing VAE Models

专知会员服务

37+阅读 · 2020年2月27日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

机器学习入门的经验与建议

机器学习入门的经验与建议

专知会员服务

94+阅读 · 2019年10月10日

热门VIP内容

开通专知VIP会员享更多权益服务

《海军陆战队感知任务引导技术评估：提升训练与作战效能的探索》最新100页

《光纤无人机技术：从通信基础到军事应用的多学科综述》

联合国：2025年军事人工智能、和平与安全对话核心要义

《无人机蜂群中继器攻击模拟研究》

相关资讯

直播 | Interpretable and Trustworthy Graph Geometric Deep Learning

直播 | Interpretable and Trustworthy Graph Geometric Deep Learning

图与推荐

2+阅读 · 2022年11月2日

GNN 新基准！Long Range Graph Benchmark

GNN 新基准！Long Range Graph Benchmark

图与推荐

0+阅读 · 2022年10月18日

KDD 2022最佳论文 | HyperSCI：在超图上学习因果效应

KDD 2022最佳论文 | HyperSCI：在超图上学习因果效应

PaperWeekly

2+阅读 · 2022年9月20日

NeurIPS'22上的GNN好文集合 (表示能力、架构设计、图对比/自监督学习、分布偏移、可解释、推荐系统等)

NeurIPS'22上的GNN好文集合 (表示能力、架构设计、图对比/自监督学习、分布偏移、可解释、推荐系统等)

图与推荐

3+阅读 · 2022年9月20日

计算机 | 入门级EI会议ICVRIS 2019诚邀稿件

计算机 | 入门级EI会议ICVRIS 2019诚邀稿件

Call4Papers

10+阅读 · 2019年6月24日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

Focal Loss for Dense Object Detection

Focal Loss for Dense Object Detection

统计学习与视觉计算组

12+阅读 · 2018年3月15日

相关论文

FilFL: Client Filtering for Optimized Client Participation in Federated Learning

Arxiv

0+阅读 · 2023年6月5日

Data Quality in Imitation Learning

Arxiv

0+阅读 · 2023年6月4日

Temporal-spatial Correlation Attention Network for Clinical Data Analysis in Intensive Care Unit

Arxiv

0+阅读 · 2023年6月3日

Subject Membership Inference Attacks in Federated Learning

Arxiv

0+阅读 · 2023年6月2日

Transformers in Medical Image Analysis: A Review

Transformers in Medical Image Analysis: A Review

Arxiv

40+阅读 · 2022年2月24日

Transformers in Medical Imaging: A Survey

Arxiv

15+阅读 · 2022年1月24日

GeomGCL: Geometric Graph Contrastive Learning for Molecular Property Prediction

Arxiv

11+阅读 · 2021年9月24日

Domain Generalization in Vision: A Survey

Arxiv

16+阅读 · 2021年7月18日

Curriculum Learning: A Survey

Arxiv

24+阅读 · 2021年1月25日

Train Large, Then Compress: Rethinking Model Size for Efficient Training and Inference of Transformers

Arxiv

12+阅读 · 2020年6月23日

相关基金

疼痛敏感性个体差异的神经机制研究

国家自然科学基金

0+阅读 · 2015年12月31日

高血压患者Corin基因变异对其蛋白结构及酶功能影响的研究

国家自然科学基金

0+阅读 · 2015年12月31日

ATP13A2基因亚型Ala746Thr和Thr12met突变与新疆维吾尔族早发型和家族型帕金森病临床的相关研究

国家自然科学基金

0+阅读 · 2014年12月31日

S3AGA样本（Spitzer-SDSS Spectral Atlas of Galaxies and AGNs)及其AGN研究

国家自然科学基金

0+阅读 · 2014年12月31日

肾癌的磁共振扩散峰度成像研究

国家自然科学基金

0+阅读 · 2013年12月31日

UGT1A新等位基因功能的研究及UGT1A基因多态性与伊立替康疗效和毒性的关系

国家自然科学基金

0+阅读 · 2013年12月31日

血管紧张素II受体基因多态性与原发性醛固酮增多症发病风险、亚型及预后的相关性研究

国家自然科学基金

0+阅读 · 2012年12月31日

线粒体相关基因遗传变异与麻风易感性关联分析

国家自然科学基金

0+阅读 · 2012年12月31日

Notch家族四种蛋白质在胰腺癌细胞基因组的靶基因谱和调控机制的比较研究

国家自然科学基金

0+阅读 · 2012年12月31日

Reality-based Interaction用户界面模型和评估方法研究

国家自然科学基金

0+阅读 · 2011年12月31日

微信扫码咨询专知VIP会员