利用重新确定的影响估计,确定培训-攻击目标 (Identifying a Training-Set Attack's Target Using Renormalized Influence Estimation) - 专知论文

会员服务 ·

0

可辨认的 · 估计/估计量 · 示例 · 训练实例 · Performer ·

2022 年 1 月 25 日

Identifying a Training-Set Attack's Target Using Renormalized Influence Estimation

翻译：利用重新确定的影响估计,确定培训-攻击目标

Zayd Hammoudeh,Daniel Lowd

Targeted training-set attacks inject malicious instances into the training set to cause a trained model to mislabel one or more specific test instances. This work proposes the task of target identification, which determines whether a specific test instance is the target of a training-set attack. This can then be combined with adversarial-instance identification to find (and remove) the attack instances, mitigating the attack with minimal impact on other predictions. Rather than focusing on a single attack method or data modality, we build on influence estimation, which quantifies each training instance's contribution to a model's prediction. We show that existing influence estimators' poor practical performance often derives from their over-reliance on instances and iterations with large losses. Our renormalized influence estimators fix this weakness; they far outperform the original ones at identifying influential groups of training examples in both adversarial and non-adversarial settings, even finding up to 100% of adversarial training instances with no clean-data false positives. Target identification then simplifies to detecting test instances with anomalous influence values. We demonstrate our method's generality on backdoor and poisoning attacks across various data domains including text, vision, and speech. Our source code is available at https://github.com/ZaydH/target_identification .

翻译：有针对性的培训设定攻击将恶意事件注入培训内容中, 使培训模式错误地标出一个或多个特定测试实例。这项工作提议了目标识别任务, 决定特定测试实例是否是培训设定攻击的目标。然后, 这可以与辨别( 排除)攻击案例的对抗- 干预性识别相结合, 减轻攻击, 对其他预测的影响最小。我们不注重单一攻击方法或数据模式, 利用影响估计, 量化每个培训实例对模型预测的贡献。我们显示, 现有影响估计者的实际表现不佳, 往往来自他们过分依赖大量损失的事件和迭代。我们的重新整顿性影响估计者修正了这一弱点; 它们远远超出最初在辨别对抗和非对抗性环境有影响力的培训案例的类别, 甚至发现高达100%的对抗性培训实例, 没有干净数据的假阳性。目标识别随后简化了检测具有反常性影响价值的测试实例。我们的方法的回溯性影响/ 图像显示我们的方法的回溯度/ 定位系统在各种数据源码上, 包括我们的语言源/ 。

0

相关内容

可辨认的

【CVPR 2022】提出一种基于Shapley value的ShapPruning后门去除算法，Few-shot Backdoor Defense Using Shapley Estimation

【CVPR 2022】提出一种基于Shapley value的ShapPruning后门去除算法，Few-shot Backdoor Defense Using Shapley Estimation

专知会员服务

7+阅读 · 2022年3月12日

不可错过！UIUC最新《对抗机器学习》课程，附PPT

专知会员服务

35+阅读 · 2020年12月28日

【Google】深度学习对抗鲁棒性，43页ppt

专知会员服务

45+阅读 · 2020年10月31日

50+篇《神经架构搜索NAS》2020论文合集

专知会员服务

61+阅读 · 2020年3月19日

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

专知会员服务

95+阅读 · 2020年3月12日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

专知会员服务

79+阅读 · 2019年10月10日

IEEE ICKG 2022: Call for Papers

IEEE ICKG 2022: Call for Papers

机器学习与推荐算法

3+阅读 · 2022年3月30日

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium9

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium9

中国图象图形学学会CSIG

0+阅读 · 2021年12月17日

强化学习三篇论文避免遗忘等

强化学习三篇论文避免遗忘等

CreateAMind

20+阅读 · 2019年5月24日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

无监督元学习表示学习

无监督元学习表示学习

CreateAMind

27+阅读 · 2019年1月4日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

vae 相关论文表示学习 1

vae 相关论文表示学习 1

CreateAMind

12+阅读 · 2018年9月6日

【SIGIR2018】五篇对抗训练文章

【SIGIR2018】五篇对抗训练文章

专知

12+阅读 · 2018年7月9日

基于串联质谱数据的多肽鉴定半监督学习并行算法研究

国家自然科学基金

0+阅读 · 2015年12月31日

考虑多源不确定因素的建筑废弃物预测与控制方法研究

国家自然科学基金

0+阅读 · 2015年12月31日

FGF-1及其 3'UTR区SNP多态性与噪声性听力损失关系及机制的研究

国家自然科学基金

0+阅读 · 2014年12月31日

TSLP联结角膜真菌感染天然免疫与获得性免疫的机制

国家自然科学基金

0+阅读 · 2014年12月31日

基于Markov博弈的计算机网络对抗行动策略分析与建模研究

国家自然科学基金

16+阅读 · 2013年12月31日

顾及监测信息不确定性的高拱坝健康诊断模型研究

国家自然科学基金

0+阅读 · 2012年12月31日

类年龄结构与免疫-传染病耦合系统建模与研究

国家自然科学基金

1+阅读 · 2012年12月31日

SPARC在强直性脊柱炎发病中的作用机制

国家自然科学基金

0+阅读 · 2011年12月31日

维吾尔族女性乳腺癌高侵袭性与肿瘤增殖基因的关联性研究

国家自然科学基金

1+阅读 · 2011年12月31日

Cystatin B缺失与Prion疾病自噬作用机制的研究

国家自然科学基金

0+阅读 · 2011年12月31日

Quantity vs Quality: Investigating the Trade-Off between Sample Size and Label Reliability

Arxiv

0+阅读 · 2022年4月20日

Indiscriminate Data Poisoning Attacks on Neural Networks

Arxiv

0+阅读 · 2022年4月19日

Training and Evaluation of Deep Policies using Reinforcement Learning and Generative Models

Arxiv

1+阅读 · 2022年4月18日

AB/BA analysis: A framework for estimating keyword spotting recall improvement while maintaining audio privacy

Arxiv

0+阅读 · 2022年4月18日

Towards Robust Neural Networks via Orthogonal Diversity

Towards Robust Neural Networks via Orthogonal Diversity

Arxiv

0+阅读 · 2022年4月18日

Target-Relevant Knowledge Preservation for Multi-Source Domain Adaptive Object Detection

Arxiv

0+阅读 · 2022年4月17日

Narcissus: A Practical Clean-Label Backdoor Attack with Limited Information

Narcissus: A Practical Clean-Label Backdoor Attack with Limited Information

Arxiv

0+阅读 · 2022年4月15日

Neural Architecture Search without Training

Neural Architecture Search without Training

Arxiv

10+阅读 · 2021年6月11日

Learning to Propagate Labels: Transductive Propagation Network for Few-shot Learning

Arxiv

21+阅读 · 2018年12月25日

Deep Anomaly Detection with Outlier Exposure

Deep Anomaly Detection with Outlier Exposure

Arxiv

17+阅读 · 2018年12月21日

VIP会员

文章信息

相关主题

估计/估计量

相关VIP内容

【CVPR 2022】提出一种基于Shapley value的ShapPruning后门去除算法，Few-shot Backdoor Defense Using Shapley Estimation

【CVPR 2022】提出一种基于Shapley value的ShapPruning后门去除算法，Few-shot Backdoor Defense Using Shapley Estimation

专知会员服务

7+阅读 · 2022年3月12日

不可错过！UIUC最新《对抗机器学习》课程，附PPT

专知会员服务

35+阅读 · 2020年12月28日

【Google】深度学习对抗鲁棒性，43页ppt

专知会员服务

45+阅读 · 2020年10月31日

50+篇《神经架构搜索NAS》2020论文合集

专知会员服务

61+阅读 · 2020年3月19日

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

专知会员服务

95+阅读 · 2020年3月12日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

专知会员服务

79+阅读 · 2019年10月10日

热门VIP内容

开通专知VIP会员享更多权益服务

《无人机蜂群避障动态传感器覆盖算法》最新38页

《扩展现实技术在军事教育中的应用：通过沉浸式体验学习疑难知识》最新30页

《战场不可信传输环境下的边缘计算与通信》48页报告

《俄乌战争：小型无人机部署经验教训》最新报告

相关资讯

IEEE ICKG 2022: Call for Papers

IEEE ICKG 2022: Call for Papers

机器学习与推荐算法

3+阅读 · 2022年3月30日

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium9

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium9

中国图象图形学学会CSIG

0+阅读 · 2021年12月17日

强化学习三篇论文避免遗忘等

强化学习三篇论文避免遗忘等

CreateAMind

20+阅读 · 2019年5月24日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

无监督元学习表示学习

无监督元学习表示学习

CreateAMind

27+阅读 · 2019年1月4日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

vae 相关论文表示学习 1

vae 相关论文表示学习 1

CreateAMind

12+阅读 · 2018年9月6日

【SIGIR2018】五篇对抗训练文章

【SIGIR2018】五篇对抗训练文章

专知

12+阅读 · 2018年7月9日

相关论文

Quantity vs Quality: Investigating the Trade-Off between Sample Size and Label Reliability

Arxiv

0+阅读 · 2022年4月20日

Indiscriminate Data Poisoning Attacks on Neural Networks

Arxiv

0+阅读 · 2022年4月19日

Training and Evaluation of Deep Policies using Reinforcement Learning and Generative Models

Arxiv

1+阅读 · 2022年4月18日

AB/BA analysis: A framework for estimating keyword spotting recall improvement while maintaining audio privacy

Arxiv

0+阅读 · 2022年4月18日

Towards Robust Neural Networks via Orthogonal Diversity

Towards Robust Neural Networks via Orthogonal Diversity

Arxiv

0+阅读 · 2022年4月18日

Target-Relevant Knowledge Preservation for Multi-Source Domain Adaptive Object Detection

Arxiv

0+阅读 · 2022年4月17日

Narcissus: A Practical Clean-Label Backdoor Attack with Limited Information

Narcissus: A Practical Clean-Label Backdoor Attack with Limited Information

Arxiv

0+阅读 · 2022年4月15日

Neural Architecture Search without Training

Neural Architecture Search without Training

Arxiv

10+阅读 · 2021年6月11日

Learning to Propagate Labels: Transductive Propagation Network for Few-shot Learning

Arxiv

21+阅读 · 2018年12月25日

Deep Anomaly Detection with Outlier Exposure

Deep Anomaly Detection with Outlier Exposure

Arxiv

17+阅读 · 2018年12月21日

相关基金

基于串联质谱数据的多肽鉴定半监督学习并行算法研究

国家自然科学基金

0+阅读 · 2015年12月31日

考虑多源不确定因素的建筑废弃物预测与控制方法研究

国家自然科学基金

0+阅读 · 2015年12月31日

FGF-1及其 3'UTR区SNP多态性与噪声性听力损失关系及机制的研究

国家自然科学基金

0+阅读 · 2014年12月31日

TSLP联结角膜真菌感染天然免疫与获得性免疫的机制

国家自然科学基金

0+阅读 · 2014年12月31日

基于Markov博弈的计算机网络对抗行动策略分析与建模研究

国家自然科学基金

16+阅读 · 2013年12月31日

顾及监测信息不确定性的高拱坝健康诊断模型研究

国家自然科学基金

0+阅读 · 2012年12月31日

类年龄结构与免疫-传染病耦合系统建模与研究

国家自然科学基金

1+阅读 · 2012年12月31日

SPARC在强直性脊柱炎发病中的作用机制

国家自然科学基金

0+阅读 · 2011年12月31日

维吾尔族女性乳腺癌高侵袭性与肿瘤增殖基因的关联性研究

国家自然科学基金

1+阅读 · 2011年12月31日

Cystatin B缺失与Prion疾病自噬作用机制的研究

国家自然科学基金

0+阅读 · 2011年12月31日

微信扫码咨询专知VIP会员