经验教训职能的易变性 (On The Fragility of Learned Reward Functions) - 专知论文

会员服务 ·

0

奖励函数 · 泛函 · Learning · Continuity · Performer ·

2023 年 1 月 9 日

On The Fragility of Learned Reward Functions

翻译：经验教训职能的易变性

Lev McKinney,Yawen Duan,David Krueger,Adam Gleave

from arxiv, 5 pages, 2 figures, presented at the NeurIPS Deep RL and ML Safety Workshops

Reward functions are notoriously difficult to specify, especially for tasks with complex goals. Reward learning approaches attempt to infer reward functions from human feedback and preferences. Prior works on reward learning have mainly focused on the performance of policies trained alongside the reward function. This practice, however, may fail to detect learned rewards that are not capable of training new policies from scratch and thus do not capture the intended behavior. Our work focuses on demonstrating and studying the causes of these relearning failures in the domain of preference-based reward learning. We demonstrate with experiments in tabular and continuous control environments that the severity of relearning failures can be sensitive to changes in reward model design and the trajectory dataset composition. Based on our findings, we emphasize the need for more retraining-based evaluations in the literature.

翻译：奖励性学习方法试图从人类的反馈和偏好中推断奖励性功能; 以往的奖励性学习工作主要侧重于与奖励性功能同时培训的政策的绩效; 但是,这种做法可能无法发现无法从零开始培训新政策从而无法捕捉预期行为的学习性奖励; 我们的工作重点是在基于优惠的奖励学习领域示范和研究这些再学习失败的原因; 我们在表格和连续的控制环境中进行实验,证明再学习失败的严重程度可能对奖赏模式设计和轨迹数据集构成的变化十分敏感。根据我们的调查结果,我们强调文献中需要更多基于再培训的评价。

0

相关内容

奖励函数

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

50+阅读 · 2019年10月17日

机器学习入门的经验与建议

机器学习入门的经验与建议

专知会员服务

94+阅读 · 2019年10月10日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

专知会员服务

65+阅读 · 2019年10月9日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

【ICIG2021】Latest News & Announcements of the Workshop

【ICIG2021】Latest News & Announcements of the Workshop

中国图象图形学学会CSIG

0+阅读 · 2021年12月20日

【ICIG2021】Latest News & Announcements of the Industry Talk1

【ICIG2021】Latest News & Announcements of the Industry Talk1

中国图象图形学学会CSIG

0+阅读 · 2021年7月28日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

Calderon问题和边界刚性问题

国家自然科学基金

0+阅读 · 2013年12月31日

基于SURE/PURE准则的图像盲反卷积算法研究

国家自然科学基金

3+阅读 · 2013年12月31日

Intraflagellar Transport运输纤毛蛋白的分子机理

国家自然科学基金

0+阅读 · 2012年12月31日

靶向调控PLCE1基因的microRNAs在新疆哈族食管癌侵袭转移中的作用与机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

Ter94在Hedgehog信号转导途径中的作用机理

国家自然科学基金

0+阅读 · 2009年12月31日

Centroid Distance Distillation for Effective Rehearsal in Continual Learning

Arxiv

0+阅读 · 2023年3月6日

Geometric Batch Optimization for the Packing Equal Circles in a Circle Problem on Large Scale

Arxiv

0+阅读 · 2023年3月5日

On the Robustness of ChatGPT: An Adversarial and Out-of-distribution Perspective

Arxiv

0+阅读 · 2023年3月2日

Sequential Attention for Feature Selection

Arxiv

0+阅读 · 2023年3月2日

From Show to Tell: A Survey on Image Captioning

Arxiv

15+阅读 · 2021年7月14日

VIP会员

文章信息

相关主题

相关VIP内容

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

50+阅读 · 2019年10月17日

机器学习入门的经验与建议

机器学习入门的经验与建议

专知会员服务

94+阅读 · 2019年10月10日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

专知会员服务

65+阅读 · 2019年10月9日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

锚定情报：合成欺骗时代的地面真相

NeurIPS 2025 | NMKE：基于神经元归因与动态稀疏掩码的终身知识编辑

《无人军用移动机器人中密码学与导航系统的集成：当前趋势与前景综述》

【MIT博士论文】弱监督学习：理论、方法与应用

相关资讯

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

【ICIG2021】Latest News & Announcements of the Workshop

【ICIG2021】Latest News & Announcements of the Workshop

中国图象图形学学会CSIG

0+阅读 · 2021年12月20日

【ICIG2021】Latest News & Announcements of the Industry Talk1

【ICIG2021】Latest News & Announcements of the Industry Talk1

中国图象图形学学会CSIG

0+阅读 · 2021年7月28日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

相关论文

Centroid Distance Distillation for Effective Rehearsal in Continual Learning

Arxiv

0+阅读 · 2023年3月6日

Geometric Batch Optimization for the Packing Equal Circles in a Circle Problem on Large Scale

Arxiv

0+阅读 · 2023年3月5日

On the Robustness of ChatGPT: An Adversarial and Out-of-distribution Perspective

Arxiv

0+阅读 · 2023年3月2日

Sequential Attention for Feature Selection

Arxiv

0+阅读 · 2023年3月2日

From Show to Tell: A Survey on Image Captioning

Arxiv

15+阅读 · 2021年7月14日

相关基金

Calderon问题和边界刚性问题

国家自然科学基金

0+阅读 · 2013年12月31日

基于SURE/PURE准则的图像盲反卷积算法研究

国家自然科学基金

3+阅读 · 2013年12月31日

Intraflagellar Transport运输纤毛蛋白的分子机理

国家自然科学基金

0+阅读 · 2012年12月31日

靶向调控PLCE1基因的microRNAs在新疆哈族食管癌侵袭转移中的作用与机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

Ter94在Hedgehog信号转导途径中的作用机理

国家自然科学基金

0+阅读 · 2009年12月31日

微信扫码咨询专知VIP会员