为基于实际交叉核实的超参数选择两个共同问题提供理论指导 (Toward Theoretical Guidance for Two Common Questions in Practical Cross-Validation based Hyperparameter Selection) - 专知论文

会员服务 ·

0

Performer · 超参数 · Guidance · 情景 · MoDELS ·

2023 年 1 月 12 日

Toward Theoretical Guidance for Two Common Questions in Practical Cross-Validation based Hyperparameter Selection

翻译：为基于实际交叉核实的超参数选择两个共同问题提供理论指导

Parikshit Ram,Alexander G. Gray,Horst C. Samulowitz,Gregory Bramble

from arxiv, Extended version of the paper appearing at the SIAM International Conference on Data Mining 2023 (SDM23)

We show, to our knowledge, the first theoretical treatments of two common questions in cross-validation based hyperparameter selection: (1) After selecting the best hyperparameter using a held-out set, we train the final model using {\em all} of the training data -- since this may or may not improve future generalization error, should one do this? (2) During optimization such as via SGD (stochastic gradient descent), we must set the optimization tolerance $\rho$ -- since it trades off predictive accuracy with computation cost, how should one set it? Toward these problems, we introduce the {\em hold-in risk} (the error due to not using the whole training data), and the {\em model class mis-specification risk} (the error due to having chosen the wrong model class) in a theoretical view which is simple, general, and suggests heuristics that can be used when faced with a dataset instance. In proof-of-concept studies in synthetic data where theoretical quantities can be controlled, we show that these heuristics can, respectively, (1) always perform at least as well as always performing retraining or never performing retraining, (2) either improve performance or reduce computational overhead by $2\times$ with no loss in predictive performance.

翻译：据我们所知,在交叉验证基于超参数的选择中,我们展示了对两个共同问题的最初理论处理方法,即交叉校准基于超分计选择最佳超参数之后:(1) 在使用一个搁置的套件选择最佳超参数之后,我们用培训数据中的[所有]来培训最后模型 -- -- 因为这样可能或不会改进未来的概括错误,我们是否应该这样做?(2) 在优化过程中,例如通过SGD(随机梯度梯度下降),我们必须设定优化容忍度$rho$ -- -- 因为它与计算成本的预测准确性相交换,如何设定它?为了解决这些问题,我们引入了`它们持有风险'(由于没有使用整个培训数据造成的错误) 和`它们类模型错误区分风险}(由于选择错误的模型类别),是否应该这样做?(2) 在使用简单、一般的理论观点和暗示在面对数据集时可以使用的超理论实例。在对可控制理论量的合成数据进行有说服力的研究中,我们表明,这些超率可以分别(1) 以美元进行至少进行经常性的再培训或从未进行再培训,或者从不进行高层再分析,至少进行2美元,至少进行两次再培训。

0

相关内容

Performer

【干货书】数据分析优化，Optimization for Modern Data Analysis，117页pdf

【干货书】数据分析优化，Optimization for Modern Data Analysis，117页pdf

专知会员服务

66+阅读 · 2023年2月15日

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

专知会员服务

76+阅读 · 2022年6月28日

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

专知会员服务

60+阅读 · 2022年4月22日

机器学习损失函数概述，Loss Functions in Machine Learning

机器学习损失函数概述，Loss Functions in Machine Learning

专知会员服务

84+阅读 · 2022年3月19日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

机器学习入门的经验与建议

机器学习入门的经验与建议

专知会员服务

94+阅读 · 2019年10月10日

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

专知会员服务

79+阅读 · 2019年10月10日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

专知会员服务

65+阅读 · 2019年10月9日

VCIP 2022 Call for Demos

VCIP 2022 Call for Demos

CCF多媒体专委会

1+阅读 · 2022年6月6日

VCIP 2022 Call for Special Session Proposals

VCIP 2022 Call for Special Session Proposals

CCF多媒体专委会

1+阅读 · 2022年4月1日

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

ACM TOMM Call for Papers

ACM TOMM Call for Papers

CCF多媒体专委会

2+阅读 · 2022年3月23日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

局部学习的特征选择：Local-Learning-Based Feature Selection

局部学习的特征选择：Local-Learning-Based Feature Selection

我爱读PAMI

14+阅读 · 2019年9月20日

强化学习三篇论文避免遗忘等

强化学习三篇论文避免遗忘等

CreateAMind

20+阅读 · 2019年5月24日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

ARB抑制miR-193a表达促进早期糖尿病肾病壁层上皮细胞-足细胞转分化研究

国家自然科学基金

0+阅读 · 2015年12月31日

高斯序列与过程的极值理论

国家自然科学基金

2+阅读 · 2015年12月31日

长链非编码RNA-VEC1340靶定KLF4在血管内皮细胞损伤中的调控及机制研究

国家自然科学基金

0+阅读 · 2014年12月31日

信号稀疏表示的广义测不准原理研究

国家自然科学基金

1+阅读 · 2014年12月31日

Calderon问题和边界刚性问题

国家自然科学基金

0+阅读 · 2013年12月31日

渐近锥流形上色散方程的研究

国家自然科学基金

0+阅读 · 2013年12月31日

关于Cayley图的若干研究

国家自然科学基金

0+阅读 · 2012年12月31日

淫羊藿总黄酮调控骨性关节炎p38MAPK信号转导通路的研究

国家自然科学基金

0+阅读 · 2010年12月31日

myostatin调控脂肪酸代谢的分子机制

国家自然科学基金

0+阅读 · 2009年12月31日

非交换最速下降法的一致渐近研究

国家自然科学基金

0+阅读 · 2008年12月31日

The Influence Function of Graphical Lasso Estimators

Arxiv

0+阅读 · 2023年3月8日

Manually Selecting The Data Function for Supervised Learning of small datasets

Arxiv

0+阅读 · 2023年3月7日

Multilevel Monte Carlo methods for stochastic convection-diffusion eigenvalue problems

Arxiv

0+阅读 · 2023年3月7日

Cross-validatory model selection for Bayesian autoregressions with exogenous regressors

Arxiv

0+阅读 · 2023年3月7日

Exploring Deep Models for Practical Gait Recognition

Arxiv

0+阅读 · 2023年3月6日

Convergence Rates for Non-Log-Concave Sampling and Log-Partition Estimation

Arxiv

0+阅读 · 2023年3月6日

Iterative Approximate Cross-Validation

Arxiv

0+阅读 · 2023年3月5日

Optimizing Low Dimensional Functions over the Integers

Arxiv

0+阅读 · 2023年3月4日

Estimating marginal treatment effects from single studies and indirect treatment comparisons: When are standardization-based methods preferable to inverse probability of treatment weighting?

Arxiv

0+阅读 · 2023年3月4日

Agent-based Collaborative Random Search for Hyper-parameter Tuning and Global Function Optimization

Arxiv

0+阅读 · 2023年3月3日

VIP会员

文章信息

相关主题

相关VIP内容

【干货书】数据分析优化，Optimization for Modern Data Analysis，117页pdf

【干货书】数据分析优化，Optimization for Modern Data Analysis，117页pdf

专知会员服务

66+阅读 · 2023年2月15日

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

专知会员服务

76+阅读 · 2022年6月28日

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

ICLR 2022杰出论文公布：7篇论文获得，清华朱军课题组摘得

专知会员服务

60+阅读 · 2022年4月22日

机器学习损失函数概述，Loss Functions in Machine Learning

机器学习损失函数概述，Loss Functions in Machine Learning

专知会员服务

84+阅读 · 2022年3月19日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

机器学习入门的经验与建议

机器学习入门的经验与建议

专知会员服务

94+阅读 · 2019年10月10日

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

专知会员服务

79+阅读 · 2019年10月10日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

专知会员服务

65+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

【博士论文】面向可扩展深度神经网络的预测编码：理论与实践

如何快速获取数百万架无人机？

EMNLP 2025 | RTQA：递归思想求解复杂的时间知识图谱问答

组合式零样本学习综述

相关资讯

VCIP 2022 Call for Demos

VCIP 2022 Call for Demos

CCF多媒体专委会

1+阅读 · 2022年6月6日

VCIP 2022 Call for Special Session Proposals

VCIP 2022 Call for Special Session Proposals

CCF多媒体专委会

1+阅读 · 2022年4月1日

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

ACM TOMM Call for Papers

ACM TOMM Call for Papers

CCF多媒体专委会

2+阅读 · 2022年3月23日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

局部学习的特征选择：Local-Learning-Based Feature Selection

局部学习的特征选择：Local-Learning-Based Feature Selection

我爱读PAMI

14+阅读 · 2019年9月20日

强化学习三篇论文避免遗忘等

强化学习三篇论文避免遗忘等

CreateAMind

20+阅读 · 2019年5月24日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

相关论文

The Influence Function of Graphical Lasso Estimators

Arxiv

0+阅读 · 2023年3月8日

Manually Selecting The Data Function for Supervised Learning of small datasets

Arxiv

0+阅读 · 2023年3月7日

Multilevel Monte Carlo methods for stochastic convection-diffusion eigenvalue problems

Arxiv

0+阅读 · 2023年3月7日

Cross-validatory model selection for Bayesian autoregressions with exogenous regressors

Arxiv

0+阅读 · 2023年3月7日

Exploring Deep Models for Practical Gait Recognition

Arxiv

0+阅读 · 2023年3月6日

Convergence Rates for Non-Log-Concave Sampling and Log-Partition Estimation

Arxiv

0+阅读 · 2023年3月6日

Iterative Approximate Cross-Validation

Arxiv

0+阅读 · 2023年3月5日

Optimizing Low Dimensional Functions over the Integers

Arxiv

0+阅读 · 2023年3月4日

Estimating marginal treatment effects from single studies and indirect treatment comparisons: When are standardization-based methods preferable to inverse probability of treatment weighting?

Arxiv

0+阅读 · 2023年3月4日

Agent-based Collaborative Random Search for Hyper-parameter Tuning and Global Function Optimization

Arxiv

0+阅读 · 2023年3月3日

相关基金

ARB抑制miR-193a表达促进早期糖尿病肾病壁层上皮细胞-足细胞转分化研究

国家自然科学基金

0+阅读 · 2015年12月31日

高斯序列与过程的极值理论

国家自然科学基金

2+阅读 · 2015年12月31日

长链非编码RNA-VEC1340靶定KLF4在血管内皮细胞损伤中的调控及机制研究

国家自然科学基金

0+阅读 · 2014年12月31日

信号稀疏表示的广义测不准原理研究

国家自然科学基金

1+阅读 · 2014年12月31日

Calderon问题和边界刚性问题

国家自然科学基金

0+阅读 · 2013年12月31日

渐近锥流形上色散方程的研究

国家自然科学基金

0+阅读 · 2013年12月31日

关于Cayley图的若干研究

国家自然科学基金

0+阅读 · 2012年12月31日

淫羊藿总黄酮调控骨性关节炎p38MAPK信号转导通路的研究

国家自然科学基金

0+阅读 · 2010年12月31日

myostatin调控脂肪酸代谢的分子机制

国家自然科学基金

0+阅读 · 2009年12月31日

非交换最速下降法的一致渐近研究

国家自然科学基金

0+阅读 · 2008年12月31日

微信扫码咨询专知VIP会员