Hölder 级带有ReLU-Sine-Exporitive 活性断线尺寸的深神经网络 (Deep Neural Networks with ReLU-Sine-Exponential Activations Break Curse of Dimensionality on Hölder Class) - 专知论文

会员服务 ·

0

维数灾难 · Continuity · Networking · Neural Networks · 近似 ·

2021 年 2 月 28 日

Deep Neural Networks with ReLU-Sine-Exponential Activations Break Curse of Dimensionality on Hölder Class

翻译：Hölder 级带有ReLU-Sine-Exporitive 活性断线尺寸的深神经网络

Yuling Jiao,Yanming Lai,Xiliang Lu,Zhijian Yang

In this paper, we construct neural networks with ReLU, sine and $2^x$ as activation functions. For general continuous $f$ defined on $[0,1]^d$ with continuity modulus $\omega_f(\cdot)$, we construct ReLU-sine-$2^x$ networks that enjoy an approximation rate $\mathcal{O}(\omega_f(\sqrt{d})\cdot2^{-M}+\omega_{f}\left(\frac{\sqrt{d}}{N}\right))$, where $M,N\in \mathbb{N}^{+}$ denote the hyperparameters related to widths of the networks. As a consequence, we can construct ReLU-sine-$2^x$ network with the depth $5$ and width $\max\left\{\left\lceil2d^{3/2}\left(\frac{3\mu}{\epsilon}\right)^{1/{\alpha}}\right\rceil,2\left\lceil\log_2\frac{3\mu d^{\alpha/2}}{2\epsilon}\right\rceil+2\right\}$ that approximates $f\in \mathcal{H}_{\mu}^{\alpha}([0,1]^d)$ within a given tolerance $\epsilon >0$ measured in $L^p$ norm $p\in[1,\infty)$, where $\mathcal{H}_{\mu}^{\alpha}([0,1]^d)$ denotes the H\"older continuous function class defined on $[0,1]^d$ with order $\alpha \in (0,1]$ and constant $\mu > 0$. Therefore, the ReLU-sine-$2^x$ networks overcome the curse of dimensionality on $\mathcal{H}_{\mu}^{\alpha}([0,1]^d)$. In addition to its supper expressive power, functions implemented by ReLU-sine-$2^x$ networks are (generalized) differentiable, enabling us to apply SGD to train.

翻译：在本文中, 我们以ReLU、 Prine 和 2 $xx 的激活功能构建神经网络。对于以 $[0, 1美元] 定义的一般连续美元, 且具有连续性的元模 $\ omega_ f(\ cddt) $, 我们建造了RLU- sine-2 $xx$ 的网络, 享有近似速率$\ mathcal{O} (\ omega_ f( sqrt})\ cdot2\\ momea} left (\ coffi2\ m) leg_ left (\ c3\ moq_ sqrent) (\\\ dqright_ right_ rentral_ $0, 0. 0, NN\\ listal\\\\ r\\\\\\\\\\ r\\\\\\\\\\\\\\\\\\\\\ ma\\\\\\\\\\ cal\ cal1美元。

0

相关内容

维数灾难

维度灾难是指在高维空间中分析和组织数据时出现的各种现象，这些现象在低维设置（例如日常体验的三维物理空间）中不会发生。

【IJCAJ 2020】多通道神经网络 Multi-Channel Graph Neural Networks

【IJCAJ 2020】多通道神经网络 Multi-Channel Graph Neural Networks

专知会员服务

25+阅读 · 2020年7月19日

【ICML2020】深度神经网络置信感知学习，Conﬁdence-Aware Learning for Deep Neural Networks

【ICML2020】深度神经网络置信感知学习，Conﬁdence-Aware Learning for Deep Neural Networks

专知会员服务

71+阅读 · 2020年7月6日

【ICML2020】用于图结构化数据的卷积核网络，Convolutional Kernel Networks for Graph-Structured Data

【ICML2020】用于图结构化数据的卷积核网络，Convolutional Kernel Networks for Graph-Structured Data

专知会员服务

42+阅读 · 2020年6月29日

神经网络与形式语言综述，12页pdf，A Survey of Neural Networks and Formal Languages

神经网络与形式语言综述，12页pdf，A Survey of Neural Networks and Formal Languages

专知会员服务

19+阅读 · 2020年6月4日

神经网络的拓扑结构，TOPOLOGY OF DEEP NEURAL NETWORKS

神经网络的拓扑结构，TOPOLOGY OF DEEP NEURAL NETWORKS

专知会员服务

30+阅读 · 2020年4月15日

【CMU】图卷积神经网络中的池化综述，Pooling in Graph Convolutional Neural Network

【CMU】图卷积神经网络中的池化综述，Pooling in Graph Convolutional Neural Network

专知会员服务

45+阅读 · 2020年4月8日

和积网络综述论文，Sum-product networks: A survey，24页pdf

和积网络综述论文，Sum-product networks: A survey，24页pdf

专知会员服务

23+阅读 · 2020年4月3日

最大均方差正则化贝叶斯神经网络，Bayesian Neural Networks With Maximum Mean Discrepancy Regularization

最大均方差正则化贝叶斯神经网络，Bayesian Neural Networks With Maximum Mean Discrepancy Regularization

专知会员服务

53+阅读 · 2020年3月5日

【ICLR2020】深度神经网络优化轨迹的平衡点，The Break-Even Point on Optimization Trajectories of Deep Neural Networks

【ICLR2020】深度神经网络优化轨迹的平衡点，The Break-Even Point on Optimization Trajectories of Deep Neural Networks

专知会员服务

32+阅读 · 2020年2月27日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

32+阅读 · 2019年10月17日

图机器学习 2.2-2.4 Properties of Networks, Random Graph

图机器学习 2.2-2.4 Properties of Networks, Random Graph

图与推荐

10+阅读 · 2020年3月28日

过参数化、剪枝和网络结构搜索

过参数化、剪枝和网络结构搜索

极市平台

16+阅读 · 2019年11月24日

分布式并行架构Ray介绍

分布式并行架构Ray介绍

CreateAMind

9+阅读 · 2019年8月9日

【TED】生命中的每一年的智慧

【TED】生命中的每一年的智慧

英语演讲视频每日一推

9+阅读 · 2019年1月29日

无监督元学习表示学习

无监督元学习表示学习

CreateAMind

26+阅读 · 2019年1月4日

误差反向传播——RNN

误差反向传播——RNN

统计学习与视觉计算组

18+阅读 · 2018年9月6日

神经网络学习率设置

神经网络学习率设置

机器学习研究会

4+阅读 · 2018年3月3日

分布式TensorFlow入门指南

分布式TensorFlow入门指南

机器学习研究会

4+阅读 · 2017年11月28日

Capsule Networks解析

Capsule Networks解析

机器学习研究会

11+阅读 · 2017年11月12日

Auto-Encoding GAN

Auto-Encoding GAN

CreateAMind

7+阅读 · 2017年8月4日

MLDS: A Dataset for Weight-Space Analysis of Neural Networks

MLDS: A Dataset for Weight-Space Analysis of Neural Networks

Arxiv

0+阅读 · 2021年4月21日

Fusing Sufficient Dimension Reduction with Neural Networks

Arxiv

0+阅读 · 2021年4月20日

The Price of Anarchy in Routing Games as a Function of the Demand

Arxiv

0+阅读 · 2021年4月20日

The Reads-From Equivalence for the TSO and PSO Memory Models

Arxiv

0+阅读 · 2021年4月19日

Neural Network Approximation: Three Hidden Layers Are Enough

Arxiv

0+阅读 · 2021年4月19日

On the approximation of functions by tanh neural networks

Arxiv

0+阅读 · 2021年4月18日

On the $Φ$-Stability and Related Conjectures

Arxiv

0+阅读 · 2021年4月18日

Barrier-Free Large-Scale Sparse Tensor Accelerator (BARISTA) For Convolutional Neural Networks

Arxiv

0+阅读 · 2021年4月18日

Stochastic Gradient Descent Optimizes Over-parameterized Deep ReLU Networks

Arxiv

8+阅读 · 2018年11月21日

A Dual Approach to Scalable Verification of Deep Networks

A Dual Approach to Scalable Verification of Deep Networks

Arxiv

3+阅读 · 2018年8月3日

VIP会员

文章信息

相关主题

Neural Networks

相关VIP内容

【IJCAJ 2020】多通道神经网络 Multi-Channel Graph Neural Networks

【IJCAJ 2020】多通道神经网络 Multi-Channel Graph Neural Networks

专知会员服务

25+阅读 · 2020年7月19日

【ICML2020】深度神经网络置信感知学习，Conﬁdence-Aware Learning for Deep Neural Networks

【ICML2020】深度神经网络置信感知学习，Conﬁdence-Aware Learning for Deep Neural Networks

专知会员服务

71+阅读 · 2020年7月6日

【ICML2020】用于图结构化数据的卷积核网络，Convolutional Kernel Networks for Graph-Structured Data

【ICML2020】用于图结构化数据的卷积核网络，Convolutional Kernel Networks for Graph-Structured Data

专知会员服务

42+阅读 · 2020年6月29日

神经网络与形式语言综述，12页pdf，A Survey of Neural Networks and Formal Languages

神经网络与形式语言综述，12页pdf，A Survey of Neural Networks and Formal Languages

专知会员服务

19+阅读 · 2020年6月4日

神经网络的拓扑结构，TOPOLOGY OF DEEP NEURAL NETWORKS

神经网络的拓扑结构，TOPOLOGY OF DEEP NEURAL NETWORKS

专知会员服务

30+阅读 · 2020年4月15日

【CMU】图卷积神经网络中的池化综述，Pooling in Graph Convolutional Neural Network

【CMU】图卷积神经网络中的池化综述，Pooling in Graph Convolutional Neural Network

专知会员服务

45+阅读 · 2020年4月8日

和积网络综述论文，Sum-product networks: A survey，24页pdf

和积网络综述论文，Sum-product networks: A survey，24页pdf

专知会员服务

23+阅读 · 2020年4月3日

最大均方差正则化贝叶斯神经网络，Bayesian Neural Networks With Maximum Mean Discrepancy Regularization

最大均方差正则化贝叶斯神经网络，Bayesian Neural Networks With Maximum Mean Discrepancy Regularization

专知会员服务

53+阅读 · 2020年3月5日

【ICLR2020】深度神经网络优化轨迹的平衡点，The Break-Even Point on Optimization Trajectories of Deep Neural Networks

【ICLR2020】深度神经网络优化轨迹的平衡点，The Break-Even Point on Optimization Trajectories of Deep Neural Networks

专知会员服务

32+阅读 · 2020年2月27日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

32+阅读 · 2019年10月17日

热门VIP内容

相关资讯

图机器学习 2.2-2.4 Properties of Networks, Random Graph

图机器学习 2.2-2.4 Properties of Networks, Random Graph

图与推荐

10+阅读 · 2020年3月28日

过参数化、剪枝和网络结构搜索

过参数化、剪枝和网络结构搜索

极市平台

16+阅读 · 2019年11月24日

分布式并行架构Ray介绍

分布式并行架构Ray介绍

CreateAMind

9+阅读 · 2019年8月9日

【TED】生命中的每一年的智慧

【TED】生命中的每一年的智慧

英语演讲视频每日一推

9+阅读 · 2019年1月29日

无监督元学习表示学习

无监督元学习表示学习

CreateAMind

26+阅读 · 2019年1月4日

误差反向传播——RNN

误差反向传播——RNN

统计学习与视觉计算组

18+阅读 · 2018年9月6日

神经网络学习率设置

神经网络学习率设置

机器学习研究会

4+阅读 · 2018年3月3日

分布式TensorFlow入门指南

分布式TensorFlow入门指南

机器学习研究会

4+阅读 · 2017年11月28日

Capsule Networks解析

Capsule Networks解析

机器学习研究会

11+阅读 · 2017年11月12日

Auto-Encoding GAN

Auto-Encoding GAN

CreateAMind

7+阅读 · 2017年8月4日

相关论文

MLDS: A Dataset for Weight-Space Analysis of Neural Networks

MLDS: A Dataset for Weight-Space Analysis of Neural Networks

Arxiv

0+阅读 · 2021年4月21日

Fusing Sufficient Dimension Reduction with Neural Networks

Arxiv

0+阅读 · 2021年4月20日

The Price of Anarchy in Routing Games as a Function of the Demand

Arxiv

0+阅读 · 2021年4月20日

The Reads-From Equivalence for the TSO and PSO Memory Models

Arxiv

0+阅读 · 2021年4月19日

Neural Network Approximation: Three Hidden Layers Are Enough

Arxiv

0+阅读 · 2021年4月19日

On the approximation of functions by tanh neural networks

Arxiv

0+阅读 · 2021年4月18日

On the $Φ$-Stability and Related Conjectures

Arxiv

0+阅读 · 2021年4月18日

Barrier-Free Large-Scale Sparse Tensor Accelerator (BARISTA) For Convolutional Neural Networks

Arxiv

0+阅读 · 2021年4月18日

Stochastic Gradient Descent Optimizes Over-parameterized Deep ReLU Networks

Arxiv

8+阅读 · 2018年11月21日

A Dual Approach to Scalable Verification of Deep Networks

A Dual Approach to Scalable Verification of Deep Networks

Arxiv

3+阅读 · 2018年8月3日

微信扫码咨询专知VIP会员