使用条件变式自动自动编码器零射击学习生成模型 (A Generative Model For Zero Shot Learning Using Conditional Variational Autoencoders)

Zero shot learning in Image Classification refers to the setting where images from some novel classes are absent in the training data but other information such as natural language descriptions or attribute vectors of the classes are available. This setting is important in the real world since one may not be able to obtain images of all the possible classes at training. While previous approaches have tried to model the relationship between the class attribute space and the image space via some kind of a transfer function in order to model the image space correspondingly to an unseen class, we take a different approach and try to generate the samples from the given attributes, using a conditional variational autoencoder, and use the generated samples for classification of the unseen classes. By extensive testing on four benchmark datasets, we show that our model outperforms the state of the art, particularly in the more realistic generalized setting, where the training classes can also appear at the test time along with the novel classes.

翻译：图像分类中的零镜头学习是指培训数据中缺少某些新类图像的设置,但有其他信息,如这些类的自然语言描述或属性矢量等。这种设置在现实世界中很重要,因为人们可能无法在培训中获得所有可能课程的图像。虽然以前的做法试图通过某种传输功能来模拟该类属性空间与图像空间之间的关系,以便模拟与无形类相对应的图像空间,但我们采取了不同的做法,试图利用一个有条件的变异自动编码器从特定属性中生成样本,并利用生成的样本对隐形类进行分类。通过对四个基准数据集的广泛测试,我们展示了我们的模型优于艺术状态,特别是在更现实的普及环境中,在测试时,培训课程也可以与新类同时出现。

相关内容

自编码器

关注 140

自动编码器是一种人工神经网络，用于以无监督的方式学习有效的数据编码。自动编码器的目的是通过训练网络忽略信号“噪声”来学习一组数据的表示（编码），通常用于降维。与简化方面一起，学习了重构方面，在此，自动编码器尝试从简化编码中生成尽可能接近其原始输入的表示形式，从而得到其名称。基本模型存在几种变体，其目的是迫使学习的输入表示形式具有有用的属性。自动编码器可有效地解决许多应用问题，从面部识别到获取单词的语义。

【ACL2020】多模态信息抽取，365页ppt

专知会员服务

150+阅读 · 2020年7月6日

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

专知会员服务

96+阅读 · 2020年3月12日

生成式对抗网络先验贝叶斯推断，Bayesian Inference with Generative Adversarial Network Priors

专知会员服务

28+阅读 · 2020年2月18日

深度强化学习策略梯度教程，53页ppt

专知会员服务

184+阅读 · 2020年2月1日