Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Yuji Tokozume

Between-class Learning for Image Classification

Apr 08, 2018

Yuji Tokozume, Yoshitaka Ushiku, Tatsuya Harada

Figure 1 for Between-class Learning for Image Classification

Figure 2 for Between-class Learning for Image Classification

Figure 3 for Between-class Learning for Image Classification

Figure 4 for Between-class Learning for Image Classification

Abstract:In this paper, we propose a novel learning method for image classification called Between-Class learning (BC learning). We generate between-class images by mixing two images belonging to different classes with a random ratio. We then input the mixed image to the model and train the model to output the mixing ratio. BC learning has the ability to impose constraints on the shape of the feature distributions, and thus the generalization ability is improved. BC learning is originally a method developed for sounds, which can be digitally mixed. Mixing two image data does not appear to make sense; however, we argue that because convolutional neural networks have an aspect of treating input data as waveforms, what works on sounds must also work on images. First, we propose a simple mixing method using internal divisions, which surprisingly proves to significantly improve performance. Second, we propose a mixing method that treats the images as waveforms, which leads to a further improvement in performance. As a result, we achieved 19.4% and 2.26% top-1 errors on ImageNet-1K and CIFAR-10, respectively.

* 11 pages, 8 figures, published as a conference paper at CVPR 2018

Via

Access Paper or Ask Questions

Learning from Between-class Examples for Deep Sound Recognition

Feb 28, 2018

Yuji Tokozume, Yoshitaka Ushiku, Tatsuya Harada

Figure 1 for Learning from Between-class Examples for Deep Sound Recognition

Figure 2 for Learning from Between-class Examples for Deep Sound Recognition

Figure 3 for Learning from Between-class Examples for Deep Sound Recognition

Figure 4 for Learning from Between-class Examples for Deep Sound Recognition

Abstract:Deep learning methods have achieved high performance in sound recognition tasks. Deciding how to feed the training data is important for further performance improvement. We propose a novel learning method for deep sound recognition: Between-Class learning (BC learning). Our strategy is to learn a discriminative feature space by recognizing the between-class sounds as between-class sounds. We generate between-class sounds by mixing two sounds belonging to different classes with a random ratio. We then input the mixed sound to the model and train the model to output the mixing ratio. The advantages of BC learning are not limited only to the increase in variation of the training data; BC learning leads to an enlargement of Fisher's criterion in the feature space and a regularization of the positional relationship among the feature distributions of the classes. The experimental results show that BC learning improves the performance on various sound recognition networks, datasets, and data augmentation schemes, in which BC learning proves to be always beneficial. Furthermore, we construct a new deep sound recognition network (EnvNet-v2) and train it with BC learning. As a result, we achieved a performance surpasses the human level.

* 13 pages, 6 figures, published as a conference paper at ICLR 2018

Via

Access Paper or Ask Questions