Learning Multiple Layers of Features from Tiny Images
Обучение нескольких слоев признаков на маленьких изображениях
2024-01-01
SCID: 54.1/sp8qt6nw
Discuss with AI
CIFAR-10 and CIFAR-100 labeled datasetsmulti-layer feature learningparallelization algorithm for distributed trainingtiny color images datasetunsupervised deep generative models
Figures from the paper
Abstract (AI)
April 8, 2009Groups at MIT and NYU have collected a dataset of millions of tiny colour images from the web. It is, in principle, an excellent dataset for unsupervised training of deep generative models, but previous researchers who have tried this have found it di cult to learn a good set of lters from the images. We show how to train a multi-layer generative model that learns to extract meaningful features which resemble those found in the human visual cortex. Using a novel parallelization algorithm to distribute the work among multiple machines connected on a network, we show how training such a model can be done in reasonable time. A second problematic aspect of the tiny images dataset is that there are no reliable class labels which makes it hard to use for object recognition experiments. We created two sets of reliable labels. The CIFAR-10 set has 6000 examples of each of 10 classes and the CIFAR-100 set has 600 examples of each of 100 non-overlapping classes. Using these labels, we show that object recognition is signi cantly
Key Findings
1
A multi-layer generative model can be trained on millions of tiny color images to learn meaningful features resembling those in human visual cortex.
2
A novel parallelization algorithm enables distributed training across multiple networked machines, making training time reasonable.
3
Previous difficulty learning good filters from tiny images can be overcome by the authors' training approach and model design.
4
The authors created two reliable labeled datasets from the tiny images: CIFAR-10 (6000 examples per 10 classes) and CIFAR-100 (600 examples per 100 classes).
5
Using the created labels, the labeled datasets enable significant object recognition experiments (implying improved evaluation capability for recognition tasks).
Research Object
Millions of tiny colour images from the web (the Tiny Images dataset, including CIFAR-10 and CIFAR-100 label sets)
Research Subject
Training a multi-layer (deep) generative model to learn meaningful visual features from these tiny images and creating reliable label sets to enable object recognition evaluation
Publication Details
Publication Date
2024-01-01
Journal
Publisher
ISSN
Cited by
25339
Open access PDF
Access Type
Author Information
Download PDF
Subscribe to digest
References available in scid.ai1
Cited by20
A Survey of Convolutional Neural Networks: Analysis, Applications, and Prospects2021
On the Opportunities and Risks of Foundation Models2021
CrossViT: Cross-Attention Multi-Scale Vision Transformer for Image Classification2021
An Empirical Study of Training Self-Supervised Vision Transformers2021
Transformer in Transformer2021
Decoupled Knowledge Distillation2022
CMT: Convolutional Neural Networks Meet Vision Transformers2022
Scaling Vision Transformers2022
Zero-Shot Out-of-Distribution Detection Based on the Pre-trained Model CLIP2022
RIAWELC: A Novel Dataset of Radiographic Images for Automatic Weld Defects Classification2023
An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale2020
Learning Filter Pruning Criteria for Deep Convolutional Neural Networks Acceleration2020
A survey on semi-supervised learning2019
Correlation Congruence for Knowledge Distillation2019
Relational Knowledge Distillation2019
Continual lifelong learning with neural networks: A review2019
Automated Machine Learning2019
Squeeze-and-Excitation Networks2018
Least Squares Generative Adversarial Networks2017
Channel Pruning for Accelerating Very Deep Neural Networks2017