**arXiv ID:** 2205.08836 **Authors:** Christoph Linse, Thomas Martinetz **Published:** 2022-05-18T10:08:28Z **Abstract:** Recent findings have shown that highly over-parameterized Neural Networks generalize without pretraining or explicit regularization. It is achieved with zero training error, i.e., complete over-fitting by memorizing the training data. This is surprising, since it is completely against traditional machine learning wisdom. In our empirical study we fortify these findings in ...
Scanned 9/11/2026
Install to Claude Code
npx -y skills add hiyenwong/ai_collection --skill large-neural-networks-learning-from-scratch-with-very-few-data-and-without-explicit-regularization --agent claude-codeInstalls into .claude/skills of the current project.
Are you the author of Large Neural Networks Learning From Scratch With Very Few Data And Without Explicit Regularization?
Add the live security badge to your README — it updates automatically with every re-scan.
[](https://www.skillsdirectory.com/skills/hiyenwong-large-neural-networks-learning-from-scratch-with-v)More formats (shields.io, HTML) on the badges page.
# Large Neural Networks Learning from Scratch with Very Few Data and without Explicit Regularization
**arXiv ID:** 2205.08836
**Authors:** Christoph Linse, Thomas Martinetz
**Published:** 2022-05-18T10:08:28Z
**Abstract:**
Recent findings have shown that highly over-parameterized Neural Networks generalize without pretraining or explicit regularization. It is achieved with zero training error, i.e., complete over-fitting by memorizing the training data. This is surprising, since it is completely against traditional machine learning wisdom. In our empirical study we fortify these findings in the domain of fine-grained image classification. We show that very large Convolutional Neural Networks with millions of weights do learn with only a handful of training samples and without image augmentation, explicit regularization or pretraining. We train the architectures ResNet018, ResNet101 and VGG19 on subsets of the difficult benchmark datasets Caltech101, CUB_200_2011, FGVCAircraft, Flowers102 and StanfordCars with 100 classes and more, perform a comprehensive comparative study and draw implications for the practical application of CNNs. Finally, we show that a randomly initialized VGG19 with 140 million weights learns to distinguish airplanes and motorbikes with up to 95% accuracy using only 20 training samples per class.
## Skill Description
This skill is generated from the arXiv paper: Large Neural Networks Learning from Scratch with Very Few Data and without Explicit Regularization (2205.08836).
## How to Use
[To be filled in by the user or by future automation]
## References
- [arXiv:2205.08836](http://arxiv.org/abs/2205.08836v2)
Is this your skill, or is something wrong with this listing? Request removal or report an issue. Author removals are honored within 72 hours.
No comments yet. Be the first to comment!