**arXiv ID:** 2605.21451 **Authors:** Soumendu Sundar Mukherjee, Himasish Talukdar **Published:** 2026-05-20T17:42:34Z **Abstract:** Universal approximation theorems provide a mathematical explanation for the expressive power of neural networks. They assert that, under mild conditions on the activation function, feedforward neural networks are dense in broad function classes, such as continuous functions on compact subsets of $\mathbb{R}^d$, $L^p$ spaces, or Sobolev spaces. Over the past four...
Scanned 9/11/2026
Install to Claude Code
npx -y skills add hiyenwong/ai_collection --skill approximation-theory-for-neural-networks-old-and-new --agent claude-codeInstalls into .claude/skills of the current project.
Are you the author of Approximation Theory For Neural Networks Old And New?
Add the live security badge to your README — it updates automatically with every re-scan.
[](https://www.skillsdirectory.com/skills/hiyenwong-approximation-theory-for-neural-networks-old-and-n)More formats (shields.io, HTML) on the badges page.
# Approximation Theory for Neural Networks: Old and New
**arXiv ID:** 2605.21451
**Authors:** Soumendu Sundar Mukherjee, Himasish Talukdar
**Published:** 2026-05-20T17:42:34Z
**Abstract:**
Universal approximation theorems provide a mathematical explanation for the expressive power of neural networks. They assert that, under mild conditions on the activation function, feedforward neural networks are dense in broad function classes, such as continuous functions on compact subsets of $\mathbb{R}^d$, $L^p$ spaces, or Sobolev spaces. Over the past four decades, these qualitative universality results have evolved into a rich quantitative theory addressing approximation rates, parameter efficiency, and the role of architectural features such as depth and width. This survey presents several glimpses into this theory. We review classical density results for single-hidden-layer networks, as well as quantitative bounds that relate approximation error to network size and smoothness assumptions on target functions. Particular emphasis is placed on depth--width trade-offs and on results demonstrating that deeper architectures can achieve superior parameter efficiency for structured function classes. In addition to standard feedforward neural networks, we also review recent developments on Kolmogorov--Arnold Networks (KANs), which offer an alternative architectural paradigm and whose approximation-theoretic properties have begun to attract significant theoretical attention.
## Skill Description
This skill is generated from the arXiv paper: Approximation Theory for Neural Networks: Old and New (2605.21451).
## How to Use
[To be filled in by the user or by future automation]
## References
- [arXiv:2605.21451](http://arxiv.org/abs/2605.21451v1)
Is this your skill, or is something wrong with this listing? Request removal or report an issue. Author removals are honored within 72 hours.
No comments yet. Be the first to comment!