Vocabulary for Universal Approximation: A Linguistic Perspective of   Mapping Compositions

Yongqiang Cai

arXiv:2305.12205·cs.LG·May 24, 2024·2 cites

Vocabulary for Universal Approximation: A Linguistic Perspective of Mapping Compositions

Yongqiang Cai

PDF

Open Access

TL;DR

This paper proves the existence of a finite set of mappings, or vocabulary, that can compose to approximate any continuous function on a compact domain, revealing new insights into the power of composition in neural networks.

Contribution

It constructs a finite vocabulary of mappings with size O(d^2) that can universally approximate any continuous function through composition, bridging deep learning and linguistic structures.

Findings

01

Finite vocabulary of size O(d^2) suffices for universal approximation.

02

Any continuous function can be approximated by compositions of vocabulary elements.

03

Results suggest a new perspective on the expressive power of compositional models.

Abstract

In recent years, deep learning-based sequence modelings, such as language models, have received much attention and success, which pushes researchers to explore the possibility of transforming non-sequential problems into a sequential form. Following this thought, deep neural networks can be represented as composite functions of a sequence of mappings, linear or nonlinear, where each composition can be viewed as a \emph{word}. However, the weights of linear mappings are undetermined and hence require an infinite number of words. In this article, we investigate the finite case and constructively prove the existence of a finite \emph{vocabulary} $V = {ϕ_{i} : R^{d} \to R^{d} ∣ i = 1, ..., n}$ with $n = O (d^{2})$ for the universal approximation. That is, for any continuous mapping $f : R^{d} \to R^{d}$ , compact domain $Ω$ and $ε > 0$ , there is a sequence of…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

Topicssemigroups and automata theory · Natural Language Processing Techniques · Constraint Satisfaction and Optimization