Gabriel Peyré: The Expressive Power of Large Language Models

Seminars - Analysis and Applied Mathematics Seminar
Speakers
Gabriel Peyré, CNRS École Normale Supérieure
12:00pm - 1:15pm
Room 3-E4-SR03, 3rd floor via Röntgen 1

Abstract: Large language models process vast sequences of input tokens by alternating between classical multi-layer perceptron layers and self-attention mechanisms. While the approximation capabilities of perceptrons are relatively well understood, those of attention mechanisms remain less explored. In this talk, I will compare the proof techniques and approximation results associated with these two types of layers, emphasizing key open questions that connect large language models with approximation theory in infinite-dimensional spaces representing input token distributions.

For further information please contact elisur.magrini@unibocconi.it