Learning Useful Representations of Recurrent Neural Network Weight Matrices

Herrmann, Vincent; Faccio, Francesco; Schmidhuber, Jürgen

Computer Science > Machine Learning

arXiv:2403.11998 (cs)

[Submitted on 18 Mar 2024 (v1), last revised 18 Jun 2024 (this version, v2)]

Title:Learning Useful Representations of Recurrent Neural Network Weight Matrices

Authors:Vincent Herrmann, Francesco Faccio, Jürgen Schmidhuber

View PDF HTML (experimental)

Abstract:Recurrent Neural Networks (RNNs) are general-purpose parallel-sequential computers. The program of an RNN is its weight matrix. How to learn useful representations of RNN weights that facilitate RNN analysis as well as downstream tasks? While the mechanistic approach directly looks at some RNN's weights to predict its behavior, the functionalist approach analyzes its overall functionality-specifically, its input-output mapping. We consider several mechanistic approaches for RNN weights and adapt the permutation equivariant Deep Weight Space layer for RNNs. Our two novel functionalist approaches extract information from RNN weights by 'interrogating' the RNN through probing inputs. We develop a theoretical framework that demonstrates conditions under which the functionalist approach can generate rich representations that help determine RNN behavior. We release the first two 'model zoo' datasets for RNN weight representation learning. One consists of generative models of a class of formal languages, and the other one of classifiers of sequentially processed MNIST digits. With the help of an emulation-based self-supervised learning technique we compare and evaluate the different RNN weight encoding techniques on multiple downstream applications. On the most challenging one, namely predicting which exact task the RNN was trained on, functionalist approaches show clear superiority.

Subjects:	Machine Learning (cs.LG)
ACM classes:	I.2.6
Cite as:	arXiv:2403.11998 [cs.LG]
	(or arXiv:2403.11998v2 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2403.11998

Submission history

From: Vincent Herrmann [view email]
[v1] Mon, 18 Mar 2024 17:32:23 UTC (6,426 KB)
[v2] Tue, 18 Jun 2024 15:27:16 UTC (6,536 KB)

Computer Science > Machine Learning

Title:Learning Useful Representations of Recurrent Neural Network Weight Matrices

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Learning Useful Representations of Recurrent Neural Network Weight Matrices

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators