Pith. sign in

REVIEW 2 cited by

Quantifying and maximizing the information flux in recurrent neural networks

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2301.12892 v2 pith:WF4EWT4O submitted 2023-01-30 q-bio.NC cs.NE

classification q-bio.NCcs.NE
keywords informationfluxnetworkssystemsmutualconstructionlargeleft
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
abstract

Free-running Recurrent Neural Networks (RNNs), especially probabilistic models, generate an ongoing information flux that can be quantified with the mutual information $I\left[\vec{x}(t),\vec{x}(t\!+\!1)\right]$ between subsequent system states $\vec{x}$. Although, former studies have shown that $I$ depends on the statistics of the network's connection weights, it is unclear (1) how to maximize $I$ systematically and (2) how to quantify the flux in large systems where computing the mutual information becomes intractable. Here, we address these questions using Boltzmann machines as model systems. We find that in networks with moderately strong connections, the mutual information $I$ is approximately a monotonic transformation of the root-mean-square averaged Pearson correlations between neuron-pairs, a quantity that can be efficiently computed even in large systems. Furthermore, evolutionary maximization of $I\left[\vec{x}(t),\vec{x}(t\!+\!1)\right]$ reveals a general design principle for the weight matrices enabling the systematic construction of systems with a high spontaneous information flux. Finally, we simultaneously maximize information flux and the mean period length of cyclic attractors in the state space of these dynamical networks. Our results are potentially useful for the construction of RNNs that serve as short-time memories or pattern generators.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Author-Specific Linguistic Patterns Unveiled: A Deep Learning Study on Word Class Distributions

    cs.CL 2025-01 reject novelty 4.0 of 10

    A small empirical study reports that bigram POS features with a CNN give higher author classification accuracy (59%) than unigram POS features with a feedforward network (44%), but the comparison is confounded and not...

  2. Exploring Narrative Clustering in Large Language Models: A Layerwise Analysis of BERT

    cs.CL 2025-01 reject novelty 4.0 of 10

    On a GPT-4-created corpus, BERT embeddings cluster by narrative content far more strongly than by authorial style, but the comparison is confounded by dataset design.

Pith tools