Passive Learning of Lattice Automata from Recurrent Neural Networks

📅 2025-09-26
📈 Citations: 0
✨ Influential: 0
📄 PDF
🤖 AI Summary
Existing automata extraction methods for recurrent neural networks (RNNs) fail to model RNN behaviors over extremely large or infinite alphabets, hindering interpretability. Method: We propose the first passive lattice automaton extraction framework for RNNs operating over **infinite alphabets**, integrating abstract interpretation (to over-approximate the state space), passive grammar inference, symbolic execution, and an equivalence query mechanism to extract semantics-preserving lattice automata from trained RNNs. Contributions/Results: (1) First automata learning approach supporting regular languages over infinite alphabets; (2) A novel infinite-alphabet extension of Tomita grammars as a principled benchmark; (3) State-of-the-art accuracy on standard Tomita tasks and successful validation on infinite-alphabet cases. This work bridges neural-symbolic learning, formal verification, and interpretable AI.

Technology Category

Natural Language Processing: Sentence-level Semantics, Textual Inference, etc.Knowledge Representation and Reasoning: Automated Reasoning and Theorem ProvingCognitive Modeling & Cognitive Systems: Conceptual Inference and Reasoning

Application Category

Semantics and Knowledge: Methods, algorithms and applications for the development of semantic models, knowledge graphs and other forms of structured data models with machine-interpretable semanticsGraph Algorithms and Modeling for the Web: Graph neural networks and deep learning approaches for Web-related graphsSearch and Retrieval-Augmented AI: Web learning to rank, online learning, and counterfactual learning for ranking
📝 Abstract
We present a passive automata learning algorithm that can extract automata from recurrent networks with very large or even infinite alphabets. Our method combines overapproximations from the field of Abstract Interpretation and passive automata learning from the field of Grammatical Inference. We evaluate our algorithm by first comparing it with the state-of-the-art automata extraction algorithm from Recurrent Neural Networks trained on Tomita grammars. Then, we extend these experiments to regular languages with infinite alphabets, which we propose as a novel benchmark.
Problem

Research questions and friction points this paper is trying to address.

Extracting automata from recurrent neural networks
Handling large or infinite alphabet automata learning
Combining abstract interpretation with grammatical inference
Innovation

Methods, ideas, or system contributions that make the work stand out.

Passive automata learning from recurrent neural networks
Combining abstract interpretation with grammatical inference
Handling infinite alphabets through overapproximation techniques
🔎 Similar Papers
No similar papers found.