Geometric Reasoning in the Embedding Space

📅 2025-04-02
📈 Citations: 0
✨ Influential: 0
📄 PDF
🤖 AI Summary
This work investigates the geometric reasoning capabilities of Graph Neural Networks (GNNs) and Transformers in embedding space, focusing on reconstructing implicit 2D geometric structures—specifically, predicting spatial coordinates and recovering underlying shapes from point sets defined by discrete geometric constraints on a 2D grid. Method: We propose a geometry-aware GNN architecture explicitly designed for geometric reasoning. Crucially, it operates without explicit coordinate supervision, relying solely on relational graph structure. Contribution/Results: We demonstrate, for the first time, that the learned node embeddings spontaneously organize into a low-dimensional subspace preserving neighborhood relationships—effectively recovering the latent 2D grid topology. Quantitatively, our GNN significantly outperforms Transformer baselines in both prediction accuracy and scalability. Qualitative analysis confirms that the embedding space faithfully encodes geometric structure, providing strong evidence of implicit geometric modeling capacity. These findings establish a novel, interpretable paradigm for spatial reasoning grounded in learned embeddings.

Technology Category

Knowledge Representation and Reasoning: Geometric, Spatial, and Temporal ReasoningMachine Learning: Graph-based Machine LearningReasoning under Uncertainty: Relational Probabilistic Models

Application Category

Graph Algorithms and Modeling for the Web: Graph embeddings and representation learning for Web-related graphsSemantics and Knowledge: Methods to enhance, augment, integrate or synergize semantic models such as knowledge graphs and LLMsSearch and Retrieval-Augmented AI: Web query analysis, representation and understanding
📝 Abstract
In this contribution, we demonstrate that Graph Neural Networks and Transformers can learn to reason about geometric constraints. We train them to predict spatial position of points in a discrete 2D grid from a set of constraints that uniquely describe hidden figures containing these points. Both models are able to predict the position of points and interestingly, they form the hidden figures described by the input constraints in the embedding space during the reasoning process. Our analysis shows that both models recover the grid structure during training so that the embeddings corresponding to the points within the grid organize themselves in a 2D subspace and reflect the neighborhood structure of the grid. We also show that the Graph Neural Network we design for the task performs significantly better than the Transformer and is also easier to scale.
Problem

Research questions and friction points this paper is trying to address.

Predict spatial positions from geometric constraints
Learn hidden figures in embedding space
Compare GNN and Transformer performance
Innovation

Methods, ideas, or system contributions that make the work stand out.

Graph Neural Networks learn geometric constraints
Transformers predict spatial positions in grid
Embeddings organize in 2D subspace reflecting structure
🔎 Similar Papers
No similar papers found.
💼 Related Jobs
No related jobs found.
J
Jan Hrula
Czech Institute of Informatics, Robotics and Cybernetics, Czech Technical University in Prague, Czechia
D
David Mojvzivsek
University of Ostrava, Ostrava, Czechia
J
Jiri Janevcek
University of Ostrava, Ostrava, Czechia
D
David Herel
Czech Institute of Informatics, Robotics and Cybernetics, Czech Technical University in Prague, Czechia
M
Mikolavs Janota
Czech Institute of Informatics, Robotics and Cybernetics, Czech Technical University in Prague, Czechia