Exactness at Inference: A Representational Criterion for Out-of-Distribution Generalization

📅 2026-09-21
📈 Citations: 0
Influential: 0
📄 PDF
🤖 AI Summary
本文提出了一种基于表示等价性的标准,用于解决模型在训练分布外的泛化问题,并通过Tensor Logic方法实现精确推理。
📝 Abstract
A model generalizes outside its training distribution only when it computes a representation structurally equivalent to the generating mechanism, not an approximation fitted to it. Such equivalence is necessary for exactness in and out of distribution, and extrapolation is governed by this exactness at inference, whatever its realization. Tensor Logic shows this: a zero-temperature contraction is equivalent to discrete logic, deducing in place with no artefact extracted, its tensors Boolean, its embeddings orthonormal, only its arithmetic continuous. Lacking infinite recursion it reaches Datalog, not Prolog, and though exact over closed domains it needs external memory to bind a novel entity. The criterion needs neither a discrete representation nor an extracted expression, and constrains inference, not training: an exact marginal in $[0,1]$ passes, a Neural Network thresholded to a hard label does not. Logic Tensor Networks fail it, while differentiable ILP and Tensor Logic at $T=0$ pass. Piecewise-affine extrapolation divergence and an inability to bind novel entities are two faces of a shortfall in exact representability. For hybrid architectures, a propagation rule follows: the output inherits the bounds of every fitted estimator on its path, explaining which axes fail in equivariant models and the ARC-AGI induction/transduction split. Only an exact hypothesis class certifies what the training data leave underdetermined: on a law-derived partition it finds the $56.3\%$ of distant queries that are answerable, which ensembles meet with false confidence and distance metrics rank backwards. Common inductive biases, from symmetries to memory, reach exactness only because humans inject them, an argument for inducing exact representations rather than fitting surrogates whose residuals, even at the arithmetic floor in training, diverge outside the data and compound under composition.
Problem

Research questions and friction points this paper is trying to address.

out-of-distribution generalization
representational criterion
exactness at inference
Innovation

Methods, ideas, or system contributions that make the work stand out.

exactness at inference
Tensor Logic
out-of-distribution generalization
representation learning
inductive biases
🔎 Similar Papers
No similar papers found.
💼 Related Jobs
No related jobs found.
F
Filipe Marinho Rocha
Faculdade de Ciências da Universidade do Porto, Porto, Portugal; INESC TEC, Campus da FEUP, Porto, Portugal
I
Inês Dutra
Faculdade de Ciências da Universidade do Porto, Porto, Portugal; CINTESIS@RISE-Health, Porto, Portugal
Vítor Santos Costa
Vítor Santos Costa
Faculdade de Ciências da Universidade do Porto, Porto, Portugal; INESC TEC, Campus da FEUP, Porto, Portugal
L
Luís Paulo Reis
Faculdade de Engenharia da Universidade do Porto, Rua Dr. Roberto Frias, Porto, Portugal; LIACC, FEUP, Porto, Portugal