REAL

The Structure of Relation Decoding Linear Operators in Large Language Models

Christ, Miranda Anna and Csiszárik, Adrián and Becsó, Gergely and Varga, Dániel (2026) The Structure of Relation Decoding Linear Operators in Large Language Models. In: Advances in Neural Information Processing Systems 38. NEURAL INFORMATION PROCESSING SYSTEMS (NIPS), [s.l.], Accepted for publication. ISBN 9798331338275

[img]
Preview
Text
NeurIPS-2025-the-structure-of-relation-decoding-linear-operators-in-large-language-models-Paper-Conference.pdf

Download (1MB) | Preview

Abstract

This paper investigates the structure of linear operators introduced in Hernandez et al. [2023] that decode specific relational facts in transformer language models. We extend their single-relation findings to a collection of relations and systematically chart their organization. We show that such collections of relation decoders can be highly compressed by simple order-3 tensor networks without significant loss in decoding accuracy. To explain this surprising redundancy, we develop a cross-evaluation protocol, in which we apply each linear decoder operator to the subjects of every other relation. Our results reveal that these linear maps do not encode distinct relations, but extract recurring, coarse-grained semantic properties (e.g., country of capital city and country of food are both in the country-of-X property). This property-centric structure clarifies both the operators’ compressibility and highlights why they generalize only to new relations that are semantically close. Our findings thus interpret linear relational decoding in transformer language models as primarily property-based, rather than relation-specific.

Item Type: Book Section
Subjects: Q Science / természettudomány > QA Mathematics / matematika
SWORD Depositor: MTMT SWORD
Depositing User: MTMT SWORD
Date Deposited: 28 Sep 2026 09:58
Last Modified: 28 Sep 2026 09:58
URI: https://real.mtak.hu/id/eprint/247899

Actions (login required)

View Item View Item