ARTFEED — Contemporary Art Intelligence

arXiv Study Tests Internal Action Maps in Qwen3-4B Without Global Affine Closure

ai-technology · 2026-08-17

A recent submission to arXiv (2608.13626) explores the potential of hidden state signals in large language models to facilitate reusable action maps without requiring global affine closure. This preprint evaluates action maps that are adjusted without a source attaining its inherent post-action activation and composition. The authors structure their experiments as an evidence lattice, confirming the geometric branch on a known affine S_5 carrier, where all held-source folds successfully navigate one-step, composition, inverse, decoding, and commutativity gates. While structured curvature and held-domain conjugacy lead to a steady increase in error, only 23 of the 30 strongest cells activate a closure gate, limiting calibration rather than universalizing it. In the post-trained Qwen/Qwen3-4B, the mean held-entity error for frozen final-token h28 affine maps is .519, compared to .398 for cross-fit within the test domain. Seven randomized entity splits and map geometry fail to endorse a purely entity-specific explanation. Earlier h4/h16 layers demonstrate superior fitting for one-step transitions, yet h4 conflict-state decoding is inadequate, and lexical controls remain unexamined. The research enhances the understanding of internal representations and action maps in transformer models, influencing interpretability and alignment.

Key facts

  • arXiv paper 2608.13626v1 tests internal action maps.
  • Action maps fitted without a source are tested for composition and activation.
  • Evidence lattice organizes the tests.
  • Geometric branch validated on affine S_5 carrier.
  • All held-source folds pass one-step, composition, inverse, decoding, and commutativity gates.
  • Structured curvature and held-domain conjugacy raise error monotonically.
  • Only 23/30 strongest cells flip a closure gate.
  • In Qwen3-4B, h28 affine maps have mean held-entity error .519 vs .398 for cross-fit.
  • Seven randomized entity splits do not support entity-specific account.
  • Earlier layers h4/h16 fit one-step transitions better but h4 decoding is weak.

Entities

Institutions

  • arXiv
  • Qwen

Sources