ARTFEED — Contemporary Art Intelligence

LLM Persona Encoding Found in Final Decoder Layers

ai-technology · 2026-07-29

A new study from arXiv (2505.24539) investigates how personas—defined by human characteristics, values, and beliefs—are encoded in large language models. Using dimension reduction and pattern recognition, researchers identified the final third of decoder layers as the region where persona representations diverge most. Within these layers, overlapping activations were observed for ethical perspectives like moral nihilism and utilitarianism, indicating polysemy, while political ideologies such as conservatism and liberalism showed distinct patterns. The findings span multiple pre-trained decoder-only LLMs.

Key facts

  • Study on persona encoding in LLMs
  • Uses dimension reduction and pattern recognition
  • Final third of decoder layers show greatest divergence
  • Overlapping activations for moral nihilism and utilitarianism
  • Distinct activations for conservatism and liberalism
  • Analyzed multiple pre-trained decoder-only LLMs
  • arXiv paper 2505.24539
  • Published as replace-cross announcement

Entities

Institutions

  • arXiv

Sources