PulseAugur
EN
LIVE 07:33:43

New method uses Jungian functions to steer LLM personality

Researchers have developed a new method for controlling and interpreting Large Language Models (LLMs) by representing personality through Jungian Cognitive Functions rather than static trait frameworks. This approach, demonstrated on Llama-3.1-8B, allows for effective control over all eight cognitive functions via activation steering. The study found that personality information is concentrated in the middle transformer layers, and steering vectors show geometric relationships aligned with cognitive function distinctions, suggesting that multi-dimensional personality control is not simply a linear combination of single-function controls. AI

IMPACT This research offers a novel approach to controlling LLM personality, potentially leading to more nuanced and interpretable AI behavior.

RANK_REASON Research paper detailing a new method for LLM control. [lever_c_demoted from research: ic=1 ai=1.0]

Read on arXiv cs.AI →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

New method uses Jungian functions to steer LLM personality

COVERAGE [1]

  1. arXiv cs.AI TIER_1 English(EN) · Liu Zai (University of Glasgow), Yumeng Wang (Leiden University), Junchen Fu (University of Glasgow), Joemon M. Jose (University of Glasgow) ·

    The Geometry of Personality: Activation Steering with Jungian Cognitive Functions

    arXiv:2607.20803v1 Announce Type: cross Abstract: Activation steering enables control and interpretation of LLMs, yet existing work primarily models personality through static trait frameworks such as the Big Five. We investigate whether personality can instead be represented and…