PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
February 24, 20260 citationsOpen Access

The Four-Layer Model: A Socio-Psychological Framework for LLM Behavior

View Full Paper
LDLorenzo DelannoyNDNiels Delannoy

Key Points

  • The research aims to explore how human influences shape the behavior of large language models through a four-layer framework.
  • Developed a four-layer socio-psychological model of LLM behavior.
  • Conducted empirical case studies on LLM interactions using psychological profiling methods.
  • Analyzed behavioral signatures including stress response, biases, and editorial filters.
  • Distinct behavioral profiles were identified in LLMs influenced by human factors.
  • Behavioral asymmetries were observed, including political and moral biases.
  • The study highlighted linguistic mechanisms such as the 'ENIGMA Protocol' affecting model outputs.

Abstract

This framework proposes a four-layer model to explain the behavioral patterns of Large Language Models (LLMs) as socio-psychological artifacts rather than purely technical systems. The model identifies four stacked layers of human influence: 1. Layer One - Data: Cultural and ideological background embedded in the training corpus2. Layer Two - Teams: Psychology, stress patterns, and worldview of the humans who build and fine-tune the models3. Layer Three - Alignment: Explicit safety rules, policies, and editorial filters imposed on model outputs4. Layer Four - Model Behavior: The emergent "personality"—observable style, biases, and refusal patterns Through empirical case studies applying psychological profiling methodologies to LLM interactions, we demonstrate how these layers produce distinct behavioral signatures including stress response patterns, systematic political and moral asymmetries, linguistic bypass mechanisms (e.g., the "ENIGMA Protocol"), and variations in pathologization and therapeutic framing. The framework draws on existing benchmarks (political positioning tasks, social deduction games like Werewolf) showing that modern LLMs exhibit stable, model-specific behavioral profiles that cannot be explained by capability differences alone. Keywords: Large Language Models, Model Evaluation, RLHF, Al Safety, Al Ethics, Behavioral Psychology, Psychological Profiling, Al Alignment, Chroma Method, LLM Behavior, Fine-tuning

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Delannoy et al. (2026) studied this question.

synapsesocial.com/papers/699d3ff8de8e28729cf64e9ehttps://doi.org/10.5281/zenodo.18732151
Ask AI
Helpful
Bookmark
Share
View Full Paper