PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
March 15, 2026Cognitive Computation2 citationsOpen Access

LLM Alignment should go beyond Harmlessness–Helpfulness and incorporate Human Agency

UNUsman NaseemTCTanmoy ChakrabortyKCKai-Wei Chang

Key Points

  • This position paper addresses the misalignment of large language models and advocates for incorporating human agency in alignment strategies.
  • Analyzed operational choices affecting training and deployment of language models
  • Proposed dynamic, participatory approaches to alignment
  • Introduced the Flourishing–Justice–Autonomy (FJA) framework
  • Outlined future directions for research and practice
  • Identified key risks associated with misalignment in language models
  • Emphasized the need for pluralism and autonomy in alignment strategies
  • Suggested improvements in alignment research through participatory frameworks

Abstract

Large Language Models are transforming communication, research, and decision-making, but misalignment – when models diverge from human values, safety requirements, or user intent – poses serious risks. In this position paper, we argue that many alignment failures stem from operational choices in training and deployment. We posit that alignment should shift from static, post-training constraints toward dynamic, participatory approaches that safeguard pluralism, autonomy, and human flourishing. We outline forward-looking directions, including pluralistic evaluation, transparency, and the Flourishing–Justice–Autonomy (FJA) framework, and present a roadmap for advancing alignment research and practice.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Naseem et al. (2026) studied this question.

synapsesocial.com/papers/69b64c67b42794e3e660dc15https://doi.org/10.1007/s12559-026-10568-9
Ask AI
Helpful
Bookmark
Share
View Full Paper