PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
February 8, 2026Proceedings of the Institution of Mechanical Engineers Part D Journal of Automobile Engineering0 citations

Drive with SVO: Decision making for autonomous vehicles considering social value orientation

View Full Paper
CSChuanliang ShenZLZirui LITYTong Yan

Key Points

  • The aim is to develop a decision-making strategy that allows autonomous vehicles to interact effectively with other road users by incorporating social value orientation.
  • Integrated social value orientation into a reinforcement learning framework.
  • Defined a quantitative method for calculating SVO values for reward shaping.
  • Employed proximal policy optimization and soft actor-critic algorithms for agent training.
  • Validated the approach using the highway-env simulator.
  • Agents completed merging tasks while exhibiting varying interaction behaviors from self-centered to altruistic.
  • The SVO-based reward function effectively guided the agents towards diverse interaction behaviors.
  • Simulation results confirmed successful integration of SVO in decision-making processes.

Abstract

Every vehicle needs to frequently handle interactions with other road users in order to travel successfully. Enabling autonomous vehicles to interact naturally, in a manner similar to human drivers, is an important challenge in the field. This paper proposes a decision-making approach for autonomous vehicles that integrates Social Value Orientation (SVO) into a reinforcement learning framework to enable interactive behaviors in complex merging scenarios. We defined a quantitative calculation method for the SVO value, used it for reward shaping, and aimed to investigate the impact of SVO on agent behavior patterns. We employed the Proximal Policy Optimization (PPO) and Soft Actor-Critic (SAC) algorithms to train our agents, and validated our findings in the highway-env simulator. The simulation results indicate that, with the proposed decision-making approach, agents complete tasks in merging scenarios while following the intended interaction behavior patterns, displaying modes ranging from self-centered to altruistic. This confirms that the SVO-based reward function is both concise and capable of effectively guiding agents to achieve a diverse range of anticipated interaction behaviors.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Shen et al. (2026) studied this question.

synapsesocial.com/papers/698828770fc35cd7a884805dhttps://doi.org/10.1177/09544070251410300
Ask AI
Helpful
Bookmark
Share
View Full Paper