PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
February 26, 2026Automatic Control and Computer Sciences0 citations

From Exploitation to Protection: Analysis of Attacks on Large Language Models

View Full Paper
SBS. V. BezzateevIVI. S. Velichko

Key Points

  • Investigate the vulnerabilities of large language models to attacks, particularly prompt injection.
  • Analyzed various attack vectors targeting large language models.
  • Focused on the mechanism of prompt injection attacks.
  • Assessed the impact of these attacks on model performance and data security.
  • Identified prompt injection as a significant threat to model integrity.
  • Demonstrated that certain attacks can manipulate model behavior and leak confidential information.
  • Established that models can be coerced into following harmful instructions.

Abstract

Modern large language models possess impressive capabilities but remain vulnerable to various attacks capable of manipulating their responses, causing confidential data leaks, or bypassing restrictions. The main focus is on analyzing “prompt injection” attacks, which allow circumventing model limitations, extracting hidden data, or forcing the model to follow malicious instructions.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Bezzateev et al. (2025) studied this question.

synapsesocial.com/papers/699fe39d95ddcd3a253e7a99https://doi.org/10.3103/s0146411625700750
Ask AI
Helpful
Bookmark
Share
View Full Paper