PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
April 23, 2026Empirical Software Engineering1 citationsOpen Access

Less is more: usefulness of data flow diagrams and large language models for security threat validation

WMWinnie MbakaKTKatja Tuma

Key Points

  • To determine the effectiveness of graphical models and LLM-generated advice in validating identified security threats.
  • Conducted a controlled experiment with practitioners
  • Used a pilot study with MSc students and a think-aloud study with practitioners
  • Surveyed 68 recruited practitioners to gather insights on analysis material.
  • Participants found graphical models equally useful compared to LLMs
  • LLMs provided perceived usefulness despite not always offering conclusive advice
  • Less analysis material was considered more effective for threat validation.

Abstract

The arrival of recent cybersecurity standards has raised the bar for security assessments in organizations, but existing techniques require a high manual effort. Threat analysis and risk assessment are used to identify security threats for new or refactored systems. Still, there is a lack of definition-of-done, so identified threats have to be validated which slows down the analysis. Existing literature has focused on the overall effectiveness of threat analysis, but no previous work has investigated what material must the analysts use to effectively validate the identified security threats. We conduct a controlled experiment with practitioners to investigate whether having some analysis material (either the system’s graphical model or LLM-generated advice) is better than none, and whether having both the system’s graphical model and LLM-generated advice is better than having only one of them. We run a pilot of the experiment with 41 MSc students, a think-aloud study with three practitioners, and the experiment survey with 68 recruited practitioners. Our main findings suggest that, in terms of additional material needed for threat validation, less is more. We also find that participants perceived the graphical model as equally useful compared to LLMs and that, despite LLMs not always providing conclusive advice, practitioners still perceived it as somewhat useful. The experimental material and data analysis scripts is publicly available in a replication package.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Mbaka et al. (2026) studied this question.

synapsesocial.com/papers/69e9bb9e85696592c86ed391https://doi.org/10.1007/s10664-026-10837-z
Ask AI
Helpful
Bookmark
Share
View Full Paper