PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
March 26, 2026Systems0 citationsOpen Access

FORESIGHT: Software Defects Prediction from Requirements Change Requests Using Machine Learning Methods

View Full Paper
HHHanan HelwaATAdel Taweel

Key Points

  • The aim is to enhance software defect prediction by utilizing information from requirements change requests through machine learning.
  • Developed the FORESIGHT representation model using contextual metrics from change-request characteristics.
  • Created three datasets from real-world industrial projects in Web, Mobile, and ASRS domains.
  • Evaluated the model using Random Forest, XGBoost, and Gradient Boosting classifiers.
  • Random Forest achieved the highest macro-F1 score (0.815–0.873 for primary defect types; 0.683–0.833 for defect manifestations).
  • The FORESIGHT model reliably predicts certain software defect types across all datasets.
  • Random Forest outperformed XGBoost and Gradient Boosting in every dataset-task combination.

Abstract

Software defect prediction is becoming key for software quality assurance. Traditional software defect prediction approaches have predominantly focused on analyzing code-level metrics, often overlooking valuable information available during the requirements phase. However, when a requirement change request (RCR) is issued, usually during the maintenance and evolution phase, predicting software defects provides an important preventative measure. Work in requirement-based software defect prediction methods typically focus on identifying requirement flaws, such as ambiguity or incompleteness, and fail to adequately predict defects that may manifest later in the operational software system. This paper proposes a context-driven representation model, named FORESIGHT, that predicts software defect types from requirements change requests using machine learning methods. The proposed model uses binary indicators to represent contextual metrics derived from change-request characteristics and supports multi-class prediction from both primary defect types and defect manifestation types. To build its representation model, three datasets were created from real-world industrial projects in different software domains (Web, Mobile, and ASRS). FORESIGHT was evaluated using Random Forest, XGBoost, and Gradient Boosting classifiers. Results show certain software defect types can be reliability predicted with Random Forest achieving the highest macro-F1 (0.815–0.873 for primary defect type prediction; 0.683–0.833 for defect manifestation prediction) across all three datasets, outperforming XGBoost and Gradient Boosting on every dataset–task combination. Findings show that contextual metrics from requirements change requests, structured within the FORESIGHT representation model, enable reliable pre-implementation prediction of specific defect types in deployed software systems.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Helwa et al. (2026) studied this question.

synapsesocial.com/papers/69c4ccc9fdc3bde44891846ahttps://doi.org/10.3390/systems14040342
Ask AI
Helpful
Bookmark
Share
View Full Paper