PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
September 16, 2025Journal of Machine and Computing0 citations

A Robust Evaluation of Bug Pre-Processing and Classification Logic using NLP Computation with Machine Learning Technique

View Full Paper
MRMamatha RacharlaLPLalitha Surya Kumari PSASharada Adepu

Key Points

  • The proposed model achieves 89% accuracy in classifying bug reports, addressing critical software maintenance challenges.
  • Utilizing both TF-IDF and BERT for feature extraction enhances performance beyond traditional methods.
  • The research involved a confidential dataset from a private company with inputs from 800 employees.
  • The approach highlights the importance of automating bug report processing for improved quality assurance.

Abstract

Enhancing the software maintenance greatly depends on the precise and prompt handing out of bug reports according to their bug-category and importance. To resolve the aforementioned problems, an automated method of classifying and ranking bug reports is required. Numerous scholars have recently looked into the automated classification and prioritization of bug reports. But not much has been accomplished in this area. During software development, the most crucial stages are testing and maintenance. In these phases of development activity, bug reports are essential. When software modules are being tested, the software quality assurance team creates a bug report. But the main issue that comes up while analysing bug data that is written in normal text. As a result, processing and extracting information from it is extremely challenging. The aforementioned requirements are the driving force for this research. The Proposed research suggested creating a hybrid model that takes advantage of machine learning models' contextual awareness as well as more conventional feature extraction methods (such as TF-IDF). A downstream classifier (such as an SVM, logistic regression) can receive these two feature sets (one from TF-IDF and the other from BERT) after they have been concatenated. This enables the model to take advantage of the extensive contextual relationships that BERT captures as well as the statistical importance of phrases (TF-IDF) These two approaches were used separately in the earlier research, which resulted in less performance. The research made use of a confidential dataset that was acquired from a private company upon request for performing testing, the data included from eight hundred employees. To aid in model training, bug keywords were first taken out of the bug description field. The results shows that proposed model achieves 89% accuracy.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Racharla et al. (2025) studied this question.

synapsesocial.com/papers/68d4508931b076d99fa58750https://doi.org/10.53759/7669/jmc202505172
Ask AI
Helpful
Bookmark
Share
View Full Paper