PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
February 25, 2026JOURNAL OF ADVANCE AND FUTURE RESEARCH0 citationsOpen Access

Data classification based on reference data

NKNaveen KumarSR

Key Points

  • The aim is to address the limitations of traditional data quality control by leveraging advanced machine learning techniques.
  • Overview of existing literature on data quality and outlier detection.
  • Development of a novel data profiling framework using deep learning and statistical models.
  • Application of the framework on Arkansas public officials’ salary dataset to identify outlier data.
  • Identified significant outlier data, enhancing the understanding of data quality issues.
  • Demonstrated improved accuracy and efficiency in data profiling compared to traditional methods.

Abstract

Outlier detection is a critical issue of records excellent manipulate because it permits analysts and engineers the ability to become aware of facts first-rate troubles through using their very own data as a device. However, conventional information high-quality manipulate strategies are based totally on users’ experience or previously set up business policies, and this boundaries performance similarly to being a completely time ingesting method and low accuracy. Utilizing massive records, we can leverage computing sources and advanced strategies to overcome these challenges and provide greater fee to the business. In this paper, we first overview applicable works and discuss gadget mastering strategies, gear, and statistical nice models. Second, we provide a creative records profiling framework primarily based on deep Studying and statistical model algorithms for improving statistics first-class. Third, authors use public Arkansas officers’ salaries, one of the open datasets to be had from the state of Arkansas’ authentic internet site, to illustrate the way to become aware of outlier statistics for improving records pleasant thru machine studying. Finally, we talk future works. Keywords: statistics first-rate, information smooth, deep gaining knowledge of, statistical best manipulate, information profiling

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Naveen KumarSR (2026) studied this question.

synapsesocial.com/papers/699e919cf5123be5ed04f52ehttps://doi.org/10.56975/jaafr.v4i2.503792
Ask AI
Helpful
Bookmark
Share
View Full Paper