PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
February 25, 2026Future Internet0 citationsOpen Access

Optimizing Collaborative Filtering for Accurate Rating Predictions in Very Sparse Datasets

View Full Paper
SLSofia-Anna LapadakiJNJohn I. NanosDMDionisis Margaris

Key Points

  • The aim is to optimize collaborative filtering settings for accurate rating predictions in very sparse datasets.
  • Conducted multiparameter experiments to investigate optimal similarity metrics and K values.
  • Evaluated user similarity based on shared ratings among users.
  • Analyzed the impact of varying K values on prediction coverage and computational efficiency.
  • Identified a small K value led to low prediction coverage, harming recommendations.
  • Found that larger K values increased memory use and prediction generation time.
  • Established optimal settings to enhance accuracy in very sparse datasets.

Abstract

Collaborative filtering is one of the most widely used methods for user rating prediction in recommender systems. To evaluate a collaborative filtering system, rating datasets are typically used, which comprise thousands to millions of records consisting of user–item–rating tuples. Initially, a similarity metric is used to quantify the closeness between each user and every other user in the dataset, typically based on the ratings that each pair of users has given to the same items. Subsequently, the K users having the largest similarity to the target user are used to produce rating predictions, which lead to recommendations. A particularly challenging case arises when the rating dataset is very sparse. In this scenario, it is difficult not only to find users with commonly rated items but also to determine the optimal similarity metric and suitable values for variable K. Setting a small value for K results in extremely low prediction coverage, leading to unsuccessful recommendations, while setting a very large K value increases memory requirements and prediction/recommendation generation time. Through a multiparameter experiment, this work aims to determine the optimal settings for rating predictions when very sparse datasets are used in collaborative filtering recommender systems.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Lapadaki et al. (2026) studied this question.

synapsesocial.com/papers/699e9177f5123be5ed04efe8https://doi.org/10.3390/fi18020114
Ask AI
Helpful
Bookmark
Share
View Full Paper