All You Need is Ratings: A Clustering Approach to Synthetic Rating Datasets Generation
The public availability of collections containing user preferences is of vital importance for performing offline evaluations in the field of recommender systems. However, the number of rating datasets is limited because of the costs required for their creation and the fear of violating the privacy of the users by sharing them. For this reason, numerous research attempts investigated the creation of synthetic collections of ratings using generative approaches. Nevertheless, these datasets are usually not reliable enough for conducting an evaluation campaign. In this paper, we propose a method for creating synthetic datasets with a configurable number of users that mimic the characteristics of already existing ones. We empirically validated the proposed approach by exploiting the synthetic datasets for evaluating different recommenders and by comparing the results with the ones obtained using real datasets.
Code (0)
등록된 구현이 없습니다.
Tasks
AllClusteringRecommendation SystemsSimilar Papers 제목 키워드 기반
Explaining reviews and ratings with PACO: Poisson Additive Co-Clustering
Understanding a user's motivations provides valuable information beyond the ability to recommend items. Quite often this can be accomplished by perusing both ratings and review texts, since it is the latter where the rea…
ClusteringCollaborative FilteringPredicting Emotional Word Ratings using Distributional Representations and Signed Clustering
Inferring the emotional content of words is important for text-based sentiment analysis, dialogue systems and psycholinguistics, but word ratings are expensive to collect at scale and across languages or domains. We deve…
ClusteringPositionSentiment AnalysisWord SimilarityA Semi-Synthetic Dataset Generation Framework for Causal Inference in Recommender Systems
Accurate recommendation and reliable explanation are two key issues for modern recommender systems. However, most recommendation benchmarks only concern the prediction of user-item ratings while omitting the underlying c…
Causal InferenceDataset GenerationDescriptiveRecommendation Systems+1A Refined SVD Algorithm for Collaborative Filtering
Collaborative filtering tries to predict the ratings of a user over some items based on opinions of other users with similar taste. The ratings are usually given in the form of a sparse matrix, the goal being to find the…
ClusteringCollaborative FilteringA Synthetic Approach for Recommendation: Combining Ratings, Social Relations, and Reviews
Recommender systems (RSs) provide an effective way of alleviating the information overload problem by selecting personalized choices. Online social networks and user-generated content provide diverse sources for recommen…
Recommendation Systems