Authors:
Nicola Barbieri
1
;
Antonio Bevacqua
2
;
Marco Carnuccio
2
;
Giuseppe Manco
3
and
Ettore Ritacco
3
Affiliations:
1
University of Calabria and Italian National Research Council, Italy
;
2
University of Calabria, Italy
;
3
Italian National Research Council, Italy
Keyword(s):
Recommender Systems, Collaborative Filtering, Probabilistic Topic Models, Performance.
Related
Ontology
Subjects/Areas/Topics:
Artificial Intelligence
;
Collaborative Filtering
;
Computational Intelligence
;
Data Mining in Electronic Commerce
;
Evolutionary Computing
;
Knowledge Discovery and Information Retrieval
;
Knowledge-Based Systems
;
Machine Learning
;
Mining High-Dimensional Data
;
Soft Computing
;
Symbolic Systems
;
User Profiling and Recommender Systems
Abstract:
Probabilistic topic models are widely used in different contexts to uncover the hidden structure in large text corpora. One of the main features of these models is that generative process follows a bag-of-words assumption, i.e each token is independent from the previous one. We extend the popular Latent Dirichlet Allocation model by exploiting a conditional Markovian assumptions, where the token generation depends on the current topic and on the previous token. The resulting model is capable of accommodating temporal correlations among tokens, which better model user behavior. This is particularly significant in a collaborative filtering context, where the choice of a user can be exploited for recommendation purposes, and hence a more realistic and accurate modeling enables better recommendations. For the mentioned model we present a fast Gibbs Sampling procedure for the parameters estimation. A thorough experimental evaluation over real-word data shows the performance advantages, in
terms of recall and precision, of the proposed sequence-modeling approach.
(More)