Off-line vs. On-line Evaluation of Recommender Systems in Small E-commerce

Peska, Ladislav; Vojtas, Peter

doi:10.1145/3372923.3404781

Computer Science > Information Retrieval

arXiv:1809.03186 (cs)

[Submitted on 10 Sep 2018 (v1), last revised 9 Jun 2020 (this version, v3)]

Title:Off-line vs. On-line Evaluation of Recommender Systems in Small E-commerce

Authors:Ladislav Peska, Peter Vojtas

View PDF

Abstract:In this paper, we present our work towards comparing on-line and off-line evaluation metrics in the context of small e-commerce recommender systems. Recommending on small e-commerce enterprises is rather challenging due to the lower volume of interactions and low user loyalty, rarely extending beyond a single session. On the other hand, we usually have to deal with lower volumes of objects, which are easier to discover by users through various browsing/searching GUIs.
The main goal of this paper is to determine applicability of off-line evaluation metrics in learning true usability of recommender systems (evaluated on-line in A/B testing). In total 800 variants of recommending algorithms were evaluated off-line w.r.t. 18 metrics covering rating-based, ranking-based, novelty and diversity evaluation. The off-line results were afterwards compared with on-line evaluation of 12 selected recommender variants and based on the results, we tried to learn and utilize an off-line to on-line results prediction model.
Off-line results shown a great variance in performance w.r.t. different metrics with the Pareto front covering 68\% of the approaches. Furthermore, we observed that on-line results are considerably affected by the novelty of users. On-line metrics correlates positively with ranking-based metrics (AUC, MRR, nDCG) for novice users, while too high values of diversity and novelty had a negative impact on the on-line results for them. For users with more visited items, however, the diversity became more important, while ranking-based metrics relevance gradually decrease.

Comments:	Submitted to ACM Hypertext 2020 Conference
Subjects:	Information Retrieval (cs.IR)
Cite as:	arXiv:1809.03186 [cs.IR]
	(or arXiv:1809.03186v3 [cs.IR] for this version)
	https://doi.org/10.48550/arXiv.1809.03186
Related DOI:	https://doi.org/10.1145/3372923.3404781

Submission history

From: Ladislav Peska [view email]
[v1] Mon, 10 Sep 2018 08:52:55 UTC (888 KB)
[v2] Fri, 7 Feb 2020 20:25:20 UTC (1,915 KB)
[v3] Tue, 9 Jun 2020 07:45:50 UTC (3,826 KB)

Computer Science > Information Retrieval

Title:Off-line vs. On-line Evaluation of Recommender Systems in Small E-commerce

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Information Retrieval

Title:Off-line vs. On-line Evaluation of Recommender Systems in Small E-commerce

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators