Learning Neural Ranking Models Online from Implicit User Feedback

Jia, Yiling; Wang, Hongning

Full-text links:

Download:

Current browse context:

cs.IR

< prev | next >

new | recent | 2201

Computer Science > Information Retrieval

Title: Learning Neural Ranking Models Online from Implicit User Feedback

Authors: Yiling Jia, Hongning Wang

(Submitted on 17 Jan 2022)

Abstract: Existing online learning to rank (OL2R) solutions are limited to linear models, which are incompetent to capture possible non-linear relations between queries and documents. In this work, to unleash the power of representation learning in OL2R, we propose to directly learn a neural ranking model from users' implicit feedback (e.g., clicks) collected on the fly. We focus on RankNet and LambdaRank, due to their great empirical success and wide adoption in offline settings, and control the notorious explore-exploit trade-off based on the convergence analysis of neural networks using neural tangent kernel. Specifically, in each round of result serving, exploration is only performed on document pairs where the predicted rank order between the two documents is uncertain; otherwise, the ranker's predicted order will be followed in result ranking. We prove that under standard assumptions our OL2R solution achieves a gap-dependent upper regret bound of $O(\log^2(T))$, in which the regret is defined on the total number of mis-ordered pairs over $T$ rounds. Comparisons against an extensive set of state-of-the-art OL2R baselines on two public learning to rank benchmark datasets demonstrate the effectiveness of the proposed solution.

Subjects:	Information Retrieval (cs.IR); Machine Learning (cs.LG)
Cite as:	arXiv:2201.06658 [cs.IR]
	(or arXiv:2201.06658v1 [cs.IR] for this version)

Submission history

From: Yiling Jia [view email]
[v1] Mon, 17 Jan 2022 23:11:39 GMT (933kb,D)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> cs > arXiv:2201.06658

Download:

Current browse context:

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

Computer Science > Information Retrieval

Title: Learning Neural Ranking Models Online from Implicit User Feedback

Submission history