Local Clustering in Contextual Multi-Armed Bandits

Ban, Yikun; He, Jingrui

Computer Science > Machine Learning

arXiv:2103.00063 (cs)

[Submitted on 26 Feb 2021 (v1), last revised 24 Mar 2023 (this version, v3)]

Title:Local Clustering in Contextual Multi-Armed Bandits

Authors:Yikun Ban, Jingrui He

View PDF

Abstract:We study identifying user clusters in contextual multi-armed bandits (MAB). Contextual MAB is an effective tool for many real applications, such as content recommendation and online advertisement. In practice, user dependency plays an essential role in the user's actions, and thus the rewards. Clustering similar users can improve the quality of reward estimation, which in turn leads to more effective content recommendation and targeted advertising. Different from traditional clustering settings, we cluster users based on the unknown bandit parameters, which will be estimated incrementally. In particular, we define the problem of cluster detection in contextual MAB, and propose a bandit algorithm, LOCB, embedded with local clustering procedure. And, we provide theoretical analysis about LOCB in terms of the correctness and efficiency of clustering and its regret bound. Finally, we evaluate the proposed algorithm from various aspects, which outperforms state-of-the-art baselines.

Comments:	13 pages
Subjects:	Machine Learning (cs.LG)
Cite as:	arXiv:2103.00063 [cs.LG]
	(or arXiv:2103.00063v3 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2103.00063

Submission history

From: Yikun Ban [view email]
[v1] Fri, 26 Feb 2021 21:59:29 UTC (567 KB)
[v2] Mon, 18 Jul 2022 04:21:21 UTC (568 KB)
[v3] Fri, 24 Mar 2023 15:05:00 UTC (568 KB)

Full-text links:

Access Paper:

view license

Current browse context:

cs.LG

< prev | next >

new | recent | 2021-03

Change to browse by:

References & Citations

DBLP - CS Bibliography

listing | bibtex

Yikun Ban
Jingrui He

export BibTeX citation

Computer Science > Machine Learning

Title:Local Clustering in Contextual Multi-Armed Bandits

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Local Clustering in Contextual Multi-Armed Bandits

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators