Fatigue-Aware Bandits for Dependent Click Models

Cao, J; Sun, W; Shen, ZM; Ett, M

File Download

There are no files associated with this item.

Links for fulltext

(May Require Subscription)

Publisher Website: 10.1609/aaai.v34i04.5735
Find via

Supplementary

Citations:
Appears in Collections:
- President's Office: Conference papers

Conference Paper: Fatigue-Aware Bandits for Dependent Click Models

Title	Fatigue-Aware Bandits for Dependent Click Models
Authors	Cao, J Sun, W Shen, ZM Ett, M
Issue Date	2020
Publisher	AAAI Press. The Journal's web site is located at https://aaai.org/Library/AAAI/aaai-library.php
Citation	Proceedings of the 34th Association for the Advancement of Artificial Intelligence (AAAI) Conference on Artificial Intelligence (AAAI-20), 32nd Conference on Innovative Applications of Artificial Intelligence & the 10th Symposium on Educational Advances in Artificial Intelligence, New York, NY, USA, 7-12 February 2020, v. 34 n. 4, p. 3341-3348 How to Cite? DOI: http://dx.doi.org/10.1609/aaai.v34i04.5735
Abstract	As recommender systems send a massive amount of content to keep users engaged, users may experience fatigue which is contributed by 1) an overexposure to irrelevant content, 2) boredom from seeing too many similar recommendations. To address this problem, we consider an online learning setting where a platform learns a policy to recommend content that takes user fatigue into account. We propose an extension of the Dependent Click Model (DCM) to describe users' behavior. We stipulate that for each piece of content, its attractiveness to a user depends on its intrinsic relevance and a discount factor which measures how many similar contents have been shown. Users view the recommended content sequentially and click on the ones that they find attractive. Users may leave the platform at any time, and the probability of exiting is higher when they do not like the content. Based on user's feedback, the platform learns the relevance of the underlying content as well as the discounting effect due to content fatigue. We refer to this learning task as “fatigue-aware DCM Bandit” problem. We consider two learning scenarios depending on whether the discounting effect is known. For each scenario, we propose a learning algorithm which simultaneously explores and exploits, and characterize its regret bound.
Description	AAAI-20 Technical Tracks 4 - Section: AAAI Technical Track: Machine Learning
Persistent Identifier	http://hdl.handle.net/10722/310146
ISSN	2159-5399

DC Field	Value	Language
dc.contributor.author	Cao, J	-
dc.contributor.author	Sun, W	-
dc.contributor.author	Shen, ZM	-
dc.contributor.author	Ett, M	-
dc.date.accessioned	2022-01-24T02:24:32Z	-
dc.date.available	2022-01-24T02:24:32Z	-
dc.date.issued	2020	-
dc.identifier.citation	Proceedings of the 34th Association for the Advancement of Artificial Intelligence (AAAI) Conference on Artificial Intelligence (AAAI-20), 32nd Conference on Innovative Applications of Artificial Intelligence & the 10th Symposium on Educational Advances in Artificial Intelligence, New York, NY, USA, 7-12 February 2020, v. 34 n. 4, p. 3341-3348	-
dc.identifier.issn	2159-5399	-
dc.identifier.uri	http://hdl.handle.net/10722/310146	-
dc.description	AAAI-20 Technical Tracks 4 - Section: AAAI Technical Track: Machine Learning	-
dc.description.abstract	As recommender systems send a massive amount of content to keep users engaged, users may experience fatigue which is contributed by 1) an overexposure to irrelevant content, 2) boredom from seeing too many similar recommendations. To address this problem, we consider an online learning setting where a platform learns a policy to recommend content that takes user fatigue into account. We propose an extension of the Dependent Click Model (DCM) to describe users' behavior. We stipulate that for each piece of content, its attractiveness to a user depends on its intrinsic relevance and a discount factor which measures how many similar contents have been shown. Users view the recommended content sequentially and click on the ones that they find attractive. Users may leave the platform at any time, and the probability of exiting is higher when they do not like the content. Based on user's feedback, the platform learns the relevance of the underlying content as well as the discounting effect due to content fatigue. We refer to this learning task as “fatigue-aware DCM Bandit” problem. We consider two learning scenarios depending on whether the discounting effect is known. For each scenario, we propose a learning algorithm which simultaneously explores and exploits, and characterize its regret bound.	-
dc.language	eng	-
dc.publisher	AAAI Press. The Journal's web site is located at https://aaai.org/Library/AAAI/aaai-library.php	-
dc.relation.ispartof	Proceedings of the AAAI Conference on Artificial Intelligence	-
dc.title	Fatigue-Aware Bandits for Dependent Click Models	-
dc.type	Conference_Paper	-
dc.identifier.email	Shen, ZM: maxshen@hku.hk	-
dc.identifier.authority	Shen, ZM=rp02779	-
dc.identifier.doi	10.1609/aaai.v34i04.5735	-
dc.identifier.hkuros	331473	-
dc.identifier.volume	34	-
dc.identifier.issue	4	-
dc.identifier.spage	3341	-
dc.identifier.epage	3348	-
dc.publisher.place	United States	-

File Download

Links for fulltext

(May Require Subscription)

Supplementary

Conference Paper: Fatigue-Aware Bandits for Dependent Click Models

Export via OAI-PMH Interface in XML Formats

OR

Export to Other Non-XML Formats