DeepAI AI Chat
Log In Sign Up

Personalized Reward Learning with Interaction-Grounded Learning (IGL)

by   Jessica Maghakian, et al.

In an era of countless content offerings, recommender systems alleviate information overload by providing users with personalized content suggestions. Due to the scarcity of explicit user feedback, modern recommender systems typically optimize for the same fixed combination of implicit feedback signals across all users. However, this approach disregards a growing body of work highlighting that (i) implicit signals can be used by users in diverse ways, signaling anything from satisfaction to active dislike, and (ii) different users communicate preferences in different ways. We propose applying the recent Interaction Grounded Learning (IGL) paradigm to address the challenge of learning representations of diverse user communication modalities. Rather than taking a fixed, human-designed reward function, IGL is able to learn personalized reward functions for different users and then optimize directly for the latent user satisfaction. We demonstrate the success of IGL with experiments using simulations as well as with real-world production traces.


page 1

page 2

page 3

page 4


Towards Learning Reward Functions from User Interactions

In the physical world, people have dynamic preferences, e.g., the same s...

A Personalized Subreddit Recommendation Engine

This paper aims to improve upon the generic recommendations that Reddit ...

Interaction-Grounded Learning

Consider a prosthetic arm, learning to adapt to its user's control signa...

Interaction-Grounded Learning with Action-inclusive Feedback

Consider the problem setting of Interaction-Grounded Learning (IGL), in ...

The Personalization Paradox: the Conflict between Accurate User Models and Personalized Adaptive Systems

Personalized adaptation technology has been adopted in a wide range of d...

What Do You Mean I'm Funny? Personalizing the Joke Skill of a Voice-Controlled Virtual Assistant

A considerable part of the success experienced by Voice-controlled virtu...

Brain Topography Adaptive Network for Satisfaction Modeling in Interactive Information Access System

With the growth of information on the Web, most users heavily rely on in...