Fair Exploration via Axiomatic Bargaining

Events Research Seminars Fair Exploration via Axiomatic Bargaining

Department of Decision Sciences and Managerial Economics

Exploration is often necessary in online learning to maximize long-term reward, but it comes at the cost of short-term ‘regret’. We study how this cost of exploration is shared across multiple groups. For example, in a clinical trial setting, patients who are assigned a sub-optimal treatment effectively incur the cost of exploration. When patients are associated with natural groups on the basis of, say, race or age, it is natural to ask whether the cost of exploration borne by any single group is ‘fair’.

So motivated, we introduce the ‘grouped’ bandit model. We leverage the theory of axiomatic bargaining, and the Nash bargaining solution in particular, to formalize what might constitute a fair division of the cost of exploration across groups. On the one hand, we show that any regret-optimal policy strikingly results in the least fair outcome: such policies will perversely leverage the most ‘disadvantaged’ groups when they can. More constructively, we derive policies that are optimally fair and simultaneously enjoy a small ‘price of fairness’. We illustrate the relative merits of our algorithmic framework with a case study on contextual bandits for warfarin dosing.

Date & Time

21 March 2023
08:30 - 10:30

Location

Zoom link: https://cuhk.zoom.us/j/97184040703

Meeting ID: 971 8404 0703 (Passcode: 714654)

Zoom Meeting

Speaker(s)

Prof. Jackie BAEK
Assistant Professor of Technology, Operations, and Statistics,
Stern School of Business,
New York University,
U.S.A.