Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:HCMD-zero: Learning Value Aligned Mechanisms from Data

Feb 21, 2022

Jan Balaguer, Raphael Koster, Ari Weinstein, Lucy Campbell-Gillingham, Christopher Summerfield, Matthew Botvinick, Andrea Tacchetti

Figure 1 for HCMD-zero: Learning Value Aligned Mechanisms from Data

Figure 2 for HCMD-zero: Learning Value Aligned Mechanisms from Data

Figure 3 for HCMD-zero: Learning Value Aligned Mechanisms from Data

Figure 4 for HCMD-zero: Learning Value Aligned Mechanisms from Data

Share this with someone who'll enjoy it:

Abstract:Artificial learning agents are mediating a larger and larger number of interactions among humans, firms, and organizations, and the intersection between mechanism design and machine learning has been heavily investigated in recent years. However, mechanism design methods make strong assumptions on how participants behave (e.g. rationality), or on the kind of knowledge designers have access to a priori (e.g. access to strong baseline mechanisms). Here we introduce HCMD-zero, a general purpose method to construct mechanism agents. HCMD-zero learns by mediating interactions among participants, while remaining engaged in an electoral contest with copies of itself, thereby accessing direct feedback from participants. Our results on the Public Investment Game, a stylized resource allocation game that highlights the tension between productivity, equality and the temptation to free-ride, show that HCMD-zero produces competitive mechanism agents that are consistently preferred by human participants over baseline alternatives, and does so automatically, without requiring human knowledge, and by using human data sparingly and effectively Our detailed analysis shows HCMD-zero elicits consistent improvements over the course of training, and that it results in a mechanism with an interpretable and intuitive policy.

View paper on

OpenReview

Share this with someone who'll enjoy it:

Title:HCMD-zero: Learning Value Aligned Mechanisms from Data

Paper and Code