PapeRman #6

作者: 朱小虎XiaohuZhu | 来源:发表于2019-03-01 14:20 被阅读31次

本文描述了一个新的推断智能体动机的方法。该方法基于影响图,这是一种图模型的类型,包含特别的决策和效用节点。图标准可以被用来确智能体观测动机和智能体干预动机**

Understanding Agent Incentives using Causal Influence Diagrams. Part I: Single Action Settings

Tom Everitt Pedro A. Ortega Elizabeth Barnes Shane Legg
Deepmind

Abstract

Agents are systems that optimize an objective function in an environment. Together, the goal and the environment induce secondary objectives, incentives. Modeling the agent-environment interaction in graphical models called influence diagrams, we can answer two fundamental questions about an agent's incentives directly from the graph: (1) which nodes is the agent incentivized to observe, and (2) which nodes is the agent incentivized to influence? The answers tell us which information and influence points need extra protection. For example, we may want a classifier for job applications to not use the ethnicity of the candidate, and a reinforcement learning agent not to take direct control of its reward mechanism. Different algorithms and training paradigms can lead to different influence diagrams, so our method can be used to identify algorithms with problematic incentives and help in designing algorithms with better incentives.

相关文章

  • PapeRman #6

    本文描述了一个新的推断智能体动机的方法。该方法基于影响图,这是一种图模型的类型,包含特别的决策和效用节点。图标准可...

  • Paperman

    最后一班地铁,所有人低着头,我看不清任何一个人的脸孔。我四处张望,无法寻找到可以交汇的目光。人们的面颊纷纷映着白炽...

  • PapeRman #3

    The Evolved Transformer Authors: David R. So, Chen Liang,...

  • PapeRman #4

    分布算法目前是强化学习的有趣的发现。以此为基础可以构造更具严格理论支持的强化学习算法。本系列给出最近 Google...

  • PapeRman #5

    对抗健壮性的研究非常具有挑战性。在众多研究方向中,存在一些相应的进展。本篇论文是一个较清楚的整理,有助于大家更好地...

  • 「动画推荐」让我忘不掉的《纸人》

    文/Jove桥薇 第一次看到《paperman》(中文译作“纸人”),那时候还在读大学,我正在去往汽车车站的途中,...

  • 2019-01-17 Paperman #1

    来自 DeepMind 的两篇重要论文,关于免模型规划和一般化的贡献分配研究。值得大家研读。感兴趣的小伙伴 可以私...

  • 2019-01-23 Paperman #2

    PROBABILISTIC SYMMETRY AND INVARIANT NEURAL NETWORKS Auth...

  • #知识体系精深营#六月+12次作业+第20组2小组+ynqj_a

    6-1 6-1 6-2 6-2 6-3 6-3 6-4 6-4 6-5 6-5 6-6 6-6 6-7 6-7 6...

  • 无标题文章

    1 2 2 3 5 6 6 6 6 6 6 8 3 6

网友评论

    本文标题:PapeRman #6

    本文链接:https://www.haomeiwen.com/subject/rbsduqtx.html