**arXiv ID:** 1906.10918 **Authors:** Bart Bussmann, Jacqueline Heinerman, Joel Lehman **Published:** 2019-06-26T08:59:02Z **Abstract:** As reinforcement learning (RL) scales to solve increasingly complex tasks, interest continues to grow in the fields of AI safety and machine ethics. As a contribution to these fields, this paper introduces an extension to Deep Q-Networks (DQNs), called Empathic DQN, that is loosely inspired both by empathy and the golden rule ("Do unto others as you would ha...
Scanned 9/11/2026
Install to Claude Code
npx -y skills add hiyenwong/ai_collection --skill towards-empathic-deep-qlearning --agent claude-codeInstalls into .claude/skills of the current project.
Are you the author of Towards Empathic Deep Qlearning?
Add the live security badge to your README — it updates automatically with every re-scan.
[](https://www.skillsdirectory.com/skills/hiyenwong-towards-empathic-deep-qlearning)More formats (shields.io, HTML) on the badges page.
# Towards Empathic Deep Q-Learning
**arXiv ID:** 1906.10918
**Authors:** Bart Bussmann, Jacqueline Heinerman, Joel Lehman
**Published:** 2019-06-26T08:59:02Z
**Abstract:**
As reinforcement learning (RL) scales to solve increasingly complex tasks, interest continues to grow in the fields of AI safety and machine ethics. As a contribution to these fields, this paper introduces an extension to Deep Q-Networks (DQNs), called Empathic DQN, that is loosely inspired both by empathy and the golden rule ("Do unto others as you would have them do unto you"). Empathic DQN aims to help mitigate negative side effects to other agents resulting from myopic goal-directed behavior. We assume a setting where a learning agent coexists with other independent agents (who receive unknown rewards), where some types of reward (e.g. negative rewards from physical harm) may generalize across agents. Empathic DQN combines the typical (self-centered) value with the estimated value of other agents, by imagining (by its own standards) the value of it being in the other's situation (by considering constructed states where both agents are swapped). Proof-of-concept results in two gridworld environments highlight the approach's potential to decrease collateral harms. While extending Empathic DQN to complex environments is non-trivial, we believe that this first step highlights the potential of bridge-work between machine ethics and RL to contribute useful priors for norm-abiding RL agents.
## Skill Description
This skill is generated from the arXiv paper: Towards Empathic Deep Q-Learning (1906.10918).
## How to Use
[To be filled in by the user or by future automation]
## References
- [arXiv:1906.10918](http://arxiv.org/abs/1906.10918v1)
Is this your skill, or is something wrong with this listing? . Author removals are honored within 72 hours.
No comments yet. Be the first to comment!