AI ALIGNMENT FORUM
AF

Counterfactuals
Personal Blog

5

Open Problems Regarding Counterfactuals: An Introduction For Beginners

by Diffractor
18th Jul 2017
1 min read
6

5

This is a linkpost for https://www.overleaf.com/read/vsmxjhbvsxky
Counterfactuals
Personal Blog
Open Problems Regarding Counterfactuals: An Introduction For Beginners
2riceissa
4Diffractor
0Vanessa Kosoy
0Vanessa Kosoy
New Comment
4 comments, sorted by
top scoring
Click to highlight new comments since: Today at 6:27 PM
[-]riceissa6y20

The link no longer works (I get "This project has not yet been moved into the new version of Overleaf. You will need to log in and move it in order to continue working on it.") Would you be willing to re-post it or move it so that it is visible?

Reply
[-]Diffractor6y40

See if this works.

Reply
[-]Vanessa Kosoy8y00

Note that the problem with exploration already arises in ordinary reinforcement learning, without going into "exotic" decision theories. Regarding the question of why humans don't seem to have this problem, I think it is a combination of

  • The universe is regular (which is related to what you said about "we can't see any plausible causal way it could happen"), so a Bayes-optimal policy with a simplicity prior has something going for it. On the other hand, sometimes you do need to experiment, so this can't be the only explanation.

  • Any individual human has parents that teach em things, including things like "touching a hot stove is dangerous." Later in life, ey can draw on much of the knowledge accumulated by human civilization. This tunnels the exploration into safe channels, analogously to the role of the advisor in my recent posts.

  • One may say that the previous point only passes the recursive buck, since we can consider all of humanity to be the "agent". From this perspective, it seems that the universe just happens to be relatively safe, in the sense that it's pretty hard for an individual human to do something that will irreparably damage all of humanity... or at least it was the case during most of human history.

  • In addition, we have some useful instincts baked in by evolution (e.g. probably some notion of existing in a three dimensional space with objects that interact mechanically). Again, you could zoom further out and say evolution works because it's hard to create a species that will wipe out all life.

Reply
[-]Vanessa Kosoy8y00

Typos on page 5:

  • "random explanation" should be "random exploration"
  • "Alpa" should be "Alpha"
Reply
Moderation Log
More from Diffractor
View more
Curated and popular this week
4Comments