Discussion about this post

User's avatar
Scott Alexander's avatar

I'm confused about the bomb example.

It seems like your example hinges upon the FDT agent picking Left.

But it also says that the predictor with a one in a trillion trillion error rate predicted you would pick Right.

If all of this is correct, it seems like you're hinging your example on this being the one time in a trillion trillion when the predictor was wrong.

But decision theories shouldn't be judged on whether they work well in unbelievably rare edge cases that you would never encounter in a million lifetimes.

Compare Lottery Decision Theory:

> you have the choice to buy a lottery ticket for $100. There is a one in a trillion trillion chance you will win. If you win you get $1 million. Should you do it? Before answering, keep in mind that, unbeknownst to you, this is the one time you would actually win.

We can use this example to prove that you should definitely play the lottery. I think the bomb situation maps to this - although it fails in the one/trillion trillion case where the superpredictor was wrong, it succeeds in the 9999999999.../trillion trillion cases where the superpredictor is right, and when you multiply out the probabilities by the utilities (eg of getting any extra $100 vs. getting bombed), FDT gets you more utility overall.

The particular way FDT succeeds is that you never (okay, 1/trillion trillion times, but this rounds to never) find yourself in this situation. So just by asking about this situation, you've already started with the assumption that FDT fails, which is why you are so easily able to prove that FDT fails.

I think this goes back to what I said last time we discussed this. You and Eliezer are optimizing for different things. You are optimizing for never finding yourself in a situation where you have to do something silly within that situation; he is optimizing for having the most utility overall if you can set your algorithm. I think the thing he's optimizing for makes more sense.

TheBorys's avatar

"You are the only person left in the universe. You have a happy life" only rationalist could write a sentence like this

197 more comments...

No posts

Ready for more?