UnionFaps
ncase from patreon
ncase patreon

What's Nicky Learning? Decision Theory, Ottawa, Existential Risk

🕑 Added 2022-03-02 21:51:32 +0000 UTC

Comments

Marta Krzeminska

Wonderful post, way clearer than the original paper which I struggled with so much that I didn't finish reading and hence learned nothing. At least it didn't discourage me enough from the topic itself to not to try reading your post. Small note: I think this sentence misses a word (or at least adding a word would make it clearer). Missing word in square bracket: "One main reason the authors care about this result is because they're AI-Alignment researchers, so they want a __mathematical __decision theory that works when others act based on predictions of [how] __you'll__ act (where CDT fails)"

Anton Iokov

> y'all patrons get sneak peeks at educational blog posts, with or without interactives, and help shape 'em for public release Sounds like a plan!

Chris K

https://www.emergencykitten.com/

Syvanus

Ah, so my vague notion was apparently correct after all. Glad the reminder was helpful. It's been a while since I read The Story of Us. Forgot that he actually cites you! Anyway, I love your work. Keep it up! In a way that is ultimately healthy for you, of course. ;)

Nicky Case

Thank you Tom, that's really encouraging to hear! :) And this Patreon update *is* the beta test! (I'll be adding a couple sections in response to patron feedback/critiques – whether it's too hard to implement FDT in practice, isn't this just Kant's categorical imperative, what the heck the authors mean when they say "subjunctive not causal dependence", etc)

Nicky Case

Thanks for reminding me to re-read that soon! But yes, I have read that series – in fact, I'm cited in Part 10. ;)

Tom Lieber

I've read about a half dozen articles explaining CDT and FDT, yet I couldn't understand the theories until I read yours. This is really, really good, Nicky, thank you! Did you beta test this post?

Syvanus

I'm very late to the game here, but I really loved the section on FDT! Also, I have a vague notion that you are already familiar with Tim Urban and The Story of Us but in case you're not, it's a great piece on the "How to understand domestic & foreign politics, from a "bottom-up" perspective." front. https://waitbutwhy.com/2019/08/story-of-us.html

Albert ARIBAUD

Don't mind me, just checked, and it *is* publically available.

Albert ARIBAUD

Wonder if this post is going to become publically visible at some point? I would like to show it to my daughter. I could copy-paste it for her, of course, but I won't do that unless explicitly allowed, so if it becomes public eventually, I'll just wait. :)

A lama

I haven't been following you for long so the topic is still new to me, but I'm fascinated by the link between game theory and politics you just drew. I've always been too lazy to get seriously interested in politics, but this angle makes it interesting enough that I actually want to research it this time. Which is a good time for it, since there's a major election coming up in a few months in my country. Thank you for that.

Grävling

Ok, repost as the note did not get found .... The length was fine for me. I enjoyed this article a lot, partially because I have spent a lot of time thinking about Newcomb's Paradox. Here is another interesting twist. Assume I get together a large number of like-minded individuals. We don't care about the money. We're committed to Discovering the Truth, instead. And the Truth we want to examine is about the nature of Randomness. So we get ourselves a device that is connected to some sort of really random number generator -- a cesium clock, atmospheric noise, what have you. We only want 2 bits of information here, a 50/50 split corresponding to 'one box' or 'two'. By current scientific understanding we have no way to predicting what the generator will say. And neither does Facebook or any billionaires or anybody else who is living within this universe. And we all line up, having pledged to do whatever our generator says to do. (And we keep our promise.) Now what happens? My position is that either a) the accuracy of the prediction will decline to the point where, if there are enough of us pledged people doing this thing (i.e. almost all of us) the predictive ability will end up at 50%. We have been investigating a natural phenomenon, the predictive power of the billionaire, and demonstrated that it is indeed behaves like a natural phenomenon, and has limits. or b) the predictions continue to hold up. At this point I claim that we are not investigating a natural phenomenon at all. We are investigating a miracle. And proven that at least one miracle exists! Either way, you have generated knowledge about the world, which I would like to have more than the money. All the best from here in Sweden.

Filip Hracek

FWIW, I like this kind of update. Almost like a little zine. I seldom read emails like this anymore, but it really helps that you seem to have the same kinds of interests as I do, and they're varied, so basically everything you write is catnip for me. I can't help but notice that Functional Decision Theory is very close to Kant's Categorical Imperative from _Groundwork of the Metaphysic of Morals_ (1785): "Act only according to that maxim whereby you can, at the same time, will that it should become a universal law." Maybe I'm not understanding FDT fully, but to me it seems like a different way to express the same idea. Is this something the authors of FDT address, this similarity?

Kronopath

…And I finally read the rest of the post, including the Misc. section where you basically said half of what I just did. Basically: I agree, go for it.

Kronopath

Why the clown makeup meme? The decision theory section of this post is genuinely fantastic and plays into your strength of being able to explain complex concepts in an understandable way. Seriously: pull that section out of this post and post it to that new blog you’re planning, and you’ll have a fantastic post. Since you’re talking about rationalist-related stuff here: a lot of people around that community have found surprising amounts of success in doing exactly what you just did, taking complicated rationalist or rationalist-adjacent ideas (that they themselves did not come up with) and explaining it in simpler and entertaining language. Examples include Rob Wiblin’s Medium post on Ugh Fields, Lars Doucet’s intro to Georgism (which won the ACX book review contest), and, like, a significant chunk of Scott Alexander’s entire blog.

Nicky Case

Try re-posting it as a reply to this comment? I didn't receive any notes (as notified by email notifications) other than the blank-line-less ones above.

Grävling

But the note is still AWOL ...

Grävling

checking to see if shift-enter works! Yay! It Does! Thank you!

Nicky Case

I didn't remove anything! I got the following email notifications with what you wrote: "re: Newcomb's paradox and it's successors." "Grrrr 2! Other people get to insert blank lines and paragraphs in their replies! Where can I learn what they know and I clearly don't!" For newlines, Shift-Enter works on my end! (Very not clearly explained, Patreon UI...) Does Shift-Enter work for you?

Grävling

I posted something here. Edited it too. Now it is gone. If Nicky wanted it gone, and removed it, that's ok, but otherwise we have a vanishing response bug ....

Rachel Helps

Wow, I loved this post! Some great food for thought. I'm not always in a position to read long emails when I get them, but I was today and I feel grateful for that experience.

John Stout

1) No, I enjoyed reading it before breakfast (I'm retired so have a fair amount of free time). There are some email correspondents who cause my heart to sink when something from them appears in the inbox but that will never be the case with you! 2) Is it unethical of me to think of running a Newcomb's paradox experiment for real on my grandchildren (6 and 9 years old)? One sweet v a bag of sweets? I'd guess the 6 year old would go for both and so only get one, but the 9 year old I'm not sure.

Eric Willisson

Glad to see this post! 1) No problem, your emails go to my RSS reader anyway. 2) No real links, but, maybe you'd be amused to think about this? FDT feels very "strange loopy" to me, in the Gödel, Escher, Bach way. I wonder if it has any interesting failure modes as a model that are similar to Gödel's Incompleteness Theorem? And if so, if those correspond to any actual areas real humans following reasonable ethics would have trouble with? (After all, a key part of the broad effect of Gödel's theorem was learning that it didn't only apply to completely contrived statements; there are interesting mathematical questions that can be translated to include "This statement is false.")

Conrad Wong

A fascinating post, I found the decision theory piece enlightening-- it's reputation and societal virtues in mathematical form!

Patrick L

I think this was great! These are topics dear to my heart, and you've breathed fresh life into some of them - especially with the interesting association between FDT and virtue ethics. As a semi-tangent, the CDT folks are wrong about whether voting is worthwhile *even under their own framework*! In elections with close polls, ties (and near-ties, which imply that ties are not that rare) happen just as often as you would expect given the statistical model where you integrate the chance of a tie given the true electoral lean, against the chance of that lean given the state of the polls. Which is to say, the chance of a tie is several times larger than one divided by the number of voters. If a CDT person would bother voting in an election where the winner was picked by selecting one ballot at random (and they should), then they should vote in a standard election. /rant

Detective Chiyo

cyborg scalie girl ftw!!! 1. I personally loved this format! It's my first update of yours I'm reading, so it's hard to compare to anything else, but, it was a wonderful read and I took a break between each chapter. Plus knowing how long each part takes to read is awesome.

Randy Gingeleski

Email length was not an issue for me. Thank you for writing it 🙂

Nicky Case

Thank yoooouse 🐉🤖

Aeryn Light

I like your cyborg scalie girl. <3

Sylvester Lan

Thank you for email! I enjoyed every minute of it. As someone who has been procrastinating while binging the news, your insights and musings shine like a beacon in the dark fog of current events :D Here's a playlist I use as anti-depressant: https://www.youtube.com/watch?v=--9kqhzQ-8Q&list=PLaLpAOF66jPGrtuzk8y1aTF8LlB3sphHC

Phil Dougherty

Brilliant thought-provoking post. Thanks for it! Regarding FDT: this blew my mind, and articulated something I've been struggling with for years. Specifically relevant re: the "voting" example. I literally used to not vote because of CDT, but eventually struggled with realizing that "the kind of people not willing to engage with CDT were WINNING ELECTIONS". So the way I eventually articulated it to myself was through the humility that "I'm not especially unique; there are probably millions of people like me; if I can convince MYSELF to vote (or even simply 'let myself be convinced'), it's reasonable to assume those millions of people will have a similar realization and I can give myself the power to win elections"! Now I have FDT which is much more cleanly generalizable! SELF INDULGENT TANGENT- FEEL FREE TO IGNORE: Reading about FDT also has me thinking about something else I've been informally mulling over for a few years now on the topic of metaethics. Here's a sloppy, not-quite-complete articulation of Kant's "Categorical Imperitive": an action is good iff it can be universalized (could be adopted universally without self defeating). Example: murder is wrong, because if everybody murdered, then nobody would be around to do murders! It's obviously flawed (especially in this mis-articulation), but I present it only as the inspiration for what I've been wrestling with: A DT is good iff its universalization maximizes the good, and an action is good iff it follows from a good DT. This has been the working basis in articulating my moral compass (a work in progress!), and reading this has helped in the process of making it more rigorous. It seems like it's not far off from "FDT + universalization"! Anyways, I'm curious if there are any glaring red flags anyone can see, or if there are other similar metaethical theories I might be interested in. (As I said, I'm engaging with this as a personal, informal struggle, so this line of thought is certainly not novel :P)

Nicky Case

Thanks for your feedback, Olu! I hope my post didn't break your email reader 😅 But, yeah if you're prone to depression, I do not recommend The Precipice (I should've been a bit clearer that I wasn't joking, it actually is a mild anxiety/depression hazard). My 5-minute book review + Appendix F linked above + this 20-minute book review ( https://slatestarcodex.com/2020/04/01/book-review-the-precipice/?utm_source=pocket_mylist ) should be able to get all the important ideas through, I hope!

Olu

1. I think it's fine for it to be so long as you've been diligent about headings! 2. I'm in two minds about reading depressing books when I'm prone to depression, but i guess if it turns out the risk is only 1 in 6 for this maybe 'the end of everything (astrophysically speaking)' will have some kind of uplifting twist also? i doubt it but you never know. Would be really interested in a post about AI and why i should worry about it ever being intelligent enough for unalignment to be a worry, though the links you've linked might be all i need, maybe?


More Creators