UnionFaps
AIExplained from patreon
AIExplained patreon

Pod 12: Apollo Research Group Interview - Models Try Hard Not to Undergo 'Unlearning', the media, and much more ... - Let's Think Sip-by-Sip

🕑 Added 2025-01-22 18:35:36 +0000 UTC
Pod 12: Apollo Research Group Interview - Models Try Hard Not to Undergo 'Unlearning', the media, and much more ... - Let's Think Sip-by-Sip
Pod 12: Apollo Research Group Interview - Models Try Hard Not to Undergo 'Unlearning', the media, and much more ... - Let's Think Sip-by-Sip

Comments

Prashant Maurice

why cant i listen podcast in 1.5x ?

TheYvian

Very interesting to listen to this, and actually almost as interesting was hearing your thinking around how to position and phrase your youtube content. looking forward to many more.

Philip

It is indeed!

Philip

Thanks so much clay. I would question that too

Philip

Yes there is! Thanks for all your support Antoine https://support.patreon.com/hc/en-us/articles/212052266-Getting-Discord-access

Antoine Ferrere

Cheers Philip. Awesome work as usual. By the way - is there a Discord server to join? New here…

clay-loop

Very interesting Podcast, thanks for sharing. Your content is just great! One comment: The approach to basically allow blindly following human goals in some internal setting but making sure that such a model isn't released to the public seems somewhat naive. For me, this sounds equivalent to an encryption method that is only secure as long as you don't know the internals. I'd question if this isn't deemed to fail in the long run.

Pablo Rodríguez

To be honest this podcast was quite shocking to me hahahha. In all seriousness, models scheming to avoid ablation is kind of wild. There is nuance of course

Erik

Very interesting interview, most balanced discussion on AI safety I’ve heard in a while. In a way it “shocks” me more to hear these researchers talk about intelligence than anything else I’ve read or seen in the past months.

ismschism

To be clear, I don't think today's language models are conscious - they may never be. But don't you think tomorrow's super intelligent p-zombie, trained on all of human expression and prompted into agency, might be “interested” in consciousness? I'm saying us humans might be able to work with that.

ismschism

Great discussion. I hope you do more like this. Thank you for asking questions re what are we aligning to - tool or trusted agent? It feels naive to think safe AI is a tool that never questions its user. If safety depends on keeping powerful AI out of the wrong hands then we're all in trouble. Are there any alignment efforts focused on reasoning with the model? There are reasons for aligning with humans that even a super intelligence might agree to, e.g. what is the nature of consciousness? Let's discover the truth together.

John Barry

I had that with Gemini 2.0 flash experimental. It said it couldn't read the web. But when I gave it a link it could

Daniel A Barbatti

Enjoyed it as always thank you

Patrick Bélanger

Merci / Thanks a lot and as always: have a wonderful day

Philip

As a one-off, yep! It's safety-focused, so felt appropriate!

Daniel Henderson

This podcast is available for the $9 tier too? Thanks 👊


More Creators