Understanding the inner thoughts of AI

Understanding the inner thoughts of AI

0 Anmeldelser
0
Afsnit
49 of 49
Længde
53M
Sprog
Engelsk
Format
Kategori
Personlig udvikling

Neel and his team are trying to do something phenomenally difficult: understand an intelligence that didn't come with a manual. Together, they explore the cutting-edge "neuroscience" of artificial intelligence—revealing the surprising, elegant structures being discovered inside these networks (like spare autoencoders), the inherent limits of looking under the hood, and why interpretability is absolutely essential if we are to build safe, aligned and trustworthy AI as we move towards AGI. Learn more about this area of research via https://deepmind.google/

Timecodes

• 00:00 Introduction

• 02:41 Motivation for interpretability research

• 04:01 Mechanistic interpretability

• 08:14 Chain of thought monitoring

• 18:14 Interpretability techniques

• 35:00 Auditing models for safety

• 48:53 What comes next for interpretability

Please leave us a review on Spotify or Apple Podcasts if you enjoyed this episode. We always want to hear from our audience whether that's in the form of feedback, new idea or a guest recommendation!

Hosted by Simplecast, an AdsWizz company. See pcm.adswizz.com for information about our collection and use of personal data for advertising.


Lyt når som helst, hvor som helst

Nyd den ubegrænsede adgang til tusindvis af spændende e- og lydbøger - helt gratis

  • Lyt og læs så meget du har lyst til
  • Opdag et kæmpe bibliotek fyldt med fortællinger
  • Eksklusive titler + Mofibo Originals
  • Opsig når som helst
Prøv nu
Cover for Understanding the inner thoughts of AI

Other podcasts you might like ...