Anthropic’s Big Charity Bill for Shareholders Save 25% to unlock this story

The Information
Sign in
Subscribe

    Data Tools

    • About Pro
    • Enterprise Software Startup Takeover List 2026
    • The AI Inflection Index
    • The Next GPs 2026
    • The Executives Leading the Data Center Race
    • The Next GPs 2025
    • The Rising Stars of AI Research
    • Leaders of the AI Shopping Revolution
    • Enterprise Software Startup Takeover List 2025
    • Org Charts
    • The Information 50 2025
    • Generative AI Takeover List
    • Generative AI Database
    • AI Chip Database
    • AI Data Center Database
    • Tech IPO Tracker
    • Tech Sentiment Tracker
    • Gigafactory Database

    Special Projects

    • The Information 50 Database
    • VC Diversity Index
    • Enterprise Tech Powerlist
  • Org Charts
  • Tech
  • Finance
  • Weekend
  • Charts
  • Events
  • TITV
    • Directory

      Search, find and engage with others who are serious about tech and business.

    • Forum

      Follow and be a part of discussions about tech, finance and media.

    • Brand Partnerships

      Premium advertising opportunities for brands

    • Group Subscriptions

      Team access to our exclusive tech news

    • Newsletters

      Journalists who break and shape the news, in your inbox

    • Video

      Catch up on conversations with global leaders in tech, media and finance

    • Partner Content

      Explore our recent partner collaborations

      XFacebookLinkedInThreadsInstagram
    • Help & Support
    • RSS Feed
    • Careers
  • About Pro
  • Enterprise Software Startup Takeover List 2026
  • The AI Inflection Index
  • The Next GPs 2026
  • The Executives Leading the Data Center Race
  • The Next GPs 2025
  • The Rising Stars of AI Research
  • Leaders of the AI Shopping Revolution
  • Enterprise Software Startup Takeover List 2025
  • Org Charts
  • The Information 50 2025
  • Generative AI Takeover List
  • Generative AI Database
  • AI Chip Database
  • AI Data Center Database
  • Tech IPO Tracker
  • Tech Sentiment Tracker
  • Gigafactory Database

SPECIAL PROJECTS

  • The Information 50 Database
  • VC Diversity Index
  • Enterprise Tech Powerlist
Deep Research
TITV
Tech
Finance
Weekend
Charts
Events
Newsletters
  • Directory

    Search, find and engage with others who are serious about tech and business.

  • Forum

    Follow and be a part of discussions about tech, finance and media.

  • Brand Partnerships

    Premium advertising opportunities for brands

  • Group Subscriptions

    Team access to our exclusive tech news

  • Newsletters

    Journalists who break and shape the news, in your inbox

  • Video

    Catch up on conversations with global leaders in tech, media and finance

  • Partner Content

    Explore our recent partner collaborations

Subscribe
  • Sign in
  • Search
  • Opinion
  • Venture Capital
  • Artificial Intelligence
  • Startups
  • Market Research
    XFacebookLinkedInThreadsInstagram
  • Help & Support
  • RSS Feed
  • Careers

In-depth insights in seconds. Ask Deep Research.

Can We Solve AI’s Alignment Problem?

October 5, 2026

Episode Description

AI Deep Dive is a show from The Information's TITV that gets to the bottom of the hardest technical problems in AI — the models, the research and the people building them. Hosted by our AI reporter, Rocket Drew.

Can an AI system do exactly what it was trained to do and still fail us? Rocket Drew joins UC Berkeley computer science professor Stuart Russell to explore where AI’s objectives come from, why human feedback can reward the wrong behavior, and the challenge of proving a powerful system is safe.

Related articles:

Exclusive: Anthropic Research Memo Shows Focus on Rogue Agents, Scheming Models: https://www.theinformation.com/articles/anthropic-research-memo-shows-focus-rogue-agents-scheming-models

Exclusive: OpenAI Technique in ‘Astra’ Model Sparks Security Concerns: https://www.theinformation.com/articles/secret-technique-behind-openais-astra-model-sparks-security-concerns

AI Safety Push Sparks Demand for Watchdog Groups. Critics Doubt Their Independence: https://www.theinformation.com/articles/ai-safety-push-sparks-demand-watchdog-groups-critics-doubt-independence

Subscribe:

YouTube: https://www.youtube.com/@theinformation

The Information: https://www.theinformation.com/subscribe_h

Sign up for the AI Agenda newsletter: https://www.theinformation.com/features/ai-agenda

Follow us:

X: https://x.com/theinformation

IG: https://www.instagram.com/theinformation/

TikTok: https://www.tiktok.com/@titv.theinformation

LinkedIn: https://www.linkedin.com/company/theinformation/

Chapters:

00:00 - Intro

00:43 - The alignment problem

07:18 - How AI misalignment shows up

16:07 - Training AI to imitate humans

26:05 - Can human feedback fix alignment?

37:47 - When humans become the obstacle

40:04 - Did AI take a wrong turn?

44:24 - AI & existential risk

51:09 - AI labs & safety evidence

58:06 - Assistance games & the off switch

Every weekday at 10am pt / 1pm et

AI Agenda Newsletter

Stay ahead of the innovation and disruption happening in AI. Stephanie Palazzolo covers the tech, startups, companies and people making headlines.

See all newsletters