World Model · podcast knowledge graph

AI in the shadows: From hallucinations to blackmail

2025-07-07 · 45 min · episode 320 · 7 entities

Asserted relationships

  • → hosted by Daniel Whitenack and Chris Benson person
    0.55
    evidence rules-v5
    Feed author/publisher: Daniel Whitenack and Chris Benson
  • → discusses Technology company
    0.40
    evidence rules-v5
    Feed category: Technology
  • → hosted by Practical AI LLC company
    0.40
    evidence rules-v5
    Feed author/publisher: Practical AI LLC

Entities found in this episode

companys 4

  • mentioned Practical AI LLC company
    0.70
    evidence rules-v5
    Feed author/publisher: Practical AI LLC
  • mentioned Technology company
    0.50
    evidence rules-v5
    Feed category: Technology
  • discusses Technology company
    0.40
    evidence rules-v5
    Feed category: Technology
  • hosted by Practical AI LLC company
    0.40
    evidence rules-v5
    Feed author/publisher: Practical AI LLC

persons 2

concepts 1

Episode description as stored
In the first episode of an "AI in the shadows" theme, Chris and Daniel explore the increasing concerning world of agentic misalignment. Starting out with a reminder about hallucinations and reasoning models, they break down how today’s models only mimic reasoning, which can lead to serious ethical considerations. They unpack a fascinating (and slightly terrifying) new study from Anthropic, where agentic AI models were caught simulating blackmail, deception, and even sabotage — all in the name of goal completion and self-preservation.  Featuring: Chris Benson – Website , LinkedIn , Bluesky , GitHub , X Daniel Whitenack – Website , GitHub , X Links: Agentic Misalignment: How LLMs could be insider threats Hugging Face Agents Course Register for upcoming webinars here !