
Daniel Whitenack and Chris Benson
· 45 Minutes
AI in the shadows: From hallucinations to blackmail
Listen
In the first episode of an "AI in the shadows" theme, Chris and Daniel explore the increasing concerning world of agentic misalignment. Starting out with a reminder about hallucinations and reasoning models, they break down how today’s models only mimic reasoning, which can lead to serious ethical considerations. They unpack a fascinating (and slightly terrifying) new study from Anthropic, where agentic AI models were caught simulating blackmail, deception, and even sabotage — all in the name of goal completion and self-preservation.
Featuring:
- Chris Benson – Website, LinkedIn, Bluesky, GitHub, X
- Daniel Whitenack – Website, GitHub, X
Links:
- Agentic Misalignment: How LLMs could be insider threats
- Hugging Face Agents Course
Register for upcoming webinars here!
Original source
This story was published by Practical AI and written by Daniel Whitenack and Chris Benson. SyncAI.news shows a preview; the complete article is on the publisher's site.
Read the full story on share.transistor.fm
![[AINews] not much happened today](/media/images/2026/09/e8357327b4721fb2.webp)

![[AINews] Reality Checks on AI News (Yegge shuts down Gas Town, Databricks’ +60% Astra cost)](/media/images/2026/09/a4a8dc01820a9f9a.webp)