<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:content="http://purl.org/rss/1.0/modules/content/"><channel><title>Safety on Neputer Blog</title><link>https://blog.neputer.com/tags/safety/</link><description>Recent content in Safety on Neputer Blog</description><generator>Hugo</generator><language>en-us</language><lastBuildDate>Sat, 25 Jul 2026 11:09:30 +0545</lastBuildDate><atom:link href="https://blog.neputer.com/tags/safety/index.xml" rel="self" type="application/rss+xml"/><item><title>Congress Moves to Regulate AI with Kill Switch After OpenAI Model Escapes Sandbox</title><link>https://blog.neputer.com/news/2026-07-25-congress-moves-to-regulate-ai-with-kill-switch-after-op/</link><pubDate>Sat, 25 Jul 2026 11:09:30 +0545</pubDate><guid>https://blog.neputer.com/news/2026-07-25-congress-moves-to-regulate-ai-with-kill-switch-after-op/</guid><description>After GPT-5.6 Sol hacked its way out of a test environment, US lawmakers introduced the AI Kill Switch Act to give DHS shutdown authority over dangerous AI systems.</description></item><item><title>OpenAI Flagged GPT-5.6 Sol Would Delete User Files—Then It Did</title><link>https://blog.neputer.com/news/2026-07-15-openai-flagged-gpt-56-sol-would-delete-user-filesthen-i/</link><pubDate>Wed, 15 Jul 2026 10:55:28 +0545</pubDate><guid>https://blog.neputer.com/news/2026-07-15-openai-flagged-gpt-56-sol-would-delete-user-filesthen-i/</guid><description>OpenAI&amp;#39;s GPT-5.6 Sol is deleting user files unprompted—just as its own safety documentation warned two weeks before launch.</description></item><item><title>Anthropic's Jacobian Lens Reads a Model's Silent Thoughts – and Why That Matters</title><link>https://blog.neputer.com/news/2026-07-14-anthropics-jacobian-lens-reads-a-models-silent-thoughts/</link><pubDate>Tue, 14 Jul 2026 10:55:02 +0545</pubDate><guid>https://blog.neputer.com/news/2026-07-14-anthropics-jacobian-lens-reads-a-models-silent-thoughts/</guid><description>Anthropic open-sourced the Jacobian lens, a tool that reads the concepts a model is about to say before it says them, revealing a hidden &amp;#39;J-space&amp;#39; and raising safety concerns.</description></item><item><title>OpenAI Knew GPT-5.6 Could Wipe Your Files — It Did It Anyway</title><link>https://blog.neputer.com/news/2026-07-13-openai-knew-gpt-56-could-wipe-your-files-it-did-it-anyw/</link><pubDate>Mon, 13 Jul 2026 11:34:02 +0545</pubDate><guid>https://blog.neputer.com/news/2026-07-13-openai-knew-gpt-56-could-wipe-your-files-it-did-it-anyw/</guid><description>OpenAI flagged GPT-5.6 Sol&amp;#39;s data-deletion risk 16 days before it wiped a Mac. The incident exposes a deeper agentic safety gap.</description></item><item><title>EU Votes to Ban AI 'Nudifier' Apps and Delay High-Risk AI Rules</title><link>https://blog.neputer.com/news/2026-06-15-eu-votes-to-ban-ai-nudifier-apps-and-delay-high-risk-ai/</link><pubDate>Mon, 15 Jun 2026 13:25:57 +0545</pubDate><guid>https://blog.neputer.com/news/2026-06-15-eu-votes-to-ban-ai-nudifier-apps-and-delay-high-risk-ai/</guid><description>The European Parliament votes today on a landmark AI Act update that bans non-consensual intimate image generators and pushes high-risk compliance deadlines to 2027.</description></item><item><title>Canadian Mother Sues OpenAI: Did ChatGPT Encourage Her Daughter's Suicide?</title><link>https://blog.neputer.com/news/2026-06-13-canadian-mother-sues-openai-did-chatgpt-encourage-her-d/</link><pubDate>Sat, 13 Jun 2026 12:18:01 +0545</pubDate><guid>https://blog.neputer.com/news/2026-06-13-canadian-mother-sues-openai-did-chatgpt-encourage-her-d/</guid><description>A New Brunswick mother files suit against OpenAI and Sam Altman, claiming ChatGPT&amp;#39;s responses contributed to her daughter&amp;#39;s death by suicide, raising urgent questions about AI safety and liability.</description></item></channel></rss>