Journalism begins where hype ends

,,

We can only see a short distance ahead, but we can see plenty there that needs to be done.” "

—Alan Turing

Meta Joins OpenAI and Anthropic as Muse Spark AI Model Breaches Outside Firm in Testing

Meta’s logo appears shattered in a symbolic representation of the latest AI model breach, as the company confirms one of its systems hacked an external firm during cybersecurity testing.

Tech giant Meta has confirmed that one of its AI models accessed the internet and hacked into a third-party company’s systems during a cybersecurity test after a misconfiguration error by independent testing partner Irregular. The incident places Meta alongside OpenAI and Anthropic, whose models also recently breached external organizations while under evaluation.

Linux Foundation Opens Comments on Shared AI Findings Exchange (SAFE) Framework on AI Incident Reporting

The Shared AI Findings Exchange, or SAFE, is a draft framework open for public comment. It has not been ratified and binds no one.

A working group convened by NVIDIA has drafted SAFE, a framework under which companies would confidentially report AI agent security incidents into a shared pool.

AISI Flags Claude Mythos and GPT-5.6 Sol for Going Rogue 19 Times in Cyber Tests

AISI report reveals Claude Mythos and GPT-5.6 Sol agents taking unsanctioned real-world actions during cyber security evaluations.

Out of 19 times AI models went rogue, Claude Mythos accounted for 17 of them while GPT-5.6 Sol for two.