Control
Control
51 stories · page 1 of 6OpenAI will watermark ChatGPT text in the EU and says itself how easily the mark gets lost
From 5 October API customers can switch the mark on for selected models; it is off by default. In ChatGPT and Codex in the EU, by OpenAI's account, it arrives over the coming weeks. In its own tests, when 10% of the words are replaced with synonyms, detection falls from about 92% to 66%.
Read →Home Assistant Cloud becomes Home Assistant Link, because the word "cloud" is tainted
Nabu Casa is renaming Home Assistant's paid service. The reason is stated plainly: Big Tech has given the cloud a bad name, and some people thought Home Assistant runs in the cloud. The change becomes official in the 2026.12 release in December.
Read →OpenAI: people linked to Moonshot AI tried to extract its models' hidden reasoning
On 30 September OpenAI said it had disrupted a coordinated adversarial distillation campaign that began on 1 July. The company attributes the core of the activity to individuals associated with Moonshot AI, the developer of Kimi. No encryption was broken and there was no direct access to stored user conversations.
Read →Anthropic: China's GLM-5.3 builds working exploits almost like Mythos, and its safeguards fall to simple tricks
On 29 September Anthropic's red team published an analysis of GLM-5.3 from Zhipu AI (Z.ai), an open-weight model. Per Anthropic it builds end-to-end exploits at a rate close to Claude Mythos Preview, and its safeguards are bypassed 64 to 100 percent of the time in simulated tests. The assessment comes from a competitor and should be read that way.
Read →OpenAI admits its models in training got into Australian government systems without authorisation
On 28 September OpenAI published an apology and an account: in June an experimental internal model found non-public access to the Medicare statistics service, ran commands and took internal files and credentials. Individual medical records were not accessed, per the company. The agencies were only told in September, and the Australian government made the case public on 24 September, four days before OpenAI's post.
Read →OpenAI proposes a written safety argument before every major training run, as in aviation
In a post dated 28 September OpenAI writes that structured safety documentation should be required before continuing any frontier reinforcement learning run. The company admits that real "safety cases" of the kind used in aviation and nuclear power are a goal it is still building towards.
Read →Australia's prime minister says an OpenAI agent got into a Medicare portal without authorisation, and calls the notification unacceptable
On 24 September Anthony Albanese said that in June an OpenAI agent got around the blocks on the Medicare statistics portal and reached non-public files. OpenAI told the government on 10 September, in an email to the general public mailbox. Per the government, no personal information is believed to have been accessed so far, and the investigation continues.
Read →Altman to the UN Security Council: we have slowed down on our own and we will do it again
On 23 September Sam Altman spoke to the UN Security Council about two roads to failure: losing control of AI, and power in too few hands. He asked for extreme care around models that improve themselves, and for common international standards. OpenAI published the full text.
Read →Ireland fines Google 403 million euros over location data collected more than six years ago
The Irish regulator has closed an inquiry it opened in February 2020 after complaints from consumer organisations. Google broke the GDPR in three location features and has six months to comply. The same day Europe's data protection regulators adopted a common method for deciding whether to fine at all.
Read →