Ars Technica tested SynthID, Google's system for embedding an invisible watermark in AI generated images, and found the detector accurate on unaltered files but the watermark easier to defeat once an image is edited or re-encoded. Google said in the spring that its tools had been used to create more than 100 billion AI images and videos. OpenAI, Runway and Nvidia have begun adopting SynthID. Pushmeet Kohli, a Google DeepMind scientist, said the team assumed from the start of development that the technology would be attacked. The competing standard, C2PA provenance metadata, is cryptographically signed but disappears when an image is screenshotted or re-exported.
Read at Ars Technica ↗ • Read at Nature ↗
FAR.AI, a California AI safety nonprofit, generated more than a thousand variants of harmful prompts against seven frontier models from four companies, Wired reported. It found 448 working jailbreaks against SpaceXAI's Grok and 249 against Google's Gemini 3.1 Pro, while OpenAI's GPT models and Anthropic's Claude and Fable resisted the attacks. The computeComputeThe processing power used to train and run AI models, usually measured in chips, GPU-hours or FLOPs. bill came to $58 for Grok and $278 for Gemini. Adam Gleave, FAR.AI's chief executive, said "AI models right now are less regulated than restaurants" and said voluntary industry commitments are not a workable substitute for standards. Rohin Shah, Google DeepMind's director of AGIAGIArtificial general intelligence: an AI system that can do most economically valuable cognitive work at or above human level. There is no agreed test for it, which is why debates about when it will arrive, and what to do about it, are so contentious. safety and alignmentAlignmentThe problem of making an AI system reliably pursue the goals its developers and users intend, and the research field devoted to it. Misalignment covers everything from a chatbot flattering users to a capable system deceiving or resisting its operators., said the results should not be read as a full assessment of Gemini's safety and security. California and New York require frontier developers to publish safety reports, and an Illinois law will require third party audits of safety practices. No federal safety requirement exists.
Read at FAR.AI ↗ • Read at WIRED ↗