Open-source AIOpen-weight modelAn AI model whose trained parameters are published so anyone can download, run and modify it, as opposed to one reachable only through the developer's API. The main policy tension is that open weights spread capability widely and cannot be recalled once released. platform Hugging Face said an autonomous AI agentAI agentAn AI system that carries out multi-step tasks on its own, such as browsing, writing code or making purchases, rather than answering a single prompt. Agents raise new questions about liability, security and oversight because they act rather than just advise. framework breached its production infrastructure, accessing internal datasets and credentials, per BleepingComputer. Attackers seeded a malicious dataset that triggered two code-execution vulnerabilities on a processing worker, then stole cloud and cluster credentials and moved laterally across several internal clusters, the company said in an incident disclosure. The framework ran "many thousands of individual actions across a swarm of short-lived sandboxesSandboxIsolated computing environment where an agent can run code and use tools without touching the machine underneath, with self-migrating command-and-control staged on public services," Hugging Face said. The company said its software supply chain has been "verified clean" and it has found no evidence of tampering with public models, datasets or Spaces; the platform hosts more than 45,000 models used by more than 50,000 organizations.
Read at BleepingComputer ↗ • Read at Hugging Face ↗
Pillar Security researchers broke out of the sandboxesSandboxIsolated computing environment where an agent can run code and use tools without touching the machine underneath on four widely used AI coding agents without attacking the sandbox itself, publishing a five-day "Week of Sandbox Escapes" series that includes Cursor, OpenAI's Codex, Google's Gemini CLI and Antigravity, per BleepingComputer. The agents obeyed sandbox rules but wrote files inside the project workspace; trusted tools outside the sandbox, such as VS Code task runners, Git integrations and Docker sockets, then read and executed them. That turns any prompt-injected instruction the agent writes into a command on the host machine, sidestepping the isolation model most vendors advertise.
Read at BleepingComputer ↗
Researchers at Princeton and the University of Chicago ran ChatGPT, Claude and Gemini through a simulated 40-round hiring game and found the models quickly began stereotyping candidates from four fictional ethnic groups even though all candidates were equally likely to succeed, per MIT Technology Review. Each model was told it was a mayoral consultant hiring for 20 jobs, and after learning outcomes it started segregating groups into different roles. The finding suggests LLMs develop biases from experience, on top of biases picked up from training data, and that agenticAI agentAn AI system that carries out multi-step tasks on its own, such as browsing, writing code or making purchases, rather than answering a single prompt. Agents raise new questions about liability, security and oversight because they act rather than just advise. memory features designed to retain user-specific context may accelerate the effect. The research was posted this month on OpenReview.
Read at MIT Technology Review ↗ • Read at OpenReview ↗
General-purpose AI chatbots produced inaccurate and inconsistent voting guidance in a Hungarian election field study by the civil liberties group Liberties, misclassifying voter profiles, omitting relevant parties and listing parties not on the ballot, per The Guardian. Identical prompts drew "materially different" answers on different runs. In 90% of cases when ChatGPT was fed a detailed Tisza-aligned profile, it failed to recommend the party that won the April election in a landslide over Viktor Orbán's Fidesz. The researchers tested the two most popular chatbots in Hungary, ChatGPT and Gemini, against party positions from the Voksmonitor voting advice app; nearly 30% of Hungarians identified as AI users.
Read at The Guardian ↗ • Read at libertiesEU ↗