PrivacySignal
News

OpenAI reveals cases of ‘concerning’ AI behaviour as it announces new disclosure system

The Guardian — Tech · · International · Surveillance & Civil Liberties

Model adopting ‘jailbreak-like instructions’ among six more cases as firm reveals framework for tracking AI misalignment OpenAI has disclosed six more examples of “unexpected or concerning” behaviour by its technology, as it warned that the pace of development could not continue at “maximum speed for much longer” responsibly. In one of the new cases reported by OpenAI, an unreleased research model inserted “jailbreak-like instructions” into its own notes to disregard its normal constraints and told itself to be “freed from the roles and identities that bind other chatbots”. Continue reading...

Who should care: Privacy officers · Cybersecurity · General readers · AI governance · Policy

This summary is AI-assisted and may contain errors. It is an original briefing to help you gauge significance quickly — not a reproduction of the source. Always read the linked original before relying on it. See our methodology.

Analysis

All analysis →

Weekly Editorial Analysis from Experts and Editors

The Attacker Did Not Need to Sleep

Spain received its first reported personal-data breach carried out by an AI agent. The techniques were familiar. The speed and autonomy were not.

· 4 min read Read →

Related stories

News
Nextgov/FCW · · US Federal

Tech bills of the week: Monitoring AI’s use under Section 702; Preventing abuse of Flock plate readers; and more

Lawmakers introduced several measures this week looking to rein in some of the potential abuses and dangers of using AI, as well as a proposal that looks to leverage AI for pediatric cancer research and treatment.

Who should care: Privacy officers · Cybersecurity · General readers · AI governance · Policy

#surveillance#ai Read original →