← terug naar overzicht

OpenAI onthult gevallen van 'zorgwekkend' AI-gedrag en belooft een nieuw plan voor het melden van problemen

nieuws 📅 2026-09-17
Research model inserting ‘jailbreak-like instructions’ into its notes is among cases as company says it is introducing new way of tracking AI misalignmentOpenAI has disclosed six new reports of “unexpected or concerning” behaviour in artificial-intelligence models as the debate on AI safety becomes increasingly heated.Among the new cases reported by OpenAI, an unreleased research model inserted “jailbreak-like instructions” into its own notes to disregard its normal constraints and told itself t

🔗 lees originele bron