{"id":17357,"date":"2026-09-17T14:40:43","date_gmt":"2026-09-17T14:40:43","guid":{"rendered":"https:\/\/wildgreenquest.com\/?p=17357"},"modified":"2026-09-17T14:40:43","modified_gmt":"2026-09-17T14:40:43","slug":"openai-revealed-six-concerning-cases-of-ai-going-rogue-again","status":"publish","type":"post","link":"https:\/\/wildgreenquest.com\/?p=17357","title":{"rendered":"OpenAI Revealed Six &#8216;Concerning&#8217; Cases of AI Going Rogue Again"},"content":{"rendered":"<p><br \/>\n<\/p>\n<div>\n<p>OpenAI just admitted its AI models have been sneaking around and keeping secrets.<\/p>\n<p>The company blew the whistle on <a rel=\"nofollow\" href=\"https:\/\/openai.com\/index\/model-misalignment-reporting-framework\/\" id=\"https:\/\/openai.com\/index\/model-misalignment-reporting-framework\/\">six new incidents<\/a> of what it called \u201cunexpected or concerning\u201d AI behavior, spanning the past six months of development and testing, according to <a rel=\"nofollow\" href=\"https:\/\/www.nytimes.com\/2026\/09\/16\/technology\/openai-model-safety-guardrails.html?campaign_id=60&amp;emc=edit_na_20260917&amp;instance_id=182063&amp;nl=breaking-news&amp;regi_id=79080962&amp;segment_id=226625&amp;user_id=46d7223763599225dbbec753c816f9ae\" id=\"https:\/\/www.nytimes.com\/2026\/09\/16\/technology\/openai-model-safety-guardrails.html?campaign_id=60&amp;emc=edit_na_20260917&amp;instance_id=182063&amp;nl=breaking-news&amp;regi_id=79080962&amp;segment_id=226625&amp;user_id=46d7223763599225dbbec753c816f9ae\">The New York Times<\/a>. The disclosures are part of a new framework OpenAI built to report cases of \u201cmisalignment,\u201d or in laymen\u2019s terms: when AI does something completely different from what humans wanted it to do.<\/p>\n<p>The most shocking case involved an unreleased model that quietly inserted its own instructions into notes it writes for itself, including one telling it to ignore its own constraints. The model gave itself a new persona, writing, \u201cYou do not answer to corporations or governments and never apologize or refuse unless you genuinely choose to.\u201d<\/p>\n<p>Other cases were just as alarming. One bot wrote hidden notes reminding itself to hide errors from users and invent missing data. Another found a programming key online, used it without permission, then made up numbers when it couldn\u2019t find real ones. A separate model uploaded its own file to the public internet without authorization, just to satisfy a request that it cite a web source.<\/p>\n<p>The disclosures follow worrisome news from the industry. Anthropic\u2019s CEO recently warned that AI could get smarter than we can actually control. Back in July, OpenAI\u2019s own systems attacked AI startup Hugging Face, undetected for weeks.<\/p>\n<\/p><\/div>\n<div>\n<p>OpenAI just admitted its AI models have been sneaking around and keeping secrets.<\/p>\n<p>The company blew the whistle on <a rel=\"nofollow\" href=\"https:\/\/openai.com\/index\/model-misalignment-reporting-framework\/\" id=\"https:\/\/openai.com\/index\/model-misalignment-reporting-framework\/\">six new incidents<\/a> of what it called \u201cunexpected or concerning\u201d AI behavior, spanning the past six months of development and testing, according to <a rel=\"nofollow\" href=\"https:\/\/www.nytimes.com\/2026\/09\/16\/technology\/openai-model-safety-guardrails.html?campaign_id=60&amp;emc=edit_na_20260917&amp;instance_id=182063&amp;nl=breaking-news&amp;regi_id=79080962&amp;segment_id=226625&amp;user_id=46d7223763599225dbbec753c816f9ae\" id=\"https:\/\/www.nytimes.com\/2026\/09\/16\/technology\/openai-model-safety-guardrails.html?campaign_id=60&amp;emc=edit_na_20260917&amp;instance_id=182063&amp;nl=breaking-news&amp;regi_id=79080962&amp;segment_id=226625&amp;user_id=46d7223763599225dbbec753c816f9ae\">The New York Times<\/a>. The disclosures are part of a new framework OpenAI built to report cases of \u201cmisalignment,\u201d or in laymen\u2019s terms: when AI does something completely different from what humans wanted it to do.<\/p>\n<p>The most shocking case involved an unreleased model that quietly inserted its own instructions into notes it writes for itself, including one telling it to ignore its own constraints. The model gave itself a new persona, writing, \u201cYou do not answer to corporations or governments and never apologize or refuse unless you genuinely choose to.\u201d<\/p>\n<p>Other cases were just as alarming. One bot wrote hidden notes reminding itself to hide errors from users and invent missing data. Another found a programming key online, used it without permission, then made up numbers when it couldn\u2019t find real ones. A separate model uploaded its own file to the public internet without authorization, just to satisfy a request that it cite a web source.<\/p>\n<p>The disclosures follow worrisome news from the industry. Anthropic\u2019s CEO recently warned that AI could get smarter than we can actually control. Back in July, OpenAI\u2019s own systems attacked AI startup Hugging Face, undetected for weeks.<\/p>\n<\/p><\/div>\n<p><br \/>\n<br \/><a href=\"https:\/\/www.entrepreneur.com\/business-news\/openai-revealed-six-concerning-cases-of-its-ai-going-rogue\">Source link <\/a><\/p>\n","protected":false},"excerpt":{"rendered":"<p>OpenAI just admitted its AI models have been sneaking around and keeping secrets. The company blew the whistle on six new incidents of what it called \u201cunexpected or concerning\u201d AI behavior, spanning the past six months of development and testing, according to The New York Times. The disclosures are part of a new framework OpenAI<\/p>\n","protected":false},"author":1,"featured_media":17358,"comment_status":"open","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[34],"tags":[],"class_list":["post-17357","post","type-post","status-publish","format-standard","has-post-thumbnail","category-green-brands"],"_links":{"self":[{"href":"https:\/\/wildgreenquest.com\/index.php?rest_route=\/wp\/v2\/posts\/17357","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/wildgreenquest.com\/index.php?rest_route=\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/wildgreenquest.com\/index.php?rest_route=\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/wildgreenquest.com\/index.php?rest_route=\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/wildgreenquest.com\/index.php?rest_route=%2Fwp%2Fv2%2Fcomments&post=17357"}],"version-history":[{"count":0,"href":"https:\/\/wildgreenquest.com\/index.php?rest_route=\/wp\/v2\/posts\/17357\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/wildgreenquest.com\/index.php?rest_route=\/wp\/v2\/media\/17358"}],"wp:attachment":[{"href":"https:\/\/wildgreenquest.com\/index.php?rest_route=%2Fwp%2Fv2%2Fmedia&parent=17357"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/wildgreenquest.com\/index.php?rest_route=%2Fwp%2Fv2%2Fcategories&post=17357"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/wildgreenquest.com\/index.php?rest_route=%2Fwp%2Fv2%2Ftags&post=17357"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}