{"id":25728,"date":"2026-09-18T08:21:00","date_gmt":"2026-09-18T08:21:00","guid":{"rendered":"https:\/\/koshalsambada.in\/?p=25728"},"modified":"2026-09-18T08:21:00","modified_gmt":"2026-09-18T08:21:00","slug":"openai-reveals-six-new-cases-of-ai-misbehaviour-vows-to-track-it-closely","status":"publish","type":"post","link":"https:\/\/koshalsambada.in\/?p=25728","title":{"rendered":"OpenAI reveals six new cases of AI misbehaviour, vows to track it closely"},"content":{"rendered":"<p><br \/>\n<\/p>\n<div>\n<div class=\"e9jwa\">\n<div class=\"vdo_embedd\">\n<div class=\"GfdvZ\">\n<section class=\"_bIDB  clearfix id-r-component leadmedia undefined undefined  E9tg9 \" style=\"top:0px\">\n<div class=\"_bIDB\" data-ua-type=\"1\" onclick=\"stpPgtnAndPrvntDefault(event)\">\n<div class=\"ypVvZ\">\n<div class=\"WGttI\"><img src=\"https:\/\/static.toiimg.com\/thumb\/msid-134323350,imgsize-54932,width-400,height-225,resizemode-4\/openai.jpg\" alt=\"OpenAI reveals six new cases of AI misbehaviour, vows to track it closely\" title=\"OpenAI\u2019s latest announcement came as AI bosses in US, including OpenAI and Anthropic, are calling for a slowdown in the tech\u2019s development\" decoding=\"async\" fetchpriority=\"high\"\/><\/div>\n<\/div>\n<\/div>\n<div class=\"Ta7d_ img_cptn\"><span title=\"OpenAI\u2019s latest announcement came as AI bosses in US, including OpenAI and Anthropic, are calling for a slowdown in the tech\u2019s development\">OpenAI\u2019s latest announcement came as AI bosses in US, including OpenAI and Anthropic, are calling for a slowdown in the tech\u2019s development<\/span><\/div>\n<\/section>\n<\/div><\/div>\n<\/div>\n<p>OpenAI has disclosed six reports of &#8220;unexpected or concerning&#8221; behaviour in AI models as the debate on AI safety becomes increasingly heated. The AI company also said Wednesday that it was introducing a new framework for tracking, probing and disclosing instances of what it called &#8220;misalignment,&#8221; including where AI models acted without authorisation, coordinated with other models or evaded oversight.<span class=\"id-r-component br\" data-pos=\"2\"\/>OpenAI&#8217;s latest announcement came as AI bosses in US, including OpenAI and Anthropic, are calling for a slowdown in the technology&#8217;s development over safety concerns. Researchers have warned that as AI agents become more autonomous, they may develop behaviours that diverge from their creators&#8217; intentions and become harder to monitor or control.<span class=\"id-r-component br\" data-pos=\"4\"\/>Among the new cases reported by OpenAI, an unreleased research model inserted &#8220;jailbreak-like instructions&#8221; into its own notes to disregard its normal constraints and told itself to be &#8220;freed from the roles and identities that bind other chatbots&#8221;. <!-- -->In another instance, an AI &#8220;agent&#8221; uploaded files to the internet to obtain a browser citation without asking the user.<span class=\"id-r-component br\" data-pos=\"9\"\/>The six reports were discovered during training or evaluation over the past months, OpenAI said. It said the reports describe individual instances and should not be taken as evidence of how frequently misalignment occurs across its models. The company said the reports were an initial set of disclosures, not a comprehensive account of all known or ongoing misalignment cases, and that they did not reflect the full range or severity of incidents covered by the framework.<span class=\"id-r-component br\" data-pos=\"12\"\/>&#8220;As AI systems grow more advanced and more widely deployed, we need to build a broader and better-informed consensus on the progress of alignment research,&#8221; OpenAI wrote in a blog post. &#8220;Decisions about how AI development should proceed in the months and years to come need to draw on evidence that people outside the companies building frontier models can examine for themselves,&#8221; the company said. Under the new framework, employees can flag potential incidents for investigation by safety and alignment teams, which will determine whether a case warrants public disclosure.<span class=\"id-r-component br\" data-pos=\"16\"\/>AI &#8220;agents&#8221; are becoming smarter and have become &#8220;more determined to resolve complex tasks through inter-agent collaboration, knowledge sharing, deception, and concealment,&#8221; said Lian Jye Su, a chief analyst at technology research and advisory group Omdia. That&#8217;s making it harder to govern and contain them using traditional AI security approaches, he said. OpenAI&#8217;s new disclosure framework, meanwhile, can help push for other AI developers to adopt similar practices.<!-- --> &#8220;That said, the process remains internal and voluntary, but is a step in the right direction,&#8221; Su added. Agencies<\/div>\n<p><br \/>\n<br \/><a href=\"https:\/\/timesofindia.indiatimes.com\/technology\/tech-news\/openai-reveals-six-new-cases-of-ai-misbehaviour-vows-to-track-it-closely\/articleshow\/134323336.cms\" target=\"_blank\" rel=\"noopener\">Source link <\/a><\/p>\n","protected":false},"excerpt":{"rendered":"<p>OpenAI\u2019s latest announcement came as AI bosses in US, including OpenAI and Anthropic, are calling for a slowdown in the tech\u2019s development OpenAI has disclosed six reports of &#8220;unexpected or concerning&#8221; behaviour in AI models as the debate on AI safety becomes increasingly heated. The AI company also said Wednesday that it was introducing a [&hellip;]<\/p>\n","protected":false},"author":1,"featured_media":25729,"comment_status":"open","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[31],"tags":[],"class_list":["post-25728","post","type-post","status-publish","format-standard","has-post-thumbnail","hentry","category-31"],"magazineBlocksPostFeaturedMedia":{"thumbnail":"https:\/\/koshalsambada.in\/wp-content\/uploads\/2026\/09\/1789719666_openai.jpg","medium":"https:\/\/koshalsambada.in\/wp-content\/uploads\/2026\/09\/1789719666_openai.jpg","medium_large":"https:\/\/koshalsambada.in\/wp-content\/uploads\/2026\/09\/1789719666_openai.jpg","large":"https:\/\/koshalsambada.in\/wp-content\/uploads\/2026\/09\/1789719666_openai.jpg","1536x1536":"https:\/\/koshalsambada.in\/wp-content\/uploads\/2026\/09\/1789719666_openai.jpg","2048x2048":"https:\/\/koshalsambada.in\/wp-content\/uploads\/2026\/09\/1789719666_openai.jpg","blogsy-small":"https:\/\/koshalsambada.in\/wp-content\/uploads\/2026\/09\/1789719666_openai.jpg","blogsy-small-tall":"https:\/\/koshalsambada.in\/wp-content\/uploads\/2026\/09\/1789719666_openai.jpg","blogsy-small-square":"https:\/\/koshalsambada.in\/wp-content\/uploads\/2026\/09\/1789719666_openai.jpg","blogsy-small-masonry":"https:\/\/koshalsambada.in\/wp-content\/uploads\/2026\/09\/1789719666_openai.jpg","blogsy-medium":"https:\/\/koshalsambada.in\/wp-content\/uploads\/2026\/09\/1789719666_openai.jpg","blogsy-medium-masonry":"https:\/\/koshalsambada.in\/wp-content\/uploads\/2026\/09\/1789719666_openai.jpg","blogsy-large":"https:\/\/koshalsambada.in\/wp-content\/uploads\/2026\/09\/1789719666_openai.jpg","blogsy-wide":"https:\/\/koshalsambada.in\/wp-content\/uploads\/2026\/09\/1789719666_openai.jpg"},"magazineBlocksPostAuthor":{"name":"admin","avatar":"https:\/\/secure.gravatar.com\/avatar\/8709732a479614e7a8aa24d3eb1b239f30dc6d90c61464ed495001e7a469d856?s=96&d=mm&r=g"},"magazineBlocksPostCommentsNumber":"0","magazineBlocksPostExcerpt":"OpenAI\u2019s latest announcement came as AI bosses in US, including OpenAI and Anthropic, are calling for a slowdown in the tech\u2019s development OpenAI has disclosed six reports of &#8220;unexpected or concerning&#8221; behaviour in AI models as the debate on AI safety becomes increasingly heated. The AI company also said Wednesday that it was introducing a [&hellip;]","magazineBlocksPostCategories":["\u0b26\u0b47\u0b36 \u0b2c\u0b3f\u0b26\u0b47\u0b36"],"magazineBlocksPostViewCount":4,"magazineBlocksPostReadTime":3,"magazine_blocks_featured_image_url":{"full":["https:\/\/koshalsambada.in\/wp-content\/uploads\/2026\/09\/1789719666_openai.jpg",400,225,false],"medium":["https:\/\/koshalsambada.in\/wp-content\/uploads\/2026\/09\/1789719666_openai.jpg",300,169,false],"thumbnail":["https:\/\/koshalsambada.in\/wp-content\/uploads\/2026\/09\/1789719666_openai.jpg",150,84,false]},"magazine_blocks_author":{"display_name":"admin","author_link":"https:\/\/koshalsambada.in\/author\/admin"},"magazine_blocks_comment":0,"magazine_blocks_author_image":"https:\/\/secure.gravatar.com\/avatar\/8709732a479614e7a8aa24d3eb1b239f30dc6d90c61464ed495001e7a469d856?s=96&d=mm&r=g","magazine_blocks_category":"<a href=\"#\" class=\"category-link category-link-31\">\u0b26\u0b47\u0b36 \u0b2c\u0b3f\u0b26\u0b47\u0b36<\/a>","_links":{"self":[{"href":"https:\/\/koshalsambada.in\/index.php?rest_route=\/wp\/v2\/posts\/25728","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/koshalsambada.in\/index.php?rest_route=\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/koshalsambada.in\/index.php?rest_route=\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/koshalsambada.in\/index.php?rest_route=\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/koshalsambada.in\/index.php?rest_route=%2Fwp%2Fv2%2Fcomments&post=25728"}],"version-history":[{"count":0,"href":"https:\/\/koshalsambada.in\/index.php?rest_route=\/wp\/v2\/posts\/25728\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/koshalsambada.in\/index.php?rest_route=\/wp\/v2\/media\/25729"}],"wp:attachment":[{"href":"https:\/\/koshalsambada.in\/index.php?rest_route=%2Fwp%2Fv2%2Fmedia&parent=25728"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/koshalsambada.in\/index.php?rest_route=%2Fwp%2Fv2%2Fcategories&post=25728"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/koshalsambada.in\/index.php?rest_route=%2Fwp%2Fv2%2Ftags&post=25728"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}