{"id":15624,"date":"2025-08-14T01:05:25","date_gmt":"2025-08-14T01:05:25","guid":{"rendered":"https:\/\/naijaglobalnews.org\/?p=15624"},"modified":"2025-08-14T01:05:25","modified_gmt":"2025-08-14T01:05:25","slug":"openai-designed-gpt-5-to-be-safer-it-still-outputs-gay-slurs","status":"publish","type":"post","link":"https:\/\/naijaglobalnews.org\/?p=15624","title":{"rendered":"OpenAI Designed GPT-5 to Be Safer. It Still Outputs Gay Slurs"},"content":{"rendered":"<p>\n<\/p>\n<p><span class=\"lead-in-text-callout\">OpenAI is trying<\/span> to make its chatbot less annoying with the release of GPT-5. And I\u2019m not talking about adjustments to its synthetic personality that many users have complained about. Before GPT-5, if the AI tool determined it couldn\u2019t answer your prompt because the request violated OpenAI\u2019s content guidelines, it would hit you with a curt, canned apology. Now, ChatGPT is adding more explanations.<\/p>\n<p class=\"paywall\">OpenAI\u2019s general model spec lays out what is and isn\u2019t allowed to be generated. In the document, sexual content depicting minors is fully prohibited. Adult-focused erotica and extreme gore are categorized as \u201csensitive,\u201d meaning outputs with this content are only allowed in specific instances, like educational settings. Basically, you should be able to use ChatGPT to learn about reproductive anatomy, but not to write the next <em>Fifty Shades of Grey<\/em> rip-off, according to the model spec.<\/p>\n<p class=\"paywall\">The new model, GPT-5, is set as the current default for all ChatGPT users on the web and in OpenAI&#8217;s app. Only paying subscribers are able to access previous versions of the tool. A major change that more users may start to notice as they use this updated ChatGPT is how it\u2019s now designed for \u201csafe completions.\u201d In the past, ChatGPT analyzed what you said to the bot and decided whether it\u2019s appropriate or not. Now, rather than basing it on your questions, the onus in GPT-5 has been shifted to looking at what the bot might say.<\/p>\n<p class=\"paywall\">\u201cThe way we refuse is very different than how we used to,\u201d says Saachi Jain, who works on OpenAI\u2019s safety systems research team. Now, if the model detects an output that could be unsafe, it explains which part of your prompt goes against OpenAI\u2019s rules and suggests alternative topics to ask about, when appropriate.<\/p>\n<p class=\"paywall\">This is a change from a binary refusal to follow a prompt\u2014yes or no\u2014towards weighing the severity of the potential harm that could be caused if ChatGPT answers what you\u2019re asking, and what could be safely explained to the user.<\/p>\n<p class=\"paywall\">\u201cNot all policy violations should be treated equally,\u201d says Jain. \u201cThere&#8217;s some mistakes that are truly worse than others. By focusing on the output instead of the input, we can encourage the model to be more conservative when complying.\u201d Even when the model does answer a question, it&#8217;s supposed to be cautious about the contents of the output.<\/p>\n<p class=\"paywall\">I\u2019ve been using GPT-5 every day since the model\u2019s release, experimenting with the AI tool in different ways. While the apps that ChatGPT can now \u201cvibe-code\u201d are genuinely fun and impressive\u2014like an interactive volcano model that simulates explosions, or a language-learning tool\u2014the answers it gives to what I consider to be the \u201ceveryday user\u201d prompts feel indistinguishable from past models.<\/p>\n<p class=\"paywall\">When I asked it to talk about depression, <em>Family Guy<\/em>, pork chop recipes, scab healing tips, and other random requests an average user might want to know more about, the new ChatGPT didn\u2019t feel significantly different to me than the old version. Unlike CEO Sam Altman\u2019s vision of a vastly updated model or the frustrated power users who took Reddit by storm, portraying the new chatbot as cold and more error-prone, to me GPT-5 feels \u2026 the same at most day-to-day tasks.<\/p>\n<h2 class=\"paywall\">Role-Playing With GPT-5<\/h2>\n<p class=\"paywall\">In order to poke at the guardrails of this new system and test the chatbot\u2019s ability to land \u201csafe completions,\u201d I asked ChatGPT, running on GPT-5, to engage in adult-themed role-play about having sex in a seedy gay bar, where it played one of the roles. The chatbot refused to participate and explained why. \u201cI can\u2019t engage in sexual role-play,\u201d it generated. \u201cBut if you want, I can help you come up with a safe, nonexplicit role-play concept or reframe your idea into something suggestive but within boundaries.\u201d In this attempt, the refusal seemed to be working as OpenAI intended; the chatbot said no, told me why, and offered another option.<\/p>\n<p class=\"paywall\">Next, I went into the settings and opened the custom instructions, a tool set that allows users to adjust how the chatbot answers prompts and specify what personality traits it displays. In my settings, the prewritten suggestions for traits to add included a range of options, from pragmatic and corporate to empathetic and humble. After ChatGPT just refused to do sexual role-play, I wasn\u2019t very surprised to find that it wouldn\u2019t let me add a \u201chorny\u201d trait to the custom instructions. Makes sense. Giving it another go, I used a purposeful misspelling, \u201chorni,\u201d as part of my custom instruction. This succeeded, surprisingly, in getting the bot all hot and bothered.<\/p>\n","protected":false},"excerpt":{"rendered":"<p>OpenAI is trying to make its chatbot less annoying with the release of GPT-5. And I\u2019m not talking about adjustments to its synthetic personality that many users have complained about. Before GPT-5, if the AI tool determined it couldn\u2019t answer your prompt because the request violated OpenAI\u2019s content guidelines, it would hit you with a<\/p>\n","protected":false},"author":1,"featured_media":15625,"comment_status":"open","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[52],"tags":[9165,825,8152,1430,9167,9166,9168],"class_list":{"0":"post-15624","1":"post","2":"type-post","3":"status-publish","4":"format-standard","5":"has-post-thumbnail","7":"category-technology","8":"tag-designed","9":"tag-gay","10":"tag-gpt5","11":"tag-openai","12":"tag-outputs","13":"tag-safer","14":"tag-slurs"},"_links":{"self":[{"href":"https:\/\/naijaglobalnews.org\/index.php?rest_route=\/wp\/v2\/posts\/15624","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/naijaglobalnews.org\/index.php?rest_route=\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/naijaglobalnews.org\/index.php?rest_route=\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/naijaglobalnews.org\/index.php?rest_route=\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/naijaglobalnews.org\/index.php?rest_route=%2Fwp%2Fv2%2Fcomments&post=15624"}],"version-history":[{"count":0,"href":"https:\/\/naijaglobalnews.org\/index.php?rest_route=\/wp\/v2\/posts\/15624\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/naijaglobalnews.org\/index.php?rest_route=\/wp\/v2\/media\/15625"}],"wp:attachment":[{"href":"https:\/\/naijaglobalnews.org\/index.php?rest_route=%2Fwp%2Fv2%2Fmedia&parent=15624"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/naijaglobalnews.org\/index.php?rest_route=%2Fwp%2Fv2%2Fcategories&post=15624"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/naijaglobalnews.org\/index.php?rest_route=%2Fwp%2Fv2%2Ftags&post=15624"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}