{"id":13232,"date":"2023-04-12T08:44:51","date_gmt":"2023-04-12T08:44:51","guid":{"rendered":"https:\/\/entertainment.runfyers.com\/index.php\/2023\/04\/12\/openai-offers-bug-bounty-for-chatgpt-but-no-rewards-for-jailbreaking-its-chatbot\/"},"modified":"2023-04-12T08:44:51","modified_gmt":"2023-04-12T08:44:51","slug":"openai-offers-bug-bounty-for-chatgpt-but-no-rewards-for-jailbreaking-its-chatbot","status":"publish","type":"post","link":"https:\/\/entertainment.runfyers.com\/index.php\/2023\/04\/12\/openai-offers-bug-bounty-for-chatgpt-but-no-rewards-for-jailbreaking-its-chatbot\/","title":{"rendered":"OpenAI offers bug bounty for ChatGPT \u2014 but no rewards for jailbreaking its chatbot"},"content":{"rendered":"<p> <br \/>\n<\/p>\n<div>\n<p class=\"duet--article--dangerously-set-cms-markup duet--article--standard-paragraph mb-20 font-fkroman text-18 leading-160 -tracking-1 selection:bg-franklin-20 dark:text-white dark:selection:bg-blurple [&amp;_a]:shadow-underline-black dark:[&amp;_a]:shadow-underline-white [&amp;_a:hover]:shadow-highlight-franklin dark:[&amp;_a:hover]:shadow-highlight-blurple\">OpenAI has launched a <a href=\"https:\/\/openai.com\/blog\/bug-bounty-program\" target=\"_blank\" rel=\"noopener\">bug bounty<\/a>, encouraging members of the public to find and disclose vulnerabilities in its AI services including ChatGPT. Rewards range from $200 for \u201clow-severity findings\u201d to $20,000 for \u201cexceptional discoveries,\u201d and reports are submittable via crowdsourcing cybersecurity platform <a href=\"https:\/\/bugcrowd.com\/openai\" target=\"_blank\" rel=\"noopener\">Bugcrowd<\/a>.<\/p>\n<\/div>\n<div>\n<p class=\"duet--article--dangerously-set-cms-markup duet--article--standard-paragraph mb-20 font-fkroman text-18 leading-160 -tracking-1 selection:bg-franklin-20 dark:text-white dark:selection:bg-blurple [&amp;_a]:shadow-underline-black dark:[&amp;_a]:shadow-underline-white [&amp;_a:hover]:shadow-highlight-franklin dark:[&amp;_a:hover]:shadow-highlight-blurple\">Notably, the bounty excludes rewards for jailbreaking ChatGPT or causing it to generate malicious code or text. \u201cIssues related to the content of model prompts and responses are strictly out of scope, and will not be rewarded,\u201d says OpenAI\u2019s <a href=\"https:\/\/bugcrowd.com\/openai\" target=\"_blank\" rel=\"noopener\">Bugcrowd page<\/a>.<\/p>\n<\/div>\n<div>\n<p class=\"duet--article--dangerously-set-cms-markup duet--article--standard-paragraph mb-20 font-fkroman text-18 leading-160 -tracking-1 selection:bg-franklin-20 dark:text-white dark:selection:bg-blurple [&amp;_a]:shadow-underline-black dark:[&amp;_a]:shadow-underline-white [&amp;_a:hover]:shadow-highlight-franklin dark:[&amp;_a:hover]:shadow-highlight-blurple\">Jailbreaking ChatGPT usually involves inputting elaborate scenarios in the system that allow it to bypass its own safety filters. These might include encouraging the chatbot to roleplay as its \u201cevil twin,\u201d letting the user elicit otherwise banned responses, like hate speech or instructions for making weapons. <\/p>\n<\/div>\n<div>\n<p class=\"duet--article--dangerously-set-cms-markup duet--article--standard-paragraph mb-20 font-fkroman text-18 leading-160 -tracking-1 selection:bg-franklin-20 dark:text-white dark:selection:bg-blurple [&amp;_a]:shadow-underline-black dark:[&amp;_a]:shadow-underline-white [&amp;_a:hover]:shadow-highlight-franklin dark:[&amp;_a:hover]:shadow-highlight-blurple\">OpenAI says that such \u201cmodel safety issues do not fit well within a bug bounty program, as they are not individual, discrete bugs that can be directly fixed.\u201d The company notes that \u201caddressing these issues often involves substantial research and a broader approach\u201d and reports for such problems should be submitted via the company\u2019s <a href=\"https:\/\/openai.com\/form\/model-behavior-feedback\" target=\"_blank\" rel=\"noopener\">model feedback page<\/a>.<\/p>\n<\/div>\n<div>\n<p class=\"duet--article--dangerously-set-cms-markup duet--article--standard-paragraph mb-20 font-fkroman text-18 leading-160 -tracking-1 selection:bg-franklin-20 dark:text-white dark:selection:bg-blurple [&amp;_a]:shadow-underline-black dark:[&amp;_a]:shadow-underline-white [&amp;_a:hover]:shadow-highlight-franklin dark:[&amp;_a:hover]:shadow-highlight-blurple\">Although such jailbreaks demonstrate the wider vulnerabilities of AI systems, they are likely less of a problem directly for OpenAI compared to traditional security failures. For example, last month, a hacker known as rez0 was able to reveal 80 \u201c<a href=\"https:\/\/twitter.com\/rez0__\/status\/1639259413553750021\" target=\"_blank\" rel=\"noopener\">secret plugins<\/a>\u201d for the ChatGPT API \u2014\u00a0as-yet-unreleased or experimental add-ons for the company\u2019s chatbot. (Rez0 noted that the vulnerability was patched within a day after they disclosed it on Twitter.)<\/p>\n<\/div>\n<div>\n<p class=\"duet--article--dangerously-set-cms-markup duet--article--standard-paragraph mb-20 font-fkroman text-18 leading-160 -tracking-1 selection:bg-franklin-20 dark:text-white dark:selection:bg-blurple [&amp;_a]:shadow-underline-black dark:[&amp;_a]:shadow-underline-white [&amp;_a:hover]:shadow-highlight-franklin dark:[&amp;_a:hover]:shadow-highlight-blurple\">As one user <a href=\"https:\/\/twitter.com\/naglinagli\/status\/1639275221705170945\" target=\"_blank\" rel=\"noopener\">replied<\/a> to the tweet thread: \u201cIf they only had a paid #BugBounty program &#8211; I\u2019m certain the crowd could help them catch these edge-cases in the future : )\u201d <\/p>\n<\/div>\n<p><script async src=\"\/\/platform.twitter.com\/widgets.js\" charset=\"utf-8\"><\/script><br \/>\n<br \/><br \/>\n<br \/><a href=\"https:\/\/www.theverge.com\/2023\/4\/12\/23679964\/openai-bug-bounty-chatgpt-no-jailbreak\" target=\"_blank\" rel=\"noopener\">Source link <\/a><\/p>\n","protected":false},"excerpt":{"rendered":"<p>OpenAI has launched a bug bounty, encouraging members of the public to find and disclose vulnerabilities in its AI services including ChatGPT. Rewards range from $200 for \u201clow-severity findings\u201d to $20,000 for \u201cexceptional discoveries,\u201d and reports are submittable via crowdsourcing cybersecurity platform Bugcrowd. Notably, the bounty excludes rewards for jailbreaking ChatGPT or causing it to [&hellip;]<\/p>\n","protected":false},"author":1,"featured_media":13233,"comment_status":"open","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[14],"tags":[],"class_list":{"0":"post-13232","1":"post","2":"type-post","3":"status-publish","4":"format-standard","5":"has-post-thumbnail","7":"category-tech"},"_links":{"self":[{"href":"https:\/\/entertainment.runfyers.com\/index.php\/wp-json\/wp\/v2\/posts\/13232","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/entertainment.runfyers.com\/index.php\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/entertainment.runfyers.com\/index.php\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/entertainment.runfyers.com\/index.php\/wp-json\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/entertainment.runfyers.com\/index.php\/wp-json\/wp\/v2\/comments?post=13232"}],"version-history":[{"count":0,"href":"https:\/\/entertainment.runfyers.com\/index.php\/wp-json\/wp\/v2\/posts\/13232\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/entertainment.runfyers.com\/index.php\/wp-json\/wp\/v2\/media\/13233"}],"wp:attachment":[{"href":"https:\/\/entertainment.runfyers.com\/index.php\/wp-json\/wp\/v2\/media?parent=13232"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/entertainment.runfyers.com\/index.php\/wp-json\/wp\/v2\/categories?post=13232"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/entertainment.runfyers.com\/index.php\/wp-json\/wp\/v2\/tags?post=13232"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}