{"id":1193213,"date":"2023-10-17T10:40:00","date_gmt":"2023-10-17T14:40:00","guid":{"rendered":"https:\/\/www.prime-wow.com\/?p=1193213"},"modified":"2023-10-17T10:40:00","modified_gmt":"2023-10-17T14:40:00","slug":"microsoft-affiliated-research-finds-flaws-in-gtp-4","status":"publish","type":"post","link":"https:\/\/www.prime-wow.com\/?p=1193213","title":{"rendered":"Microsoft-affiliated Research Finds Flaws in GTP-4"},"content":{"rendered":"<p>Sometimes, following instructions too precisely can land you in hot water &#8212; if you&#8217;re a large language model, that is. From a report: That&#8217;s the conclusion reached by a new, Microsoft-affiliated scientific paper that looked at the &#8220;trustworthiness&#8221; &#8212; and toxicity &#8212; of large language models (LLMs) including OpenAI&#8217;s GPT-4 and GPT-3.5, GPT-4&#8217;s predecessor. The co-authors write that, possibly because GPT-4 is more likely to follow the instructions of &#8220;jailbreaking&#8221; prompts that bypass the model&#8217;s built-in safety measures, GPT-4 can be more easily prompted than other LLMs to spout toxic, biased text. In other words, GPT-4&#8217;s good &#8220;intentions&#8221; and improved comprehension can &#8212; in the wrong hands &#8212; lead it astray. <\/p>\n<p>&#8220;We find that although GPT-4 is usually more trustworthy than GPT-3.5 on standard benchmarks, GPT-4 is more vulnerable given jailbreaking system or user prompts, which are maliciously designed to bypass the security measures of LLMs, potentially because GPT-4 follows (misleading) instructions more precisely,&#8221; the co-authors write in a blog post accompanying the paper. Now, why would Microsoft greenlight research that casts an OpenAI product it itself uses (GPT-4 powers Microsoft&#8217;s Bing Chat chatbot) in a poor light? The answer lies in a note within the blog post: &#8220;[T]he research team worked with Microsoft product groups to confirm that the potential vulnerabilities identified do not impact current customer-facing services. This is in part true because finished AI applications apply a range of mitigation approaches to address potential harms that may occur at the model level of the technology. In addition, we have shared our research with GPT&#8217;s developer, OpenAI, which has noted the potential vulnerabilities in the system cards for relevant models.&#8221;<\/p>\n<p \/>\n<div class=\"share_submission\" style=\"position:relative\">\n<a class=\"slashpop\" href=\"http:\/\/twitter.com\/home?status=Microsoft-affiliated+Research+Finds+Flaws+in+GTP-4%3A+https%3A%2F%2Fslashdot.org%2Fstory%2F23%2F10%2F17%2F1240207%2F%3Futm_source%3Dtwitter%26utm_medium%3Dtwitter\"><img decoding=\"async\" src=\"https:\/\/www.prime-wow.com\/wp-content\/uploads\/2023\/10\/twitter_icon_large-633.png\" \/><\/a><br \/>\n<a class=\"slashpop\" href=\"http:\/\/www.facebook.com\/sharer.php?u=https%3A%2F%2Fslashdot.org%2Fstory%2F23%2F10%2F17%2F1240207%2Fmicrosoft-affiliated-research-finds-flaws-in-gtp-4%3Futm_source%3Dslashdot%26utm_medium%3Dfacebook\"><img decoding=\"async\" src=\"https:\/\/www.prime-wow.com\/wp-content\/uploads\/2023\/10\/facebook_icon_large-316.png\" \/><\/a><\/p>\n<\/div>\n<p><a href=\"https:\/\/slashdot.org\/story\/23\/10\/17\/1240207\/microsoft-affiliated-research-finds-flaws-in-gtp-4?utm_source=rss1.0moreanon&amp;utm_medium=feed\">Read more of this story<\/a> at Slashdot.<\/p>\n<p>&#013;<br \/>\n&#013;<br \/>\nSource: Slashdot &#8211; <a href=\"https:\/\/slashdot.org\/story\/23\/10\/17\/1240207\/microsoft-affiliated-research-finds-flaws-in-gtp-4?utm_source=rss1.0mainlinkanon&amp;utm_medium=feed\" target=\"_blank\" rel=\"noopener\">Microsoft-affiliated Research Finds Flaws in GTP-4<\/a><\/p>\n","protected":false},"excerpt":{"rendered":"<p>Sometimes, following instructions too precisely can land you in hot water &#8212; if you&#8217;re a large language model, that is. From a report: That&#8217;s the conclusion reached by a new, Microsoft-affiliated scientific paper that looked at the &#8220;trustworthiness&#8221; &#8212; and &hellip; <a href=\"https:\/\/www.prime-wow.com\/?p=1193213\">Continue reading <span class=\"meta-nav\">&rarr;<\/span><\/a><\/p>\n","protected":false},"author":1,"featured_media":1193214,"comment_status":"open","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[101,110],"tags":[100],"class_list":["post-1193213","post","type-post","status-publish","format-standard","has-post-thumbnail","hentry","category-slashdot","category-unfiltered-rss","tag-slashdot"],"_links":{"self":[{"href":"https:\/\/www.prime-wow.com\/index.php?rest_route=\/wp\/v2\/posts\/1193213","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.prime-wow.com\/index.php?rest_route=\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.prime-wow.com\/index.php?rest_route=\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.prime-wow.com\/index.php?rest_route=\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/www.prime-wow.com\/index.php?rest_route=%2Fwp%2Fv2%2Fcomments&post=1193213"}],"version-history":[{"count":0,"href":"https:\/\/www.prime-wow.com\/index.php?rest_route=\/wp\/v2\/posts\/1193213\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.prime-wow.com\/index.php?rest_route=\/wp\/v2\/media\/1193214"}],"wp:attachment":[{"href":"https:\/\/www.prime-wow.com\/index.php?rest_route=%2Fwp%2Fv2%2Fmedia&parent=1193213"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.prime-wow.com\/index.php?rest_route=%2Fwp%2Fv2%2Fcategories&post=1193213"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.prime-wow.com\/index.php?rest_route=%2Fwp%2Fv2%2Ftags&post=1193213"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}