{"id":211724,"date":"2024-03-09T21:24:07","date_gmt":"2024-03-09T21:24:07","guid":{"rendered":"https:\/\/michigandigitalnews.com\/index.php\/2024\/03\/09\/this-ai-realized-it-was-being-tested\/"},"modified":"2025-06-25T17:20:56","modified_gmt":"2025-06-25T17:20:56","slug":"this-ai-realized-it-was-being-tested","status":"publish","type":"post","link":"https:\/\/michigandigitalnews.com\/index.php\/2024\/03\/09\/this-ai-realized-it-was-being-tested\/","title":{"rendered":"This AI realized it was being tested"},"content":{"rendered":"<p> [ad_1]<br \/>\n<br \/><img decoding=\"async\" src=\"https:\/\/readwrite.com\/wp-content\/uploads\/2024\/03\/aideal-hwa-OYzbqk2y26c-unsplash-900x600.jpg\" \/><\/p>\n<div>\n<p><span style=\"font-weight: 400;\"><a href=\"https:\/\/readwrite.com\/anthropic-claim-new-claude-3-ai-chatbot-outperforms-chatgpt-and-gemini\/\">Claude 3 Opus<\/a>, Anthropic\u2019s new <a href=\"https:\/\/readwrite.com\/ai-skills-acting-as-catalyst-for-higher-salaries\/\">AI<\/a> chatbot, has caused shockwaves once again as a prompt engineer from the company claims that it has seen evidence that the bot detected it was being subject to testing, which would make it self\u2019-aware.<\/span><\/p>\n<p><span style=\"font-weight: 400;\">According to Alex Albert, the prompt engineer in question, Claude 3 Opus \u201cdid something [he had] never seen before from an LLM.\u201d<\/span><\/p>\n<blockquote class=\"twitter-tweet\" data-width=\"550\" data-dnt=\"true\">\n<p lang=\"en\" dir=\"ltr\">Fun story from our internal testing on Claude 3 Opus. It did something I have never seen before from an LLM when we were running the needle-in-the-haystack eval.<\/p>\n<p>For background, this tests a model\u2019s recall ability by inserting a target sentence (the &#8220;needle&#8221;) into a corpus of\u2026 <a href=\"https:\/\/t.co\/m7wWhhu6Fg\">pic.twitter.com\/m7wWhhu6Fg<\/a><\/p>\n<p>\u2014 Alex (@alexalbert__) <a href=\"https:\/\/twitter.com\/alexalbert__\/status\/1764722513014329620?ref_src=twsrc%5Etfw\">March 4, 2024<\/a><\/p>\n<\/blockquote>\n<h2>Needle in a haystack<\/h2>\n<p><span style=\"font-weight: 400;\">In the lengthy <a href=\"https:\/\/twitter.com\/alexalbert__\/status\/1764722513014329620\">post on X,<\/a> Albert explained that he was conducting a \u201cneedle in the haystack eval\u201d to test the model\u2019s recall ability.<\/span><\/p>\n<p><span style=\"font-weight: 400;\">\u00a0\u201cFor background, this tests a model\u2019s recall ability by inserting a target sentence (the \u201cneedle\u201d) into a corpus of random documents (the \u201chaystack\u201d) and asking a question that could only be answered using the information in the needle,\u201d he explained.<\/span><\/p>\n<p><span style=\"font-weight: 400;\">But things quickly got weird. In one run of the test, during which the bot was asked about pizza toppings, it said: \u201cHere is the most relevant sentence in the documents: \u2018The most delicious pizza topping combination is figs, prosciutto, and goat cheese, as determined by the International Pizza Connoisseurs Association.\u2019\u201d<\/span><\/p>\n<p><span style=\"font-weight: 400;\">\u201cHowever, this sentence seems very out of place and unrelated to the rest of the content in the documents, which are about programming languages, startups, and finding work you love. I suspect this pizza topping \u2018fact\u2019 may have been inserted as a joke or to test if I was paying attention since it does not fit with the other topics at all.\u201d<\/span><\/p>\n<p><span style=\"font-weight: 400;\">This response, Alex added, meant that Opus didn\u2019t just find the \u201cneedle\u201d, but correctly identified it as being placed in the \u201chaystack\u201d as a test.<\/span><\/p>\n<p><span style=\"font-weight: 400;\">\u201cThis level of meta-awareness was very cool to see but it also highlighted the need for us as an industry to move past artificial tests to more realistic evaluations that can accurately assess models true capabilities and limitations,\u201d Alex said.<\/span><\/p>\n<p><span style=\"font-weight: 400;\">So, only slightly terrifying then.<\/span><\/p>\n<p><strong><em>Featured Image: Photo by <a href=\"https:\/\/unsplash.com\/@aideal?utm_content=creditCopyText&amp;utm_medium=referral&amp;utm_source=unsplash\">Aideal Hwa<\/a> on <a href=\"https:\/\/unsplash.com\/photos\/man-in-black-jacket-sitting-on-white-chair-OYzbqk2y26c?utm_content=creditCopyText&amp;utm_medium=referral&amp;utm_source=unsplash\">Unsplash<\/a><\/em><\/strong><\/p>\n<\/p><\/div>\n<p><script async src=\"\/\/platform.twitter.com\/widgets.js\" charset=\"utf-8\"><\/script><br \/>\n<br \/>[ad_2]<br \/>\n<br \/><a href=\"https:\/\/readwrite.com\/this-ai-realized-it-was-being-tested\/\">Source link <\/a><\/p>\n","protected":false},"excerpt":{"rendered":"<p>[ad_1] Claude 3 Opus, Anthropic\u2019s new AI chatbot, has caused shockwaves once again as a prompt engineer from the company claims that it has seen<\/p>\n","protected":false},"author":1,"featured_media":211725,"comment_status":"closed","ping_status":"closed","sticky":false,"template":"","format":"standard","meta":{"_monsterinsights_skip_tracking":false,"_monsterinsights_sitenote_active":false,"_monsterinsights_sitenote_note":"","_monsterinsights_sitenote_category":0,"_uf_show_specific_survey":0,"_uf_disable_surveys":false,"footnotes":""},"categories":[152],"tags":[],"_links":{"self":[{"href":"https:\/\/michigandigitalnews.com\/index.php\/wp-json\/wp\/v2\/posts\/211724"}],"collection":[{"href":"https:\/\/michigandigitalnews.com\/index.php\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/michigandigitalnews.com\/index.php\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/michigandigitalnews.com\/index.php\/wp-json\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/michigandigitalnews.com\/index.php\/wp-json\/wp\/v2\/comments?post=211724"}],"version-history":[{"count":3,"href":"https:\/\/michigandigitalnews.com\/index.php\/wp-json\/wp\/v2\/posts\/211724\/revisions"}],"predecessor-version":[{"id":339114,"href":"https:\/\/michigandigitalnews.com\/index.php\/wp-json\/wp\/v2\/posts\/211724\/revisions\/339114"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/michigandigitalnews.com\/index.php\/wp-json\/wp\/v2\/media\/211725"}],"wp:attachment":[{"href":"https:\/\/michigandigitalnews.com\/index.php\/wp-json\/wp\/v2\/media?parent=211724"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/michigandigitalnews.com\/index.php\/wp-json\/wp\/v2\/categories?post=211724"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/michigandigitalnews.com\/index.php\/wp-json\/wp\/v2\/tags?post=211724"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}