{"id":134072,"date":"2024-11-21T21:46:13","date_gmt":"2024-11-21T21:46:13","guid":{"rendered":"https:\/\/entertainment.runfyers.com\/index.php\/2024\/11\/21\/openai-accidentally-erases-potential-evidence-in-training-data-lawsuit\/"},"modified":"2024-11-21T21:46:13","modified_gmt":"2024-11-21T21:46:13","slug":"openai-accidentally-erases-potential-evidence-in-training-data-lawsuit","status":"publish","type":"post","link":"https:\/\/entertainment.runfyers.com\/index.php\/2024\/11\/21\/openai-accidentally-erases-potential-evidence-in-training-data-lawsuit\/","title":{"rendered":"OpenAI accidentally erases potential evidence in training data lawsuit"},"content":{"rendered":"<p> <br \/>\n<\/p>\n<div>\n<p class=\"duet--article--dangerously-set-cms-markup duet--article--standard-paragraph mb-20 font-fkroman text-18 leading-160 -tracking-1 selection:bg-franklin-20 dark:text-white dark:selection:bg-blurple [&amp;_a:hover]:shadow-highlight-franklin dark:[&amp;_a:hover]:shadow-highlight-blurple [&amp;_a]:shadow-underline-black dark:[&amp;_a]:shadow-underline-white\">In a stunning misstep, OpenAI engineers accidentally erased critical evidence gathered by <em>The New York Times<\/em> and other major newspapers in their lawsuit over AI training data, according to a court filing Wednesday.<\/p>\n<\/div>\n<div>\n<p class=\"duet--article--dangerously-set-cms-markup duet--article--standard-paragraph mb-20 font-fkroman text-18 leading-160 -tracking-1 selection:bg-franklin-20 dark:text-white dark:selection:bg-blurple [&amp;_a:hover]:shadow-highlight-franklin dark:[&amp;_a:hover]:shadow-highlight-blurple [&amp;_a]:shadow-underline-black dark:[&amp;_a]:shadow-underline-white\">The newspapers\u2019 legal teams had spent over 150 hours searching through OpenAI\u2019s AI training data to find instances where their news articles were included, the filing claims. But it doesn\u2019t explain how this mistake occurred or what precisely the data included. While the filing says OpenAI admitted to the error and tried to recover the data, what it managed to salvage was incomplete and unreliable \u2014 so what was recovered cannot help properly trace how the news organizations\u2019 articles were used in building OpenAI\u2019s AI models. While OpenAI\u2019s lawyers <a href=\"https:\/\/storage.courtlistener.com\/recap\/gov.uscourts.nysd.612697\/gov.uscourts.nysd.612697.210.2.pdf\" target=\"_blank\" rel=\"noopener\">characterized<\/a> the data erasure as a \u201cglitch,\u201d The New York Times\u2019 attorneys noted they had \u201cno reason to believe\u201d it was intentional.<\/p>\n<\/div>\n<div>\n<p class=\"duet--article--dangerously-set-cms-markup duet--article--standard-paragraph mb-20 font-fkroman text-18 leading-160 -tracking-1 selection:bg-franklin-20 dark:text-white dark:selection:bg-blurple [&amp;_a:hover]:shadow-highlight-franklin dark:[&amp;_a:hover]:shadow-highlight-blurple [&amp;_a]:shadow-underline-black dark:[&amp;_a]:shadow-underline-white\">The New York Times Company <a href=\"https:\/\/www.theverge.com\/2023\/12\/27\/24016212\/new-york-times-openai-microsoft-lawsuit-copyright-infringement\" target=\"_blank\" rel=\"noopener\">launched this landmark battle<\/a> last December, claiming OpenAI and its partner Microsoft had built their AI tools by \u201ccopying and using millions\u201d of the publication\u2019s articles and now \u201cdirectly compete\u201d with its content as a result. The publication is asking for OpenAI to be held liable for \u201cbillions of dollars in statutory and actual damages\u201d for allegedly copying its works.\u00a0<\/p>\n<\/div>\n<div>\n<p class=\"duet--article--dangerously-set-cms-markup duet--article--standard-paragraph mb-20 font-fkroman text-18 leading-160 -tracking-1 selection:bg-franklin-20 dark:text-white dark:selection:bg-blurple [&amp;_a:hover]:shadow-highlight-franklin dark:[&amp;_a:hover]:shadow-highlight-blurple [&amp;_a]:shadow-underline-black dark:[&amp;_a]:shadow-underline-white\">The Times has already spent <a href=\"https:\/\/www.theverge.com\/2024\/5\/9\/24152893\/the-new-york-times-spent-1-million-so-far-in-its-openai-lawsuit\" target=\"_blank\" rel=\"noopener\">more than $1 million<\/a> battling OpenAI in court \u2014 a significant fee few publishers can match. Meanwhile, OpenAI has struck deals with major outlets like Axel Springer, Conde Nast, and <em>The Verge\u2019s<\/em> parent company Vox Media, suggesting many publishers would rather partner than fight.<\/p>\n<\/div>\n<div>\n<p class=\"duet--article--dangerously-set-cms-markup duet--article--standard-paragraph mb-20 font-fkroman text-18 leading-160 -tracking-1 selection:bg-franklin-20 dark:text-white dark:selection:bg-blurple [&amp;_a:hover]:shadow-highlight-franklin dark:[&amp;_a:hover]:shadow-highlight-blurple [&amp;_a]:shadow-underline-black dark:[&amp;_a]:shadow-underline-white\">OpenAI declined to join The New York Times in filing the update to the court. This declaration was filed by Jennifer Maisel, an attorney representing the news organizations, to formally notify the court about what happened.<\/p>\n<\/div>\n<div>\n<p class=\"duet--article--dangerously-set-cms-markup duet--article--standard-paragraph mb-20 font-fkroman text-18 leading-160 -tracking-1 selection:bg-franklin-20 dark:text-white dark:selection:bg-blurple [&amp;_a:hover]:shadow-highlight-franklin dark:[&amp;_a:hover]:shadow-highlight-blurple [&amp;_a]:shadow-underline-black dark:[&amp;_a]:shadow-underline-white\">In an email to <em>The Verge<\/em>, OpenAI spokesperson Jason Deutrom said that the company disagrees with the characterizations made, and will file its own response soon. The New York Times declined <em>The Verge<\/em>\u2019s request for comment. <\/p>\n<\/div>\n<p><br \/>\n<br \/><a href=\"https:\/\/www.theverge.com\/2024\/11\/21\/24302606\/openai-erases-evidence-in-training-data-lawsuit\" target=\"_blank\" rel=\"noopener\">Source link <\/a><\/p>\n","protected":false},"excerpt":{"rendered":"<p>In a stunning misstep, OpenAI engineers accidentally erased critical evidence gathered by The New York Times and other major newspapers in their lawsuit over AI training data, according to a court filing Wednesday. The newspapers\u2019 legal teams had spent over 150 hours searching through OpenAI\u2019s AI training data to find instances where their news articles [&hellip;]<\/p>\n","protected":false},"author":1,"featured_media":134073,"comment_status":"open","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[14],"tags":[],"class_list":{"0":"post-134072","1":"post","2":"type-post","3":"status-publish","4":"format-standard","5":"has-post-thumbnail","7":"category-tech"},"_links":{"self":[{"href":"https:\/\/entertainment.runfyers.com\/index.php\/wp-json\/wp\/v2\/posts\/134072","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/entertainment.runfyers.com\/index.php\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/entertainment.runfyers.com\/index.php\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/entertainment.runfyers.com\/index.php\/wp-json\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/entertainment.runfyers.com\/index.php\/wp-json\/wp\/v2\/comments?post=134072"}],"version-history":[{"count":0,"href":"https:\/\/entertainment.runfyers.com\/index.php\/wp-json\/wp\/v2\/posts\/134072\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/entertainment.runfyers.com\/index.php\/wp-json\/wp\/v2\/media\/134073"}],"wp:attachment":[{"href":"https:\/\/entertainment.runfyers.com\/index.php\/wp-json\/wp\/v2\/media?parent=134072"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/entertainment.runfyers.com\/index.php\/wp-json\/wp\/v2\/categories?post=134072"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/entertainment.runfyers.com\/index.php\/wp-json\/wp\/v2\/tags?post=134072"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}