{"id":148822,"date":"2026-08-23T19:17:11","date_gmt":"2026-08-23T19:17:11","guid":{"rendered":"https:\/\/www.europesays.com\/ai\/148822\/"},"modified":"2026-08-23T19:17:11","modified_gmt":"2026-08-23T19:17:11","slug":"artists-built-a-site-to-escape-ai-scrapers-are-coming-for-it-anyway","status":"publish","type":"post","link":"https:\/\/www.europesays.com\/ai\/148822\/","title":{"rendered":"Artists Built A Site To Escape AI. Scrapers Are Coming For It Anyway"},"content":{"rendered":"<p><img decoding=\"async\" class=\" top-image\" src=\"https:\/\/www.europesays.com\/ai\/wp-content\/uploads\/2026\/08\/1787512631_870_0x0.jpg\" alt=\"cara-app-users---jingna-zhang\" data-height=\"2616\" data-width=\"3508\" fetchpriority=\"high\" style=\"position:absolute;top:0\"\/><\/p>\n<p>User page from Cara, a social network for artists to share their work where AI usage and AI data scraping is explicitly forbidden.<\/p>\n<p>Courtesy of Cara<\/p>\n<p><a class=\"color-link\" href=\"https:\/\/cara.app\/explore\" target=\"_blank\" rel=\"nofollow noopener noreferrer\" data-ga-track=\"ExternalLink:https:\/\/cara.app\/explore\" aria-label=\"Cara\">Cara<\/a>, an image sharing social network built specifically for artists who do not consent to have their work used to train generative AI models, suffered a third scraping attack in the past ten days, with copyrighted images and metadata taken from the site without permission appearing on the open source data network Academic Torrents on August 22. This follows previous scrapes of the dataset posted on Reddit (since taken down by the original poster) and the AI model clearinghouse Hugging Face.<\/p>\n<p>\u201cWe have been scraped for a third time. This is now a targeted attack meant to cause artists pain,&#8221; wrote Cara founder Jingna Zhang <a class=\"color-link\" href=\"https:\/\/x.com\/zemotion\/status\/2091590260438569394?s=20\" target=\"_blank\" rel=\"nofollow noopener noreferrer\" data-ga-track=\"ExternalLink:https:\/\/x.com\/zemotion\/status\/2091590260438569394?s=20\" aria-label=\"in a thread on X\">in a thread on X <\/a>Sunday morning. \u201cScrapers know we\u2019re a volunteer project with no financial means to respond. So they say it\u2019s legal, believing themselves to be untouchable.\u201d<\/p>\n<p>\u201cI believe artists who want to share their work should get to do so without violating the principle of consent. I won\u2019t be bullied into making Cara members-only because some people think \u2018consent isn\u2019t always valid.\u2019 You don\u2019t blame victims after harming them.\u201d <\/p>\n<p>She concludes with a plea for community support to raise funds for a legal defense via a <a class=\"color-link\" href=\"https:\/\/urldefense.proofpoint.com\/v2\/url?u=https-3A__www.gofundme.com_f_help-2Dcara-2Dlegal-2Dfund&amp;d=DwMFaQ&amp;c=euGZstcaTDllvimEN8b7jXrwqOf-v5A_CdpgnVfiiMM&amp;r=lfRzqcaWvGwkELwfGOwzXi2YdN-usPhi7djPcyP_FRk&amp;m=dc7eLnyCXUyzXo9sKQYNz4wMI7U5N5Uhtv__-14ToSizqHZ_SRH8xGsoinK5w65D&amp;s=CA0o227whGJ8lVQ1sOM_oL0UlF5_ktqm6ixsGEifHaU&amp;e=\" target=\"_blank\" rel=\"nofollow noopener noreferrer\" data-ga-track=\"ExternalLink:https:\/\/urldefense.proofpoint.com\/v2\/url?u=https-3A__www.gofundme.com_f_help-2Dcara-2Dlegal-2Dfund&amp;d=DwMFaQ&amp;c=euGZstcaTDllvimEN8b7jXrwqOf-v5A_CdpgnVfiiMM&amp;r=lfRzqcaWvGwkELwfGOwzXi2YdN-usPhi7djPcyP_FRk&amp;m=dc7eLnyCXUyzXo9sKQYNz4wMI7U5N5Uhtv__-14ToSizqHZ_SRH8xGsoinK5w65D&amp;s=CA0o227whGJ8lVQ1sOM_oL0UlF5_ktqm6ixsGEifHaU&amp;e=\" aria-label=\"GoFundMe\">GoFundMe<\/a>.<\/p>\n<p>Zhang, a fashion and fine art photographer who was named to the <a class=\"color-link\" href=\"https:\/\/www.forbes.com\/profile\/jingna-zhang\/\" data-ga-track=\"InternalLink:https:\/\/www.forbes.com\/profile\/jingna-zhang\/\" target=\"_self\" aria-label=\"Forbes 30 Under 30 list\" rel=\"nofollow noopener\">Forbes 30 Under 30 list <\/a>in 2018 and whose work has appeared in Vogue, Elle, Harper\u2019s Bazaar and Time, founded Cara, a public benefit corporation, in late 2022. The network shares some of the same functionality as popular commercial applications like Meta\u2019s Instagram, but has explicit terms of service and features to protect the rights of artists from having their work incorporated into AI training sets without compensation, control or consent.<\/p>\n<p>Artists Seek Refuge from AI<\/p>\n<p>Since it launched, the service has attracted over 1.5 million users, including many professional and hobbyist artists who see AI as a <a class=\"color-link\" href=\"https:\/\/www.forbes.com\/sites\/robsalkowitz\/2022\/09\/16\/ai-is-coming-for-commercial-art-jobs-can-it-be-stopped\/\" data-ga-track=\"InternalLink:https:\/\/www.forbes.com\/sites\/robsalkowitz\/2022\/09\/16\/ai-is-coming-for-commercial-art-jobs-can-it-be-stopped\/\" target=\"_self\" aria-label=\"threat to their livelihoods\" rel=\"nofollow noopener\">threat to their livelihoods<\/a> and, often, as a<a class=\"color-link\" href=\"https:\/\/www.forbes.com\/sites\/robsalkowitz\/2026\/04\/09\/ai-infiltration-into-the-arts-has-fans-seeing-slop-everywhere\/\" data-ga-track=\"InternalLink:https:\/\/www.forbes.com\/sites\/robsalkowitz\/2026\/04\/09\/ai-infiltration-into-the-arts-has-fans-seeing-slop-everywhere\/\" target=\"_self\" aria-label=\"basic affront\" rel=\"nofollow noopener\"> basic affront<\/a> to the ideals of art. As working artists, they need to share and promote their artwork to get work from clients and collectors, but they do not want their work used to train systems designed to replace them.<\/p>\n<p>When the first wave of generative image LLMs like Midjourney, Stable Diffusion and Dall-E appeared, they were trained used a huge archive of images, including many under copyright, scraped from the Internet by automated bots. Artists realized that participating in public platforms, even their own personal websites, exposed their work to these systems. Zhang built and launched Cara on a shoestring as a place where artists could share work in an environment where no consent could plausibly be implied. While no site can be perfectly secured in today\u2019s online world, it provides a modicum of legal and technological countermeasures and a clear, human-centric ethical stance.<\/p>\n<p>Jingna Zhang, photographer and founder of the art site Cara<\/p>\n<p>Courtesy of Jingna Zhang<\/p>\n<p>\u201cIn a world where it feel like nobody is actually doing anything, Cara is trying to make a difference,\u201d said Zhang in a phone interview. \u201cIt might not be perfect, it might not solve all the problems or change how the internet is designed, but we are trying, and that makes a huge difference.\u201d<\/p>\n<p>Scrapers defend their actions<\/p>\n<p>Artists\u2019 claims to own and control their own work online are disputed by individuals, groups and commercial entities who believe that advancing the progress of AI entitles them to any and all data they can obtain, regardless of consent. This was apparently the motivation of the original scraper, who posted an archive of nearly 12 million images from Cara to the sub-Reddit r\/DefendingAIArt  under the handle \u201cMandarinDrawnPoppy994.\u201d<\/p>\n<p>\u201cScraping is necessary to develop good models. It\u2019s like building a highway \u2013 some houses must be demolished, but in the end everyone benefits,\u201d the poster <a href=\"https:\/\/www.reddit.com\/r\/DefendingAIArt\/comments\/1vmzv8t\/comment\/p3emgjb\/?utm_source=share&amp;utm_medium=web3x&amp;utm_name=web3xcss&amp;utm_term=1\" target=\"_blank\" rel=\"nofollow noopener noreferrer\" data-ga-track=\"ExternalLink:https:\/\/www.reddit.com\/r\/DefendingAIArt\/comments\/1vmzv8t\/comment\/p3emgjb\/?utm_source=share&amp;utm_medium=web3x&amp;utm_name=web3xcss&amp;utm_term=1\" aria-label=\"wrote in a thread titled\">wrote in a thread titled<\/a> \u201c[AMA] I scraped all of Cara.\u201d<\/p>\n<p>Zhang says she and others reached out to the original scraper and eventually prevailed on him to take down the post. She adds, \u201cnot only did the first scraper delete the dataset, but he has turned around now to offer help, and we\u2019re now co-creating an <a class=\"color-link\" href=\"https:\/\/lanternite.org\/ \" target=\"_blank\" rel=\"nofollow noopener noreferrer\" data-ga-track=\"ExternalLink:https:\/\/lanternite.org\/\" aria-label=\"open source tool\">open source tool <\/a>separate from Cara that will help people check if they&#8217;ve been scraped in new datasets in the future.\u201d<\/p>\n<p>Unfortunately, that was not the end of the problem. Several days later, the site was scraped again by a different actor. This time the data was posted on Hugging Face, a hub of resources for AI developers rooted in the open source community, by a <a href=\"https:\/\/huggingface.co\/CaptiveDreamer\" target=\"_blank\" rel=\"nofollow noopener noreferrer\" data-ga-track=\"ExternalLink:https:\/\/huggingface.co\/CaptiveDreamer\" aria-label=\"poster under the name \u201cIoannis\/Captive Dreamer.\u201d\">poster under the name \u201cIoannis\/Captive Dreamer.\u201d<\/a> <\/p>\n<p>After some people reported the post to Hugging Face Trust and Safety, <a href=\"https:\/\/huggingface.co\/datasets\/CaptiveDreamer\/CaraArchive\/discussions\/57\" target=\"_blank\" rel=\"nofollow noopener noreferrer\" data-ga-track=\"ExternalLink:https:\/\/huggingface.co\/datasets\/CaptiveDreamer\/CaraArchive\/discussions\/57\" aria-label=\"the team responded\">the team responded<\/a> that \u201cwe have reviewed these [copyright reports] carefully. Because no copies of the artworks are hosted here, and because the URLs [in the dataset] point to the copies the artists published on Cara, there is nothing hosted on Hugging Face that we can disable through our notice and takedown process. This is not a judgement about who owns the works (the artists do); it is about what is  stored on our servers. Further copyright reports on the same basis will not change this outcome.\u201d<\/p>\n<p>Hugging Face did not respond to a request for further comment for this story.<\/p>\n<p>Third attack in 10 days<\/p>\n<p>Now on Saturday, August 22, Zhang says the site was scraped for a third time, with the perpetrator taking just 123K images, but also a second data set that includes users\u2019 text posts, bio and information they share on the site. That archive has been <a href=\"https:\/\/academictorrents.com\/details\/77d4e2a65852258420a8e47bf29c8f3a5043a446\" target=\"_blank\" rel=\"nofollow noopener noreferrer\" data-ga-track=\"ExternalLink:https:\/\/academictorrents.com\/details\/77d4e2a65852258420a8e47bf29c8f3a5043a446\" aria-label=\"posted on Academic Torrents\">posted on Academic Torrents<\/a>, a site that \u201cwas established to meet the demands of science in the age of big data\u201d by providing data for researchers, according to its <a href=\"https:\/\/academictorrents.com\/docs\/about.html\" target=\"_blank\" rel=\"nofollow noopener noreferrer\" data-ga-track=\"ExternalLink:https:\/\/academictorrents.com\/docs\/about.html\" aria-label=\"\u201cAbout\u201d page.\">\u201cAbout\u201d page.<\/a><\/p>\n<p>The series of attacks has shaken the members of the community, despite most users recognizing there was not much that the Cara team could do in the face of escalating capabilities of scrapers and hackers.<\/p>\n<p>\u201cI never had an expectation that scraping couldn\u2019t happen,\u201d wrote a poster under the name Venoregard, representing the sentiments of many commenters <a href=\"https:\/\/cara.app\/post\/43fa4225-5e19-455a-a7b3-7f92f89412cd\" target=\"_blank\" rel=\"nofollow noopener noreferrer\" data-ga-track=\"ExternalLink:https:\/\/cara.app\/post\/43fa4225-5e19-455a-a7b3-7f92f89412cd\" aria-label=\"on the thread\">on the thread<\/a>. \u201cI was just happy to see an art community free from genAI images so I didn&#8217;t have to wonder as a non-artist if something was human made. I mean if you can screenshot something you can probably feed it into AI. Art isn&#8217;t unstealable, it&#8217;s bound to happen. But Cara isn&#8217;t doing the feeding and nor are they hosting genAI images on their site, which is a huge step up from the largest social media sites right now imo.\u201d<\/p>\n<p>Zhang says the incursions not only represent an attack on the values of Cara artists and the entire notion that consent matters, they also drain the company\u2019s scarce financial resources. She says the scrapes mean thousands of extra dollars in server fees to handle the automated traffic, at a moment when Cara was hoping to break even or come out slightly ahead for the year. Zhang sjust launched a GoFundMe to help defray the unforeseen expenses.<\/p>\n<p>Zhang says she believes the two most recent attacks are motivated by animus against the Cara community simply for taking a stand against generative AI in the arts. There is ample evidence to suggest that point of view is fairly widespread in communities like r\/DefendingAIArt, who believe themselves persecuted by human artists for using image generation tools and posting the outputs as their personal creations. <\/p>\n<p>\u201cThey go through the trouble of harassing us for using AI and making such a big deal about how they don\u2019t want their art being used to train AI, so to get back at them [original poster, MandarinDrawnPoppy994] took their art to use to train models. An eye for an eye type of situation. Doesn&#8217;t make it right of course, but that&#8217;s the purpose here,\u201d wrote a redditor going by Dragin410 in the AMA thread.<\/p>\n<p>\u201cSomebody really wants to highlight how they can troll us,\u201d she said. \u201cThey\u2019re like, \u2018you say you want to protect artists, you say you\u2019re opting out. So we\u2019re going to violate your consent, violate your saying no.\u2019 At this point this is nothing but malicious, targeted harassment meant to inflict harm.&#8221;<\/p>\n","protected":false},"excerpt":{"rendered":"User page from Cara, a social network for artists to share their work where AI usage and AI&hellip;\n","protected":false},"author":2,"featured_media":148823,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[2],"tags":[24,25,72922,72920,18044,72921],"class_list":["post-148822","post","type-post","status-publish","format-standard","has-post-thumbnail","category-ai","tag-ai","tag-artificial-intelligence","tag-artists-against-ai","tag-cara","tag-hugging-face","tag-jingna-zhang"],"_links":{"self":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/148822","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/comments?post=148822"}],"version-history":[{"count":0,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/148822\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media\/148823"}],"wp:attachment":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media?parent=148822"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/categories?post=148822"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/tags?post=148822"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}