{"id":68028,"date":"2026-06-09T21:15:08","date_gmt":"2026-06-09T21:15:08","guid":{"rendered":"https:\/\/www.europesays.com\/ai\/68028\/"},"modified":"2026-06-09T21:15:08","modified_gmt":"2026-06-09T21:15:08","slug":"same-gatekeepers-new-tollbooths-in-the-ai-content-licensing-market","status":"publish","type":"post","link":"https:\/\/www.europesays.com\/ai\/68028\/","title":{"rendered":"Same gatekeepers, new tollbooths in the AI content licensing market"},"content":{"rendered":"<p>When Google began indexing news websites at the turn of the century, publishers <a href=\"https:\/\/www.techpolicy.press\/cloudflare-wades-into-the-battle-over-ai-consent-and-compensation\/\" rel=\"nofollow noopener\" target=\"_blank\">exchanged<\/a> free listings of their content in exchange for referral traffic, at a rate of 2-to-1. That was before Google <a href=\"https:\/\/www.techpolicy.press\/doj-vs-google-back-to-court-for-remedies-to-break-digital-ads-monopoly\/\" rel=\"nofollow noopener\" target=\"_blank\">monopolized<\/a> digital advertising. Within a decade, referral traffic from search had become the dominant driver of publisher revenue even as Google crawled their sites at a greater rate than it returned traffic, and the economic logic of online journalism was set by the infrastructure of a company that now used snippets and high-quality photos to index results and established an <a href=\"https:\/\/www.journalismliberty.org\/google-search-monopoly\" rel=\"nofollow noopener\" target=\"_blank\">illegal<\/a> monopoly in search (and would turn out to have an illegal monopoly in the <a href=\"https:\/\/www.journalismliberty.org\/google-adtech-monopoly\" rel=\"nofollow noopener\" target=\"_blank\">digital advertising<\/a> market as well).<\/p>\n<p>The story of artificial intelligence (AI) and journalism now follows the same arc, only faster. A small-town reporter who finishes a story about the local school board tonight may find that by morning an AI system has crawled it, synthesized its facts into an answer for a user query, and served that answer without sending anyone back to the publisher\u2019s website. This results in waning reader traffic, advertising revenue, and subscriber conversion for the publisher. While the AI company gets to offer a product that better meets the needs of its user in part through the availability of journalistic information, it comes at a cost\u2014often the models <a href=\"https:\/\/arxiv.org\/html\/2507.05301v1\" rel=\"nofollow noopener\" target=\"_blank\">do not attribute<\/a> where it sourced its information nor return any of the value generated by the user to the source. Now Google is reportedly crawling exponentially <a href=\"https:\/\/www.openmarketsinstitute.org\/publications\/report-mapping-the-ai-content-licensing-market\" rel=\"nofollow noopener\" target=\"_blank\">more times<\/a> per referral, with the other AI engines <a href=\"https:\/\/blog.cloudflare.com\/ai-search-crawl-refer-ratio-on-radar\/\" rel=\"nofollow noopener\" target=\"_blank\">offering<\/a> even worse referral rates, with no way to opt out without also harming one\u2019s search visibility. Generative AI is replicating many of the same dynamics at a scope and speed that makes the search and social media moments feel quaint.<\/p>\n<p>It is in this context that a market for AI content licensing has begun to take shape\u2014the subject of a <a href=\"https:\/\/www.openmarketsinstitute.org\/publications\/report-mapping-the-ai-content-licensing-market\" rel=\"nofollow noopener\" target=\"_blank\">new report<\/a>, \u201cSame Gatekeepers, New Tollbooths: Mapping the AI Content Licensing Market,\u201d which I co-authored with Karina Montoya at the Center for Journalism and Liberty at the Open Markets Institute. The report reveals that the conditions under which the publishing market is forming closely resembles those that structurally damaged journalism and the public interest. <a href=\"https:\/\/www.openmarketsinstitute.org\/publications\/report-mapping-the-ai-content-licensing-market\" rel=\"nofollow noopener\" target=\"_blank\">Our report<\/a> also comes on the heels of two major independent analyses of AI and copyright\u2014studies published by the <a href=\"https:\/\/committees.parliament.uk\/committee\/170\/communications-and-digital-committee\/news\/212361\/uk-creative-industries-face-a-clear-and-present-danger-from-generative-ai\/\" rel=\"nofollow noopener\" target=\"_blank\">U.K. House of Lords<\/a> and the <a href=\"https:\/\/www.europarl.europa.eu\/RegData\/etudes\/STUD\/2025\/778859\/IUST_STU(2025)778859_EN.pdf\" rel=\"nofollow noopener\" target=\"_blank\">European Parliament<\/a>, which arrive at conclusions strikingly convergent with our own. This means three independent bodies working according to different methodologies and different institutional vantage points reached the same diagnosis, suggesting the window for policy intervention may be open to mitigate the harms.<\/p>\n<p>Over the course of my career as a journalist and scholar, I\u2019ve been writing about the need for a more nuanced approach to valuing the press-platform value exchange since <a href=\"https:\/\/papers.ssrn.com\/sol3\/papers.cfm?abstract_id=4192026\" rel=\"nofollow noopener\" target=\"_blank\">well before<\/a> generative AI became widely accessible. My <a href=\"https:\/\/www.openmarketsinstitute.org\/publications\/value-of-journalism-to-ai\" rel=\"nofollow noopener\" target=\"_blank\">research<\/a> has long underscored that the journalism sector was going to need a more <a href=\"https:\/\/www.openmarketsinstitute.org\/publications\/value-of-journalism-to-ai\" rel=\"nofollow noopener\" target=\"_blank\">sophisticated framework<\/a> for valuing its contribution across the full AI value chain and not just the referral traffic layer that platforms had always used to define what news was worth to them. That <a href=\"https:\/\/niemanreports.org\/the-battle-over-using-journalism-to-build-ai-models-is-just-starting\/\" rel=\"nofollow noopener\" target=\"_blank\">narrow framing<\/a> had already cost publishers enormously in the search and social media eras, and accepting it again as AI\u2019s governing logic would compound the damage.<\/p>\n<p>Instead, we must keep in mind that the leverage of technology <a href=\"https:\/\/washingtonmonthly.com\/2024\/10\/29\/ai-needs-us-more-than-we-need-it\/\" rel=\"nofollow noopener\" target=\"_blank\">actually runs<\/a> both ways. AI systems depend on a continuous supply of high-quality human content to remain useful. Degrade that supply\u2014by destroying the economic conditions under which it\u2019s produced\u2014and you degrade the AI itself. Publishers and creators have more bargaining power than they tend to act like they do.<\/p>\n<p>                      3 tiers, 1 structural problem<\/p>\n<p>The AI content licensing market operates across three distinct tiers, each with its own dynamics and each with significant limitations.<\/p>\n<p>First tier<\/p>\n<p>The most visible tier is bilateral deals: confidential agreements between major AI companies and select publishers typically with national or global brand recognition (and the ones most likely to be capable of suing). But while there have been real sums of money flowing to real newsrooms, <a href=\"https:\/\/cmdg.tech\/publications\/ai-content-licensing-report\" rel=\"nofollow noopener\" target=\"_blank\">our research<\/a> shows they are not doing what publishers may have hoped.<\/p>\n<p>Publishers with direct AI licensing agreements initially enjoyed a substantial click-through advantage from AI interfaces. By the fourth quarter of 2025, that \u201cdeal premium\u201d had essentially evaporated\u2014amid a six-fold collapse in click-through rates from AI systems. Publishers without deals fared worse in absolute terms but experienced a smaller proportional drop. But both groups lost. The bilateral deal market is not insulating anyone from the broader erosion of AI-driven referrals, and it is structurally inaccessible to the vast majority of publishers around the world or at the local level.<\/p>\n<p>For example, the Lords Committee rightfully <a href=\"https:\/\/committees.parliament.uk\/committee\/170\/communications-and-digital-committee\/news\/212361\/uk-creative-industries-face-a-clear-and-present-danger-from-generative-ai\/\" rel=\"nofollow noopener\" target=\"_blank\">observed<\/a> that limited disclosure makes it difficult for rights-holders to know whether their works have been used or to enforce their rights. As a result, publishers negotiating bilateral deals are doing so without visibility into how their content is being used, at what frequency, or to what commercial effect, which means they are negotiating blind. Our interviews with AI licensing startup founders confirm this is precisely the information gap that intermediaries have found a commercial foothold in trying to address.<\/p>\n<p>Second tier<\/p>\n<p>The second tier is the intermediary layer, a field that expanded from a handful of Silicon Valley startups to more than a dozen companies since 2024, including startups like TollBit, Sphere AI, ScalePost, Created by Humans, ProRata, Miso.ai, and increasingly Big Tech firms like <a href=\"https:\/\/washingtonmonthly.com\/wp-content\/uploads\/2026\/05\/Cloudflare-Whitepaper.pdf\" rel=\"nofollow noopener\" target=\"_blank\">Cloudflare<\/a> and <a href=\"https:\/\/www.google.com\/url?sa=i&amp;source=web&amp;rct=j&amp;url=https:\/\/www.youtube.com\/watch?v%3DrCOcPNZw8iM&amp;ved=2ahUKEwijmurzycWUAxXGL1kFHb6RMwQQqYcPegoIAggACAAIEBAH&amp;opi=89978449&amp;cd&amp;psig=AOvVaw3pnAsg5hQvFUdUBXB12QR0&amp;ust=1779287851134000\" rel=\"nofollow noopener\" target=\"_blank\">Microsoft<\/a>. As our report explains, these companies offer bot detection and blocking, content marketplaces with pay-per-use pricing, and attribution-based revenue distribution. The analytics and publisher control they offer are meaningful improvements over the opacity of one-on-one deals but carries structural vulnerabilities.<\/p>\n<p>Most startups are venture-backed and therefore exposed to acquisition by the same large technology companies from which they nominally protect publishers (furthermore, their ability to attract capital is influenced by the lack of clarity with respect to the dozens of outstanding copyright lawsuits facing AI firms). I traced the contours of this risk when Cloudflare moved <a href=\"https:\/\/www.techpolicy.press\/cloudflare-wades-into-the-battle-over-ai-consent-and-compensation\/\" rel=\"nofollow noopener\" target=\"_blank\">to block<\/a> AI crawlers by default in mid-2025 and launched its pay-per-crawl marketplace, which was a significant shift that offered publishers more control while also raising serious questions about what it means for an <a href=\"https:\/\/washingtonmonthly.com\/wp-content\/uploads\/2026\/05\/Cloudflare-Whitepaper.pdf\" rel=\"nofollow noopener\" target=\"_blank\">infrastructural gatekeeper<\/a> of Cloudflare\u2019s scope and scale. The independent ad tech ecosystem went through something strikingly similar over the previous decade that resulted in <a href=\"https:\/\/www.journalismliberty.org\/google-adtech-monopoly\" rel=\"nofollow noopener\" target=\"_blank\">Google\u2019s illegal monopoly<\/a>.<\/p>\n<p>Third tier<\/p>\n<p>The third tier is the long tail of media and content producers: local newspapers, regional broadcasters, ethnic and indigenous media, non-English language publishers, and specialized outlets whose loss would cause the greatest civic harm\u2014not to mention individual journalists and creators. They are effectively absent from the AI licensing market entirely. This is a structural feature of how market power is distributed that fails to value how journalism\u2019s civic value is <a href=\"https:\/\/washingtonmonthly.com\/2024\/10\/29\/ai-needs-us-more-than-we-need-it\/\" rel=\"nofollow noopener\" target=\"_blank\">actually distributed<\/a>\u2014much less its role in the accuracy, safety, and integrity of large language models (LLMs). The European Parliament study <a href=\"https:\/\/www.europarl.europa.eu\/RegData\/etudes\/STUD\/2025\/778859\/IUST_STU(2025)778859_EN.pdf\" rel=\"nofollow noopener\" target=\"_blank\">flags<\/a> the same distributional risk in economic terms: Voluntary licensing leads to fragmented coverage and selective deals that produce biased and incomplete datasets, undermining both AI performance and overall welfare. A market that compensates only the publishers large enough to attract bilateral deal interest is not a market that will sustain a <a href=\"https:\/\/taicollaborative.org\/what-makes-for-a-healthy-information-ecosystem-new-visual-tool\" rel=\"nofollow noopener\" target=\"_blank\">healthy information ecosystem<\/a>.<\/p>\n<p>What makes the emerging market structure particularly corrosive is what I call the publisher double bind. The same Big Tech firms whose AI products are eroding website traffic are now building and controlling the licensing infrastructure those publishers must turn to. Traffic erosion pushes publishers toward licensing revenue. Licensing revenue increasingly runs through the corporations that caused the traffic erosion. Google and Microsoft occupy both ends of the value chain simultaneously, not through formal exclusivity arrangements that regulators could easily challenge (for example, via antitrust laws), but through standardization lock-in, data asymmetry, and the magnitude of platform scale (recall we are talking about trillion-dollar intermediaries in these cases).<\/p>\n<p>This dynamic does not require bad intent to produce bad outcomes. It is simply how <a href=\"https:\/\/papers.ssrn.com\/sol3\/papers.cfm?abstract_id=4397263\" rel=\"nofollow noopener\" target=\"_blank\">platform capture<\/a> works and is a <a href=\"https:\/\/www.journalismfestival.com\/programme\/2026\/an-autopsy-for-ai-lessons-for-journalism-from-the-platform-era\" rel=\"nofollow noopener\" target=\"_blank\">recurring pattern<\/a> in the sector\u2019s relationships with platforms. The Lords Committee put the strategic stakes plainly, observing that the continued drift toward tacit acceptance of large-scale, unlicensed use of creative content and long-term dependence on opaque models trained overseas is a \u201cpoor bet\u201d that would sacrifice creative capacity for speculative AI gains expected to accrue largely to a few U.S.-based developers. This framing applies with equal force to publishers operating within any jurisdiction where dominant AI firms are not domestically based (i.e., most of the world).<\/p>\n<p>                      The valuation problem runs deeper than traffic<\/p>\n<p>A recurring problem in the media sector, and journalism in particular, is valuation. The compensation logic governing most negotiations rests on a narrow conception of value: referral traffic lost. But this framing is radically incomplete.<\/p>\n<p>My <a href=\"https:\/\/www.journalismliberty.org\/publications\/value-of-journalism-to-ai\" rel=\"nofollow noopener\" target=\"_blank\">previous research<\/a> on AI and journalism valuation found that publishers\u2019 contributions to AI <a href=\"https:\/\/niemanreports.org\/the-battle-over-using-journalism-to-build-ai-models-is-just-starting\/\" rel=\"nofollow noopener\" target=\"_blank\">extend across<\/a> a much wider range of dimensions: training and fine-tuning; linguistic and reasoning capacity; factual grounding; temporal currency; and civic legitimacy. Accepting referral traffic as the governing benchmark ignores most of these entirely. The European Parliament study corroborates the incentive logic: When creators discover their works have been used for AI training without compensation, they tend to reduce output, risking a degradation in the quality and representativeness of future training data and ultimately harming the performance of AI systems themselves. This is the economic formalization of what we call \u201ccontent cannibalization,\u201d and it closes the loop on the leverage argument. AI companies have a direct interest in the economic sustainability of the content they depend on, yet the market is not currently structured or governed to reflect that interest.<\/p>\n<p>Similarly, licensing for inference or what is known as retrieval-augmented generation (RAG) versus training data should not be considered separate markets. They are layers of the same value stack. The foundation model whose capacity to synthesize information and produce coherent prose was built with publisher content scraped without <a href=\"https:\/\/www.brookings.edu\/articles\/the-case-for-consent-in-the-ai-data-gold-rush\/\" rel=\"nofollow noopener\" target=\"_blank\">consent<\/a>. Without that training foundation, the retrieval layer is useless. Pricing only the retrieval layer while treating the underlying model as a cost publishers have already donated for free is a category error that serves AI companies\u2019 bottom line, not the media\u2019s and much less the public interest.<\/p>\n<p>There is also a legal dimension that has not received adequate attention in this debate, and to which we propose a revised way of interpreting in our paper. Copyright discourse around AI training has largely treated content-scraping as a discrete historical event. But with RAG, the model trained on publisher content is activated anew with every inference call. Publishers who accept current deal terms as final may be foreclosing substantially larger claims that courts have not yet adjudicated and locking in precedents that will be very difficult to revise once normalized and standardized.<\/p>\n<p>                      What a different path requires<\/p>\n<p>It should be crystal clear now that voluntary commitments, platform goodwill, and industry self-regulation have consistently failed to level the playing field. A system in which platform intermediaries control the infrastructure, information flows, and monetization systems while leaving publishers to gather up the shards of traffic and audience left behind is a broken system.<\/p>\n<p>Several fixes are \u00a0achievable within existing legislative traditions, including statutory licensing frameworks with set rates, collective licensing and sectoral bargaining, mandatory transparency on deal terms and data usage, attribution systems at the model inference layer, and explicit inclusion requirements for local and independent media (akin to <a href=\"https:\/\/papers.ssrn.com\/sol3\/papers.cfm?abstract_id=6035354\" rel=\"nofollow noopener\" target=\"_blank\">must carry requirements<\/a> that are already present in many countries). Australia and Canada have demonstrated that bargaining code frameworks are legislatively viable. Music publishers and the industry more broadly has demonstrated that collective licensing works at scale. The architectural blueprints exist, and policymakers can work to facilitate a fairer business climate. What\u2019s missing is the political will to apply it before the market structures calcify and the journalism industry further withers away.<\/p>\n<p><a href=\"https:\/\/www.brookings.edu\/articles\/the-case-for-consent-in-the-ai-data-gold-rush\/\" rel=\"nofollow noopener\" target=\"_blank\">Explicit consent<\/a> from rights-holders must be a precondition for AI training data collection in the absence of a statutory framework. The Lords Committee has now given that position formal parliamentary backing with its (non-binding) recommendation that the government rule out any exceptions for AI model training on copyrighted works and focus instead on strengthening licensing, transparency, and enforcement. The European Parliament study goes further, recommending statutory licensing as the primary framework to ensure broad access to works with regulator-determined royalties balancing the interests of rightsholders, AI developers, and users, while maintaining incentives for ongoing creative output. An opt-out statutory system where individual publishers that did not want to take part could nonetheless pursue their own deals would be optimal, based on my assessment of the market.<\/p>\n<p>Publishers have continuously missed opportunities to value their products sustainably with search engines. They missed it again with social media. The difference this time is that the legal claims are more mature, the extraction is more visible and blatant, and the sector has had two decades to observe how prior platform relationships developed. Whether that accumulated experience translates into collective action and effective policy before market structures settle is unclear, but if it doesn\u2019t, it will not be because of a lack of knowledge about what is happening. AI firms will lobby for voluntary fixes, which is how Big Tech has always promised to fix the problems it\u2019s created. \u00a0<\/p>\n<p>The deal structures, price precedents, intermediary take rates, and governance norms being established right now will be difficult to dislodge once they\u2019re normalized. The window for intervention is narrowing. The terms, if policymakers don\u2019t set them, will be set by the largest tech firms with the biggest budgets sufficient to withstand the increasing slew of lawsuits and bend Congress, parliaments, and regulators to their will.<\/p>\n<p>The Brookings Institution is committed to quality, independence, and impact.<br \/>We are supported by a <a href=\"https:\/\/www.brookings.edu\/about-us\/annual-report\/\" rel=\"nofollow noopener\" target=\"_blank\">diverse array of funders<\/a>. In line with our <a href=\"https:\/\/www.brookings.edu\/about-us\/research-independence-and-integrity-policies\/\" rel=\"nofollow noopener\" target=\"_blank\">values and policies<\/a>, each Brookings publication represents the sole views of its author(s).<\/p>\n","protected":false},"excerpt":{"rendered":"When Google began indexing news websites at the turn of the century, publishers exchanged free listings of their&hellip;\n","protected":false},"author":2,"featured_media":68029,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[2],"tags":[24,1673,25,2771,2778,2524,1426,2780,38058,7457,2776,2777,5354,6140],"class_list":["post-68028","post","type-post","status-publish","format-standard","has-post-thumbnail","category-ai","tag-ai","tag-article","tag-artificial-intelligence","tag-business-workforce","tag-center-for-technology-innovation-cti","tag-commentary","tag-corporations","tag-governance-studies","tag-media-journalism","tag-regulatory-policy","tag-technology-information","tag-technology-policy-regulation","tag-techtank","tag-u-s-economy"],"_links":{"self":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/68028","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/comments?post=68028"}],"version-history":[{"count":0,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/68028\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media\/68029"}],"wp:attachment":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media?parent=68028"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/categories?post=68028"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/tags?post=68028"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}