Creative Commons Backs Pay-to-Crawl Systems for AI Training

scraping to offset declining search traffic.

TAGS: Creative Commons, Artificial Intelligence, Web Scraping, Digital Publishing, AI Licensing

CONTEUDO:
Creative Commons announces tentative support for AI ‘pay-to-crawl’ systems

The nonprofit Creative Commons (CC) has officially signaled its cautious support for “pay-to-crawl” technology. This emerging framework aims to automate compensation for websites whenever their content is accessed by AI web crawlers for training purposes.

Known for its role in copyright and open licensing, the organization is pivoting to address the economic fallout publishers face as AI-generated answers replace traditional search engine traffic. By implementing these systems, the nonprofit suggests publishers could sustain content creation without resorting to more restrictive, closed-off paywalls.

The Shift in Web Economics

Historically, websites welcomed crawlers from search giants like Google, as indexing led to increased traffic and clicks. The rise of AI chatbots has disrupted this model, as users often receive answers directly from the AI, bypassing the original source entirely. This shift has placed significant financial pressure on publishers, leading many to seek alternative revenue streams through content licensing.

While major media entities like The New York Times, Gannett, and Axel Springer have secured individual deals with AI providers, smaller publishers lack the leverage to negotiate similar terms. Pay-to-crawl systems offer a potential path to monetization for these smaller players.

Risks and Proposed Principles

Despite its support, Creative Commons outlined several concerns regarding the implementation of these systems, noting that they could centralize power on the web or hinder access for academic and public-interest research. To mitigate these risks, the organization advocates for specific standards:

  • Pay-to-crawl should not be a default setting for all websites.
  • Systems must allow for “throttling” access rather than strictly blocking it.
  • Access must be preserved for nonprofits, educators, and cultural institutions.
  • Technology should remain open, interoperable, and built on standardized components.

The Growing AI Monetization Landscape

The push for standardized AI access is gaining momentum across the industry. Companies like Microsoft are developing publisher marketplaces, while startups like TollBit and ProRata.ai are entering the space. Furthermore, the Really Simple Licensing (RSL) standard—backed by entities like Cloudflare, Fastly, and Akamai—has emerged to help publishers manage crawler access.

Creative Commons, which recently introduced its own “CC Signals” project, has also announced its support for the RSL specification as part of its broader initiative to establish a social contract for the age of artificial intelligence.

Deixe um comentário

O seu endereço de e-mail não será publicado. Campos obrigatórios são marcados com *