Anna’s Archive Claims Massive Spotify Data Scraping
The shadow library and pirate activist collective Anna’s Archive has announced the acquisition of a vast portion of Spotify’s music library. The group claims to have scraped approximately 86 million music files, a collection they assert represents nearly 99.6% of all content consumed on the streaming platform.

While Spotify’s total catalog is estimated at around 256 million tracks, the data harvested by the group includes metadata for roughly 99.9% of those entries. The total volume of the gathered data reaches nearly 300 terabytes. Currently, the organization has only made the metadata public, refraining from releasing the actual audio files.
Preservation or Piracy?
In a blog post detailing the operation, Anna’s Archive framed the move as an effort to establish a “preservation archive” for music. The group stated that while they acknowledge Spotify does not host every piece of music in existence, they consider the platform a comprehensive starting point for their broader goal of archiving human knowledge and culture, regardless of the media format.
Spotify’s Response
The streaming giant has moved to mitigate the impact of the incident. According to a company spokesperson, Spotify identified the accounts responsible for the scraping and has since disabled them. The company emphasized its commitment to protecting intellectual property:
- Implementation of new security safeguards to prevent future copyright-related attacks.
- Continuous monitoring of the platform for suspicious user behavior.
- Active collaboration with industry partners to defend the rights of creators.
Spotify maintained that its stance against piracy has been a core principle since its inception, and it continues to work alongside the artist community to combat unauthorized data extraction.