OTHER

Strava Stands Firm Against Scrapers Ahead of Upcoming IPO

AI companies have transformed into data-centric entities, necessitating ever-growing datasets for training. In response to this craving, many AI startups are ignoring foundational internet protocols—like adhering to robots.txt files that specify which areas of a website are off-limits to automated crawlers—resorting instead to aggressive data scraping tactics. Consequently, many websites are restricting data access and, in some cases, negotiating licensing agreements with AI companies. Strava, a fitness and social running platform, is moving in this direction by limiting access to its site and introducing fees for developer access.

To address scraping issues, Strava is bolstering its platform security and will limit the visibility of certain data to authenticated users exclusively. Previously, public profiles and fitness club listings were accessible without a login. Now, users will need to authenticate to view this data, thereby preventing unauthorized scraping by AI.

Regarding the API, developers used to create applications on Strava through a complimentary, tiered access model—initially applying for basic access and later requesting more as their application scaled. Now, the company is implementing a uniform fee of $11.99 per month for all developers, although this rate may vary by region.

Strava has reported that its developer community has expanded from 185,000 last year to 241,000 this year, and the company is dedicated to supporting them. As part of this commitment, Strava plans to endorse the Model Context Protocol (MCP), an emerging standard allowing AI assistants and applications to systematically access external data, giving Strava better control over what information is shared and how.

The company also plans to deactivate certain API endpoints—specific access points that let external applications retrieve designated data, such as information about clubs—to protect user data. Previously, Strava tightened API regulations in 2024, restricting its use for AI training and placing limits on third-party applications displaying others’ data. These changes faced backlash from developers, who expressed concerns that their applications would be adversely impacted.

While some developers may accept the subscription model, the deprecation of certain API endpoints could still negatively affect applications that rely on them. Strava is offering a 90-day grace period to developers before implementing these changes.

In a discussion with TechCrunch, Strava’s CEO, Michael Martin, voiced concerns that rampant AI scraping could threaten the future of the public internet.

“AI companies are aggressively scraping public websites due to their insatiable need for training data, universally damaging site performance,” Martin expressed. “We have recently encountered multiple instances where performance has deteriorated significantly. Aside from scraping public sites, they are also trying to access our data through our API, neglecting our API terms.”

He revealed that Strava has turned down requests from major AI labs for data licensing agreements. Martin specifically highlighted Perplexity, noting that the AI search startup disguised its scraping through aggregator services to obscure its origin, even after being declined. This incident resonates with Perplexity’s prior behavior in similar contexts.

Martin also pointed out server overload issues caused by inefficient API requests from collaboration applications, which frequently impose excess strain on Strava’s infrastructure. This pattern echoes Meta’s rationale when it prohibited third-party chatbots from WhatsApp last year due to system overload concerns.

The timing of these initiatives is likely deliberate. Strava filed confidentially for an IPO earlier this year, and its data protection measures may aim to showcase data responsibility to potential investors. Martin quickly acknowledged comparisons to Reddit’s API access limitations from 2024, clarifying that, unlike Reddit—which based API access charges on call volume (making it unaffordable for numerous developers)—Strava’s flat fee structure seeks to foster a thriving developer ecosystem.

“We want users to feel they possess their data and have confidence in how we manage and protect it. At the same time, we strive for developers to continue thriving and growing,” Martin concluded.

When you purchase through links in our articles, we may earn a small commission. This doesn’t affect our editorial independence.