logoalt Hacker News

codemonkey-zetayesterday at 6:05 PM1 replyview on HN

Indeed, the article mentions Wikipedia experiencing similar scraping pains, even though they already DO have bulk data available.


Replies

HeatrayEnjoyeryesterday at 6:53 PM

Who are running these bots? I presume developers at all of the frontier labs know (or at least would know to look for) Wikipedia has bulk APIs for automated access. Unnecessary scraping increases their workload/costs too, so why in 2026 is this still a problem?

show 1 reply