Skip to content

Comment on Wikipedia is struggling with voracious AI bot crawlersparent

Comments

I think those crawlers are just very generic: they basically operate like wget scripts, without much logic for avoiding sites that already offer clean data dumps.

That is not an excuse. Wikipedia isn't just any site.

Not an excuse, a plausible explanation of what's actually happening.

Also plausibly they are trying to kill the site via soft ddos. Then they can sell a service based on all the data they scraped + unauditable censoring.

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.