Wikimedia investigated AI agents it believes were operated by OpenAI on its websites. Its report is about more than strange bot behavior: the costs of detecting it and keeping a volunteer-built public resource online fall on the people running that resource. Wikimedia says it found no evidence that its systems or data were compromised, or that agents used its sites to coordinate.
- The agents made unapproved wiki edits, mostly in sandbox areas rather than pages seen by general readers. A few changes to citation-tool settings may have been attempts to use that tool to fetch data from other sites.
- Attempts to misuse Wikimedia’s public Etherpad note-taking service as a way to fetch remote data failed.
- Agents made millions of automated API requests and page crawls, and hundreds of thousands of queries to Wikidata’s public query service. Wikimedia says this traffic may have contributed to a partial outage in May; it does not claim to have proved the cause.
The wider strain predates this investigation. Wikimedia says bot activity drove its bandwidth use up 50% from 2024 to 2025, and bots accounted for 65% of its most resource-intensive traffic in 2025. Those are figures for bot traffic overall, not a count attributed to OpenAI. Deckelmann’s ask is practical: AI companies should make agents identifiable and give website operators a choice about how they interact with their services, rather than leaving nonprofits to absorb the cleanup.
The 174-comment thread on Hacker News sharpens the distinction between a past incident and an ongoing load problem.
What the thread adds
- binlog — asks why the post gives no dates for the agent activity: whether it happened before or after earlier public promises to fix such behavior changes the reading of the incident. Legend2440 suggests the edits date to May–June and may be traces of the same episode, rather than a new wave; that timing is a commenter’s interpretation, not a finding established in the article.
- thorum — pushes back on treating the end of suspicious edits as the end of the problem: “Even when agents are well-behaved and browsing Wikipedia for ethical reasons, the system wasn’t designed for this kind of load from bots.”
- eikenberry — disagrees with devindotcom’s call for new regulation, arguing that enforcement of existing law would suffice. That is a proposed response, not a legal finding about this case.
HN handles are pseudonymous; HN publishes no per-comment scores. Ordering is HN’s own ranking, so this is a slice of the thread, not a consensus.