Page MenuHomePhabricator

[Scraper] Disable search engine exposure of any metadata-wikibase pages/links
Closed, ResolvedPublic5 Estimated Story Points

Description

Not sure if this is a bug or just a functionality that needs an update.

Current situation:

  • When using Google search and entering "Clayarena Wikibase" the Google search gives a result that links to a CSV download
  • We can not allow those links to appear in Google search

Desired situation:

  • All links and pages/links that are generated by the metadata-wikibase need a 'no index' function to avoid unexpected exposure in search engines

Example:
https://www.google.com/url?sa=t&source=web&rct=j&opi=89978449&url=https://wikibase-metadata.wmcloud.org/csv/metrics%3Fauthorization%3D0c528a8e-026f-4fd2-ae8b-bd832083256e&ved=2ahUKEwjxx7uf642UAxXVRfEDHXe8C-cQFnoECBoQAQ&usg=AOvVaw1-Lf3_Co8ZvDrrOwxOznET

Bildschirmfoto vom 2026-04-27 12-46-17.png (1,124×528 px, 135 KB)

AC:

  • update authorization guid; distribute to need-to-know parties" to the implementation
  • update URL first
  • add robot.txt file

Event Timeline

Leif_WMDE renamed this task from [Discovery] Disable search egine tracking of metadata-wikibase to [Scraper] Disable search engine tracking of metadata-wikibase.Apr 28 2026, 9:04 AM
Leif_WMDE renamed this task from [Scraper] Disable search engine tracking of metadata-wikibase to [Scraper] Disable search engine exposure of any metadata-wikibase pages/links.Apr 28 2026, 10:00 AM
Leif_WMDE updated the task description. (Show Details)
Leif_WMDE triaged this task as Medium priority.

closing this but need to wait until new URL is connected