The crawler behind Spagify's price comparison, and how to keep it off your site.
Spagify runs Google Shopping campaigns for Shopify stores. When one of those stores advertises a product that other retailers also sell, Spagify reads the public price of that product on the other retailers' storefronts so the merchant can see how their own price compares. SpagifyPriceBot is the crawler that does that reading. It runs from Spagify's servers, on behalf of one Spagify customer at a time, and identifies itself with this user agent:
SpagifyPriceBot/1.0 (+https://spagify.com/bot; price comparison for a retailer that sells the same products)
It never signs in, never submits a form, and never reads anything that is not already public.
robots.txt first, every visit./products.json.Product and Offer markup, with Open Graph price tags as the fallback.From a product page it keeps the product name, brand, GTIN, MPN, SKU, price, currency, availability, condition and URL. Nothing else on the page is stored, and no page is stored.
robots.txt Disallow and Allow rules: the SpagifyPriceBot group if there is one, otherwise *.Add this to your robots.txt. The crawler reads the file on every visit, so the rule takes effect within a day.
User-agent: SpagifyPriceBot Disallow: /
To block only part of a site, use a Disallow path instead of /. If you would rather not changerobots.txt, email legal@spagify.com with your domain and it is added to the crawler's own exclusion list.
Questions about the crawler or about Spagify: legal@spagify.com. How Spagify handles data in general is in the Privacy Policy.