Crawl and log pipeline
The crawl comes first: up to 50,000 URLs in the audit, rendered where your templates need JavaScript, with the status code, canonical, directives and internal links recorded for each. On Retainer and Embedded we add your raw access logs, usually the last thirty to ninety days, and separate verified Googlebot from the fake crawlers that borrow its name. Verification is by reverse DNS lookup, not by user-agent string, because on most sites a share of the “Googlebot” traffic is not Google at all.


