screaming frog
Bot User-Agent:screaming-frog
🤖 Overview
Screaming Frog SEO Spider is a desktop-based web crawler developed by Screaming Frog Ltd, a UK-based company, first released in 2010. Its primary purpose is to perform technical SEO audits, site architecture analysis, and content inventory by crawling websites exactly like a search engine bot would, but under manual user control. The software is widely used by SEO professionals, digital marketers, and web developers to identify broken links, duplicate content, missing meta tags, redirect chains, and other on-page issues. It is not a cloud service but a locally installed tool that can be configured to run on-demand or scheduled tasks via a paid license (Screaming Frog SEO Spider) which offers unlimited URLs, while the free version caps at 500 URLs.
🌐 Technical Behavior
The Screaming Frog SEO Spider crawls by sending HTTP/HTTPS requests sequentially from the user’s own IP address, meaning it does not use any fixed IP ranges belonging to the vendor. The default crawl speed is aggressive — it can send thousands of requests per minute depending on the user's bandwidth and configuration. The tool supports crawling via sitemap, seed URLs, or list imports. It respects HTTP status codes, follows redirects (configurable), and can parse JavaScript-rendered content using an integrated Chromium-based renderer (since version 14.0). The tool does not operate on a distributed bot network; each instance is a single-threaded or multi-threaded crawler controlled by the user. It supports both GET and HEAD requests, the latter being faster for URL validation. The user can also set custom request headers and cookies. Importantly, the tool pauses when the server returns a 429 (Too Many Requests) or 503 status, provided the user enables the “Respect Crawl Delay” option (enabled by default when reading robots.txt).
📋 robots.txt Compliance
By default, Screaming Frog SEO Spider respects robots.txt directives, including Disallow rules and Crawl-Delay directives. The user can disable robots.txt compliance in the configuration menu, but this is not the default behavior. The tool also supports custom User-Agent strings; by default it uses “Mozilla/5.0 (compatible; Screaming Frog SEO Spider/19.0; +https://www.screamingfrog.co.uk/seo-spider/)” (version varies). Screaming Frog explicitly recommends in its documentation that webmasters allow the crawler, but for testing purposes, they advise limiting it via robots.txt if server load is a concern. The tool will not crawl any URL explicitly disallowed in robots.txt if the “Respect Robots.txt” checkbox remains checked.
🔍 Detection Indicators
The primary detection method is the User-Agent string: “Screaming Frog SEO Spider” followed by version number and the vendor URL. A secondary indicator is the rapid, sequential nature of requests — often hitting hundreds of URLs in a short burst from a single IP. The tool does not include any special headers beyond standard browser-like headers (Accept, Accept-Language, etc.). Network administrators can look for patterns of consecutive GET requests with a consistent interval (user-configurable crawl delay). The tool may also send a “Screaming Frog” or “SF” value in the User-Agent field when the user customizes it, but the default is unambiguous. There is no official public IP range list since it runs from the user’s own connection.
📊 Data Usage
The data collected by Screaming Frog SEO Spider is used exclusively by the user for SEO auditing and site analysis — it is not transmitted to Screaming Frog Ltd or any third party. The tool generates reports (CSV, XML, XLSX) containing URLs, status codes, page titles, meta descriptions, response times, size, content type, and link relationships. The user retains full ownership of all collected data. Screaming Frog does not aggregate or train AI models using crawled content; the software is strictly a local analysis tool. It is not a search engine or a data brokerage product. The official website (screamingfrog.co.uk) clearly states that no data leaves the user’s machine without explicit action (e.g., exporting reports).
⚙️ Rate Limiting Policy
Webmasters are advised to rate-limit the Screaming Frog SEO Spider because its default crawl speed can overwhelm small or under-resourced servers. The tool is not designed for continuous crawling — it executes a single audit and then stops. Threshold-based blocking (e.g., returning 429 after 100 requests per minute from the same IP) is recommended to prevent accidental denial of service. The rationale is that the tool’s behavior mimics a modest but rapid single-user crawl, and excessive load can degrade performance for other visitors.
Similar Threats
🛡️
Stop Bots. Save Bandwidth. Protect Revenue.
Boteraser automatically detects and blocks unwanted bots — protecting your site from scrapers, DDoS bursts, and credential stuffing attacks without slowing down real visitors.
✅ Start Free ProtectionSetup takes under a minute · Free trial available
ⓘ Data Notice: The information presented above has been compiled from publicly available internet sources. Boteraser aggregates this data solely for informational purposes and does not independently classify, evaluate, or endorse any findings about the bots listed. The accuracy and completeness of this information is the sole responsibility of the original publishers. Boteraser and its operators accept no liability for any decisions made based on this data.