dataforseobot
Bot User-Agent:dataforseobot
🤖 Overview
DataForSeoBot is a web crawler operated by DataForSEO, a company headquartered in Cyprus that provides SEO data APIs (search engine results, backlinks, rankings) to digital marketing professionals and enterprises. The bot is documented in DataForSEO’s official documentation as a legitimate, public crawler used to collect publicly accessible web content for their On‑Page API and Search Engine Results Page (SERP) API services. According to DataForSEO’s website, the crawler indexes pages to supply real‑time SEO metrics and competitive analysis data to subscribers.
🌐 Technical Behavior
The bot crawls using HTTP/1.1 and HTTP/2 protocols and typically sends requests from IP addresses belonging to DataForSEO’s cloud infrastructure, which includes ranges managed by AWS and other providers. Official documentation states that DataForSeoBot respects a crawl delay of 5 seconds between consecutive requests to the same domain, but this delay is configurable by webmasters via the Crawl‑Delay directive in robots.txt. The crawler identifies itself with the User‑Agent: DataForSeoBot string and does not rotate user agents. It fetches both HTML and Ajax‑rendered content, mimicking a standard desktop browser (Chrome on Windows) to obtain JavaScript‑rendered data. Published IP ranges include subnets such as 94.130.35.0/24 and 89.234.186.0/24, though these are subject to change. DataForSEO advises webmasters to whitelist these ranges if they wish to allow the crawler.
📋 robots.txt Compliance
DataForSeoBot explicitly honors the Disallow and Crawl‑Delay directives defined in a site’s robots.txt file, as confirmed by the company’s official crawler policy page. In addition, it supports the User‑Agent: DataForSeoBot directive and will cease crawling any path listed under Disallow. Webmasters can block specific directories by adding the appropriate rule. The bot also respects the noindex meta tag if encountered in page headers, but this is not its primary compliance mechanism.
🔍 Detection Indicators
The definitive identification string is Mozilla/5.0 (compatible; DataForSeoBot/1.0; +https://dataforseo.com/dataforseo-bot). The bot sends a Connection: keep‑alive header and a Accept‑Language: en‑US,en;q=0.9 header. Unlike many other SEO crawlers, it does not appear to spoof its User‑Agent; it consistently presents the same string. With X‑Forwarded‑For headers absent, webmasters can rely on the User‑Agent and source IP ranges to identify requests. Behavioral fingerprints include a regular interval of roughly 5 seconds between page requests (unless a lower Crawl‑Delay is set) and a predictable request pattern with no random jitter.
📊 Data Usage
The data collected by DataForSeoBot is fed directly into DataForSEO’s commercial APIs, including the On‑Page API which returns technical SEO audits (meta tags, headings, images, etc.) and the SERP API which provides real‑time search engine rankings for target keywords. The content is aggregated, anonymized, and stored for up to 30 days to power customers’ competitive analysis, rank tracking, and content optimization workflows. No data is used for third‑party AI model training or resold as raw content.
⚙️ Rate Limiting Policy
Despite its compliance, DataForSeoBot may still generate a noticeable crawl rate on high‑traffic or resource‑constrained servers, especially on smaller sites. A recommended rate‑limiting threshold of 10 requests per minute per IP is advised by security practitioners to prevent excessive load while still allowing legitimate data collection. The policy rationale is that the bot’s bursts can temporarily consume bandwidth and CPU, potentially degrading user experience if not throttled.
Similar Threats
🛡️
Stop Bots. Save Bandwidth. Protect Revenue.
Boteraser automatically detects and blocks unwanted bots — protecting your site from scrapers, DDoS bursts, and credential stuffing attacks without slowing down real visitors.
✅ Start Free ProtectionSetup takes under a minute · Free trial available
ⓘ Data Notice: The information presented above has been compiled from publicly available internet sources. Boteraser aggregates this data solely for informational purposes and does not independently classify, evaluate, or endorse any findings about the bots listed. The accuracy and completeness of this information is the sole responsibility of the original publishers. Boteraser and its operators accept no liability for any decisions made based on this data.