WebCopier
Bot User-Agent:webcopier
🤖 Overview
WebCopier is a legitimate offline browsing and website mirroring tool developed by Maximumsoft (formerly known as WebCopier by Soft Experience). Its primary purpose is to download entire websites or specific sections for offline viewing, allowing users to browse content without an active internet connection. The product feeds data into the local file system, enabling archival, research, and personal use cases. According to the official documentation at maximumsoft.com, WebCopier has been available since the early 2000s and is widely used for educational and archival purposes.
🌐 Technical Behavior
WebCopier operates over standard HTTP/HTTPS protocols and performs recursive crawling to retrieve linked pages, images, CSS, JavaScript, and other embedded resources. By default, it respects the robots.txt exclusion standard and limits its crawl depth based on user configuration. The bot does not have a fixed set of IP ranges; it runs on the client machine, so the source IP is the user’s own address. Users can adjust the number of concurrent connections (typically 1–10) and the delay between requests via the application’s settings. The tool supports authentication, cookie handling, and proxy configuration, making it flexible for mirrored sites behind login walls. Its crawl pattern is deterministic, following the link structure of the starting URL, and it does not scrape external domains unless explicitly configured.
📋 robots.txt Compliance
Documentation from Maximumsoft confirms that WebCopier honors robots.txt directives by default, reading the file before each crawl and skipping disallowed paths. Users have the option to disable this behavior in the advanced settings, but the standard operation is compliant. The tool also respects meta robots tags on individual pages when configured to do so, further aligning with webmaster expectations for polite crawling.
🔍 Detection Indicators
The most reliable detection indicator is the User-Agent string, which typically follows the format WebCopier/4.0 or WebCopier/4.0.0.5 depending on the version. Additionally, the requests often include a header Via: WebCopier and may carry a custom From header containing the user’s email address if configured. The bot’s default request rate is moderate, but aggressive settings can generate rapid successive requests to the same domain.
📊 Data Usage
All data collected by WebCopier is stored locally on the user’s machine for offline viewing and archival purposes. It is not used for any centralized AI training, indexing, analytics, or advertising. The tool is purely a client-side utility, and no data is transmitted back to Maximumsoft or any third party. This makes it a low-risk crawler from a data privacy standpoint, though site owners may still wish to rate-limit it if it generates excessive load.
⚙️ Rate Limiting Policy
Rate limiting is recommended for WebCopier because users can configure it to run many concurrent threads and minimal delays, potentially overwhelming a server. A threshold-based blocking approach (e.g., limiting requests per second from a single IP) ensures that legitimate archival activities remain possible while preventing accidental denial-of-service from misconfigured instances.
Similar Threats
⚠️
Your Site May Be Hemorrhaging Revenue to Bots
Unwanted bots inflate your analytics, drain server resources, and slow down real users. Check if your site is affected — completely free.
Check My Site for FreeFree to start · Cancel anytime
ⓘ Data Notice: The information presented above has been compiled from publicly available internet sources. Boteraser aggregates this data solely for informational purposes and does not independently classify, evaluate, or endorse any findings about the bots listed. The accuracy and completeness of this information is the sole responsibility of the original publishers. Boteraser and its operators accept no liability for any decisions made based on this data.