anemone
Anemone is a research-oriented web crawler operated by the University of Washington’s Web Research Group, originally developed for academic studies on web graph structure, link analysis, and crawling efficiency. It feeds data into public research datasets and supports projects like the WebBase repository.
Anemone employs a politeness policy with a default crawl delay of 1–2 seconds between requests, though it may be configured to lower intervals during bulk academic crawls. It uses HTTP/1.1 with standard GET requests and follows redirects (up to 10 hops). Its IP ranges are typically drawn from University of Washington’s autonomous system (AS73) and may also appear from other academic institutions when the crawler is distributed. It supports both breadth-first and focused crawling strategies, and it announces itself via the User-Agent header.
Anemone consistently honors robots.txt directives, including Disallow and Crawl-delay instructions, as documented in the crawler’s source code available on the GitHub repository (https://github.com/uw-web-research/anemone). It does not override explicit exclusions and will abort crawl paths that are disallowed.
The primary User-Agent string is Anemone/1.0 (+http://webresearch.cs.washington.edu/anemone). Additional fingerprints include a consistent lack of Accept-Encoding negotiation and a small window of concurrent connections (typically 2–4). No custom X-Forwarded-For or referrer headers are injected.
Collected data is used exclusively for academic research — including web graph analysis, link prediction studies, and language modeling datasets that are made publicly available under permissive licenses. No commercial indexing or AI model training occurs without explicit consent.
Although Anemone is legitimate, its ability to perform large-scale crawls may overwhelm server resources; therefore it is rate-limited to protect application performance. Threshold-based blocking (e.g., >20 requests/minute from the same IP) is recommended without permanent ban.
Similar Threats
— Imperva Bad Bot Report 2026
How much of your traffic is automated? Get your personal bot traffic report and see exactly what's hitting your server — completely free.
📊 Get My Bot ReportSign up in seconds · No card required
ⓘ Data Notice: The information presented above has been compiled from publicly available internet sources. Boteraser aggregates this data solely for informational purposes and does not independently classify, evaluate, or endorse any findings about the bots listed. The accuracy and completeness of this information is the sole responsibility of the original publishers. Boteraser and its operators accept no liability for any decisions made based on this data.
Stay up to date with the latest from Boteraser.
We use cookies to improve your experience on our site. By using our site, you consent to cookies.
Manage your cookie preferences below:
Essential cookies enable basic functions and are necessary for the proper function of the website.
CloudFlare provides web performance and security solutions, enhancing site speed and protecting against threats.
Service URL: developers.cloudflare.com (opens in a new window)
These cookies are needed for adding comments on this website.
These cookies are used for managing login functionality on this website.
Statistics cookies collect information anonymously. This information helps us understand how visitors use our website.
Google Analytics is a powerful tool that tracks and analyzes website traffic for informed marketing decisions.
Service URL: policies.google.com (opens in a new window)
You can find more information in our Cookie Policy and Privacy Policy.