searchmee
searchmee! is a web crawler operated by Searchmee Ltd., a UK-based company that provides a privacy-focused search engine at searchmee.com. Its primary purpose is to index publicly accessible web pages to populate the Searchmee search index, which emphasizes user privacy by not tracking searches or storing personal data. The crawler was first introduced publicly in 2019 and has since been documented on the company’s official website and in forum posts about webmaster best practices.
The searchmee! bot follows a broad crawl pattern, starting from seed URLs submitted by webmasters or discovered through sitemaps. It sends HTTP GET requests at a configurable rate, typically between 5 and 20 requests per second per IP, but can burst higher during initial discovery. The bot primarily uses HTTP/1.1 and HTTPS, supports gzip compression, and identifies itself via a distinct User-Agent string. IP addresses used by searchmee! are allocated from ranges owned by Searchmee Ltd. and its cloud hosting providers (e.g., AWS and DigitalOcean), though the company does not publish a comprehensive IP list. The crawler respects Cache-Control and Last-Modified headers to avoid re‑downloading unchanged content, and it uses E‑Tag validation when available. It also follows nofollow and noindex meta tags as part of its crawling discipline.
Searchmee! strictly honors robots.txt directives, as stated in its official documentation on the Searchmee webmaster portal. The bot checks the robots.txt file before each crawl and caches it for up to 24 hours. It also respects Allow and Disallow rules, including wildcard patterns. Webmasters have reported in community forums that searchmee! reliably stops crawling paths marked with “Disallow,” and there are no known incidents of the bot ignoring these instructions.
The primary User‑Agent string is Mozilla/5.0 (compatible; searchmee!/1.0; +https://searchmee.com/). Some variations include a version number (e.g., “1.1”) and the bot may also add a From header with a contact email address found in its documentation. Another identifying header is X-Robots-Tag: noindex which the bot includes in its requests to signal its purpose. Behavioral fingerprints include a high frequency of HEAD requests before GET, and a preference for low‑latency responses.
Collected content is used exclusively to build the Searchmee search index, which powers the search engine’s results. The company explicitly states that it does not sell user data or use crawled content for AI training without separate permission. The data is stored temporarily for indexing and then aggregated into a public search database that respects the noindex and nofollow signals from publishers.
Rate limiting for searchmee! is recommended because its bursty crawl pattern can temporarily overwhelm smaller websites, even though it is a legitimate bot. A threshold of 10 requests per second per IP, with a 60‑second ban window for exceeding that rate, is a common webmaster practice to protect server resources while still allowing the bot to complete its indexing.
⚠️
Unwanted bots inflate your analytics, drain server resources, and slow down real users. Check if your site is affected — completely free.
Check My Site for FreeFree to start · Cancel anytime
ⓘ Data Notice: The information presented above has been compiled from publicly available internet sources. Boteraser aggregates this data solely for informational purposes and does not independently classify, evaluate, or endorse any findings about the bots listed. The accuracy and completeness of this information is the sole responsibility of the original publishers. Boteraser and its operators accept no liability for any decisions made based on this data.
Stay up to date with the latest from Boteraser.
We use cookies to improve your experience on our site. By using our site, you consent to cookies.
Manage your cookie preferences below:
Essential cookies enable basic functions and are necessary for the proper function of the website.
CloudFlare provides web performance and security solutions, enhancing site speed and protecting against threats.
Service URL: developers.cloudflare.com (opens in a new window)
These cookies are needed for adding comments on this website.
These cookies are used for managing login functionality on this website.
Statistics cookies collect information anonymously. This information helps us understand how visitors use our website.
Google Analytics is a powerful tool that tracks and analyzes website traffic for informed marketing decisions.
Service URL: policies.google.com (opens in a new window)
You can find more information in our Cookie Policy and Privacy Policy.