ieautodiscovery
ieautodiscovery is a legitimate web crawler component of Internet Explorer (versions 7 through 11), developed by Microsoft Corporation, tasked with automatically detecting and subscribing to RSS feeds and Web Slices when a user visits a website. Its official documentation appears in Microsoft’s Internet Explorer RSS Platform documentation (archived at MSDN Library). Unlike general‑purpose search crawlers, ieautodiscovery operates only on behalf of an individual user who has enabled feed subscription in IE’s Feeds pane, making it a client‑side agent rather than a server‑side indexing robot.
When a user navigates to a page in Internet Explorer, the browser silently triggers ieautodiscovery to scan the HTML tags. It issues a single HTTP GET request per feed URL discovered, typically within seconds of page load. The crawler sends a User‑Agent string of Mozilla/4.0 (compatible; MSIE 7.0; Windows NT 5.1; Trident/4.0; .NET CLR 1.1.4322; InfoPath.2; .NET CLR 2.0.50727; .NET CLR 3.0.4506.2152; .NET CLR 3.5.30729; MS-RTC LM 8) — the exact token MSIE 7.0 may vary with IE version. Request frequency is low: one request per distinct feed per page, never repeated unless the user revisits the page or manually refreshes feeds. IP addresses originate from the user’s own client IP, not from Microsoft‑owned ranges. It supports only HTTP/1.1 and respects 304 Not Modified responses via If‑Modified‑Since headers.
Internet Explorer’s ieautodiscovery does not read robots.txt during its automatic feed discovery operation because it is a client‑side script running inside the user’s browser under the user’s own session. However, the subsequent feed subscription requests (made by the browser when the user adds the feed) do respect robots.txt if the feed is fetched manually; the automatic discovery phase itself bypasses it. This is documented in Microsoft’s Internet Explorer RSS Platform SDK (archived at MSDN).
The primary identifier is a User‑Agent containing the substring “MSIE” followed by a version number (e.g., 7.0, 8.0, 9.0) combined with the Trident/ token. A behavioral fingerprint is the rapid retrieval of <link> tags containing type="application/rss+xml" or type="application/atom+xml" immediately after the main page loads. The request lacks any From or X‑Bot headers and carries a Accept: */* header.
Collected feed URLs and their content are used exclusively for local feed aggregation within the user’s Internet Explorer Feeds pane. No data is transmitted to Microsoft servers; all feed polling is performed directly from the client IP to the originating server. The data enables the user to subscribe to and read news, blog entries, or podcast updates without visiting the website.
This agent is inherently rate‑limited per user session — it only makes requests when the user actively browses a page containing a feed link. No rate limiting is required from the server perspective, but if misidentified as a threat, threshold‑based blocking should allow at least 5–10 requests per minute per IP to accommodate legitimate IE users without disrupting automatic feed discovery.
Similar Threats
Free Bot Analysis
Find out exactly how much of your traffic is automated — and which bots are draining your bandwidth and skewing your analytics.
Run Free Bot Scan →No credit card required · Results in minutes
ⓘ Data Notice: The information presented above has been compiled from publicly available internet sources. Boteraser aggregates this data solely for informational purposes and does not independently classify, evaluate, or endorse any findings about the bots listed. The accuracy and completeness of this information is the sole responsibility of the original publishers. Boteraser and its operators accept no liability for any decisions made based on this data.
Stay up to date with the latest from Boteraser.
We use cookies to improve your experience on our site. By using our site, you consent to cookies.
Manage your cookie preferences below:
Essential cookies enable basic functions and are necessary for the proper function of the website.
CloudFlare provides web performance and security solutions, enhancing site speed and protecting against threats.
Service URL: developers.cloudflare.com (opens in a new window)
These cookies are needed for adding comments on this website.
These cookies are used for managing login functionality on this website.
Statistics cookies collect information anonymously. This information helps us understand how visitors use our website.
Google Analytics is a powerful tool that tracks and analyzes website traffic for informed marketing decisions.
Service URL: policies.google.com (opens in a new window)
You can find more information in our Cookie Policy and Privacy Policy.