Feedfetcher is how FeedArmy grabs data for importing product data. Feedfetcher collects and periodically refreshes these user-initiated feeds. Find answers below to some of the most commonly asked questions about how this user-controlled feed grabber works.
When users add a service or app that uses Feedfetcher data, FeedArmy’s Feedfetcher attempts to obtain the content of the feed in order to display it. Since Feedfetcher requests come from explicit action by human users, and not from automated crawlers, Feedfetcher does not follow robots.txt guidelines.
If your feed is publicly available, FeedArmy can’t restrict users from accessing it. One solution is to configure your site to serve a 404, 410, or other error status message to user-agent Feedfetcher-FeedArmy.
If your feed is provided by a hosting service, please work directly with that service to restrict access to your feed.
Feedfetcher shouldn’t retrieve feeds from most sites more than once every day on average. Some frequently updated sites may be refreshed more often. Note, however, that due to network delays, it’s possible that Feedfetcher may briefly appear to retrieve your feeds more frequently.
Feedfetcher retrieves feeds at the request of services or apps installed by users. It is possible that a user has requested a feed URL location that does not exist.
Feedfetcher retrieves feeds at the request of services or apps installed by users. It is possible that the request came from a user who knows about your “secret” server or typed it in by mistake.
Feedfetcher retrieves feeds only after users have explicitly started a service or app that requests data from the feed. Feedfetcher behaves as a direct agent of the human user, not as a robot, so it ignores robots.txt entries. Feedfetcher does have one special advantage, though: because it’s acting as the agent of multiple users, it conserves bandwidth by making requests for common feeds only once for all users.
You can prevent Feedfetcher from crawling your site by configuring your server to serve a 404, 410, or other error status message to user-agent
Feedfetcher was designed to be distributed on several machines to improve performance and scale as the web grows. To cut down on bandwidth usage, the machines used are often located near the sites that they’re retrieving in the network.
The IP addresses used by Feedfetcher change from time to time. The best way to identify accesses by Feedfetcher is to use its identifiable user-agent: Feedfetcher-FeedArmy.
In general, Feedfetcher should only download one copy of each file from your site during a given feed retrieval. Very occasionally, the machines are stopped and restarted, which may cause it to again retrieve pages that it’s recently visited.
Unlike normal web crawlers, Feedfetcher isn’t following links at all; instead, it follows the requests given to it by users of a service or app that uses Feedfetcher.
If you’re still having trouble, you can contact us below.