Verify Amazonbot IP Address

Verify if an IP address truly belongs to Amazon, using official verification methods. Enter both IP address and User-Agent from your logs for the most accurate bot verification.

[Amazon Bots can take upto 30 days to read your Robots.txt updates.] Amazonbot is Amazon’s official web crawler, used to discover and fetch webpage content for applications such as Alexa, product-related features, and Amazon’s AI and search systems. Crawl activity varies based on Amazon services that rely on external web content, but it is generally moderate and focused on structured data, text content, and page metadata. Its purpose is to enhance Amazon’s search, AI models, and user-facing features. It ignores the global user agent (*) rule. RobotSense.io verifies Amazonbot using Amazon’s official validation methods, ensuring only genuine Amazonbot traffic is identified.

This bot does not honor Crawl-Delay rule.

User Agent Examples

Mozilla/5.0 AppleWebKit/537.36 (KHTML, like Gecko; compatible; Amazonbot/0.1; +https://developer.amazon.com/support/amazonbot) Chrome/119.0.6045.214 Safari/537.36
Example user agent strings for Amazonbot

Robots.txt Configuration for Amazonbot

Robots.txt User-Agent:Amazonbot

Use this identifier in your robots.txt User-agent directive to target Amazonbot.

Recommended Configuration

Our recommended robots.txt configuration for Amazonbot:

User-agent: Amazonbot
Allow: /

Completely Block Amazonbot

Prevent this bot from crawling your entire site:

User-agent: Amazonbot
Disallow: /

Completely Allow Amazonbot

Allow this bot to crawl your entire site:

User-agent: Amazonbot
Allow: /

Block Specific Paths

Block this bot from specific directories or pages:

User-agent: Amazonbot
Disallow: /private/
Disallow: /admin/
Disallow: /api/

Allow Only Specific Paths

Block everything but allow specific directories:

User-agent: Amazonbot
Disallow: /
Allow: /public/
Allow: /blog/

Set Crawl Delay

Limit how frequently Amazonbot can request pages (in seconds):

User-agent: Amazonbot
Allow: /
Crawl-delay: 10

Note: This bot does not officially mention about honoring Crawl-Delay rule.

Frequently Asked Questions

What is Amazonbot, and why is it visiting my website?
Amazonbot is a web crawler operated by Amazon that collects publicly accessible web content. Its visits are typically related to services such as search indexing, content discovery for Amazon services, and datasets used across Amazon’s technology ecosystem, including AI-related research. The crawler automatically discovers pages through links and other standard web discovery methods, so it may appear in server logs when it encounters publicly accessible pages. For most public websites, occasional Amazonbot traffic is normal.
Is Amazonbot a legitimate bot, or is it commonly spoofed?
Amazonbot is an officially operated crawler run by Amazon. However, like most well-known bots, its user-agent string can be spoofed by malicious actors attempting to disguise automated traffic. Attackers may imitate the Amazonbot user-agent to bypass basic bot filters or appear as legitimate crawler traffic in logs. Because of this, the User-Agent string alone is not sufficient to verify that a request actually originates from Amazonbot. You can use Amazon's recommended methods mentioned below to verify a legitimate visit, or use RobotSense.io API to easily verify Amazonbot visits.
How can I verify that a request is really coming from Amazonbot?
You can use Amazon's recommended official methods to verify Amazonbot visits, these include: - IP range checks Do not use User-Agent based detection as that can be easily spoofed. Alternatively, you can use RobotSense.io API to easily verify Amazonbot and other bots from Amazon.
Should I allow or block Amazonbot on my website?
Allowing Amazonbot is generally optional and depends on whether you want Amazon services to access your publicly available content. Allowing it may help Amazon-powered systems discover and analyze publicly available web content. If you are suddenly seeing too many visits, you can consider adding a small crawl-delay in your robots.txt before completely disallowing. Blocking it may make sense if: - your server experiences excessive automated traffic - pages contain sensitive or restricted information - the site hosts internal tools, APIs, or staging environments For most public informational websites, Amazonbot traffic is typically low-impact and not harmful.
How can I control or block Amazonbot using robots.txt or other methods?
You can add a rule in your robots.txt, as given above to control (crawl-delay) or disallow Amazonbot. Amazonbot honors robots.txt directives, but it may take up to 30 days for your recent robots.txt changes to reflect properly. Also, you can use further controls in your WAF, or in RobotSense enforcement settings to manage the bot behavior.
How often does Amazonbot crawl websites, and can it impact server performance?
Amazonbot typically performs automated crawling that varies depending on site visibility, link discovery, and crawl scheduling. For most websites, request rates are modest and distributed over time rather than aggressive bursts. On large or highly linked sites, crawl frequency may increase as the bot discovers more URLs. In most cases the performance impact is minimal, though smaller servers or dynamically generated pages may notice additional request load during active crawl periods. Some administrators choose to rate-limit or restrict it.
What happens if I block Amazonbot? SEO, visibility, and feature impact explained.
Blocking Amazonbot does not affect traditional search engine rankings, since it is not the primary crawler for a public search engine. However, blocking it may limit how your content appears within Amazon-related services. Possible effects include: - Reduced visibility in Amazon-powered discovery or knowledge systems - Limited inclusion in Amazon data analysis or indexing datasets - Reduced availability of your content for Amazon-related previews or integrations For many sites, blocking Amazonbot has no direct SEO impact.
Does Amazonbot collect, scrape, or use my content for training or reuse?
Amazonbot collected data may be used for purposes such as indexing, metadata extraction, and building datasets used across Amazon services, including machine learning research and AI systems.

Other Amazon Bots

Amazon operates other crawlers you may also need to configure.

amazon-QBusiness

AI Service

[Amazon Bots can take upto 30 days to read your Robots.txt updates.] amazon-QBusiness is a crawler associated with Amazon’s Q Business and enterprise AI services, used to retrieve webpage content that organizations reference within Q-based workflows. It performs targeted, purpose-specific fetches triggered by users or automated enterprise integrations. The bot gathers text, metadata, and structural information to support search, summarization, and knowledge enrichment within Amazon’s AI tools. It is not a broad web crawler and does not index content for public Amazon services. It's crawl volume is typically low, reflecting only the URLs explicitly accessed through Q Business environments. RobotSense.io verifies amazon-QBusiness using Amazon’s official validation methods, ensuring only genuine amazon-QBusiness traffic is identified.

Amzn-SearchBot

Search

[Amazon Bots can take upto 30 days to read your Robots.txt updates.] Amzn-SearchBot is Amazon’s web crawler used to discover and retrieve publicly available content for Amazon search and AI-related services. It fetches webpages to analyze text, metadata, and structured information that can support Amazon’s search features and machine learning systems. Crawl activity is typically moderate and focused on publicly accessible pages. Its purpose is to help Amazon improve content discovery, relevance, and information retrieval across its ecosystem. It ignores the global user agent (*) rule. RobotSense.io verifies Amzn-SearchBot using Amazon’s official validation methods, ensuring only genuine Amzn-SearchBot traffic is identified.

Amzn-User

Others

[Amazon Bots can take upto 30 days to read your Robots.txt updates.] Amzn-User is a bot associated with Amazon services that fetch webpage content on behalf of end users or Amazon applications rather than acting as a general-purpose crawler. It typically appears when Amazon apps, devices, or internal systems request metadata, previews, or content needed for features like link expansion, in-app browsing, or contextual analysis. The traffic is user-driven, not designed for large-scale indexing or scraping. Amzn-User usually performs lightweight, targeted fetches limited to specific URLs users interact with. Its purpose is to support Amazon product experiences by retrieving just enough page data to power user-facing functionality. It ignores the global user agent (*) rule. RobotSense.io verifies Amzn-User using Amazon’s official validation methods, ensuring only genuine Amzn-User traffic is identified.

Similar Bots

Other Others bots from different operators.

BingPreview

Others

by Microsoft

BingPreview is Microsoft’s rendering and compatibility crawler used to evaluate how webpages appear in browsers and Bing search features. It fetches pages to test layout, mobile responsiveness, JavaScript rendering, and visual elements. These checks help Bing understand how content will display in search results and improve snippet generation and ranking signals tied to user experience. Crawl activity is moderate and often concentrated on pages important to Bing’s index. Its purpose is to simulate real-browser behavior and refine Bing’s presentation quality. RobotSense.io verifies BingPreview using Microsoft’s official validation methods, ensuring only genuine BingPreview traffic is identified.

DuplexWeb-Google

Others

by Google

[This crawler is officially retired as per Google] DuplexWeb-Google is a Google crawler associated with Duplex and Assistant-related technologies that fetch web content to help generate conversational responses and perform task-oriented actions. It retrieves page information needed to understand structured data, business details, menus, appointment flows, and other interactive elements. Crawl activity is selective and generally tied to user-initiated tasks or systems that prepare content for automated assistance. Its purpose is to support natural-language interactions by ensuring Google’s assistant technologies can interpret and use real-time webpage information accurately. It ignores the global user agent (*) rule. RobotSense.io verifies DuplexWeb-Google using Google’s official validation methods, ensuring only genuine DuplexWeb-Google traffic is identified.

FacebookExternalHit

Others

by Meta / Facebook

FacebookExternalHit is Facebook’s (Meta’s) crawler used to fetch webpage content for link previews across Facebook, Messenger, Instagram, and other Meta surfaces. It retrieves metadata such as Open Graph tags, titles, descriptions, images, and structured data. These requests are user-triggered, occurring when someone shares or pastes a URL on a Meta platform. The bot does not index or rank websites and has no connection to search algorithms. Blocking it may prevent accurate link previews. Crawl activity is lightweight and focused on fetching just enough content to generate rich social previews. It ignores the global user agent (*) rule. RobotSense.io verifies FacebookExternalHit using Meta’s official validation methods, ensuring only genuine FacebookExternalHit traffic is identified.

Google Messages

Others

by Google

Google Messages uses a fetcher to retrieve webpage data for generating link previews when users send URLs in chat conversations. It fetches metadata such as titles, descriptions, images, and structured tags (e.g., Open Graph) to render rich previews inside messages. This traffic is strictly user-triggered, occurring only when a link is shared. It is not a crawler for indexing or discovery and has no impact on Google Search rankings. Blocking it may prevent previews from displaying correctly. Activity is lightweight and targeted, focused solely on enhancing the messaging experience with accurate link previews. Google Messages bot does not respect robots.txt rules. RobotSense.io verifies Google Messages using Google’s official validation methods, ensuring only genuine Google Messages traffic is identified.

GoogleOther

Others

by Google

GoogleOther is a general-purpose crawler used by Google for internal research, large-scale data analysis, and non–Search-related fetching. It is part of Google’s secondary crawling infrastructure, designed to offload tasks that don’t require the full capabilities or strict policies of Googlebot. GoogleOther typically performs broad but lower-priority fetches, such as machine learning dataset generation or internal experiments. Its activity is generally lightweight compared to Googlebot and is separate from indexing operations that directly influence Google Search results. RobotSense.io verifies GoogleOther using Google’s official validation methods, ensuring only genuine GoogleOther traffic is identified.

GoogleOther-Image

Others

by Google

GoogleOther-Image is a specialized image-focused variant of the GoogleOther crawler, used for internal research, large-scale image analysis, and non–Search-related processing. The bot fetches image files and surrounding metadata but does not directly influence Google Images or Search rankings. Activity is usually lightweight and broad, supporting tasks such as dataset generation, model training, or experimental visual analysis within Google’s internal systems. RobotSense.io verifies GoogleOther-Image using Google’s official validation methods, ensuring only genuine GoogleOther-Image traffic is identified.