Amzn-User
OthersVerify Amzn-User IP Address
Verify if an IP address truly belongs to Amazon, using official verification methods. Enter both IP address and User-Agent from your logs for the most accurate bot verification.
[Amazon Bots can take upto 30 days to read your Robots.txt updates.] Amzn-User is a bot associated with Amazon services that fetch webpage content on behalf of end users or Amazon applications rather than acting as a general-purpose crawler. It typically appears when Amazon apps, devices, or internal systems request metadata, previews, or content needed for features like link expansion, in-app browsing, or contextual analysis. The traffic is user-driven, not designed for large-scale indexing or scraping. Amzn-User usually performs lightweight, targeted fetches limited to specific URLs users interact with. Its purpose is to support Amazon product experiences by retrieving just enough page data to power user-facing functionality. It ignores the global user agent (*) rule. RobotSense.io verifies Amzn-User using Amazon’s official validation methods, ensuring only genuine Amzn-User traffic is identified.
User Agent Examples
Mozilla/5.0 AppleWebKit/537.36 (KHTML, like Gecko; compatible; Amzn-User/0.1) Chrome/119.0.6045.214 Safari/537.36Robots.txt Configuration for Amzn-User
Amzn-UserUse this identifier in your robots.txt User-agent directive to target Amzn-User.
Recommended Configuration
Our recommended robots.txt configuration for Amzn-User:
User-agent: Amzn-User
Allow: /Completely Block Amzn-User
Prevent this bot from crawling your entire site:
User-agent: Amzn-User
Disallow: /Completely Allow Amzn-User
Allow this bot to crawl your entire site:
User-agent: Amzn-User
Allow: /Block Specific Paths
Block this bot from specific directories or pages:
User-agent: Amzn-User
Disallow: /private/
Disallow: /admin/
Disallow: /api/Allow Only Specific Paths
Block everything but allow specific directories:
User-agent: Amzn-User
Disallow: /
Allow: /public/
Allow: /blog/Set Crawl Delay
Limit how frequently Amzn-User can request pages (in seconds):
User-agent: Amzn-User
Allow: /
Crawl-delay: 10Note: This bot does not officially mention about honoring Crawl-Delay rule.
Put these rules to work
Frequently Asked Questions
- What is Amzn-User, and why is it visiting my website?
- Amzn-User is a web crawler operated by Amazon that retrieves publicly accessible pages from websites. It is generally used by Amazon services to access web content for features such as link previews, integrations, and other service-level content retrieval tasks. Visits are typically triggered when Amazon systems need to fetch a page URL that appears in user activity or service workflows. Seeing Amzn-User in website logs is normal for publicly accessible pages that are referenced within Amazon platforms.
- Is Amzn-User a legitimate bot, or is it commonly spoofed?
- Amzn-User is an official crawler operated by Amazon. However, its user-agent string can be spoofed by automated scripts or malicious traffic attempting to appear as legitimate crawler activity. Attackers sometimes imitate well-known bots in order to bypass simple filters or hide scraping activity in server logs. Because of this, the User-Agent string alone is not sufficient to confirm that traffic is genuinely from Amzn-User. You can use Amazon's recommended methods mentioned below to verify a legitimate visit, or use RobotSense.io API to easily verify Amzn-User visits.
- How can I verify that a request is really coming from Amzn-User?
- You can use Amazon's recommended official methods to verify Amzn-User visits, these include: - IP range checks Do not use User-Agent based detection as that can be easily spoofed. Alternatively, you can use RobotSense.io API to easily verify Amzn-User and other bots from Amazon.
- Should I allow or block Amzn-User on my website?
- Allowing Amzn-User is generally optional and depends on whether you want Amazon services to access your publicly available pages. Allowing it may help ensure that links shared or referenced within Amazon services can retrieve page content correctly. If you are suddenly seeing too many visits, you can consider adding a small crawl-delay in your robots.txt before completely disallowing. Blocking may be appropriate in situations such as: - high server load caused by automated traffic - pages containing sensitive or restricted information - internal systems, APIs, or staging environments not intended for public access For most public websites, Amzn-User traffic is typically limited and does not cause operational issues.
- How can I control or block Amzn-User using robots.txt or other methods?
- You can add a rule in your robots.txt, as given above to control (crawl-delay) or disallow Amzn-User bot. Amzn-User crawler honors robots.txt directives, but it may take up to 30 days for your recent robots.txt changes to reflect properly. Also, you can use further controls in your WAF, or in RobotSense enforcement settings to manage the bot behavior.
- How often does Amzn-User crawl websites, and can it impact server performance?
- Amzn-User typically retrieves pages on demand rather than performing continuous large-scale crawling. Requests often occur when Amazon services need to access a specific URL, such as when a link is shared or referenced within a platform workflow. For most sites, the request rate is low and distributed over time. Performance impact is generally minimal, though dynamic pages or smaller servers may notice occasional additional requests. Some administrators choose to rate-limit or restrict it.
- What happens if I block Amzn-User? SEO, visibility, and feature impact explained.
- Blocking Amzn-User does not affect traditional search engine rankings because it is not a primary search engine crawler. However, blocking it may prevent Amazon services from retrieving page content when needed. Possible effects include: - Links shared within Amazon services may not generate previews - Certain Amazon integrations may not be able to retrieve page metadata - Content may not be accessible to Amazon systems that fetch external URLs For most websites, blocking Amzn-User has no direct impact on general web search visibility.
- Does Amzn-User collect, scrape, or use my content for training or reuse?
- No, Amzn-User bot has no officially documented AI purpose or republishing use-case. Amzn-User collected data may be used for purposes such as indexing, metadata extraction, and building datasets used across Amazon services.
Other Amazon Bots
Amazon operates other crawlers you may also need to configure.
amazon-QBusiness
AI Service[Amazon Bots can take upto 30 days to read your Robots.txt updates.] amazon-QBusiness is a crawler associated with Amazon’s Q Business and enterprise AI services, used to retrieve webpage content that organizations reference within Q-based workflows. It performs targeted, purpose-specific fetches triggered by users or automated enterprise integrations. The bot gathers text, metadata, and structural information to support search, summarization, and knowledge enrichment within Amazon’s AI tools. It is not a broad web crawler and does not index content for public Amazon services. It's crawl volume is typically low, reflecting only the URLs explicitly accessed through Q Business environments. RobotSense.io verifies amazon-QBusiness using Amazon’s official validation methods, ensuring only genuine amazon-QBusiness traffic is identified.
Amazonbot
Others[Amazon Bots can take upto 30 days to read your Robots.txt updates.] Amazonbot is Amazon’s official web crawler, used to discover and fetch webpage content for applications such as Alexa, product-related features, and Amazon’s AI and search systems. Crawl activity varies based on Amazon services that rely on external web content, but it is generally moderate and focused on structured data, text content, and page metadata. Its purpose is to enhance Amazon’s search, AI models, and user-facing features. It ignores the global user agent (*) rule. RobotSense.io verifies Amazonbot using Amazon’s official validation methods, ensuring only genuine Amazonbot traffic is identified.
Amzn-SearchBot
Search[Amazon Bots can take upto 30 days to read your Robots.txt updates.] Amzn-SearchBot is Amazon’s web crawler used to discover and retrieve publicly available content for Amazon search and AI-related services. It fetches webpages to analyze text, metadata, and structured information that can support Amazon’s search features and machine learning systems. Crawl activity is typically moderate and focused on publicly accessible pages. Its purpose is to help Amazon improve content discovery, relevance, and information retrieval across its ecosystem. It ignores the global user agent (*) rule. RobotSense.io verifies Amzn-SearchBot using Amazon’s official validation methods, ensuring only genuine Amzn-SearchBot traffic is identified.
Similar Bots
Other Others bots from different operators.
BingPreview
Othersby Microsoft
BingPreview is Microsoft’s rendering and compatibility crawler used to evaluate how webpages appear in browsers and Bing search features. It fetches pages to test layout, mobile responsiveness, JavaScript rendering, and visual elements. These checks help Bing understand how content will display in search results and improve snippet generation and ranking signals tied to user experience. Crawl activity is moderate and often concentrated on pages important to Bing’s index. Its purpose is to simulate real-browser behavior and refine Bing’s presentation quality. RobotSense.io verifies BingPreview using Microsoft’s official validation methods, ensuring only genuine BingPreview traffic is identified.
DuplexWeb-Google
Othersby Google
[This crawler is officially retired as per Google] DuplexWeb-Google is a Google crawler associated with Duplex and Assistant-related technologies that fetch web content to help generate conversational responses and perform task-oriented actions. It retrieves page information needed to understand structured data, business details, menus, appointment flows, and other interactive elements. Crawl activity is selective and generally tied to user-initiated tasks or systems that prepare content for automated assistance. Its purpose is to support natural-language interactions by ensuring Google’s assistant technologies can interpret and use real-time webpage information accurately. It ignores the global user agent (*) rule. RobotSense.io verifies DuplexWeb-Google using Google’s official validation methods, ensuring only genuine DuplexWeb-Google traffic is identified.
FacebookExternalHit
Othersby Meta / Facebook
FacebookExternalHit is Facebook’s (Meta’s) crawler used to fetch webpage content for link previews across Facebook, Messenger, Instagram, and other Meta surfaces. It retrieves metadata such as Open Graph tags, titles, descriptions, images, and structured data. These requests are user-triggered, occurring when someone shares or pastes a URL on a Meta platform. The bot does not index or rank websites and has no connection to search algorithms. Blocking it may prevent accurate link previews. Crawl activity is lightweight and focused on fetching just enough content to generate rich social previews. It ignores the global user agent (*) rule. RobotSense.io verifies FacebookExternalHit using Meta’s official validation methods, ensuring only genuine FacebookExternalHit traffic is identified.
Google Messages
Othersby Google
Google Messages uses a fetcher to retrieve webpage data for generating link previews when users send URLs in chat conversations. It fetches metadata such as titles, descriptions, images, and structured tags (e.g., Open Graph) to render rich previews inside messages. This traffic is strictly user-triggered, occurring only when a link is shared. It is not a crawler for indexing or discovery and has no impact on Google Search rankings. Blocking it may prevent previews from displaying correctly. Activity is lightweight and targeted, focused solely on enhancing the messaging experience with accurate link previews. Google Messages bot does not respect robots.txt rules. RobotSense.io verifies Google Messages using Google’s official validation methods, ensuring only genuine Google Messages traffic is identified.
GoogleOther
Othersby Google
GoogleOther is a general-purpose crawler used by Google for internal research, large-scale data analysis, and non–Search-related fetching. It is part of Google’s secondary crawling infrastructure, designed to offload tasks that don’t require the full capabilities or strict policies of Googlebot. GoogleOther typically performs broad but lower-priority fetches, such as machine learning dataset generation or internal experiments. Its activity is generally lightweight compared to Googlebot and is separate from indexing operations that directly influence Google Search results. RobotSense.io verifies GoogleOther using Google’s official validation methods, ensuring only genuine GoogleOther traffic is identified.
GoogleOther-Image
Othersby Google
GoogleOther-Image is a specialized image-focused variant of the GoogleOther crawler, used for internal research, large-scale image analysis, and non–Search-related processing. The bot fetches image files and surrounding metadata but does not directly influence Google Images or Search rankings. Activity is usually lightweight and broad, supporting tasks such as dataset generation, model training, or experimental visual analysis within Google’s internal systems. RobotSense.io verifies GoogleOther-Image using Google’s official validation methods, ensuring only genuine GoogleOther-Image traffic is identified.