BingVideoPreview
SearchVerify BingVideoPreview IP Address
Verify if an IP address truly belongs to Microsoft, using official verification methods. Enter both IP address and User-Agent from your logs for the most accurate bot verification.
BingVideoPreview is Microsoft’s crawler for fetching video-related content to generate previews, thumbnails, and metadata for Bing’s video search experiences. It retrieves video files, poster images, structured data, captions, and surrounding context. This crawler does not perform full-site indexing; instead, it focuses specifically on video assets and the information required to power Bing’s video carousels and preview interfaces. Activity is targeted and relatively low-volume, driven by pages that contain or reference video content. RobotSense.io verifies BingVideoPreview using Microsoft’s official validation methods, ensuring only genuine BingVideoPreview traffic is identified.
User Agent Examples
Mozilla/5.0 AppleWebKit/537.36 (KHTML, like Gecko; compatible; BingVideoPreview/1.0; +https://aka.ms/microsoftbots) Chrome/W.X.Y.Z Safari/537.36
Mozilla/5.0 (compatible; BingVideoPreview/1.0; +https://aka.ms/microsoftbots) W.X.Y.Z Safari/537.36
Mozilla/5.0 (Linux; Android 6.0.1; Nexus 5X Build/MMB29P) AppleWebKit/537.36 (KHTML, like Gecko) Chrome/W.X.Y.Z Mobile Safari/537.36 (compatible; BingVideoPreview/1.0; +https://aka.ms/microsoftbots)Robots.txt Configuration for BingVideoPreview
BingVideoPreviewUse this identifier in your robots.txt User-agent directive to target BingVideoPreview.
Recommended Configuration
Our recommended robots.txt configuration for BingVideoPreview:
User-agent: BingVideoPreview
Allow: /Completely Block BingVideoPreview
Prevent this bot from crawling your entire site:
User-agent: BingVideoPreview
Disallow: /Completely Allow BingVideoPreview
Allow this bot to crawl your entire site:
User-agent: BingVideoPreview
Allow: /Block Specific Paths
Block this bot from specific directories or pages:
User-agent: BingVideoPreview
Disallow: /private/
Disallow: /admin/
Disallow: /api/Allow Only Specific Paths
Block everything but allow specific directories:
User-agent: BingVideoPreview
Disallow: /
Allow: /public/
Allow: /blog/Set Crawl Delay
Limit how frequently BingVideoPreview can request pages (in seconds):
User-agent: BingVideoPreview
Allow: /
Crawl-delay: 10Note: This bot officially honors the Crawl-delay directive.
Put these rules to work
Other Microsoft Bots
Microsoft operates other crawlers you may also need to configure.
AdIdxBot
AdsAdIdxBot is Microsoft’s advertising-focused crawler used to evaluate landing pages associated with Microsoft Advertising (formerly Bing Ads). It performs targeted fetches to assess page load speed, relevance, redirects, content quality, and policy compliance. These checks help determine ad quality scores, eligibility, and overall user experience. Blocking AdIdxBot, may limit Microsoft’s ability to review landing pages, potentially affecting ad performance. Crawl activity is selective and low-volume, typically triggered when advertisers create or modify campaigns. Its purpose is to ensure that ad destinations meet Microsoft’s standards for safety, usability, and relevance. RobotSense.io verifies AdIdxBot using Microsoft’s official validation methods, ensuring only genuine AdIdxBot traffic is identified.
Bingbot
SearchBingbot is Microsoft’s primary web crawler, responsible for discovering and indexing content for Bing Search and other Microsoft services. The crawler fetches HTML, structured data, images, and metadata to understand page relevance and ranking signals. Crawl activity varies based on site authority, update frequency, and sitemap signals. Its purpose is to keep Bing’s search index fresh, accurate, and aligned with user search intent across Microsoft platforms. RobotSense.io verifies Bingbot using Microsoft’s official validation methods, ensuring only genuine Bingbot traffic is identified.
BingPreview
OthersBingPreview is Microsoft’s rendering and compatibility crawler used to evaluate how webpages appear in browsers and Bing search features. It fetches pages to test layout, mobile responsiveness, JavaScript rendering, and visual elements. These checks help Bing understand how content will display in search results and improve snippet generation and ranking signals tied to user experience. Crawl activity is moderate and often concentrated on pages important to Bing’s index. Its purpose is to simulate real-browser behavior and refine Bing’s presentation quality. RobotSense.io verifies BingPreview using Microsoft’s official validation methods, ensuring only genuine BingPreview traffic is identified.
MicrosoftPreview
OthersMicrosoftPreview is a Microsoft crawler used to render webpages in a browser-like environment for testing, feature evaluation, and content understanding across Microsoft services. It performs fetches that simulate modern browser behavior, including JavaScript execution, layout rendering, and metadata extraction. Unlike Bingbot, it is not used for core indexing but for assessing how pages display in Microsoft products. Activity is moderate and focused on pages relevant to rendering quality checks. Its purpose is to enhance visual accuracy and user experience across Microsoft’s platforms. RobotSense.io verifies MicrosoftPreview using Microsoft’s official validation methods, ensuring only genuine MicrosoftPreview traffic is identified.
Similar Bots
Other Search bots from different operators.
Amzn-SearchBot
Searchby Amazon
[Amazon Bots can take upto 30 days to read your Robots.txt updates.] Amzn-SearchBot is Amazon’s web crawler used to discover and retrieve publicly available content for Amazon search and AI-related services. It fetches webpages to analyze text, metadata, and structured information that can support Amazon’s search features and machine learning systems. Crawl activity is typically moderate and focused on publicly accessible pages. Its purpose is to help Amazon improve content discovery, relevance, and information retrieval across its ecosystem. It ignores the global user agent (*) rule. RobotSense.io verifies Amzn-SearchBot using Amazon’s official validation methods, ensuring only genuine Amzn-SearchBot traffic is identified.
Applebot
Searchby Apple
Applebot is Apple's official web crawler used to power search and content features across Apple services such as Siri, Spotlight Suggestions, and Safari. It crawls webpages to discover content, metadata, and structured information that enhance on-device and cloud-based search experiences. Crawl activity is generally moderate and focused on high-quality, publicly accessible content. Its purpose is to improve search relevance, answers, and suggestions across Apple’s ecosystem without operating a standalone public web search engine. Data crawled by Applebot may be utilized by Apple for foundational model training. Apple allows site owners to opt-out of having their content used for generative model training by disallowing Applebot-Extended in the robots.txt file. RobotSense.io verifies Applebot using Apple's official validation methods, ensuring only genuine Applebot traffic is identified.
Google Favicon
Searchby Google
[This crawler is officially retired as per Google] Google Favicon is a specialized Google crawler that retrieves website favicons for use across Google Search, Chrome, and other Google products. It fetches small icon files such as favicon.ico or declared alternative icons in HTML. This bot does not index page content or affect Search rankings; its role is purely to collect icons that visually represent sites in SERPs and browser surfaces. Most sites allow it since its requests are lightweight. Crawl activity is minimal and typically occurs when Google detects new or updated favicon assets. RobotSense.io verifies Google Favicon using Google’s official validation methods, ensuring only genuine Google Favicon traffic is identified.
Google Publisher Center / GoogleProducer
Searchby Google
Google Publisher Center is a platform that allows news publishers to manage how their content appears across Google News surfaces. When publishers submit feeds, sections, or site updates, Google may fetch associated URLs using Publisher Center–related user-agents to verify content, metadata, and feed accuracy. These fetches are not broad crawls; they are targeted checks tied to publisher actions such as updating feeds, article structures, or publication settings. Blocking it can disrupt feed validation or delay updates in Google News. Activity is typically light, triggered by publisher configuration changes or system refresh cycles. It ignores robots.txt rules. RobotSense.io verifies Google Publisher Center / GoogleProducer using Google’s official validation methods, ensuring only genuine Google Publisher Center / GoogleProducer traffic is identified.
Google StoreBot
Searchby Google
Google StoreBot is Google’s crawler responsible for fetching and validating data related to product listings, merchant feeds, and eCommerce pages used across Google Shopping surfaces. It helps Google evaluate product availability, pricing, structured data, and landing page quality. StoreBot works alongside Merchant Center systems to ensure product information is accurate, up-to-date, and compliant with Google’s listing requirements. Crawl behavior is focused, lightweight, and typically triggered by updates to product feeds or changes detected on merchant landing pages. RobotSense.io verifies Google StoreBot using Google’s official validation methods, ensuring only genuine Google StoreBot traffic is identified.
Googlebot
Searchby Google
Googlebot is Google’s primary web crawler, responsible for discovering, fetching, and updating content across the public internet for inclusion in Google Search. It operates at massive scale, continuously revisiting sites based on their importance, freshness, and user demand. Googlebot uses a distributed crawling infrastructure that intelligently balances crawl frequency with server load, aiming to gather the most useful and up-to-date information without overwhelming websites. It identifies itself with the Googlebot user-agent family and is fully transparent about its behavior. Genuine Googlebot traffic can be verified through Google’s published reverse-DNS method, which confirms whether an IP truly belongs to Google’s crawling network. Beyond standard HTML pages, Googlebot is capable of rendering JavaScript, interpreting structured data, and evaluating mobile friendliness, which directly influences how pages appear in search results. Googlebot has 2 internal variants i.e., Googlebot Smartphone and Googlebot Desktop. Google increasingly uses Googlebot Smartphone for content crawling. RobotSense.io verifies Googlebot using Google’s official validation methods, ensuring only genuine Googlebot traffic is identified.