How To Control Googlebot’s Interaction With Your Website
· 2023-06-09

Google's Search Relations team provides insights into controlling Googlebot's interactions with webpages on the latest 'Search Off The Record' podcast.

c5dbb20d085834cb175c1df224f46c42.png


Google’s Search Relations answered several questions regarding webpage indexing on the latest episode of the ‘Search Off The Record’ podcast.


The topics discussed were how to block Googlebot from crawling specific sections of a page and how to prevent Googlebot from accessing a site altogether.


Google’s John Mueller and Gary Illyes answered the questions examined in this article.


Blocking Googlebot From Specific Web Page Sections


Mueller says it’s impossible when asked how to stop Googlebot from crawling specific web page sections, such as “also bought” areas on product pages.


“The short version is that you can’t block crawling of a specific section on an HTML page,” Mueller said.


He went on to offer two potential strategies for dealing with the issue, neither of which, he stressed, are ideal solutions.


Mueller suggested using the data-nosnippet HTML attribute to prevent text from appearing in a search snippet.


Alternatively, you could use an iframe or JavaScript with the source blocked by robots.txt, although he cautioned that’s not a good idea.


“Using a robotted iframe or JavaScript file can cause problems in crawling and indexing that are hard to diagnose and resolve,” Mueller stated.


He reassured everyone listening that if the content in question is being reused across multiple pages, it’s not a problem that needs fixing.


“There’s no need to block Googlebot from seeing that kind of duplication,” he added.


Blocking Googlebot From Accessing A Website


In response to a question about preventing Googlebot from accessing any part of a site, Illyes provided an easy-to-follow solution.


“The simplest way is robots.txt: if you add a disallow: / for the Googlebot user agent, Googlebot will leave your site alone for as long you keep that rule there,” Illyes explained.


For those seeking a more robust solution, Illyes offers another method:


“If you want to block even network access, you’d need to create firewall rules that load our IP ranges into a deny rule,” he said.


See Google’s official documentation for a list of Googlebot’s IP addresses.


In Summary


Though it’s impossible to prevent Googlebot from accessing specific sections of an HTML page, methods such as using the data-nosnippet attribute can offer control.


When considering blocking Googlebot from your site entirely, a simple disallow rule in your robots.txt file will do the trick. However, more extreme measures like creating specific firewall rules are also available.













熱門文章
亞洲遊戲市場觀察:15大市場熱門遊戲與用戶趨勢
網路遊戲
2027 Global Game Connect(GGC)斯里蘭卡招商全面啟動!業務人脈盡在掌握!
灰度頭條
JILI 宣佈與全球板球傳奇 AB de Villiers(ABD)達成重磅戰略合作
體育遊戲
超級PAC籌資4800萬美元:體育博彩勢力加碼
合規與政策
西班牙監管機構警告在線賭博平臺存在身份盜竊行為
合規與政策
越南在線博彩業政策收緊 催生市場新機遇
東南亞資訊
新澤西州7月博彩收入創6.06億美元新高,頒布禁令
合規與政策
印第安納州在線賭場法案在眾議院委員會停滯不前
合規與政策
菲律賓網絡賭博和加密貨幣仍構成持續的洗錢風險
東南亞資訊
橫跨全球6個城市,灰度8場派對邀你共看世界盃,重塑高質量社交新場景
灰度頭條
越南博彩管控逐步放寬,惟本土需求仍顯乏力
東南亞資訊
菲律賓博彩技術賽道迎來新變局,B2B 供應模式加速滲透
東南亞資訊
斯里蘭卡博弈產業大轉型,官方:劍指南亞拉斯維加斯
合規與政策
印度最高法院受理公益訴訟,要求全國禁封「偽裝」成社交遊戲的賭博平台
合規與政策
哈薩克計劃對線上賭場促銷活動進行處罰
合規與政策
首頁
遊戲
合作
發現
我的