apache、iis規則屏蔽攔截蜘蛛抓取如果是正常的搜索引擎蜘蛛訪問,不建議對蜘蛛進行禁止,否則網站在百度等搜索引擎中的收錄和排名將會丟失,造成客戶流失等損失。 可以優先考慮升級虛擬主機型號以獲得更多的流量或升級為云服務器(不限流量)。 更多詳情請訪問: http://www.shinetop.cn/faq/list.asp?unid=626
Linux下規則文件.htaccess(手工創建.htaccess文件到站點根目錄) <IfModule mod_rewrite.c>
RewriteEngine On
#Block spider
RewriteCond %{HTTP_USER_AGENT} "Apache-HttpClient|SemrushBot|Webdup|AcoonBot|AhrefsBot|Ezooms|EdisterBot|EC2LinkFinder|jikespider|Purebot|MJ12bot|WangIDSpider|WBSearchBot|Wotbox|xbfMozilla|Yottaa|YandexBot|Jorgee|SWEBot|spbot|TurnitinBot-Agent|mail.RU|curl|perl|Python|Wget|Xenu|ZmEu" [NC]
RewriteRule !(^robots\.txt$) - [F]
</IfModule>Windows2008、2012或更高系統下規則文件web.config (手工創建web.config文件到站點根目錄) <?xml version="1.0" encoding="UTF-8"?>
<configuration>
<system.webServer>
<rewrite>
<rules>
<rule name="Block spider">
<match url="(^robots.txt$)" ignoreCase="false" negate="true" />
<conditions>
<add input="{HTTP_USER_AGENT}" pattern="Apache-HttpClient|SemrushBot|Webdup|AcoonBot|AhrefsBot|Ezooms|EdisterBot|EC2LinkFinder|jikespider|Purebot|MJ12bot|WangIDSpider|WBSearchBot|Wotbox|xbfMozilla|Yottaa|YandexBot|Jorgee|SWEBot|spbot|TurnitinBot-Agent|mail.RU|curl|perl|Python|Wget|Xenu|ZmEu" ignoreCase="true" />
</conditions>
<action type="AbortRequest"/>
</rule>
</rules>
</rewrite>
</system.webServer>
</configuration>注:“{HTTP_USER_AGENT}”所在行中是不明蜘蛛名稱,根據需要添加以"|"為分割。 請參照工單,根據實際情況設置。 如需我司幫助設置,請提交工單,根據實際問題收費30元-200元
|
|||||
| >> 相關文章 | |||||
|
|
|||||