hockeysyte.com
robots.txt

Robots Exclusion Standard data for hockeysyte.com

Resource Scan

Scan Details

Site Domain hockeysyte.com
Base Domain hockeysyte.com
Scan Status Failed
Failure StageFetching resource.
Failure ReasonServer returned a client error.
Last Scan2024-10-07T07:19:45+00:00
Next Scan 2024-12-06T07:19:45+00:00

Last Successful Scan

Scanned2024-07-17T07:18:28+00:00
URL http://hockeysyte.com/robots.txt
Domain IPs 13.236.224.120, 13.54.152.216, 3.104.234.42
Response IP 13.236.224.120
Found Yes
Hash 5a639a0e32438f3ba8b03e4cb3e4968925ba6647ee8f9e66e0769b9274a801a2
SimHash 5294f150ea9a

Groups

cliqzbot/3.0

Rule Path
Disallow /

cliqzbot

Rule Path
Disallow /

dotbot

Rule Path
Disallow /

dotbot/1.1

Rule Path
Disallow /

seekport crawler

Rule Path
Disallow /

paracrawl

Rule Path
Disallow /

scrapy/1.5.0

Rule Path
Disallow /

scrapy

Rule Path
Disallow /

velenpublicwebcrawler (velen.io)

Rule Path
Disallow /

velenpublicwebcrawler

Rule Path
Disallow /

semrushbot/2~bl

Rule Path
Disallow /

semrushbot

Rule Path
Disallow /

pcore-http

Rule Path
Disallow /

the knowledge ai

Rule Path
Disallow /

mauibot

Rule Path
Disallow /

crawler.feedback+wc@gmail.com

Rule Path
Disallow /

cyotekwebcopy/1.0

Rule Path
Disallow /

centurybot9@gmail.com

Rule Path
Disallow /

crawler (crawler.feedback@gmail.com)

Rule Path
Disallow /

crawler

Rule Path
Disallow /

barkrowler/0.7 (+http://www.exensa.com/crawl)

Rule Path
Disallow /

go-http-client/1.1

Rule Path
Disallow /

test crawl

Rule Path
Disallow /

scalaj-http/1.0

Rule Path
Disallow /

bubing

Rule Path
Disallow /

wotbox/2.01

Rule Path
Disallow /

ccbot/2.0

Rule Path
Disallow /

ccbot

Rule Path
Disallow /

ebibot

Rule Path
Disallow /

pcore-http/v0.24.5

Rule Path
Disallow /

testitest1

Rule Path
Disallow /

vegi bot

Rule Path
Disallow /

istellabot/t.1

Rule Path
Disallow /

istellabot/t.1.13

Rule Path
Disallow /

istellabot

Rule Path
Disallow /

ltx71

Rule Path
Disallow /

ltx71 - (http://ltx71.com/)

Rule Path
Disallow /

mj12bot

Rule Path
Disallow /

booglebot2

Rule Path
Disallow /

booglebot

Rule Path
Disallow /

booglebot 2.0

Rule Path
Disallow /

booglebot/2.0

Rule Path
Disallow /

mj12bot/v1.0.5

Rule Path
Disallow /

sitebot

Rule Path
Disallow /

baiduspider

Rule Path
Disallow /

baiduspider/2.0

Rule Path
Disallow /

influencebo

Rule Path
Disallow /

searchmetricsbot

Rule Path
Disallow /

sosospider

Rule Path
Disallow /

acoonbot

Rule Path
Disallow /

easouspider

Rule Path
Disallow /

businessdbbot

Rule Path
Disallow /

superfeedr

Rule Path
Disallow /

flipboardproxy

Rule Path
Disallow /

flipboard

Rule Path
Disallow /

yandexbot

Rule Path
Disallow /

jikespider

Rule Path
Disallow /

rogerbot

Rule Path
Disallow /

rogerbot/1.0

Rule Path
Disallow /

flipboardproxy

Rule Path
Disallow /

swebot

Rule Path
Disallow /

swebot

Rule Path
Disallow /

seznambot

Rule Path
Disallow /

www.80legs.com

Rule Path
Disallow /

nerdbynature.bot

Rule Path
Disallow /

comodospider

Rule Path
Disallow /

comodospider/nutch-1.2

Rule Path
Disallow /

daumoa

Rule Path
Disallow /

ahrefsbot

Rule Path
Disallow /

beetlebot

Rule Path
Disallow /

niki-bot

Rule Path
Disallow /

riddler

Rule Path
Disallow /

spbot

Rule Path
Disallow /

icarus6

Rule Path
Disallow /

icarus6

Rule Path
Disallow /

icarus

Rule Path
Disallow /

icarus

Rule Path
Disallow /

icarus6j

Rule Path
Disallow /

siteexplorer

Rule Path
Disallow /

seokicks-robot

Rule Path
Disallow /

knelson

Rule Path
Disallow /

knelson/0.9

Rule Path
Disallow /

wotbox/2.01

Rule Path
Disallow /

blexbot/1.0

Rule Path
Disallow /

yisouspider

Rule Path
Disallow /

qwantify

Rule Path
Disallow /

nimbostratus-bot/v1.3.2

Rule Path
Disallow /

megaindex.ru

Rule Path
Disallow /

megaindex.com

Rule Path
Disallow /

semanticscholarbot

Rule Path
Disallow /

petalbot

Rule Path
Disallow /

yandex

Rule Path
Disallow /

trendictionbot

Rule Path
Disallow /

*

Rule Path
Allow /themes/*.css$
Allow /themes/*.css?
Allow /themes/*.js$
Allow /themes/*.js?
Allow /themes/*.gif
Allow /themes/*.jpg
Allow /themes/*.jpeg
Allow /themes/*.png
Disallow /assets/
Disallow /js/
Disallow /css/
Disallow /CHANGELOG.md
Disallow /cron.php
Disallow /INSTALL.mysql.txt
Disallow /INSTALL.pgsql.txt
Disallow /INSTALL.sqlite.txt
Disallow /install.php
Disallow /INSTALL.txt
Disallow /LICENSE.txt
Disallow /MAINTAINERS.txt
Disallow /update.php
Disallow /UPGRADE.txt
Disallow /xmlrpc.php
Disallow /admin/
Disallow /comment/reply/
Disallow /filter/tips/
Disallow /node/add/
Disallow /search/
Disallow /user/register/
Disallow /user/password/
Disallow /user/login/
Disallow /user/logout/
Disallow /?q=admin%2F
Disallow /?q=comment%2Freply%2F
Disallow /?q=filter%2Ftips%2F
Disallow /?q=node%2Fadd%2F
Disallow /?q=search%2F
Disallow /?q=user%2Fpassword%2F
Disallow /?q=user%2Fregister%2F
Disallow /?q=user%2Flogin%2F
Disallow /?q=user%2Flogout%2F

Other Records

Field Value
crawl-delay 10

Comments

  • CSS, JS, Images
  • Directories to avoid
  • Files to avoid
  • Paths (clean URLs)
  • Paths (no clean URLs)

Warnings

  • 10 invalid lines.