beerclub.ro
robots.txt

Robots Exclusion Standard data for beerclub.ro

Resource Scan

Scan Details

Site Domain beerclub.ro
Base Domain beerclub.ro
Scan Status Ok
Last Scan2024-10-16T11:34:55+00:00
Next Scan 2024-10-23T11:34:55+00:00

Last Scan

Scanned2024-10-16T11:34:55+00:00
URL https://beerclub.ro/robots.txt
Redirect https://www.beerclub.ro/robots.txt
Redirect Domain www.beerclub.ro
Redirect Base beerclub.ro
Domain IPs 93.114.43.47
Redirect IPs 93.114.43.47
Response IP 93.114.43.47
Found Yes
Hash 738465c8c26e5655ee8288ad3e1c24877947ea2fc30be5ae1c3fa16e4fe02f6b
SimHash e6161b486dcf

Groups

googlebot

Rule Path
Disallow

adsbot-google

Rule Path
Disallow

googlebot-image

Rule Path
Disallow

*

Rule Path
Disallow /admin/
Disallow /cache/
Disallow /lang/
Disallow /lib/
Disallow /log/
Disallow /temp/

yandex

Rule Path
Disallow /

yandexbot

Rule Path
Disallow /

ahrefsbot

Rule Path
Disallow /

serpstatbot

Rule Path
Disallow /

seznambot

Rule Path
Disallow /

cliqzbot

Rule Path
Disallow /

baiduspider

Rule Path
Disallow /

coccocbot-web

Rule Path
Disallow /

ccbot

Rule Path
Disallow /

grapeshot

Rule Path
Disallow /

adbeat_bot

Rule Path
Disallow /

yisouspider

Rule Path
Disallow /

mojeekbot

Rule Path
Disallow /

qwantify

Rule Path
Disallow /

bleriot

Rule Path
Disallow /

sogou spider

Rule Path
Disallow /

sogou spider

Rule Path
Disallow /

ubicrawler
doc
zao
sitecheck.internetseer.com
zealbot
msiecrawler
sitesnagger
webstripper
webcopier
fetch
offline explorer
teleport
teleportpro
webzip
linko
httrack
microsoft.url.control
xenu
larbin
libwww
zyborg
download ninja
alexibot
alvinetspider
antenne hatena
apocalxexplorerbot
asterias
backdoorbot/1.0
bizinformation
black hole
blowfish/1.0
botalot
builtbottough
bullseye/1.0
bunnyslippers
cegbfeieh
cheesebot
cherrypicker
cherrypickerelite/1.0
cherrypickerse/1.0
copyrightcheck
cosmos
crescent
crescent internet toolpak http ole control v.1.0
disco pump 3.1
dittospyder
dotbot
emailcollector
emailsiphon
emailwolf
erocrawler
extractorpro
flamingo_searchengine
foobot
harvest/1.5
hloader
httplib
httrack
httrack 3.0
humanlinks
igentia
infonavirobot
jennybot
jikespider
kenjin spider
lexibot
libweb/clshttp
linkextractorpro
linkscan/8.1a unix
linkwalker
lwp-trivial
lwp-trivial/1.34
mata hari
microsoft url control - 5.01.4511
microsoft url control - 6.00.8169
miixpc
miixpc/4.2
mister pix
mlbot
moget
moget/2.1
ms search 4.0 robot
ms search 5.0 robot
naverbot
netants
netattache
netmechanic
nicerspro
offline explorer
openfind
openindexspider
propowerbot/2.14
prowebwalker
psbot
quepasacreep
queryn metasearch
repomonkey
rma
semrushbot
sightupbot
sitebot
sitesnagger
sitesucker
sogou web spider
sosospider
spankbot
spanner
speedy
suggybot
superbot
superbot/2.6
suzuran
szukacz/1.4
teleport
telesoft
the intraformant
thenomad
tighttwatbot
titan
tocrawl/urldispatcher
toscrawler
true_robot
true_robot/1.0
turingos
turnitinbot
urlpouls
urly warning
vci
web image collector
webauto
webbandit
webbandit/3.50
webcopier
webcopy
webenhancer
webmasterworldforumbot
webmirror
webreaper
websauger
website extractor
website quester
webster pro
webstripper
webstripper/2.02
webzip
wget
wikiofeedbot
winhttrack
www-collector-e
xenu link sleuth/1.3.8
yacy
yrspider
zeus
zookabot

Rule Path
Disallow /

mj12bot

Rule Path
Disallow /

israbot

Rule Path
Disallow

orthogaffe

Rule Path
Disallow

ubicrawler

Rule Path
Disallow /

doc

Rule Path
Disallow /

zao

Rule Path
Disallow /

sitecheck.internetseer.com

Rule Path
Disallow /

zealbot

Rule Path
Disallow /

msiecrawler

Rule Path
Disallow /

sitesnagger

Rule Path
Disallow /

webstripper

Rule Path
Disallow /

webcopier

Rule Path
Disallow /

fetch

Rule Path
Disallow /

offline explorer

Rule Path
Disallow /

teleport

Rule Path
Disallow /

teleportpro

Rule Path
Disallow /

webzip

Rule Path
Disallow /

linko

Rule Path
Disallow /

httrack

Rule Path
Disallow /

microsoft.url.control

Rule Path
Disallow /

xenu

Rule Path
Disallow /

larbin

Rule Path
Disallow /

libwww

Rule Path
Disallow /

zyborg

Rule Path
Disallow /

download ninja

Rule Path
Disallow /

fast

Rule Path
Disallow /

k2spider

Rule Path
Disallow /

npbot

Rule Path
Disallow /

webreaper

Rule Path
Disallow /

Comments

  • Disallow: /css/
  • Disallow: /js/
  • All Yandex bots
  • Yandex bot...
  • Observed spamming large amounts of https://en.wikipedia.org/?curid=NNNNNN
  • and ignoring 429 ratelimit responses, claims to respect robots:
  • http://mj12bot.com/
  • advertising-related bots:
  • User-agent: Mediapartners-Google*
  • Disallow: /
  • Wikipedia work bots:
  • Crawlers that are kind enough to obey, but which we'd rather not have
  • unless they're feeding search engines.
  • Some bots are known to be trouble, particularly those designed to copy
  • entire sites. Please obey robots.txt.
  • Misbehaving: requests much too fast:
  • Doesn't follow robots.txt anyway, but...
  • Hits many times per second, not acceptable
  • http://www.nameprotect.com/botinfo.html
  • A capture bot, downloads gazillions of pages with no public benefit
  • http://www.webreaper.net/

Warnings

  • 1 invalid line.