mondopiante.com
robots.txt

Robots Exclusion Standard data for mondopiante.com

Resource Scan

Scan Details

Site Domain mondopiante.com
Base Domain mondopiante.com
Scan Status Ok
Last Scan2024-05-26T05:06:17+00:00
Next Scan 2024-06-25T05:06:17+00:00

Last Scan

Scanned2024-05-26T05:06:17+00:00
URL https://mondopiante.com/robots.txt
Domain IPs 86.107.32.93
Response IP 86.107.32.93
Found Yes
Hash 5178d733b864b03cfbddfab2d76206bb48e920998a1b494d29995e7d003734c2
SimHash db1e72d8c3b6

Groups

*

Rule Path
Allow */modules/*.css
Allow */modules/*.js
Allow */modules/*.png
Allow */modules/*.jpg
Allow /js/jquery/*
Disallow /*?order=
Disallow /*?tag=
Disallow /*?id_currency=
Disallow /*?search_query=
Disallow /*?back=
Disallow /*?n=
Disallow /*%26order%3D
Disallow /*%26tag%3D
Disallow /*%26id_currency%3D
Disallow /*%26search_query%3D
Disallow /*%26back%3D
Disallow /*%26n%3D
Disallow /*controller%3Daddresses
Disallow /*controller%3Daddress
Disallow /*controller%3Dauthentication
Disallow /*controller%3Dcart
Disallow /*controller%3Ddiscount
Disallow /*controller%3Dfooter
Disallow /*controller%3Dget-file
Disallow /*controller%3Dheader
Disallow /*controller%3Dhistory
Disallow /*controller%3Didentity
Disallow /*controller%3Dimages.inc
Disallow /*controller%3Dinit
Disallow /*controller%3Dmy-account
Disallow /*controller%3Dorder
Disallow /*controller%3Dorder-slip
Disallow /*controller%3Dorder-detail
Disallow /*controller%3Dorder-follow
Disallow /*controller%3Dorder-return
Disallow /*controller%3Dorder-confirmation
Disallow /*controller%3Dpagination
Disallow /*controller%3Dpassword
Disallow /*controller%3Dpdf-invoice
Disallow /*controller%3Dpdf-order-return
Disallow /*controller%3Dpdf-order-slip
Disallow /*controller%3Dproduct-sort
Disallow /*controller%3Dsearch
Disallow /*controller%3Dstatistics
Disallow /*controller%3Dattachment
Disallow /*controller%3Dguest-tracking
Disallow /app/
Disallow /cache/
Disallow /classes/
Disallow /config/
Disallow /controllers/
Disallow /download/
Disallow /js/
Disallow /localization/
Disallow /log/
Disallow /mails/
Disallow /modules/
Disallow /override/
Disallow /pdf/
Disallow /src/
Disallow /tools/
Disallow /translations/
Disallow /upload/
Disallow /var/
Disallow /vendor/
Disallow /webservice/
Disallow /it/app/
Disallow /it/cache/
Disallow /it/classes/
Disallow /it/config/
Disallow /it/controllers/
Disallow /it/download/
Disallow /it/js/
Disallow /it/localization/
Disallow /it/log/
Disallow /it/mails/
Disallow /it/modules/
Disallow /it/override/
Disallow /it/pdf/
Disallow /it/src/
Disallow /it/tools/
Disallow /it/translations/
Disallow /it/upload/
Disallow /it/var/
Disallow /it/vendor/
Disallow /it/webservice/
Disallow /en/app/
Disallow /en/cache/
Disallow /en/classes/
Disallow /en/config/
Disallow /en/controllers/
Disallow /en/download/
Disallow /en/js/
Disallow /en/localization/
Disallow /en/log/
Disallow /en/mails/
Disallow /en/modules/
Disallow /en/override/
Disallow /en/pdf/
Disallow /en/src/
Disallow /en/tools/
Disallow /en/translations/
Disallow /en/upload/
Disallow /en/var/
Disallow /en/vendor/
Disallow /en/webservice/
Disallow /*it/indirizzo
Disallow /*it/indirizzi
Disallow /*it/login
Disallow /*it/carrello
Disallow /*it/buoni-sconto
Disallow /*it/tracciatura-ospite
Disallow /*it/cronologia-ordini
Disallow /*it/dati-personali
Disallow /*it/account
Disallow /*it/ordine
Disallow /*it/conferma-ordine
Disallow /*it/segui-ordine
Disallow /*it/buono-ordine
Disallow /*it/recupero-password
Disallow /*it/ricerca
Disallow /*en/address
Disallow /*en/addresses
Disallow /*en/login
Disallow /*en/cart
Disallow /*en/discount
Disallow /*en/guest-tracking
Disallow /*en/order-history
Disallow /*en/identity
Disallow /*en/my-account
Disallow /*en/order
Disallow /*en/order-confirmation
Disallow /*en/order-follow
Disallow /*en/credit-slip
Disallow /*en/password-recovery
Disallow /*en/search
Allow */js/jquery/*
Allow */modules/*.gif
Allow */modules/*.jpeg
Allow */modules/*.woff
Allow */modules/*.woff2
Allow */modules/*.ttf
Allow */themes/*/assets/cache/*.js
Allow */themes/*/assets/cache/*.css
Allow */themes/*/assets/css/*
Allow */themes/*/assets/js/*
Allow */modules/revsliderprestashop/public/assets/fonts/font-awesome/fonts/*.ttf$
Allow */modules/revsliderprestashop/public/assets/fonts/font-awesome/fonts/*.ttf?
Allow */modules/revsliderprestashop/public/assets/fonts/font-awesome/fonts/*.woff?
Allow */modules/revsliderprestashop/public/assets/fonts/font-awesome/fonts/*.woff2?
Allow /img/*.gif
Allow /img/*.jpg
Allow /img/*.jpeg
Allow /img/*.png
Allow /*/*.jpg
Allow /*/*.png

yandex
yandexturbo
yandexbot
yandexbot/3.0
baiduspider

Rule Path
Disallow /

uptimebot
dotbot
dotbot/1.1
screaming frog seo spider
screaming frog seo spider/2,50
mj12bot
ahrefsbot
semrushbot
semrushbot-bm
semrushbot-sa
semrushbot-si
semrushbot-si/0.97
rogerbot/1.0
yandexmetrika
yandexmetrika/2.0
xenu's
xenu's link sleuth 1.1c
sistrix
ahrefsbot/5.2
ahrefsbot
seokicks
seokicks-robot
yisouspider
qwantify

Rule Path
Disallow /

ruby
python-requests/2.11.0
python-requests
libwww-perl/6.26
dotbot/1.1
mail.ru_bot/2.0
mail.ru_bot
barkrowler/0.5.1 (experimenting / debugging - sorry for your logs ) http://www.exensa.com/crawl - admin@exensa.com -- based on bubing
alexibot
aqua_products
asterias
b2w/0.1
backdoorbot/1.0
becomebot
bl.uk_lddc_bot
blexbot
bloglovin
blowfish/1.0
bookmark search tool
bot
botalot
builtbottough
bullseye/1.0
bunnyslippers
calculon spider
cheesebot
cherrypicker
cherrypickerelite/1.0
cherrypickerse/1.0
coccoc
copernic
copyrightcheck
cosmos
crescent
crescent internet toolpak http ole control v.1.0
daum
dittospyder
dumbot
emailcollector
emailsiphon
emailwolf
enterprise_search
enterprise_search/1.0
erocrawler
es
exabot
extractorpro
ezooms
fairad client
fatbot
flaming attackbot
foobot
freefind
gaisbot
getright/4.2
grub
grub-client
haosouspider
harvest/1.5
hatena antenna
heritrix
hloader
httplib
humanlinks
ia_archiver
idg
infonavirobot
inoreader.com
iron33/1.0.2
istellabot
jennybot
jetbot
jetbot/1.0
jikespider
kenjin spider
keyword density/0.9
larbin
lexibot
libweb/clshttp
linkextractorpro
linkscan/8.1a unix
linkwalker
lnspiderguy
ltx71
lwp-trivial
lwp-trivial/1.34
mata hari
megaindex.ru
megaindex.ru/
megaindex.ru/2.0
microsoft url control
microsoft url control - 5.01.4511
microsoft url control - 6.00.8169
miixpc
miixpc/4.2
mister pix
moget
moget/2.1
mozilla/5.0 (compatible; megaindex.ru/2.0; +https://www.megaindex.ru/?tab=linkanalyze)
mozilla/5.0 (compatible; spbot/5.0.3; +http://openlinkprofiler.org/bot )
msiecrawler
naver
netants
netmechanic
nicerspro
nutch
obot
offline explorer
omniexplorer_bot
openbot
openfind
openfind data gathere
optimizer
oracle ultra search
parser
paperlibot
pcore-http/v0.25.0
perman
php
propowerbot/2.14
prowebwalker
psbot
pu_in crawler
python-urllib
pyton-requests
qwantify
queryn metasearch
radiation retriever 1.1
repomonkey
repomonkey bait & tackle/v1.01
rma
searchpreview
semvisubot 2.0
seznambot
seznam screenshot-generator 2.1
sitebot
sitesnagger
smtbot
sogou
sootle
sosospider
spankbot
spanner
spinn3r
sqlmap
stanford
stanford comp sci
suzuran
swiftbot
szukacz/1.4
teleport
teleportpro
telesoft
the intraformant
thenomad
tighttwatbot
tocrawl/urldispatcher
trident
true_robot
true_robot/1.0
turingos
updownerbot
url control
url_spider_pro
urly warning
vci
vci webviewer vci webviewer win32
voilabot
web image collector
webauto
webbandit
webbandit/3.50
webcopier
webenhancer
webmasterworldforumbot
websauger
website quester
webster pro
webstripper
webvac
webzip
webzip/4.0
wget
wget/1.5.3
wget/1.6
willybot
wordpress
wordpress/4.7.2
wordpress/4.6.3
wotbox
www-collector-e
yodaobot
yak
yisouspider
zeus
zeus 32297 webster pro v2.9 win32
zeus link scout
zoominfobot
zumbot

Rule Path
Disallow /

Other Records

Field Value
sitemap https://www.mondopiante.com/1_index_sitemap.xml

Comments

  • robots.txt automatically generated by PrestaShop e-commerce open-source solution
  • http://www.prestashop.com - http://www.prestashop.com/forums
  • This file is to prevent the crawling and indexing of certain parts
  • of your site by web crawlers and spiders run by sites like Yahoo!
  • and Google. By telling these "robots" where not to go on your site,
  • you save bandwidth and server resources.
  • For more information about the robots.txt standard, see:
  • http://www.robotstxt.org/robotstxt.html
  • Allow Directives
  • Private pages
  • Directories for www.mondopiante.com
  • Files
  • By INGEMATIC - Manage Error form Google Search Console
  • Allow Directives
  • End
  • INIZIO DISALLOW BOT - Data 15-05-2019
  • SE
  • User-agent: Slurp
  • SEOTOOLS
  • SOCIAL
  • User-agent: Twitterbot
  • Disallow: /
  • OTHER
  • User-agent: TwengaBot
  • End
  • Sitemap

Warnings

  • 3 invalid lines.