modlitewnik.pl
robots.txt

Robots Exclusion Standard data for modlitewnik.pl

Archived Snapshots

Resource Scan

Scan Details

Site Domain	modlitewnik.pl
Base Domain	modlitewnik.pl
Scan Status	Ok
Last Scan	2025-10-14T12:21:38+00:00
Next Scan	2025-10-21T12:21:38+00:00

Last Scan

Scanned	2025-10-14T12:21:38+00:00
URL	https://modlitewnik.pl/robots.txt
Domain IPs	104.21.25.153, 172.67.134.87, 2606:4700:3035::ac43:8657, 2606:4700:3037::6815:1999
Response IP	172.67.134.87
Found	Yes
Hash	9503641b35a1c79a87f789bb61e0b054162de9222abac302895c2e11aee728e4
SimHash	56955243cdb0

Groups

*

Rule	Path
Allow	/

Rule

Path

Allow

amazonbot

Rule	Path
Disallow	/

Rule

Path

Disallow

applebot-extended

Rule	Path
Disallow	/

Rule

Path

Disallow

bytespider

Rule	Path
Disallow	/

Rule

Path

Disallow

ccbot

Rule	Path
Disallow	/

Rule

Path

Disallow

claudebot

Rule	Path
Disallow	/

Rule

Path

Disallow

google-extended

Rule	Path
Disallow	/

Rule

Path

Disallow

gptbot

Rule	Path
Disallow	/

Rule

Path

Disallow

meta-externalagent

Rule	Path
Disallow	/

Rule

Path

Disallow

*

Rule	Path
Allow	/wp-content/uploads/
Disallow	/wp-content/plugins/
Disallow	/wp-admin/
Disallow	/wp-json/
Disallow	/feed/
Disallow	/readme.html
Disallow	/ads.txt
Disallow	/refer/
Disallow	/wp-login.php
Disallow	/wp-admin/
Disallow	/*.pdf$
Disallow	/*.doc$

Rule

Path

Allow

/wp-content/uploads/

Disallow

/wp-content/plugins/

Disallow

/wp-admin/

Disallow

/wp-json/

Disallow

/feed/

Disallow

/readme.html

Disallow

/ads.txt

Disallow

/refer/

Disallow

/wp-login.php

Disallow

/wp-admin/

Disallow

/*.pdf$

Disallow

/*.doc$

Other Records

Field	Value
crawl-delay	600

Field

Value

crawl-delay

600

bingbot

No rules defined. All paths allowed.

Other Records

Field	Value
crawl-delay	10

Field

Value

crawl-delay

yandex

No rules defined. All paths allowed.

Other Records

Field	Value
crawl-delay	8

Field

Value

crawl-delay

mj12bot

No rules defined. All paths allowed.

Other Records

Field	Value
crawl-delay	10

Field

Value

crawl-delay

ahrefsbot
ahrefssiteaudit
adbeat_bot
alexibot
appengine
aqua_products
archive.org_bot
archive
asterias
b2w/0.1
backdoorbot/1.0
becomebot
blekkobot
blexbot
blowfish/1.0
bookmark search tool
botalot
builtbottough
bullseye/1.0
bunnyslippers
ccbot
cheesebot
cherrypicker
cherrypickerelite/1.0
cherrypickerse/1.0
chroot
copernic
copyrightcheck
cosmos
crescent
crescent internet toolpak http ole control v.1.0
dittospyder
dotbot
dumbot
emailcollector
emailsiphon
emailwolf
enterprise_search
enterprise_search/1.0
erocrawler
es
exabot
extractorpro
fairad client
flaming attackbot
foobot
gaisbot
getright/4.2
gigabot
grub
grub-client
go-http-client
harvest/1.5
hatena antenna
hloader
http://www.searchengineworld.com bot
http://www.webmasterworld.com bot
httplib
humanlinks
ia_archiver
ia_archiver/1.6
infonavirobot
iron33/1.0.2
jamesbot
jennybot
jetbot
jetbot/1.0
jorgee
kenjin spider
keyword density/0.9
larbin
lexibot
libweb/clshttp
linkextractorpro
linkpadbot
linkscan/8.1a unix
linkwalker
lnspiderguy
looksmart
lwp-trivial
lwp-trivial/1.34
mata hari
megalodon
microsoft url control
microsoft url control - 5.01.4511
microsoft url control - 6.00.8169
miixpc
miixpc/4.2
mister pix
mj12bot
moget
moget/2.1
mozilla
mozilla
mozilla/3
mozilla/4
mozilla/4.0 (compatible; bullseye; windows 95)
mozilla/4.0 (compatible; msie 4.0; windows 2000)
mozilla/4.0 (compatible; msie 4.0; windows 95)
mozilla/4.0 (compatible; msie 4.0; windows 98)
mozilla/4.0 (compatible; msie 4.0; windows nt)
mozilla/4.0 (compatible; msie 4.0; windows xp)
mozilla/5
msiecrawler
naver
nerdybot
netants
netmechanic
nicerspro
nutch
offline explorer
openbot
openfind
openfind data gathere
oracle ultra search
perman
propowerbot/2.14
prowebwalker
psbot
python-urllib
queryn metasearch
radiation retriever 1.1
repomonkey
repomonkey bait & tackle/v1.01
rma
rogerbot
scooter
screaming frog seo spider
searchpreview
semrushbot
semrushbot
semrushbot-sa
seokicks-robot
sitesnagger
sootle
spankbot
spanner
spbot
stanford
stanford comp sci
stanford compclub
stanford compsciclub
stanford spiderboys
surveybot
surveybot_ignoreip
suzuran
szukacz/1.4
szukacz/1.4
teleport
teleportpro
telesoft
teoma
the intraformant
thenomad
tocrawl/urldispatcher
true_robot
true_robot/1.0
turingos
typhoeus
url control
url_spider_pro
urly warning
vci
vci webviewer vci webviewer win32
web image collector
webauto
webbandit
webbandit/3.50
webcopier
webenhancer
webmasterworld extractor
webmasterworldforumbot
websauger
website quester
webster pro
webstripper
webvac
webzip
webzip/4.0
wget
wget/1.5.3
wget/1.6
www-collector-e
xenu's
xenu's link sleuth 1.1c
zeus
zeus 32297 webster pro v2.9 win32
zeus link scout

Rule	Path
Disallow	/

Rule

Path

Disallow

Other Records

Field	Value
sitemap	/wp-sitemap.xml

Field

Value

sitemap

/wp-sitemap.xml

Comments

As a condition of accessing this website, you agree to abide by the following
content signals:
(a) If a content-signal = yes, you may collect content for the corresponding
use.
(b) If a content-signal = no, you may not collect content for the
corresponding use.
(c) If the website operator does not include a content signal for a
corresponding use, the website operator neither grants nor restricts
permission via content signal with respect to the corresponding use.
The content signals and their meanings are:
search: building a search index and providing search results (e.g., returning
hyperlinks and short excerpts from your website's contents). Search does not
include providing AI-generated search summaries.
ai-input: inputting content into one or more AI models (e.g., retrieval
augmented generation, grounding, or other real-time taking of content for
generative AI search answers).
ai-train: training or fine-tuning AI models.
ANY RESTRICTIONS EXPRESSED VIA CONTENT SIGNALS ARE EXPRESS RESERVATIONS OF
AND RELATED RIGHTS IN THE DIGITAL SINGLE MARKET.
BEGIN Cloudflare Managed content
END Cloudflare Managed Content

Warnings

`content-signal` is not a known field.

modlitewnik.plrobots.txt

Resource Scan

Scan Details

Last Scan

Groups

*

amazonbot

applebot-extended

bytespider

ccbot

claudebot

google-extended

gptbot

meta-externalagent

*

Other Records

bingbot

Other Records

yandex

Other Records

mj12bot

Other Records

Other Records

Comments

Warnings

modlitewnik.pl
robots.txt