findsalesrep.com
robots.txt

Robots Exclusion Standard data for findsalesrep.com

Archived Snapshots

Resource Scan

Scan Details

Site Domain	findsalesrep.com
Base Domain	findsalesrep.com
Scan Status	Ok
Last Scan	2024-06-25T14:59:00+00:00
Next Scan	2024-07-25T14:59:00+00:00

Last Scan

Scanned	2024-06-25T14:59:00+00:00
URL	https://findsalesrep.com/robots.txt
Redirect	https://www.findsalesrep.com/robots.txt
Redirect Domain	www.findsalesrep.com
Redirect Base	findsalesrep.com
Domain IPs	104.21.64.74, 172.67.178.117, 2606:4700:3032::ac43:b275, 2606:4700:3035::6815:404a
Redirect IPs	104.21.64.74, 172.67.178.117, 2606:4700:3032::ac43:b275, 2606:4700:3035::6815:404a
Response IP	104.21.64.74
Found	Yes
Hash	248d4d4e8a92159cbc2add9d5cc67e02a484399b03a99a909df12e7e2692e6a1
SimHash	38941d184f54

Groups

*

Rule	Path
Disallow	/includes/
Disallow	/misc/
Disallow	/modules/
Disallow	/profiles/
Disallow	/scripts/
Disallow	/r
Disallow	/r/
Disallow	/find
Disallow	/cart
Disallow	/cart/
Disallow	/CHANGELOG.txt
Disallow	/cron.php
Disallow	/INSTALL.mysql.txt
Disallow	/INSTALL.pgsql.txt
Disallow	/install.php
Disallow	/INSTALL.txt
Disallow	/LICENSE.txt
Disallow	/MAINTAINERS.txt
Disallow	/update.php
Disallow	/UPGRADE.txt
Disallow	/xmlrpc.php
Disallow	/admin/
Disallow	/comment/reply/
Disallow	/contact/
Disallow	/filter/tips/
Disallow	/logout/
Disallow	/node/add/
Disallow	/search/
Disallow	/user/register/
Disallow	/user/password/
Disallow	/user/login/
Disallow	/?q=admin%2F
Disallow	/?q=comment%2Freply%2F
Disallow	/?q=contact%2F
Disallow	/?q=filter%2Ftips%2F
Disallow	/?q=logout%2F
Disallow	/?q=node%2Fadd%2F
Disallow	/?q=search%2F
Disallow	/?q=user%2Fpassword%2F
Disallow	/?q=user%2Fregister%2F
Disallow	/?q=user%2Flogin%2F

Rule

Path

Disallow

/includes/

Disallow

/misc/

Disallow

/modules/

Disallow

/profiles/

Disallow

/scripts/

Disallow

/r/

Disallow

/find

Disallow

/cart

Disallow

/cart/

Disallow

/CHANGELOG.txt

Disallow

/cron.php

Disallow

/INSTALL.mysql.txt

Disallow

/INSTALL.pgsql.txt

Disallow

/install.php

Disallow

/INSTALL.txt

Disallow

/LICENSE.txt

Disallow

/MAINTAINERS.txt

Disallow

/update.php

Disallow

/UPGRADE.txt

Disallow

/xmlrpc.php

Disallow

/admin/

Disallow

/comment/reply/

Disallow

/contact/

Disallow

/filter/tips/

Disallow

/logout/

Disallow

/node/add/

Disallow

/search/

Disallow

/user/register/

Disallow

/user/password/

Disallow

/user/login/

Disallow

/?q=admin%2F

Disallow

/?q=comment%2Freply%2F

Disallow

/?q=contact%2F

Disallow

/?q=filter%2Ftips%2F

Disallow

/?q=logout%2F

Disallow

/?q=node%2Fadd%2F

Disallow

/?q=search%2F

Disallow

/?q=user%2Fpassword%2F

Disallow

/?q=user%2Fregister%2F

Disallow

/?q=user%2Flogin%2F

mediapartners-google

Rule	Path
Allow	/
Disallow	/admin/

Rule

Path

Allow

Disallow

/admin/

ia_archiver

Rule	Path
Disallow	/

Rule

Path

Disallow

baiduspider

Rule	Path
Disallow	/

Rule

Path

Disallow

httrack

Rule	Path
Disallow	/

Rule

Path

Disallow

yandex

Rule	Path
Disallow	/

Rule

Path

Disallow

bsalsa

Rule	Path
Disallow	/

Rule

Path

Disallow

blexbot

Rule	Path
Disallow	/

Rule

Path

Disallow

flamingo_searchengine

Rule	Path
Disallow	/

Rule

Path

Disallow

mauibot

Rule	Path
Disallow	/

Rule

Path

Disallow

semrushbot

Rule	Path
Disallow	/

Rule

Path

Disallow

ahrefsbot

Rule	Path
Disallow	/

Rule

Path

Disallow

mj12bot

Rule	Path
Disallow	/

Rule

Path

Disallow

bluechipbacklinks

Rule	Path
Disallow	/

Rule

Path

Disallow

garlikcrawler

Rule	Path
Disallow	/

Rule

Path

Disallow

sentibot

Rule	Path
Disallow	/

Rule

Path

Disallow

wadoo

Rule	Path
Disallow	/

Rule

Path

Disallow

Comments

$Id$
robots.txt
This file is to prevent the crawling and indexing of certain parts
of your site by web crawlers and spiders run by sites like Yahoo!
and Google. By telling these "robots" where not to go on your site,
you save bandwidth and server resources.
This file will be ignored unless it is at the root of your host:
Used: http://example.com/robots.txt
Ignored: http://example.com/site/robots.txt
For more information about the robots.txt standard, see:
http://www.robotstxt.org/wc/robots.html
For syntax checking, see:
http://www.sxw.org.uk/computing/robots/check.html
Directories
Files
Paths (clean URLs)
Paths (no clean URLs)

findsalesrep.comrobots.txt

Resource Scan

Scan Details

Last Scan

Groups

*

mediapartners-google

ia_archiver

baiduspider

httrack

yandex

bsalsa

blexbot

flamingo_searchengine

mauibot

semrushbot

ahrefsbot

mj12bot

bluechipbacklinks

garlikcrawler

sentibot

wadoo

Comments

findsalesrep.com
robots.txt