l-expert-comptable.com
robots.txt

Robots Exclusion Standard data for l-expert-comptable.com

Archived Snapshots

Resource Scan

Scan Details

Site Domain	l-expert-comptable.com
Base Domain	l-expert-comptable.com
Scan Status	Failed
Failure Stage	Fetching resource.
Failure Reason	Couldn't establish SSL connection.
Last Scan	2024-11-14T19:58:43+00:00
Next Scan	2025-02-12T19:58:43+00:00

Last Successful Scan

Scanned	2023-06-28T21:13:55+00:00
URL	https://l-expert-comptable.com/robots.txt
Redirect	https://www.l-expert-comptable.com/robots.txt
Redirect Domain	www.l-expert-comptable.com
Redirect Base	l-expert-comptable.com
Domain IPs	217.182.130.205
Redirect IPs	217.182.130.205
Response IP	217.182.130.205
Found	Yes
Hash	c42b13c53d0096babd4c37660be7148bf64f0d15986b11b48288b35a5c9c79f4
SimHash	7796b549cb40

Groups

baiduspider-favo
baiduspider-cpro
baiduspider-ads
baidu
baiduspider-news
baiduspider-video
baiduspider-image
baiduspider
blexbot
alexibot
alvinetspider
antenne hatena
apocalxexplorerbot
asterias
backdoorbot/1.0
bizinformation
black hole
blowfish/1.0
botalot
builtbottough
bullseye/1.0
bunnyslippers
cegbfeieh
cheesebot
cherrypicker
cherrypickerelite/1.0
cherrypickerse/1.0
copyrightcheck
cosmos
crescent
crescent internet toolpak http ole control v.1.0
disco pump 3.1
dittospyder
dotbot
emailcollector
emailsiphon
emailwolf
erocrawler
exabot
extractorpro
flamingo_searchengine
foobot
harvest/1.5
hloader
httplib
httrack
httrack 3.0
humanlinks
igentia
infonavirobot
jennybot
jikespider
kenjin spider
lexibot
libweb/clshttp
linkextractorpro
linkscan/8.1a unix
linkwalker
lwp-trivial
lwp-trivial/1.34
mata hari
microsoft url control - 5.01.4511
microsoft url control - 6.00.8169
miixpc
miixpc/4.2
mister pix
mlbot
moget
moget/2.1
ms search 4.0 robot
ms search 5.0 robot
naverbot
netants
netattache
netmechanic
nicerspro
offline explorer
openfind
openindexspider
propowerbot/2.14
prowebwalker
psbot
quepasacreep
queryn metasearch
repomonkey
rma
sightupbot
sitebot
sitesnagger
sitesucker
sogou web spider
sosospider
spankbot
spanner
speedy
suggybot
superbot
superbot/2.6
suzuran
szukacz/1.4
teleport
telesoft
the intraformant
thenomad
tighttwatbot
titan
tocrawl/urldispatcher
toscrawler
trendictionbot
true_robot
true_robot/1.0
turingos
turnitinbot
urlpouls
urly warning
vci
web image collector
webauto
webbandit
webbandit/3.50
webcopier
webcopy
webenhancer
webmasterworldforumbot
webmirror
webreaper
websauger
website extractor
website quester
webster pro
webstripper
webstripper/2.02
webzip
wget
wikiofeedbot
winhttrack
www-collector-e
xenu link sleuth/1.3.8
yacy
yandex
yrspider
zeus
zookabot

Rule	Path
Disallow	/

Rule

Path

Disallow

/

*

Rule	Path
Allow	/core/*.css$
Allow	/core/*.css?
Allow	/core/*.js$
Allow	/core/*.js?
Allow	/core/*.gif
Allow	/core/*.jpg
Allow	/core/*.jpeg
Allow	/core/*.png
Allow	/core/*.svg
Allow	/profiles/*.css$
Allow	/profiles/*.css?
Allow	/profiles/*.js$
Allow	/profiles/*.js?
Allow	/profiles/*.gif
Allow	/profiles/*.jpg
Allow	/profiles/*.jpeg
Allow	/profiles/*.png
Allow	/profiles/*.svg
Disallow	/core/
Disallow	/profiles/
Disallow	/README.txt
Disallow	/web.config
Disallow	/sites/default/files/modele/document/lettre-de-demission_128318051872464300.doc
Disallow	action%3D
Disallow	hubs_post
Disallow	?script
Disallow	?hsCta
Disallow	?utm_source
Disallow	?utm_campaign
Disallow	?__hstc
Disallow	?fbclid
Disallow	?q
Disallow	/search/node
Disallow	/admin/
Disallow	/comment/reply/
Disallow	/filter/tips
Disallow	/node/add/
Disallow	/search/
Disallow	/user/register/
Disallow	/user/password/
Disallow	/user/login/
Disallow	/user/logout/
Disallow	/index.php/admin/
Disallow	/index.php/comment/reply/
Disallow	/index.php/filter/tips
Disallow	/index.php/node/add/
Disallow	/index.php/search/
Disallow	/index.php/user/password/
Disallow	/index.php/user/register/
Disallow	/index.php/user/login/
Disallow	/index.php/user/logout/
Disallow	/recherche
Disallow	/recherche?
Disallow	/node/
Disallow	/article/
Disallow	/notion/
Disallow	/annuaire/
Disallow	/user/
Disallow	/besoins/
Disallow	/cibles/
Disallow	/localite/
Disallow	/notion/
Disallow	/mots-cles/
Disallow	/types-d039organisme/
Disallow	/interviews/
Disallow	/membre/
Disallow	/taxonomy/
Disallow	/node/
Disallow	/*.doc$
Disallow	/comment/
Disallow	/image-captcha
Disallow	sid%3D
Disallow	src%3D
Disallow	form_id%3D
Disallow	items_per_page%3D
Disallow	gclid%3D
Disallow	/d/
Disallow	/adresse/
Disallow	/expert-comptable/
Disallow	action%3D%22

Rule

Path

Allow

/core/*.css$

Allow

/core/*.css?

Allow

/core/*.js$

Allow

/core/*.js?

Allow

/core/*.gif

Allow

/core/*.jpg

Allow

/core/*.jpeg

Allow

/core/*.png

Allow

/core/*.svg

Allow

/profiles/*.css$

Allow

/profiles/*.css?

Allow

/profiles/*.js$

Allow

/profiles/*.js?

Allow

/profiles/*.gif

Allow

/profiles/*.jpg

Allow

/profiles/*.jpeg

Allow

/profiles/*.png

Allow

/profiles/*.svg

Disallow

/core/

Disallow

/profiles/

Disallow

/README.txt

Disallow

/web.config

Disallow

/sites/default/files/modele/document/lettre-de-demission_128318051872464300.doc

Disallow

*action%3D*

Disallow

*hubs_post*

Disallow

*?script*

Disallow

*?hsCta*

Disallow

*?utm_source*

Disallow

*?utm_campaign*

Disallow

*?__hstc*

Disallow

*?fbclid*

Disallow

*?q*

Disallow

/search/node

Disallow

/admin/

Disallow

/comment/reply/

Disallow

/filter/tips

Disallow

/node/add/

Disallow

/search/

Disallow

/user/register/

Disallow

/user/password/

Disallow

/user/login/

Disallow

/user/logout/

Disallow

/index.php/admin/

Disallow

/index.php/comment/reply/

Disallow

/index.php/filter/tips

Disallow

/index.php/node/add/

Disallow

/index.php/search/

Disallow

/index.php/user/password/

Disallow

/index.php/user/register/

Disallow

/index.php/user/login/

Disallow

/index.php/user/logout/

Disallow

/recherche

Disallow

/recherche?

Disallow

/node/

Disallow

/article/

Disallow

/notion/

Disallow

/annuaire/

Disallow

/user/

Disallow

/besoins/

Disallow

/cibles/

Disallow

/localite/

Disallow

/notion/

Disallow

/mots-cles/

Disallow

/types-d039organisme/

Disallow

/interviews/

Disallow

/membre/

Disallow

/taxonomy/

Disallow

/node/

Disallow

/*.doc$

Disallow

/comment/

Disallow

/image-captcha

Disallow

*sid%3D*

Disallow

*src%3D*

Disallow

*form_id%3D*

Disallow

*items_per_page%3D*

Disallow

*gclid%3D*

Disallow

/d/

Disallow

/adresse/

Disallow

/expert-comptable/

Disallow

*action%3D*%22

chatgpt-user

Rule	Path
Disallow	/
Allow	/sitemap.xml?page=1
Allow	/sitemap.xml?page=2
Allow	/sitemap.xml?page=3

Rule

Path

Disallow

/

Allow

/sitemap.xml?page=1

Allow

/sitemap.xml?page=2

Allow

/sitemap.xml?page=3

Back to top

Other Records

Field	Value
sitemap	https://www.l-expert-comptable.com/sitemap.xml

Field

Value

sitemap

https://www.l-expert-comptable.com/sitemap.xml

Back to top

Comments

robots.txt
This file is to prevent the crawling and indexing of certain parts
of your site by web crawlers and spiders run by sites like Yahoo!
and Google. By telling these "robots" where not to go on your site,
you save bandwidth and server resources.
This file will be ignored unless it is at the root of your host:
Used: http://example.com/robots.txt
Ignored: http://example.com/site/robots.txt
For more information about the robots.txt standard, see:
http://www.robotstxt.org/robotstxt.html
CSS, JS, Images
Directories
Files
Paths (clean URLs)
Paths (no clean URLs)
ExpertComptable
JRB 2016-08-19
ChatGPT

Back to top

Warnings

1 invalid line.

Back to top

l-expert-comptable.comrobots.txt

Resource Scan

Scan Details

Last Successful Scan

Groups

*

chatgpt-user

Other Records

Comments

Warnings

l-expert-comptable.com
robots.txt