isyou-43.com
robots.txt

Robots Exclusion Standard data for isyou-43.com

Resource Scan

Scan Details

Site Domain isyou-43.com
Base Domain isyou-43.com
Scan Status Ok
Last Scan2024-10-27T23:17:09+00:00
Next Scan 2024-11-26T23:17:09+00:00

Last Scan

Scanned2024-10-27T23:17:09+00:00
URL https://isyou-43.com/robots.txt
Redirect https://isyou-44.com/robots.txt
Redirect Domain isyou-44.com
Redirect Base isyou-44.com
Domain IPs 104.26.6.2, 104.26.7.2, 172.67.70.182, 2606:4700:20::681a:602, 2606:4700:20::681a:702, 2606:4700:20::ac43:46b6
Redirect IPs 198.143.133.98
Response IP 198.143.133.98
Found Yes
Hash 96a0bb869452a6c3d47f42fde4ab553abebfe2623146782b198fbe9ece1417e5
SimHash 6304df43803b

Groups

*

Rule Path
Allow /wp-admin/admin-ajax.php
Allow /*/*.css
Allow /*/*.js
Allow .js
Allow /wp-includes/js/jquery/jquery.js
Disallow /wp-admin/
Disallow /wp-includes/
Disallow /readme.html
Disallow /license.txt
Disallow /xmlrpc.php
Disallow /wp-login.php
Disallow /wp-register.php
Disallow */disclaimer/*
Disallow *?attachment_id=
Disallow /privacy-policy

googlebot

Rule Path
Allow /

googlebot-image

Rule Path
Allow /wp-content/uploads/

mediapartners-google

Rule Path
Allow /

adsbot-google

Rule Path
Allow /

adsbot-google-mobile

Rule Path
Allow /

bingbot

Rule Path
Allow /

msnbot

Rule Path
Allow /

msnbot-media

Rule Path
Allow /wp-content/uploads/

applebot

Rule Path
Allow /

yandex

Rule Path
Allow /

yandeximages

Rule Path
Allow /wp-content/uploads/

slurp

Rule Path
Allow /

duckduckbot

Rule Path
Allow /

qwantify

Rule Path
Allow /

baiduspider

Rule Path
Allow /

baiduspider/2.0

Rule Path
Allow /

baiduspider-video

Rule Path
Allow /

baiduspider-image

Rule Path
Allow /

sogou spider

Rule Path
Allow /

sogou web spider

Rule Path
Allow /

sosospider

Rule Path
Allow /

sosospider+

Rule Path
Allow /

sosospider/2.0

Rule Path
Allow /

yodao

Rule Path
Allow /

youdao

Rule Path
Allow /

youdaobot

Rule Path
Allow /

youdaobot/1.0

Rule Path
Allow /

ia_archiver

Rule Path
Disallow /

archive.org_bot

Rule Path
Disallow /

siteexplorer

Rule Path
Disallow /

spbot

Rule Path
Disallow /

wbsearchbot

Rule Path
Disallow /

linkdexbot

Rule Path
Disallow /

screaming frog seo spider

Rule Path
Disallow /

netestate ne crawler

Rule Path
Disallow /

moreover

Rule Path
Disallow /

sentibot

Rule Path
Disallow /

aboundexbot

Rule Path
Disallow /

proximic

Rule Path
Disallow /

obot

Rule Path
Disallow /

meanpathbot

Rule Path
Disallow /

nutch

Rule Path
Disallow /

turnitinbot

Rule Path
Disallow /

zoominfobot

Rule Path
Disallow /

zmeu

Rule Path
Disallow /

grapeshot

Rule Path
Disallow /

python-requests

Rule Path
Disallow /

go-http-client

Rule Path
Disallow /

apache-httpclient

Rule Path
Disallow /

libwww-perl

Rule Path
Disallow /

curl

Rule Path
Disallow /

wget

Rule Path
Disallow /

gptbot

Rule Path
Disallow /

*

Rule Path
Allow /ads.txt

*

Rule Path
Allow /app-ads.txt

Other Records

Field Value
crawl-delay 5

Comments

  • Popular chinese search engines
  • Block Bad Bots. "AI recommended setting" by ChatGPT
  • ChatGPT Bot Blocker - Block ChatGPT Bot from scrapping your content
  • Allow/Disallow Ads.txt
  • Allow/Disallow App-ads.txt
  • This robots.txt file was created by Better Robots.txt (Index & Rank Booster by Pagup) Plugin. https://www.better-robots.com/