#
# robots.txt
#
# This file is to prevent the crawling and indexing of certain parts
# of your site by web crawlers and spiders run by sites like Yahoo!
# and Google. By telling these "robots" where not to go on your site,
# you save bandwidth and server resources.
#
# This file will be ignored unless it is at the root of your host:
# Used: http://example.com/robots.txt
# Ignored: http://example.com/site/robots.txt
#
# For more information about the robots.txt standard, see:
# http://www.robotstxt.org/robotstxt.html
User-agent: *
# Files
Disallow: /*.rss
Disallow: /*.atom
Disallow: /espace-debat/*
Disallow: /redis
Disallow: /comments
Disallow: /shareCount
Disallow: /automobile/mag
Disallow: /*.asp
Disallow: /*/breve.html
Disallow: /*/article.html
Disallow: /*/article_p*.html
Disallow: /*.php
Disallow: /search/*
Disallow: /*/commentaires
Disallow: /user.json*
# Allow sitemap access before blocking rankings
Allow: /rankings/sitemap.xml
Disallow: /rankings/*
Disallow: /auth/*
Disallow: /partners/header
Disallow: /partners/footer
Sitemap: https://www.challenges.fr/sitemap.news.xml
Sitemap: https://www.challenges.fr/sitemap.xml
Sitemap: https://www.challenges.fr/rankings/sitemap.xml
User-agent: Riddler
Disallow: /
User-agent: GPTBot
Disallow: /
User-agent: ChatGPT-User
Disallow: /
User-agent: CCBot
Disallow: /
User-agent: Google-Extended
Disallow: /
User-Agent: omgilibot
Disallow: /
User-Agent: omgili
Disallow: /
User-Agent: Claude-Web
Disallow: /
User-Agent: Claude-SearchBot
Disallow: /
User-Agent: ClaudeBot
Disallow: /
User-Agent: claudebot
Disallow: /
User-agent: *
Disallow: /_Incapsula_Resource?*