Machine Readiness
Stored receipt and evidence
20
65
0
0
0
Samples
No stored offer samples.
Samples
No stored action samples.
Samples
No stored product samples.
Document
# As a condition of accessing this website, you agree to abide by the following # content signals: # (a) If a Content-Signal = yes, you may collect content for the corresponding # use. # (b) If a Content-Signal = no, you may not collect content for the # corresponding use. # (c) If the website operator does not include a Content-Signal for a # corresponding use, the website operator neither grants nor restricts # permission via Content-Signal with respect to the corresponding use. # The content signals and their meanings are: # search: building a search index and providing search results (e.g., returning # hyperlinks and short excerpts from your website's contents). Search does not # include providing AI-generated search summaries. # ai-input: inputting content into one or more AI models (e.g., retrieval # augmented generation, grounding, or other real-time taking of content for # generative AI search answers). # ai-train: training or fine-tuning AI models. # ANY RESTRICTIONS EXPRESSED VIA CONTENT SIGNALS ARE EXPRESS RESERVATIONS OF # RIGHTS UNDER ARTICLE 4 OF THE EUROPEAN UNION DIRECTIVE 2019/790 ON COPYRIGHT # AND RELATED RIGHTS IN THE DIGITAL SINGLE MARKET. # BEGIN Cloudflare Managed content User-agent: * Content-Signal: search=yes,ai-train=no Allow: / User-agent: Amazonbot Disallow: / User-agent: Applebot-Extended Disallow: / User-agent: Bytespider Disallow: / User-agent: CCBot Disallow: / User-agent: ClaudeBot Disallow: / User-agent: CloudflareBrowserRenderingCrawler Disallow: / User-agent: Google-Extended Disallow: / User-agent: GPTBot Disallow: / User-agent: meta-externalagent Disallow: / # END Cloudflare Managed Content # Bloqueo basico para todos los bots y crawlers # puede dar problemas por bloqueo de recursos en GWT User-agent: * Allow: /wp-content/uploads/* Allow: /wp-content/*.js Allow: /wp-content/*.css Allow: /wp-includes/*.js Allow: /wp-includes/*.css Disallow: /cgi-bin Disallow: /wp-content/plugins/ Disallow: /wp-content/themes/ Disallow: /wp-includes/ Disallow: /*/attachment/ Disallow: /tag/*/page/ Disallow: /tag/*/feed/ Disallow: /page/ Disallow: /comments/ Disallow: /xmlrpc.php Disallow: /?attachment_id* # Bloqueo de las URL dinamicas #Disallow: /*? #Bloqueo de busquedas #Disallow: /?s= #Disallow: /search # Bloqueo de trackbacks User-agent: * Disallow: /trackback Disallow: /*trackback Disallow: /*trackback* Disallow: /*/trackback # Bloqueo de feeds para crawlers User-agent: * Allow: /feed/$ Disallow: /feed/ Disallow: /comments/feed/ Disallow: /*/feed/$ Disallow: /*/feed/rss/$ Disallow: /*/trackback/$ Disallow: /*/*/feed/$ Disallow: /*/*/feed/rss/$ Disallow: /*/*/trackback/$ Disallow: /*/*/*/feed/$ Disallow: /*/*/*/feed/rss/$ Disallow: /*/*/*/trackback/$ # Ralentizamos algunos bots que se suelen volver locos User-agent: noxtrumbot Crawl-delay: 20 User-agent: msnbot Crawl-delay: 20 User-agent: Slurp Crawl-delay: 20 # Bloqueo de bots y crawlers poco utiles User-agent: MSIECrawler Disallow: / User-agent: WebCopier Disallow: / User-agent: HTTrack Disallow: / User-agent: Microsoft.URL.Control Disallow: / User-agent: libwww Disallow: / User-agent: Orthogaffe Disallow: / User-agent: UbiCrawler Disallow: / User-agent: DOC Disallow: / User-agent: Zao Disallow: / User-agent: sitecheck.internetseer.com Disallow: / User-agent: Zealbot Disallow: / User-agent: MSIECrawler Disallow: / User-agent: SiteSnagger Disallow: / User-agent: WebStripper Disallow: / User-agent: WebCopier Disallow: / User-agent: Fetch Disallow: / User-agent: Offline Explorer Disallow: / User-agent: Teleport Disallow: / User-agent: TeleportPro Disallow: / User-agent: WebZIP Disallow: / User-agent: linko Disallow: / User-agent: HTTrack Disallow: / User-agent: Microsoft.URL.Control Disallow: / User-agent: Xenu Disallow: / User-agent: larbin Disallow: / User-agent: libwww Disallow: / User-agent: ZyBORG Disallow: / User-agent: Download Ninja Disallow: / User-agent: wget Disallow: / User-agent: grub-client Disallow: / User-agent: k2spider Disallow: / User-agent: NPBot Disallow: / User-agent: WebReaper Disallow: / User-agent: Mediapartners-Google Allow: / User-agent: Googlebot Allow: / User-agent: Googlebot-News Allow: / User-agent: Googlebot-Image Allow: / User-agent: Googlebot-Video Allow: / User-agent: AdsBot-Google Allow: / #Evitar IAs User-agent: Amazonbot Disallow: / User-agent: magpie-crawler Disallow: / User-agent: CCBot Disallow: / User-Agent: omgili Disallow: / User-Agent: omgilibot Disallow: / User-agent: Claude-Web Disallow: / User-agent: ClaudeBot Disallow: / User-agent: anthropic-ai Disallow: / User-agent: cohere-ai Disallow: / User-agent: Bytespider Disallow: / User-agent: PetalBot Disallow: / User-agent: Scrapy Disallow: / User-agent: Applebot-Extended Disallow: / User-agent: GPTBot Disallow: / User-agent: ChatGPT-User Disallow: / User-agent: Google-Extended Disallow: / User-Agent: PerplexityBot Disallow: / User-agent: Perplexity-User Disallow: / User-agent: Google-CloudVertexBot Disallow: / User-agent: meta-externalagent Disallow: / User-agent: OAI-SearchBot Disallow: / User-agent: YandexAdditional Disallow: / User-agent: YandexAdditionalBot Disallow: / User-agent: TurnitinBot Disallow: / # Previene problemas de recursos bloqueados en Google Webmaster Tools User-Agent: Googlebot Allow: /*.css$ Allow: /*.js$ # En condiciones normales este es el sitemap Sitemap: https://cultoro.es/sitemap_index.xml
Document
# Cultoro\.es: Noticias de toros y toreros, ferias, cultura y campo bravo > Las últimas noticias de toros\. Actualidad, reportajes, noticias de toros, toreros, festejos, plazas, ganaderías, campo bravo y televisión\. Generated by Yoast SEO v27.4, this is an llms.txt file, meant for consumption by LLMs. ## Páginas - [Noticias de toros \- Toda la actualidad taurina](https://cultoro.es/) - [TOROS EN TELEVISIÓN](https://cultoro.es/toros-en-television-agenda) - [Feria de Abril 2026](https://cultoro.es/ferias/feria-de-abril-2026) - [Arles Pascua 2026](https://cultoro.es/ferias/arles-arroz-2026) - [Fallas Valencia 2026](https://cultoro.es/ferias/fallas-valencia-2026) ## Entradas - [El quite con el que Roca Rey sorprendió en Illescas: mezcló el 'puente de la muerte' con chicuelinas y gaoneras](https://cultoro.es/actualidad/roca-rey-quite-illescas-146635) - [Un toro recién enfundado la 'emprende' contra una cámara de Canal Extremadura y acaba con los pitones 'enredados'](https://cultoro.es/actualidad/toro-fundas-canal-extremadura-145394) - [El naturalísimo concepto de Aguado, en una personal obra al tercero en la que acabó cortándose con el acero](https://cultoro.es/actualidad/naturalisimo-concepto-de-aguado-376435) - [La capacidad y el pundonor de Luque, en una faena de premio al cuarto](https://cultoro.es/actualidad/capacidad-pundonor-luque-oreja-376453) - [Juan Pedro chafa el Viernes de Farolillos: la oreja para la fe de Luque, oasis de un festejo con dos bellas faenas de Ortega y Aguado](https://cultoro.es/festejos/avance-sevilla-juan-pedro-376388) ## Categorías - [ACTUALIDAD](https://cultoro.es/actualidad) - [FESTEJOS](https://cultoro.es/festejos) - [Televisión](https://cultoro.es/actualidad/television) - [Ferias](https://cultoro.es/ferias) - [Carteles](https://cultoro.es/actualidad/carteles) ## Etiquetas - [Portada](https://cultoro.es/temas/portada) - [toros](https://cultoro.es/temas/toros) - [Toreros](https://cultoro.es/temas/toreros) - [cultoro](https://cultoro.es/temas/cultoro) - [temporada](https://cultoro.es/temas/temporada) ## Optional - [Sitemap index](https://cultoro.es/sitemap_index.xml)
Document
Not stored for this site.