Novi – strona główna

Novi / Knowledge base A–Z

Technology

Robots.txt

A file with instructions for search engine crawlers about which parts of the site they may visit.

Robots.txt can, for example, exclude the admin area from indexing. A wrong configuration can accidentally block the whole site from Google, so it requires care.

What it means in practice

Robots.txt is a simple text file in a website’s root (yoursite.pl/robots.txt) telling search engine and AI crawlers which parts of the site they may visit and which to skip. It usually also contains the sitemap address.

Why it matters for your business

A well-configured robots.txt directs crawlers to important content and protects e.g. the admin area. One wrong line (“Disallow: /”) can, however, hide the whole website from Google.

Examples and good practice

  • Block only technical areas, e.g. /wp-admin/.
  • Include your XML sitemap address.
  • Don’t block CSS and JS files needed to render the page.
  • Decide consciously about access for AI crawlers.

How Novi handles it

Novi websites have a correctly configured robots.txt with a sitemap and access for search engines and AI crawlers. Test environments are blocked so they don’t end up in Google.

Tip for Novi online advisors

Related terms

SSL Firewall (WAF) Uptime WordPress Cookie banner WebP Image optimisation CDN

All of this comes as standard with a Novi website.

A complete Novi website: PLN 1,995. And if you enjoy talking about things like this with business owners — join the Novi team of online advisors.