#Robots.txt
Robots.txt is a plain text file at the root of a website that tells crawlers which paths they may request. It is part of the Robots Exclusion Protocol, which is now documented as an internet standard. The file controls crawling, not indexing, so a blocked URL can still appear in search results if other pages link to it. Articles under this tag explain how to read, write and test the file for a blog or content site. They also cover the common mistakes that hide pages from search engines without any visible error.
30 вер. 2026 р. · Час читання: 9 хвRobots.txt for a Blog: What to Allow and What to Block
What robots.txt does on a blog, which paths are worth blocking, which must stay open, and the mistakes that quietly hide posts from search…
27 вер. 2026 р. · Час читання: 9 хвCrawl Budget: Does It Matter for a Small Blog?
Crawl budget is rarely the problem for a blog with a few hundred posts. What it is, when it starts to matter, and the…