Robots.txt setup
We set up robots.txt: a file controlling which site sections crawlers should crawl and which not. Honestly and bluntly upfront: robots.txt is important technical hygiene, but it is a file with a HIGH risk of error: one wrong line can accidentally close the whole site from crawling and crash traffic. And an important honesty nuance: robots.txt controls CRAWLING, not indexing directly — a page disallowed in robots can still get into the index via external links (other tools are needed to forbid indexing). robots.txt does not raise positions — it is hygiene, not a boost. We will set it up carefully and honestly, understanding the cost of an error.
Robots.txt setup — overview

Robots.txt is a text file in the site root that gives crawlers instructions on which paths to crawl and which not (Disallow/Allow), and also points to the sitemap. It is used to not waste crawl budget on service/unneeded sections, close junk parameters, manage crawling. Setup is correct rules for the site structure without harmful errors. Honestly about 'high risk of error', this is key: robots.txt is a powerful and dangerous tool. One wrong directive (e.g. Disallow: /) can close the whole site from crawling, and pages will drop out of search, crashing traffic. This is a frequent and painful mistake. So robots.txt requires care and checking, not edits 'on the fly'. We treat it with this caution. Honestly about 'crawling, not indexing', this is an important nuance: a common misconception is that Disallow in robots.txt removes a page from the index. In fact robots.txt controls CRAWLING: it tells the robot not to visit the page. But if there are external links to a disallowed page, it can still get into the index (without content, but in results). To genuinely forbid indexing, a noindex meta tag or other means are needed — and for this the page must, on the contrary, be open for crawling. We honestly explain this difference so you do not close it with the 'wrong' tool. Honestly about 'does not raise positions': robots.txt is crawl hygiene, not a ranking factor. It does not raise positions; correct setup only helps robots spend resource on what is needed. Honestly about the effect: it gives correct crawl management (crawl-budget savings, cutting off junk) without harmful errors, but it is technical hygiene, not indexing directly and not position growth. Honestly about access: access to the site root is needed. An important boundary: this is robots.txt (crawling); forbidding indexing (noindex) — different; robots for AI bots — 383; sitemap — 381; crawl budget — 394; indexing — 396. The base price starts from 30,000 ₽ (depends on rule complexity).
Problems we solve
- Robots waste crawl budget on service/junk sections.
- You need to manage site crawling but fear making a mistake (risk of closing everything).
- A misconception that robots.txt removes pages from the index (it is about crawling).
- An erroneous robots.txt already closed needed sections or the whole site from crawling.
What's included in the Robots.txt setup service
- Setting up robots.txt for the site structure (Disallow/Allow, sitemap pointer)
- Careful rules without harmful errors (protection against accidentally closing the site)
- Explaining the difference: crawling (robots) vs indexing (noindex)
- Cutting off service/junk paths to save crawl budget
- Correctness check (that the needed is open, the unneeded is closed)
- Honest boundaries (high risk of error — the cost is great; controls crawling, not indexing directly; does NOT raise positions; the disallowed can be indexed via links)
- A link with the sitemap (381) and crawl budget (394)
- Documentation and handover
What you get
- A correct robots.txt: the needed is open, junk is closed, the sitemap is pointed to
- Crawl-budget savings without harmful errors
- An honest understanding: crawling vs indexing (what to close with what)
- Honest boundaries (risk of error; not indexing directly; not position growth)
How the work goes: steps
- We analyze the structure and current robots.txt, identify risks
- We set up careful rules, check that the needed is accessible
- We honestly set boundaries (crawling vs indexing) and hand over to you
Why PDV Expert
- Fixed price and timeline — no surprises on the invoice.
- Report and recommendations in plain language — clear without a technical background.
- In touch at every step and answering questions about the result.
FAQ
If I disallow a page in robots.txt, will it disappear from search?
Not necessarily, honestly, and this is a common misconception: robots.txt controls CRAWLING (tells the robot not to visit the page), not indexing directly. If there are external links to a disallowed page, it can still get into the index (in results, without content). To genuinely forbid indexing, a noindex meta tag is needed — and for this the page must, on the contrary, be open for crawling. We honestly explain this difference so you do not close it with the 'wrong' tool.
Will robots.txt help me rise in results?
No, honestly: robots.txt is crawl hygiene, not a ranking factor. It does not raise positions; correct setup only helps robots not waste resource on the unneeded (crawl-budget savings). Confusing crawl management with position growth is a mistake. We honestly set it up as technical hygiene, not a top-growth tool.
It is just a simple file, what is the risk?
The risk is high, honestly: one wrong line (e.g. Disallow: /) can close the whole site from crawling, pages drop out of search and traffic crashes. This is a frequent and painful mistake. So robots.txt requires care, checking and an understanding of consequences, not edits 'on the fly'. We treat it with caution and verify that the needed is open and only the excess is closed.
About the provider
The «Robots.txt setup» service is provided by PDV Expert — a team specialising in «Content & acquisition». We work under contract and deliver a written report with recommendations.