Bot traffic audit
I audit your website from the outside. The report shows how to protect your data from scrapers, how to make the site show up in AI-generated answers, and where you might profit from your data.
Automated traffic is no longer just background noise. Some bots help people find a business. Others collect its content, compare its products, train AI systems, or act on behalf of potential customers.
The right response is not to block them all. A useful strategy decides which bots to welcome, which to limit, which to charge, and which to keep out.
Crawlers, scrapers, and why the difference matters
Search engines such as Google use crawlers to discover pages by following links and building an index of what they find. Scrapers have a different goal: they extract selected information such as prices, listings, product details, images, or contact information for analysis or reuse.
The same automated system may crawl a site to find its pages and then scrape the information it needs. This is why the technical shape of a request does not tell the whole story. What matters is who is making it, what they collect, and what they do with the result.
From SEO to GEO
Search engine optimization, or SEO, helps search engines understand a website and helps people find it through organic search results. Legitimate search crawlers need access to the public parts of the site, so blocking every bot can make a business harder to find.
At the same time, organic traffic from search is declining for many publishers. More questions are being answered before a person ever reaches the source website. Generative Engine Optimization, or GEO, aims to make content easier for AI systems to find, understand, cite, and include in their answers.
Not every bot should be blocked
A search crawler indexing public pages is different from a scraper copying an entire database. An assistant gathering information for a potential customer is different again. The practical goal is control: identify the visitor where possible, understand its purpose, and decide whether to allow, limit, charge, or block it.
When scraping becomes harmful
Uncontrolled scraping can help a competitor reproduce years of valuable work without making the same investment. The goal is rarely to make extraction impossible. It is to make unauthorized extraction unreliable, expensive, or no longer worthwhile. For the technical version of that argument, see a field guide to how scrapers map a website, find the underlying data, and keep collection economical.
Charging AI crawlers for access
Pay Per Crawl and licensed data feeds can turn unwanted demand into recurring income while giving the owner control over available fields, update frequency, request volume, and permitted uses.
What the audit covers
- Only publicly available information is examined; no access to server-side code is required
- Public pages, files, and interfaces exposed to automated visitors are mapped
- robots.txt, sitemaps, indexing directives, and existing bot controls are reviewed
- Legitimate crawlers are checked for reach to the content intended for them
- A controlled proof of concept extracts representative data
- Clarity for automated systems that understand and cite the content is assessed
- Risks, opportunities, and recommended improvements are documented in priority order
info@bakosbence.com