Log file crawl budget analysis
Use when server logs are available and you want to know where crawl budget is going.
Fill in before running
Replace each placeholder with your own detail. The more specific you are, the less the model invents.
- {{SITE_URL}}
- {{LOG_SUMMARY}}
- {{SITE_STRUCTURE}}
- {{INDEXABLE_COUNT}}
Getting a better result
- Aggregate by directory or pattern before pasting - raw log lines waste the context window.
- Use at least 30 days of logs so weekly crawl cycles do not distort the picture.
- Pair this with an indexation report to see which under-crawled sections are actually missing.
Questions about this prompt
When is log analysis worth doing?
On large sites where indexation is incomplete and you cannot see why. On a small site crawl budget is rarely the constraint, and the effort is better spent elsewhere. The signal you are looking for is Google spending its time on pages that do not matter.
How much log data do I need?
At least thirty days, aggregated by directory or pattern before you paste it. Shorter windows let weekly crawl cycles distort the picture, and raw log lines waste the context window without adding anything the aggregate does not show.
What does a bad result look like?
Crawl concentrated on parameter URLs, pagination or filtered pages while your commercial templates are visited rarely. That is budget being spent on pages you never wanted indexed, and it is usually fixable with the faceted navigation rules rather than with more content.
What should I pair it with?
An indexation report. Logs tell you what was crawled and indexation tells you what made it in. Under-crawled sections that are also missing from the index are the priority; under-crawled sections that are indexed fine are not a problem.