Internal links tool

LivvuxBot

LivvuxBot is the on-demand crawler for this website’s internal link audit. It runs only when someone submits a website and their own TypeSafe API key. It does not continuously index the web.

LivvuxBot/1.0 (+https://livvux.dev/tools/crawler)

Crawl behaviour

The crawler reads robots.txt before HTML pages or sitemaps. It honours matching Allow, Disallow and Crawl-delay rules. Requests start at least 200 ms apart per origin unless a longer Crawl-delay is set. At most four pages are fetched concurrently.

Only public HTTP(S) URLs on default ports are supported. Crawling stays on the entered hostname and its www alias. Redirect destinations are validated, including DNS and robots rules. Private networks, login credentials in URLs, non-HTML pages, and JavaScript rendering are not supported.

Block this crawler

User-agent: LivvuxBot
Disallow: /

A 401 or 403 from robots.txt blocks the origin. Network errors and server errors stop crawling rather than bypassing rules. A missing robots.txt (404 or 410) is treated as having no rules.

Firewall troubleshooting

Check your firewall logs for the user-agent above. There is no fixed IP allowlist published here. Do not disable your firewall or trust a user-agent alone: user-agent strings can be spoofed.

GitHub
LinkedIn
X
youtube