Network
Generate an llms.txt file (llmstxt.org v2 format: H1, blockquote summary, H2 file lists) from a live site crawl or a pasted page inventory, and audit robots.txt policies for GPTBot, ClaudeBot, PerplexityBot and Bytespider using the RFC 9309 matching algorithm — with allow/block previews and ready-to-paste directives.
Call this tool from your code in three languages.
curl -X POST 'http://127.0.0.1:3003/en/api/tools/llms-txt-and-ai-crawler-audit' \
-H 'Content-Type: application/json' \
-d '{"siteUrl":"https://example.com — homepage links, plus sitemap.xml fallback for JS-rendered sites (max 25 pages)","pagesInput":"\nhttps://example.com/docs/quickstart | Quickstart | Install the CLI and run your first job\nhttps://example.com/docs/api | API Reference | Endpoint catalog with auth examples\nhttps://example.com/blog/cli-tips | CLI Tips | Twenty shortcuts for daily use\nhttps://example.com/blog/release-2026 | Release 2026 | What changed this quarter","robotsInput":"User-agent: GPTBot\nDisallow: /\n\nUser-agent: ClaudeBot\nDisallow:\n\nUser-agent: *\nDisallow: /private/","siteName":"Example Platform Docs","summary":"Example Platform lets teams ship scheduled data jobs without infrastructure.","extraNotes":"All endpoints require an API key; the quickstart covers sandbox setup.","grouping":"auto"}'Send a POST request with your inputs as JSON. File parameters require a separate upload first.
POST http://127.0.0.1:3003/en/api/tools/llms-txt-and-ai-crawler-audit| Name | Type | Required | Description |
|---|---|---|---|
| siteUrl | text | No | — |
| pagesInput | textarea | No | — |
| robotsInput | textarea | No | — |
| siteName | text | No | — |
| summary | text | No | — |
| extraNotes | textarea | No | — |
| grouping | select | Yes | — |
HTML result
{
"result": "<div>Processed HTML content</div>",
"error": "Error message (optional)",
"message": "Notification message (optional)",
"metadata": {
"key": "value"
}
}Add this tool to your Model Context Protocol server so AI agents can list and call it.
Add this block to your MCP client configuration:
{
"mcpServers": {
"elysiatools-llms-txt-and-ai-crawler-audit": {
"name": "llms-txt-and-ai-crawler-audit",
"description": "Generate an llms.txt file (llmstxt.org v2 format: H1, blockquote summary, H2 file lists) from a live site crawl or a pasted page inventory, and audit robots.txt policies for GPTBot, ClaudeBot, PerplexityBot and Bytespider using the RFC 9309 matching algorithm — with allow/block previews and ready-to-paste directives.",
"baseUrl": "http://127.0.0.1:3003/mcp/sse?toolId=llms-txt-and-ai-crawler-audit",
"command": "",
"args": [],
"env": {},
"isActive": true,
"type": "sse"
}
}
}After connecting to the SSE endpoint, list the exposed tools:
{
"jsonrpc": "2.0",
"id": 1,
"method": "tools/list"
}Invoke the tool by its id, passing arguments built from its parameters:
{
"jsonrpc": "2.0",
"id": 2,
"method": "tools/call",
"params": {
"name": "llms-txt-and-ai-crawler-audit",
"arguments": {
"siteUrl": "https://example.com — homepage links, plus sitemap.xml fallback for JS-rendered sites (max 25 pages)",
"pagesInput": "\nhttps://example.com/docs/quickstart | Quickstart | Install the CLI and run your first job\nhttps://example.com/docs/api | API Reference | Endpoint catalog with auth examples\nhttps://example.com/blog/cli-tips | CLI Tips | Twenty shortcuts for daily use\nhttps://example.com/blog/release-2026 | Release 2026 | What changed this quarter",
"robotsInput": "User-agent: GPTBot\nDisallow: /\n\nUser-agent: ClaudeBot\nDisallow:\n\nUser-agent: *\nDisallow: /private/",
"siteName": "Example Platform Docs",
"summary": "Example Platform lets teams ship scheduled data jobs without infrastructure.",
"extraNotes": "All endpoints require an API key; the quickstart covers sandbox setup.",
"grouping": "auto"
}
}
}Questions or issues? Contact [email protected]