scanned Jun 2, 2026

Sitedex

sitedex.dev

Sitedex audits and scores websites based on how well they answer questions that AI agents ask, providing a fix list to improve the site's AI-readiness.

84/100

Tier 5 · Agent-Native

Content answers80/100
Protocol plumbing94/10015 of 16 checks pass

Scored by asking 15 questions a buyer of a ai-ml product asks, then grading this site’s own pages: answered, hedged (partial or vague), or silent (no page answers it). How scoring works

The fix queue

16 points sit between sitedex.dev and 100: 5 open questions and 1 missing protocol check, ordered by estimated payoff.

Point estimates are per fix under scoring v2. They are not additive to a promised total.

01migration · importance lowGoes silent+7 content pts est.

I'm currently using a different AI-readiness tool — can I import my historical audit data to maintain trend continuity, or do I start fresh with your scoring methodology?

What the pages say

No page on the site addresses this.

The fix

Add a FAQ or documentation page explaining whether users can import historical audit data from other AI-readiness tools, or clarify that they must start fresh with Sitedex's scoring methodology.

confidence high · grounding world-knowledge · weight 0.00 · Absent

02support · importance lowGoes silent+7 content pts est.

If my MCP endpoint goes down during a critical product launch, what's your guaranteed response time for enterprise customers versus the standard plan?

What the pages say

No page on the site addresses this.

The fix

Publish a dedicated Support SLA or status page that specifies guaranteed response times for enterprise and standard plans during critical incidents.

confidence high · grounding world-knowledge · weight 0.00 · Absent

Hedged · 3 of 15 questions

A buyer gets something, then has to guess the rest. Expand any row for the evidence and the fix.

03If I delete a site from my dashboard, do you purge all historical audit data immediately or is there a retention period where I could still recover past scores?operationsHedges+2 pts est.

If I delete a site from my dashboard, do you purge all historical audit data immediately or is there a retention period where I could still recover past scores?

What the pages say

Sitedex says it keeps audit data 'only as long as it's useful for running the service,' and that you can email them to have something specific removed. The chunks do not say what happens to historical audit data when you delete a site from your dashboard, nor do they mention a retention period or recovery window after such a deletion.

The fix

Add a help article or FAQ entry that explicitly states what happens to historical audit data when a site is deleted from the dashboard, including whether there is a grace period for recovery.

Where we looked: sitedex.dev

confidence medium · grounding stated · weight 0.75 · Vague

04Your MCP endpoint mentions serving content to agents — does this work with Claude's tool use, OpenAI's function calling, or do I need to implement a custom client for each LLM provider?integrationHedges+2 pts est.

Your MCP endpoint mentions serving content to agents — does this work with Claude's tool use, OpenAI's function calling, or do I need to implement a custom client for each LLM provider?

What the pages say

The chunks show that Sitedex's MCP endpoint provides ready-made setup snippets for Claude Code (`claude mcp add ... --transport http`) and Cursor / Continue (JSON config for `mcp.json`), and supports WebMCP-capable browsers via a widget script and discovery manifest. None of the chunks mention OpenAI function calling or address whether you need to implement a custom client for each LLM provider.

The fix

Add explicit documentation clarifying which LLM providers and protocols are supported (e.g., OpenAI function calling vs. Anthropic tool use vs. standard MCP consumers) and whether the endpoint works with any MCP-compatible client or requires custom implementation per provider.

Where we looked: /sites/sitedex.dev/report, /sites/browserbase.com

confidence high · grounding synthesized · weight 0.70 · Vague

05For white-labeling purposes, can I host the MCP endpoint on my own subdomain instead of api.sitedex.dev, or is the hosted endpoint strictly tied to your infrastructure?operationsHedges+2 pts est.

For white-labeling purposes, can I host the MCP endpoint on my own subdomain instead of api.sitedex.dev, or is the hosted endpoint strictly tied to your infrastructure?

What the pages say

The hosted MCP endpoints are served from Sitedex's infrastructure at `https://mcp.sitedex.dev/s/{slug}/mcp` (e.g., `https://mcp.sitedex.dev/s/sitedex-dev/mcp`), not `api.sitedex.dev` (which is the REST API base URL). A local MCP server is also available via `npx -y @sitedex/cli mcp`. The retrieved chunks do not mention any option to host the MCP endpoint on a custom subdomain or to white-label the endpoint URL.

The fix

Clarify in the MCP or hosting documentation whether customers can use a custom subdomain or CNAME for their MCP endpoint, or explicitly confirm that all hosted endpoints are restricted to mcp.sitedex.dev.

Where we looked: /llms-full.txt, /sites/sitedex.dev/report, /docs

confidence high · grounding synthesized · weight 0.70 · Vague

What the pages already answer cleanly

10 of 15. Answered with enough specificity for a buyer to act on. This is the bar the queue above is chasing.

technicalAnswered

I'm building a batch processing pipeline — what are the specific rate limit headers your API returns so I can implement proper backoff without hitting hard limits?

What the pages say

Every rate-limited response carries four standard headers: `X-RateLimit-Limit` (total requests allowed in the current window), `X-RateLimit-Remaining` (requests remaining, clamped at 0), `X-RateLimit-Reset` (Unix timestamp in seconds when the window resets), and `Retry-After` (seconds to wait before retrying, RFC 6585 §4, set only on 429 responses). For backoff, the docs recommend preferring `Retry-After` on a 429 if your client supports it; otherwise compute the wait from `X-RateLimit-Reset`. For sustained throughput, watch `X-RateLimit-Remaining` on every 2xx and slow down before the cap.

Source: /openapi.json, /pricing

confidence high · grounding stated · weight 1.00 · Answered

getting-startedAnswered

Before I commit to a paid plan, what can I actually do on the free tier — is it limited to one audit, read-only access to public site data, or something else entirely?

What the pages say

Something else entirely. The free tier is not limited to one audit, nor is it read-only. Without paying you can: search every indexed site, read any page as markdown, view full diagnoses (every score and fix) for any site, use the hosted MCP, REST API, and CLI, and index new sites without even creating an account. The first crawl of any site is free and unlimited—'the first crawl of any site is on us.' The single credit you receive on signup is specifically for re-crawls ('your first re-crawl is on us'). Free crawls are rate-limited (10 audits per hour and 30 per day per IP) and stop at 500 pages per crawl.

Source: /pricing, /pricing, /pricing, /docs, /llms-full.txt

confidence high · grounding synthesized · weight 0.95 · Answered

integrationAnswered

Your MCP endpoint is hosted — if you update the protocol version, do you maintain backward compatibility or will my agent integration break without warning?

What the pages say

The hosted MCP endpoint is backward compatible. Sitedex adds capabilities without breaking the ones you already use, and they won't remove or rename a tool out from under you, so an integration you ship today keeps working as they evolve.

Source: /docs

confidence high · grounding stated · weight 1.00 · Answered

pricingAnswered

If I buy the Starter credit pack, how many site audits does that actually translate to — is it 1 credit per audit or does the scoring depth affect the cost?

What the pages say

The Starter pack costs $50 and includes 15 credits, which translates to 15 site audits. Each audit costs exactly 1 credit, and scoring depth does not affect the cost. The site states that “same product everyone gets” and “Bigger packs cost less per audit. That’s the only difference,” meaning the only variable is the per-credit price based on pack size.

Source: /llms.txt, /pricing, /pricing, /llms.txt

confidence high · grounding synthesized · weight 0.95 · Answered

securityAnswered

Can I create API keys with restricted scopes — like read-only access to audit results without the ability to trigger new crawls or manage billing?

What the pages say

No. Sitedex's documentation explicitly states that read-only and scope-limited keys are not available yet: "A key carries your organization's full access — the same as signing in. Read-only and scope-limited keys aren't available yet, so treat every `sdx_…` key like a password: keep it server-side, and rotate it if it leaks."

Source: /docs

confidence high · grounding stated · weight 1.00 · Answered

technicalAnswered

When you show me where my site ranks against competitors, are you comparing against all indexed sites or only those in my selected category — and how do you handle sites with no declared category?

What the pages say

Indexed sites appear on a global leaderboard — top sites by score — and on per-category boards for any category with enough indexed sites to compare fairly. If your site doesn't fall into a clear category yet, it still ranks globally; it just won't show in a category board until there's a cohort to rank it against.

Source: /docs

confidence high · grounding stated · weight 1.00 · Answered

limitsAnswered

I need to track how my site's AI-readiness changes weekly — is there a hard limit on how often I can re-audit the same domain, or can I schedule automatic recrawls?

What the pages say

There is no hard limit or cooldown on re-auditing the same domain—each re-audit costs one credit. Sitedex does not auto-schedule re-crawls itself, but the API lets you wire up automatic recrawls on whatever cadence you like, including weekly.

Source: sitedex.dev

confidence high · grounding stated · weight 1.00 · Answered

technicalAnswered

You grade sites on AI-readiness, but what exactly weights the score — is it just schema markup coverage, or do you factor in actual LLM retrieval success rates from your agent testing?

What the pages say

The AI-readiness score is weighted 70% to content and 30% to protocol, so it is not just schema markup coverage. The content score (70%) comes from 15 category-specific buyer questions that a real LLM answers from the site’s indexed content; a judge grades each answer on a diagnosis taxonomy (answered, vague claim, collapsed content, missing page, contradictory pages, genuinely absent). This functions as actual LLM retrieval-success testing. The protocol score (30%) checks 14 technical signals—including `llms.txt`, sitemap, robots-for-AI-crawlers, MCP card, WebMCP, and JSON-LD—that determine whether agents can read the site at all. The composite formula is `0.7 × content + 0.3 × protocol`, capped so protocol cannot more than double a weak content score.

Source: /llms.txt, /sites/category/open-source, /sites/category/infrastructure

confidence high · grounding synthesized · weight 0.95 · Answered

technicalAnswered

My site is a single-page application — does your crawler execute JavaScript to index dynamically rendered content, or do I need to pre-render for Sitedex to see my full content?

What the pages say

Yes. Sitedex renders every page in a real browser before reading it, so JavaScript-heavy sites and SPAs get indexed the way a visitor sees them, “not as an empty shell.” No pre-rendering is required: “Nothing to pre-render, nothing to change.”

Source: sitedex.dev

confidence high · grounding stated · weight 1.00 · Answered

technicalAnswered

For the SEO optimization scoring, do you require specific schema.org types like Organization and WebSite, or do you also evaluate custom JSON-LD that might be more relevant to my industry?

What the pages say

Sitedex requires a valid Organization JSON-LD block with a name on your homepage. It does not require WebSite or industry-specific schema types, and extra custom JSON-LD does not move the score.

Source: /docs

confidence high · grounding stated · weight 1.00 · Answered

Protocol plumbing · 94/10015 of 16 checks pass · each fix +6 protocol pts est.

The other half of the score: 16 checks for the files and headers agents look for. The 1 below are installs, not judgment calls, and most are an afternoon. Expand any for the snippet and the standard it follows. They sit after the queue because none of them changes what your pages say.

Markdown negotiationRendering+6 pts est.

StandardRFC 9110 + 7763IETF RFC

Already passing 15 of 16: robots.txt, sitemap.xml, llms.txt, AI crawler access, Content signal, Clean crawl, Server-rendered content, MCP card, OpenAPI spec, WebMCP widget, Canonical URLs, Meta descriptions, HTML lang attribute, Organization schema, Sitemap lastmod.

Ask this site’s index

Sitedex already serves sitedex.dev as an MCP endpoint. Ask sitedex.dev anything an AI agent might ask, and see what its index returns. (To score your own site, use the form below.)

Snippets & configs

For developers and the engineer-on-call: copy these into your tools or your site.

Files from this audit

Built from this crawl. Download or copy each, then install it at the path noted.

llms.txt

Built from this crawl. Install at /llms.txt so agents start here.

organization.json

Organization JSON-LD, pre-filled from this crawl. Wrap in a ld+json script.

server-card.json

MCP server card built from this crawl. Host at /.well-known/mcp/server-card.json.

webmcp.json

WebMCP discovery manifest built from this crawl. Host at /.well-known/webmcp.json.

MCP endpoint

https://mcp.sitedex.dev/s/sitedex-dev/mcp

The URL anyone's agent points at. Read-only; safe to share.

Claude Code

claude mcp add sitedex --transport http https://mcp.sitedex.dev/s/sitedex-dev/mcp

One command, then the agent has it.

Cursor / Continue

{
  "mcpServers": {
    "sitedex": {
      "url": "https://mcp.sitedex.dev/s/sitedex-dev/mcp"
    }
  }
}

Drop into mcp.json.

WebMCP: two parts

WebMCP-capable browsers run the widget at runtime. Crawlers without JS rendering need the discovery manifest to find your tool surface. Install both.

1 · Widget script

<script async src="https://sitedex.dev/widget.js"></script>

Drop in <head>. WebMCP-capable browsers (Chrome 146+ Origin Trial) call navigator.modelContext.provideContext() via this script.

2 · Discovery manifest

{
  "$schema": "https://wellknownmcp.org/schemas/webmcp.json",
  "name": "sitedex.dev",
  "tools": [
    { "name": "search", "description": "Search sitedex.dev's indexed content." },
    { "name": "get_page", "description": "Fetch a page from sitedex.dev as markdown." }
  ]
}

Host alongside the script at /.well-known/webmcp.json. Crawlers that don't render JS rely on this.

Your turn

See which of these questions your site goes silent on.

Free, about 5 minutes. We crawl your site, test it against the buyer questions your category asks, and name what’s vague, contradictory, or missing, plus the files AI agents look for.

ComingEmbeddable grade badgeScore history and deltasOpt-in public board