scanned Jul 14, 2026

Crunchbase

crunchbase.com

Crunchbase provides predictive company intelligence and private market analysis

43/100

Tier 3 · Agent-Accessible

Content answers50/100
Protocol plumbing27/1004 of 15 checks pass

Scored by asking 15 questions a buyer of a analytics product asks, then grading this site’s own pages: answered, hedged (partial or vague), or silent (no page answers it). How scoring works

This report is public. Own crunchbase.com? Claiming is free: crawl every page, re-audit as you fix, and track your score over time.

Sign in to claim

The fix queue

57 points sit between crunchbase.com and 100: 13 open questions and 11 missing protocol checks, ordered by estimated payoff.

Point estimates are per fix under scoring v2. They are not additive to a promised total.

01limits · importance highGoes silent+7 content pts est.

When using the bulk enrichment API endpoint, what's the hard limit on how many company domains or names I can submit in a single batch request?

What the pages say

No page on the site addresses this.

The fix

Add documentation for the bulk enrichment API endpoint that explicitly states the hard limit on company domains or names per batch request.

confidence high · grounding world-knowledge · weight 0.00 · Absent

02technical · importance mediumGoes silent+7 content pts est.

If two company profiles are merged, does the API return a redirect or deprecation notice for the old UUID, or do I need to handle deduplication myself?

What the pages say

No page on the site addresses this.

The fix

Create a dedicated API behavior or data integrity page that explains what happens when company profiles are merged—specifically whether the API returns a redirect, deprecation notice, or empty response for the old UUID, and whether consumers must handle deduplication.

confidence high · grounding world-knowledge · weight 0.00 · Absent

03security · importance mediumGoes silent+7 content pts est.

What specific browser permissions does the Crunchbase Chrome extension require, and does it read data from all tabs or only active ones?

What the pages say

No page on the site addresses this.

The fix

Create a dedicated security or help-center page documenting the Crunchbase Chrome extension's required browser permissions and data-access scope (e.g., activeTab vs. all tabs).

confidence high · grounding world-knowledge · weight 0.00 · Absent

04limits · importance mediumGoes silent+7 content pts est.

What's the upper limit on how many organizations I can save to one custom list or portfolio?

What the pages say

No page on the site addresses this.

The fix

Add a help center article or FAQ entry documenting the maximum number of organizations allowed per custom list or portfolio.

Where we looked: /product-updates

confidence high · grounding synthesized · weight 0.00 · Absent

05security · importance lowGoes silent+7 content pts est.

For Enterprise API access, how many distinct IP address ranges can I whitelist for a single API key?

What the pages say

No page on the site addresses this.

The fix

Add documentation about IP whitelisting capabilities and limits for Enterprise API keys.

confidence high · grounding world-knowledge · weight 0.00 · Absent

Hedged · 8 of 15 questions

A buyer gets something, then has to guess the rest. Expand any row for the evidence and the fix.

06How many API requests can I make per day on a standard Crunchbase Pro plan before I hit the cap?limitsHedges+2 pts est.

How many API requests can I make per day on a standard Crunchbase Pro plan before I hit the cap?

What the pages say

The retrieved chunks do not mention a daily API request cap for a "Crunchbase Pro" plan, nor do they reference a "Pro" tier for API access. The only rate limit specified is 200 calls per minute across API endpoints. The API access tiers mentioned are Crunchbase Basic, Advanced, Commercial, Enterprise, and Applications.

The fix

Add a dedicated pricing or limits page that specifies daily API request caps for each plan tier, including whether a Pro plan includes API access and what its limits are.

Where we looked: /docs/using-the-api, /docs/crunchbase-basic-using-api

confidence medium · grounding synthesized · weight 0.65 · Vague

07What's the maximum number of company rows I can download in a single CSV export from the web interface?limitsHedges+2 pts est.

What's the maximum number of company rows I can download in a single CSV export from the web interface?

What the pages say

Crunchbase Pro allows users to export up to 2,000 rows per month from the web interface. The retrieved chunks do not specify whether this 2,000-row figure is also the maximum allowed in a single CSV export, or if it is an aggregate monthly cap across multiple exports.

The fix

Clarify whether the 2,000-row monthly export limit for Crunchbase Pro is a per-export maximum or an aggregate monthly cap, and whether users can perform multiple smaller exports.

Where we looked: /products/crunchbase-pro

confidence medium · grounding stated · weight 0.75 · Vague

08Does the Salesforce integration refresh company data automatically on a schedule, or do I have to trigger updates manually from the web app?integrationHedges+2 pts est.

Does the Salesforce integration refresh company data automatically on a schedule, or do I have to trigger updates manually from the web app?

What the pages say

Crunchbase states that its data enrichment involves ‘continuous refreshment,’ ‘automated data hygiene,’ and ‘real-time company updates’ that ‘keep your tools fresh,’ and the retrieved chunks confirm a Salesforce integration exists. However, none of the chunks specify whether the Salesforce integration refreshes company data automatically on a schedule or if users must trigger updates manually from the web app.

The fix

Add a dedicated section or FAQ to the Salesforce integration or CRM Enrichment documentation that explicitly states whether data syncs automatically on a schedule (and what the interval is) or requires manual triggering from the web app.

Where we looked: /products/data-enrichment, /blog/sales-intelligence-tools, /docs/crm-enrichment

confidence medium · grounding synthesized · weight 0.65 · Vague

09If I hit the API rate limit and get a 429 error, does the response include a Retry-After header with the exact wait time, or is the backoff window fixed?technicalHedges+2 pts est.

If I hit the API rate limit and get a 429 error, does the response include a Retry-After header with the exact wait time, or is the backoff window fixed?

What the pages say

The Crunchbase API documentation states that endpoints have a rate limit of 200 calls per minute and that hitting the limit causes the API to respond with an error message, but the retrieved chunks do not specify whether this is a 429 error, whether the response includes a Retry-After header, or whether the backoff window is fixed. Chunk [7] documents an MD403 error for exceeding search result limits, showing specific error codes are documented in other contexts.

The fix

Expand the Rate Limit section to specify the exact HTTP status code returned (e.g., 429), whether the response includes a Retry-After header with a wait time, and whether clients should implement a fixed or exponential backoff.

Where we looked: /docs/using-the-api, /docs/using-search-apis

confidence medium · grounding synthesized · weight 0.65 · Vague

10How far back does Crunchbase’s historical funding round data go — are pre-2000 deals included in the standard dataset?dataHedges+2 pts est.

How far back does Crunchbase’s historical funding round data go — are pre-2000 deals included in the standard dataset?

What the pages say

Crunchbase’s website confirms it provides historical funding round data (described as “round-by-round funding data” and “fundamental historical records”), but the retrieved chunks do not specify how far back this data extends or whether pre-2000 deals are included in the standard dataset.

The fix

Add a clear statement on the Data page or in the FAQs specifying the temporal coverage of the historical funding dataset, including whether pre-2000 deals are included.

Where we looked: /data, api.crunchbase.com

confidence medium · grounding synthesized · weight 0.65 · Vague

11Is it possible to schedule automated data exports to run weekly, or must I generate every CSV download manually?operationsHedges+2 pts est.

Is it possible to schedule automated data exports to run weekly, or must I generate every CSV download manually?

What the pages say

The retrieved documentation describes a manual download workflow for CSV and NDJSON exports: you enter a URL into your browser and press Enter to initiate the download. The Terms of Service explicitly prohibit using “automated means or tools” to access, download, or export Content. Crunchbase states that updated export files are generated each morning, but no built-in scheduling feature for customer-initiated weekly exports is described. Consequently, the available site content indicates that downloads must be generated manually rather than scheduled automatically.

The fix

Add an explicit section to the Export documentation stating whether automated scheduling is supported, or clearly confirm that exports must be downloaded manually and cannot be automated.

Where we looked: /docs/crunchbase-basic-export, /docs/legacy-export, /terms-of-service

confidence medium · grounding synthesized · weight 0.65 · Vague

12Does Crunchbase support SAML-based single sign-on for team accounts, or is password login the only option?securityHedges+2 pts est.

Does Crunchbase support SAML-based single sign-on for team accounts, or is password login the only option?

What the pages say

Password login is not the only option. Chunk [3] shows that Crunchbase’s login page offers Email/Password, Google, LinkedIn, and a separate “Log In With Single Sign-on” option. However, none of the retrieved chunks specify whether this Single Sign-on uses SAML, nor do they confirm whether it is available for team accounts.

The fix

Publish a dedicated security or authentication page (or expand the login/help docs) that explicitly lists supported SSO protocols such as SAML 2.0 and clarifies which account tiers or team plans can use them.

Where we looked: /account, /terms-of-service

confidence high · grounding synthesized · weight 0.70 · Vague

13Can I retrieve deleted or inactive company profiles through the API, and if so, which response field indicates the profile status?technicalHedges+2 pts est.

Can I retrieve deleted or inactive company profiles through the API, and if so, which response field indicates the profile status?

What the pages say

You can retrieve deleted organization profiles via the Deleted Entities API endpoint (`/v4/data/deleted_entities/organizations`), but the chunks do not indicate whether inactive profiles are retrievable, nor do they specify which response field indicates profile status.

The fix

Document the response schema for the Deleted Entities API, including which field indicates an entity's deleted or inactive status.

Where we looked: /docs/using-deleted-entities-api

confidence medium · grounding stated · weight 0.75 · Vague

What the pages already answer cleanly

2 of 15. Answered with enough specificity for a buyer to act on. This is the bar the queue above is chasing.

technicalAnswered

Do I need to set up an OAuth 2.0 app to authenticate with the Crunchbase API, or can I just use a simple API key?

What the pages say

No, you do not need to set up an OAuth 2.0 app. The Crunchbase API uses token-based authentication, which means developers must pass an individual API key with every request. You can pass the key either as a `user_key` parameter in the request URL or as an `X-cb-user-key` header. All API calls require a user key, which is emailed to you following registration.

Source: /docs/using-the-api

confidence high · grounding stated · weight 1.00 · Answered

technicalAnswered

Can I create or edit company profiles programmatically through the API, or is it strictly read-only?

What the pages say

The Crunchbase API is strictly read-only. According to the documentation, "The Crunchbase API is a read-only RESTful service that enables approved developers to leverage the same data that powers www.crunchbase.com." You cannot create or edit company profiles programmatically through the API.

Source: /docs/using-the-api

confidence high · grounding stated · weight 1.00 · Answered

Protocol plumbing · 27/1004 of 15 checks pass · each fix +7 protocol pts est.

The other half of the score: 15 checks for the files and headers agents look for. The 11 below are installs, not judgment calls, and most are an afternoon. Expand any for the snippet and the standard it follows. They sit after the queue because none of them changes what your pages say.

sitemap.xmlDiscoverability+7 pts est.

Standardsitemaps.orgIndustry standard

llms.txtDiscoverability+7 pts est.

Sitedex generates this file from your crawl. Grab it in Files from this audit below.

Standardllmstxt.orgCommunity spec

AI crawler accessAccess+7 pts est.
Install snippet
User-agent: GPTBot
Allow: /

User-agent: ClaudeBot
Allow: /

User-agent: anthropic-ai
Allow: /

User-agent: PerplexityBot
Allow: /

User-agent: OAI-SearchBot
Allow: /

User-agent: Googlebot-Extended
Allow: /

StandardRFC 9309IETF RFC

Content signalAccess+7 pts est.
Install snippet
User-agent: *
Content-Signal: search=yes, ai-input=yes, ai-train=no
Allow: /

StandardCloudflare proposalVendor proposal

Markdown negotiationRendering+7 pts est.

StandardRFC 9110 + 7763IETF RFC

MCP cardInteraction+7 pts est.

Sitedex generates this file from your crawl. Grab it in Files from this audit below.

StandardModel Context ProtocolCommunity spec

OpenAPI specInteraction+7 pts est.

StandardOpenAPI SpecIndustry standard

WebMCP widgetInteraction+7 pts est.

Sitedex generates this file from your crawl. Grab it in Files from this audit below.

StandardW3C WebMCP draftW3C / WHATWG

Canonical URLsHygiene+7 pts est.
Install snippet
<link rel="canonical" href="https://crunchbase.com/" />

StandardRFC 6596IETF RFC

Meta descriptionsHygiene+7 pts est.
Install snippet
<meta name="description" content="crunchbase.com: [outcome you deliver] for [who you help]. One sentence, 50-160 characters." />

StandardHTML Living StandardW3C / WHATWG

Organization schemaIdentity+7 pts est.

Sitedex generates this file from your crawl. Grab it in Files from this audit below.

StandardSchema.org + JSON-LDIndustry standard

Already passing 4 of 15: robots.txt, Clean crawl, Server-rendered content, HTML lang attribute.

Ask this site’s index

Sitedex already serves crunchbase.com as an MCP endpoint. Ask crunchbase.com anything an AI agent might ask, and see what its index returns. (To score your own site, use the form below.)

Snippets & configs

For developers and the engineer-on-call: copy these into your tools or your site.

Files from this audit

Built from this crawl. Download or copy each, then install it at the path noted.

llms.txt

Built from this crawl. Install at /llms.txt so agents start here.

organization.json

Organization JSON-LD, pre-filled from this crawl. Wrap in a ld+json script.

server-card.json

MCP server card built from this crawl. Host at /.well-known/mcp/server-card.json.

webmcp.json

WebMCP discovery manifest built from this crawl. Host at /.well-known/webmcp.json.

MCP endpoint

https://mcp.sitedex.dev/s/crunchbase-com/mcp

The URL anyone's agent points at. Read-only; safe to share.

Claude Code

claude mcp add crunchbase --transport http https://mcp.sitedex.dev/s/crunchbase-com/mcp

One command, then the agent has it.

Cursor / Continue

{
  "mcpServers": {
    "crunchbase": {
      "url": "https://mcp.sitedex.dev/s/crunchbase-com/mcp"
    }
  }
}

Drop into mcp.json.

WebMCP: two parts

WebMCP-capable browsers run the widget at runtime. Crawlers without JS rendering need the discovery manifest to find your tool surface. Install both.

1 · Widget script

<script async src="https://sitedex.dev/widget.js"></script>

Drop in <head>. WebMCP-capable browsers (Chrome 146+ Origin Trial) call navigator.modelContext.provideContext() via this script.

2 · Discovery manifest

{
  "$schema": "https://wellknownmcp.org/schemas/webmcp.json",
  "name": "crunchbase.com",
  "tools": [
    { "name": "search", "description": "Search crunchbase.com's indexed content." },
    { "name": "get_page", "description": "Fetch a page from crunchbase.com as markdown." }
  ]
}

Host alongside the script at /.well-known/webmcp.json. Crawlers that don't render JS rely on this.

Your turn

See which of these questions your site goes silent on.

Free, about 5 minutes. We crawl your site, test it against the buyer questions your category asks, and name what’s vague, contradictory, or missing, plus the files AI agents look for.

ComingEmbeddable grade badgeScore history and deltasOpt-in public board