crunchbase.com

agent readiness score · scanned Aug 18, 2026 · 7s · Data & Analytics

#686 of 979 · #31 in category

36/ 100F

Crunchbase provides private company data and predictive intelligence to help various teams like GTM, investors, product developers, analysts, wealth managers, and entrepreneurs make smarter decisions and identify opportunities.

badge

Discovery 12.5/20

Access 11.5/30

Usability 12.2/40

Payments 0/10

Top fixes

  1. Pricing discoverable

    Expose a crawlable /pricing page with literal plan prices and link it from your homepage nav.

    +10
  2. AI crawler policy

    robots.txt blocks GPTBot, ClaudeBot, anthropic-ai, Google-Extended, CCBot, meta-externalagent, Bytespider. Add explicit `User-agent: <bot>` / `Allow: /` groups for the AI agents you welcome; a blanket Disallow makes you invisible to them.

    +5.8
  3. MCP server discovery

    Publish /.well-known/mcp/server-card.json describing your MCP endpoint so agents can autodiscover it. Evidence found: none.

    +5.6
  4. Sitemap present & fresh

    Publish /sitemap.xml, reference it from robots.txt, and include <lastmod> dates so agents can tell what's current. Freshest lastmod seen: none.

    +5
  5. JSON-LD structured data

    Embed Organization + WebSite JSON-LD on your homepage (name, url, logo, sameAs) and Product/Offer JSON-LD on pricing.

    +4.6
  6. llms.txt

    Add /llms.txt: an H1 with your name, a one-line blockquote summary, and H2 sections of curated links to docs, pricing, and API.

    +4.6

1.Can an agent discover and trust you?

4/8

Whether agents can crawl you, find you in the registries and searches they check, and trust what they find.

  • Sitemap present & freshrequired0/4

    No valid sitemap found: robots.txt listed 1 Sitemap URL(s), and GET /sitemap.xml returned HTTP 403

    fix → Publish /sitemap.xml, reference it from robots.txt, and include <lastmod> dates so agents can tell what's current. Freshest lastmod seen: none.

    Resolved the sitemap from robots.txt or /sitemap.xml, validated XML, and checked lastmod freshness

    spec ↗
  • Brand search discoverabilityrequired4/4

    Both searches ("Crunchbase" and "Crunchbase data analytics") cited crunchbase.com: https://about.crunchbase.com/research-reports, https://about.crunchbase.com/media-and-entertainment-report-2025, https://about.crunchbase.com/media-entertainment-report-2024/?ereport=, https://news.crunchbase.com/tag/q4-2017/, https://news.crunchbase.com/venture/billion-dollar-seed-ai-biotech-mcdonald-bison/

    Ran clean brand-name web searches and checked whether your domain is cited in the results

  • MCP registry listingsrecommendedna

    No MCP server detected for this product

    Queried the official MCP registry, Smithery, Glama, and PulseMCP for servers matching your domain

    spec ↗

2.Do you welcome agents?

4/16

Whether your robots policy, bot protection, and agent guidance actively admit AI agents instead of blocking them.

  • robots.txt present & parseablerequired2/2

    /robots.txt parses: 3 User-agent group(s), 1 Sitemap line(s)

    Fetched /robots.txt and validated it parses as a robots policy

    spec ↗
  • AI crawler policyrequired0/5

    robots.txt blocks 7 of 12 AI crawler user-agents from /: GPTBot, ClaudeBot, anthropic-ai, Google-Extended, CCBot, meta-externalagent, Bytespider

    fix → robots.txt blocks GPTBot, ClaudeBot, anthropic-ai, Google-Extended, CCBot, meta-externalagent, Bytespider. Add explicit `User-agent: <bot>` / `Allow: /` groups for the AI agents you welcome; a blanket Disallow makes you invisible to them.

    Evaluated robots.txt groups for the major AI agent user-agents (GPTBot, ClaudeBot, PerplexityBot, …)

  • Content Signals directivesemerging0/2

    No Content-Signal lines in robots.txt

    fix → Declare Content Signals in robots.txt (e.g. `Content-Signal: search=yes, ai-train=no`) to express AI usage preferences machine-readably.

    Looked for Content-Signal lines in robots.txt

    spec ↗
  • Agent user-agent paritybetarequired2/4

    GPTBot UA gets HTTP 429 while the browser UA gets HTTP 200

    fix → Your WAF serves HTTP 429 to GPTBot while browsers get 200. Allowlist verified AI crawler UA/IP ranges in your bot-protection rules.

    Compared responses served to browser and AI-agent user-agents (informational while in beta)

  • agents.mdrecommended0/3

    No agents.md: https://crunchbase.com/agents.md → HTTP 403, https://crunchbase.com/AGENTS.md → HTTP 403

    fix → Add /agents.md: what agents may do on your site, key URLs, auth, rate limits, and who to contact.

    Fetched /agents.md and checked for substantive agent guidance

    spec ↗

3.Does an agent understand who you are and what you do?

3/18

Whether your pages carry machine-readable identity: structured data, llms.txt, clear copy an agent can quote.

  • Homepage states what you arerequired3/3

    Homepage clarity rated 5/5 — an agent can confidently say what this company does (model's one-sentence read: "Crunchbase provides private company data and predictive intelligence to help various teams like GTM, investors, product developers, analysts, wealth managers, and entrepreneurs make smarter decisions and identify opportunities.")

    An LLM read your homepage as an agent would and rated how confidently it could say what you do

  • OpenGraph / social metadatarecommended0/2

    Homepage HTML contains none of og:title, og:description, og:image (twitter:card absent)

    fix → Add og:title, og:description, and og:image meta tags to your homepage; agents and link unfurlers use them as your canonical summary.

    Parsed homepage og:title / og:description / og:image and twitter:card meta tags

    spec ↗
  • llms.txtrecommended0/4

    GET /llms.txt returned HTTP 404

    fix → Add /llms.txt: an H1 with your name, a one-line blockquote summary, and H2 sections of curated links to docs, pricing, and API.

    Fetched /llms.txt and validated it against the llmstxt.org shape (H1, summary, curated links)

    spec ↗
  • llms-full.txtbonusrecommended0/2

    GET /llms-full.txt returned HTTP 404

    fix → Add /llms-full.txt with expanded inline docs content so agents can load everything in one fetch.

    Fetched /llms-full.txt and checked for substantial inline markdown content

    spec ↗
  • JSON-LD structured datarequired0/4

    No application/ld+json blocks found on homepage

    fix → Embed Organization + WebSite JSON-LD on your homepage (name, url, logo, sameAs) and Product/Offer JSON-LD on pricing.

    Extracted and validated application/ld+json blocks on the homepage and pricing page

    spec ↗
  • Markdown content negotiationrecommended0/3

    GET / with `Accept: text/markdown` returned HTTP 200 text/html; no .md twins found (2 probed)

    fix → Serve text/markdown when clients send `Accept: text/markdown` (or expose .md twins of key pages) — agents get far more signal per token.

    Requested key pages with Accept: text/markdown and probed .md twin URLs

4.Can an agent integrate with you?

watch →10.5/21

Whether the artifacts an agent needs to build on you — docs, API specs, SDKs, MCP servers — exist and are findable.

  • Developer resource discoverabilityrecommended3/3

    Search "Crunchbase API documentation" cited your own domain: https://data.crunchbase.com/docs/using-the-api, https://data.crunchbase.com/docs/using-search-apis, https://data.crunchbase.com/docs/using-entity-lookup-apis, https://data.crunchbase.com/reference, https://data.crunchbase.com/llms.txt

    Searched for your brand with developer-keyword suffixes and checked which official resources are cited

  • OpenAPI spec discoverablerequired2.5/5

    An OpenAPI spec URL was referenced but returned HTTP 401/403 — not fetchable without authentication

    fix → Expose your OpenAPI spec unauthenticated at /openapi.json; agents cannot integrate against a spec they cannot read.

    Probed standard OpenAPI locations and docs links for a fetchable, parseable spec

    spec ↗
  • Docs discoverablerequired4/4

    Docs found at https://about.crunchbase.com/build-your-product (via homepage link, 4142 chars of readable text)

    Followed homepage nav/footer links and probed /docs, /developers, docs.{domain}

  • MCP server discoveryrequired0/5

    No MCP evidence found: probed /.well-known/mcp/server-card.json, /.well-known/mcp.json, /mcp.json and scanned llms.txt, agents.md, and docs for endpoints or install commands

    fix → Publish /.well-known/mcp/server-card.json describing your MCP endpoint so agents can autodiscover it. Evidence found: none.

    Probed /.well-known/mcp/server-card.json, mcp.json, and docs mentions for an MCP endpoint

    spec ↗
  • npm SDKrecommended1/2

    Official npm package "crunchbase" (verified via README mentions the domain) but latest 0.1.2 is stale — published 2012-12-05, over 18 months ago

    fix → Publish a fresh release of "crunchbase" — agents treat SDKs untouched for 18+ months as abandoned.

    Searched the npm registry for an official, domain-verified SDK package

  • PyPI SDKrecommended0/2

    No official PyPI SDK found: probed 4 candidate(s) (crunchbase, crunchbase-sdk, crunchbaseapi, crunchbase-python)

    fix → Publish an official Python SDK (suggest "crunchbase") with your domain in the project URLs so agents can verify it's official.

    Searched PyPI for an official, domain-verified SDK package

5.Is your integration well-built?

watch →0/5

Whether your specs, docs, and tools are complete and descriptive enough for an agent to use them without guessing.

  • OpenAPI validity & qualityrecommendedna

    No OpenAPI spec found — nothing to lint

    Linted the spec: descriptions, operationIds, securitySchemes, servers

  • Docs qualityrecommended0/3

    Docs at https://about.crunchbase.com/build-your-product scored clarity 1/5, completeness 0/5, runnable examples 0/5, agent-friendliness 0/5 (mean 0.3/5)

    fix → Docs scored 0/5 on completeness. The documentation provided is a marketing page for Crunchbase products and does not contain any technical API documentation, making it impossible for an agent to integrate.

    An LLM rated your docs for clarity, completeness, runnable examples, and agent-friendliness

  • Quickstart / getting startedrecommended0/2

    No quickstart/getting-started link or heading found on docs page https://about.crunchbase.com/build-your-product (searched 81 links)

    fix → Add a copy-pasteable quickstart (auth → first API call) — agents follow it literally.

    Looked for a quickstart/getting-started guide containing code blocks

  • MCP tool qualityrecommendedna

    No MCP endpoint known — tool lint requires a tools/list

    Linted listed tools for descriptions, typed input schemas, and naming

6.Can an agent use you reliably in production?

watch →3/4

Response hygiene, TLS and redirect discipline, and security contact channels agents depend on at runtime.

  • TLS & redirect hygienerequired2/2

    http://crunchbase.com/ upgrades to https in 1 hop(s), 1 redirect(s) total, final HTTP 200; HSTS present

    Checked http→https redirect behavior, chain length, and TLS health

  • Response speed & weightrequired1/1

    Homepage TTFB 18ms, payload 77KB, Content-Encoding: br

    Measured homepage TTFB, payload size, and compression

  • security.txtrecommended0/1

    https://crunchbase.com/.well-known/security.txt returned HTTP 429

    fix → Publish RFC 9116 /.well-known/security.txt with a Contact and a future Expires.

    Fetched /.well-known/security.txt and validated Contact + Expires

    spec ↗

7.Can an agent authenticate to you?

watch →1.5/7

Whether agents can discover your auth model machine-readably (OAuth metadata) and follow documented steps to credentials.

  • OAuth authorization server metadatarecommended1.5/3

    Valid RFC 8414 metadata at https://crunchbase.com/.well-known/oauth-authorization-server (issuer https://www.crunchbase.com, PKCE S256) but no registration_endpoint

    fix → Add a registration_endpoint (dynamic client registration) — agents cannot pre-register OAuth clients by hand.

    Fetched /.well-known/oauth-authorization-server (and openid-configuration fallback)

    spec ↗
  • OAuth protected resource metadatarecommended0/2

    No RFC 9728 metadata: https://crunchbase.com/.well-known/oauth-protected-resource → HTTP 429

    fix → Serve RFC 9728 metadata at /.well-known/oauth-protected-resource naming your authorization servers so agents can discover how to authenticate.

    Fetched /.well-known/oauth-protected-resource

    spec ↗
  • Auth documentationrecommended0/2

    No auth/API-key link found among 81 links on docs page https://about.crunchbase.com/build-your-product, and /auth.md is absent

    fix → Document auth end-to-end (key creation → header format → example call), or ship /auth.md.

    Looked for /auth.md or an authentication docs page with code examples

8.Can an agent transact with you?

watch →0/10

Whether pricing is discoverable and machine-readable, and whether you support agent payment protocols.

  • Pricing discoverablerequired0/5

    No pricing page found: no homepage link matching pricing/plans/billing, and probes of /pricing and /plans returned no HTML page.

    fix → Expose a crawlable /pricing page with literal plan prices and link it from your homepage nav.

    Followed nav/footer links and probed /pricing for a page with legible price signals

  • Machine-readable pricingrecommendedna

    No pricing page found to evaluate

    Looked for Offer JSON-LD, then had an LLM attempt structured extraction of your plans

    spec ↗
  • x402 payment supportbonusrecommended0/2

    /.well-known/x402 returned HTTP 429. No x402 support detected (bonus check — absence costs nothing).

    fix → Support x402: serve payment requirements at /.well-known/x402, or answer unauthenticated API calls with HTTP 402 plus an x402 payment-requirements payload (x402Version, accepts[]).

    Probed /.well-known/x402 and API endpoints for HTTP 402 payment-required envelopes

    spec ↗
  • AP2 readinessemerging0/1

    No AP2 hints found — scanned docs page (https://about.crunchbase.com/build-your-product); /.well-known/ap2 returned HTTP 429.

    fix → Adopt AP2 to accept delegated agent payments — see https://ap2-protocol.org.

    Scanned fetched artifacts and well-known paths for AP2 hints

    spec ↗
  • Agentic Commerce Protocolemerging0/1

    No Agentic Commerce Protocol hints found — scanned docs page (https://about.crunchbase.com/build-your-product); /.well-known/acp returned HTTP 429.

    fix → Adopt the Agentic Commerce Protocol so agent checkouts can complete against your store — see https://developers.openai.com/commerce.

    Scanned docs and specs for agentic checkout endpoints

    spec ↗
  • Other agent payment protocolsbonusemerging0/1

    No UCP/MPP hints found — scanned docs page (https://about.crunchbase.com/build-your-product).

    fix → Track emerging agent payment protocols (UCP, MPP) and adopt the ones your buyers' agents use.

    Scanned fetched artifacts for UCP/MPP protocol hints

9.Can a user act through an agent?

Whether an end user's agent can operate on their behalf: working MCP tools, published skills, agent configs.

  • Agent skill publishedbonusemergingna

    Requires the analysis phase — not yet evaluated

    Checked /skill.md, /.well-known/skills/, and skills.sh for published agent skills

  • MCP handshake & tools listrequiredna

    No MCP endpoint known — nothing to handshake with

    Performed a streamable-HTTP initialize + tools/list against the MCP endpoint

    spec ↗
  • Agent configs in public repobonusemergingna

    Requires the analysis phase — not yet evaluated

    Checked your public GitHub org's main repos for AGENTS.md / .claude / .cursor configs

10.Can an agent operate your website directly?

7/9

Whether the site itself is legible to non-rendering and browser agents: semantic HTML, no JS walls, accessibility.

  • NLWeb endpointemerging0/1

    No NLWeb endpoint detected: /.well-known/nlweb.json returned HTTP 404 (text/html); GET /ask?query=hello returned HTTP 403 (text/html)

    fix → Consider exposing an NLWeb /ask endpoint (and /.well-known/nlweb.json) for conversational access to your content.

    Probed /.well-known/nlweb.json and the /ask endpoint

    spec ↗
  • Semantic HTML structurerequired3/3

    5/6 semantic signals present on the homepage (missing: text-to-markup ratio ≥ 0.10)

    Scored landmark elements, heading hierarchy, and text-to-markup ratio on the homepage

  • Content readable without JavaScriptrequired2/2

    Homepage raw HTML contains 5477 chars of visible text without JavaScript

    Measured visible text in the raw, unrendered homepage HTML

  • Accessibility basicsrequired2/2

    Homepage accessibility: 5/5 checks passed

    Static checks: lang attribute, title, alt coverage, labeled inputs, landmarks

  • WebMCPemerging0/1

    No WebMCP evidence on homepage: no application/webmcp script and no navigator.modelContext reference

    fix → Consider WebMCP to expose page actions as tools to browser agents.

    Looked for WebMCP script declarations on the homepage

    spec ↗