Sameness Index · 3 sites compared

Your page scores 53 out of 100 for sameness against the 2 competitors you named.

https://www.kapa.ai/

01

Your verdict

53
Sameness · vs 2 named
How is this calculated?

Your page is half shared.

About half of what your page says, Exa also says. You're less same than 201 of the 366 SaaS sites scored (SaaS avg 56).

No category benchmark matched this set.
Sameness Index on a 0 to 100 scale, from distinctive at 0 to interchangeable at 100. This page scores 53. Inkeep scores 28. Exa scores 43. The SaaS avg is 56.

Each named site is scored the same way, against the other 2 in this set, so its tick means the same as your marker. The dashed line is the frozen benchmark average.

  1. Less same than average

    The average SaaS site scores 56; you scored 53.

    You are 3 points less same than the average SaaS site.

  2. Closest overlap: Exa

    Of the 2 sites you named, Exa echoes the most of what your page says. Scored the same way against the rest of the set, Exa sits at 43.

    Exa is the competitor you sound most like.

  3. Room to own more

    18% of your claim space is ownable: unique, relevant, and hard to copy. 1 of those claims sits in body copy, where few readers reach it.

    18% is ownable, and 1 buried opportunity could help you stand out more.

Do this first

Three changes worth testing first.

Chosen by rule from the comparison with Inkeep and Exa: the shared claim taking your most prominent space, then the claims only you make that sit too low on the page to be read. Each one links to its claim card.

  1. Connect your unstructured data and give your agents accurate, cited context from your knowledge base via API or MCP

    Inkeep and Exa all say it too. Buyers may still need it, but shared ground cannot carry your hero — move it lower and give that space to something only you can say.

    Table stakes
    Most of the set says this too.
  2. Every source's auth, pagination, and rate limits, handled

    Nobody in the set says this. It sits in body copy, where few readers reach it — worth testing higher up the page; only buyers can tell you whether it lands.

    Surface
    Yours alone. Test it higher up.
  3. Pre-built connectors for Web crawls, Zendesk, GitHub, Slack, PDFs, and more

    Keep the fact, lose the position: inkeep and Exa all say it too, and your section is spending its first impression on the same territory as theirs.

    Table stakes
    Most of the set says this too.

Try these changes, then test them with real buyers.

This measures overlap. Whether buyers notice is a different question, and only they can answer it.

02

What you can own

Claims only you make, that buyers weigh, and that competitors can’t easily copy.

18% of your page’s claim space is yours to keep.

18%Yours to keep32%Unique but weak50%A competitor says it too
Why this is not 100 minus the Sameness Index

The index is a weighted composite across six categories, including page structure and visuals. This bar is measured on your claims alone, weighted by where each one sits on the page. Different denominators, so the two never add to 100 and are not meant to.

Already leading with · 3

  1. Managed pipeline: chunking, embedding, hybrid search, reranking, evals

    Managed retrieval pipeline: Chunking, embedding, hybrid search, reranking and evals

  2. One knowledge base serves every agent across teams

    One knowledge base grounds every agent across your product, support, and engineering stack in your technical documentation

  3. Pre-built connectors to many data sources

    Pre-built connectors for Web crawls, Zendesk, GitHub, Slack, PDFs, and more

Buried in body copy · 1

  1. Handles source auth, pagination, rate limits

    Every source's auth, pagination, and rate limits, handled

03

Where you blend in

Territory you spend prominent space on that the set also occupies. Not every line is one to delete — the question is whether it has earned the space, or whether something only you can say should be there instead.

  • Commodity · 100%Table stakesHero
    retrieval/search API grounds AI agents
    • Connect your unstructured data and give your agents accurate, cited context from your knowledge base via API or MCP
    • Knowledge retrieval API for AI agents

    You say this 2 different ways.

    Inkeep and Exa all cover this territory. Buyers may need to hear it, but in your hero it spends the first impression on shared ground.

    They say
    • InkeepDeploy AI teammates that understand your data, product, knowledge, and tools
    • ExaThe world's data for AI agents
  • Commodity · 100%Table stakesSection
    connects to many data sources out of the box
    • Pre-built connectors for Web crawls, Zendesk, GitHub, Slack, PDFs, and more

    Common ground with Inkeep and Exa. Say it if buyers need it — lower on the page, where it is not the thing they read first. Nobody else uses your exact words, but everyone is on the ground. Nobody else makes this exact claim, but everyone occupies the territory — lead with the specific, not the generic.

    They say
    • InkeepQuery customer-specific information with Customer Intelligence
    • ExaBest-in-class across company search, people search, and code — not just general web queries
  • Commodity · 100%Table stakesSection
    named customers and customer proof points
    • Trusted by category-defining startups and Fortune 500 enterprises

    Inkeep and Exa make the same claim. It cannot set you apart, so it should not carry the section.

    They say
    • InkeepTrusted by leading teams
    • ExaJoin 500,000 developers building AI on top of the world's information
  • Contested · 50%SharpenHero
    access via API, MCP, or SDK components
    • via API or MCP

    Inkeep is on this territory too (50% of the set). It narrows the field without winning it — make it specific enough that it cannot be said of them. Nobody else uses your exact words, but everyone is on the ground. Nobody else makes this exact claim, but everyone occupies the territory — lead with the specific, not the generic.

    They say
    • InkeepOpen Agent Builder & Developer SDK
  • Contested · 50%SharpenHero
    accurate, cited answers from retrieval
    • accurate, cited context

    Shared with Exa. Sharpen it to the thing only you do here, or it reads as a claim any of you could make. Nobody else uses your exact words, but everyone is on the ground. Nobody else makes this exact claim, but everyone occupies the territory — lead with the specific, not the generic.

    They say
    • ExaHighest quality search at every latency
  • Contested · 50%SharpenSection
    customer-facing AI agent deflects support work
    • Support agent: Most support tickets ask something you've already answered
    • Documentation agent: Embed an agent in your documentation so users can self-serve
    • In-product agent: AI agents embedded directly in your application

    You say this 3 different ways.

    Shared with Inkeep. Sharpen it to the thing only you do here, or it reads as a claim any of you could make.

    They say
    • InkeepAI assistant trained on documentation, help articles, and other content
  • Contested · 50%SharpenSection
    keeps index fresh with incremental sync
    • Kapa detects and re-processes only what changed, and updates land in minutes

    Contested ground: Inkeep claim it as well. The version that wins names a mechanism, a number or a scope that theirs cannot match.

    They say
    • InkeepKeep your knowledge base always up to date
  • Contested · 50%SharpenSection
    managed end-to-end retrieval pipeline
    • Managed retrieval pipeline: Chunking, embedding, hybrid search, reranking and evals
    • Kapa prunes the rest to give agents context that actually matters

    You say this 2 different ways.

    Exa is on this territory too (50% of the set). It narrows the field without winning it — make it specific enough that it cannot be said of them.

    They say
    • ExaSpecialized model trained to extract the most relevant excerpts from the web
04

Claim-by-claim evidence

Every claim on your page (22)
What the columns mean
Claim
The grouped claim, then your exact line beneath it.
Type
What kind of claim it is: category, segment, outcome, capability, quality or proof.
Placement
Where it sits on your page: hero, section or body copy. Hero claims weigh most in the index.
Same claim
Share of the competitors making this exact claim. Drives ownership and ownable share.
Same territory
Share of the competitors with any claim in the same buyer-facing territory. This is what the index is scored on.
Sayability
Whether a competitor could truthfully make the same claim: anyone could, copyable with effort, or hard to copy.
Relevant
Whether buyers decide on this. A unique claim nobody buys on is not ownable.
Ownership
Commodity: 60% or more of the set says it. Contested: 20–59%. Unique: under 20%, owned when it is also hard to copy.

Tap a column to sort by it; tap again to reverse. Sorted by Same claim, highest first.

  • Trusted by well-known/large customers
    Trusted by category-defining startups and Fortune 500 enterprises
    proofSectionsame claim 100%same territory 100%Copyable with effort
    Commodity
    Also on Inkeep, Exa
  • Connects unstructured company data for agent context
    Connect your unstructured data and give your agents accurate, cited context from your knowledge base via API or MCP
    capabilityHerosame claim 50%same territory 100%Anyone could say it
    Contested
    Also on Inkeep
  • Retrieval/search API that grounds AI agents
    Knowledge retrieval API for AI agents
    categoryHerosame claim 50%same territory 100%Anyone could say it
    Contested
    Also on Exa
  • Deflects support tickets with an AI support agent
    Support agent: Most support tickets ask something you've already answered
    capabilitySectionsame claim 50%same territory 50%Anyone could say it
    Contested
    Also on Inkeep
  • Documentation agent lets users self-serve
    Documentation agent: Embed an agent in your documentation so users can self-serve
    capabilitySectionsame claim 50%same territory 50%Anyone could say it
    Contested
    Also on Inkeep
  • Incremental re-sync keeps index fresh in minutes
    Kapa detects and re-processes only what changed, and updates land in minutes
    capabilitySectionsame claim 50%same territory 50%Copyable with effort
    Contested
    Also on Inkeep
  • Prunes irrelevant context before sending to the model
    Kapa prunes the rest to give agents context that actually matters
    capabilitySectionsame claim 50%same territory 50%Copyable with effort
    Contested
    Also on Exa
  • Large reduction in tokens consumed
    Accurate retrieval at scale for 68% fewer tokens
    outcomeBodysame claim 50%same territory 50%Copyable with effort
    Contested
    Also on Exa
  • Prebuilt SDK components to build agents
    Build one on Kapa's Retrieval API, or use the pre-built components in the Kapa Agent SDK
    capabilityBodysame claim 50%same territory 50%Anyone could say it
    Contested
    Also on Inkeep
  • Retrieval models tuned by in-house research team
    tuned continuously by our world-class research team for you
    proofBodysame claim 50%same territory 50%Copyable with effort
    Contested
    Also on Exa
  • Accessible via API or MCP
    via API or MCP
    capabilityHerosame claim 0%same territory 50%Anyone could say it
    Unique for now
  • Returns accurate, cited context
    accurate, cited context
    qualityHerosame claim 0%same territory 50%Anyone could say it
    Unique for now
  • Embeddable in-product AI agent
    In-product agent: AI agents embedded directly in your application
    capabilitySectionsame claim 0%same territory 50%Anyone could say it
    Unique for now
  • Managed pipeline: chunking, embedding, hybrid search, reranking, evals
    Managed retrieval pipeline: Chunking, embedding, hybrid search, reranking and evals
    capabilitySectionsame claim 0%same territory 50%Copyable with effort
    Unique and owned
  • One knowledge base serves every agent across teams
    One knowledge base grounds every agent across your product, support, and engineering stack in your technical documentation
    outcomeSectionsame claim 0%same territory 0%Copyable with effort
    Unique and owned
  • Pre-built connectors to many data sources
    Pre-built connectors for Web crawls, Zendesk, GitHub, Slack, PDFs, and more
    capabilitySectionsame claim 0%same territory 100%Copyable with effort
    Unique and owned
  • Handles source auth, pagination, rate limits
    Every source's auth, pagination, and rate limits, handled
    capabilityBodysame claim 0%same territory 100%Copyable with effort
    Unique and owned
  • Ingests API reference/OpenAPI specs
    API reference with 206 endpoints
    capabilityBodysame claim 0%same territory 100%Anyone could say itnot a buying criterion
    Unique for now
  • Ingests code repositories and issues
    GitHub repos with 412 files ingested / 12 repos ingested
    capabilityBodysame claim 0%same territory 100%Anyone could say itnot a buying criterion
    Unique for now
  • Ingests large volumes of documentation pages
    Documentation source with 1,284 pages / 2,107 pages ingested
    proofBodysame claim 0%same territory 100%Anyone could say itnot a buying criterion
    Unique for now
  • Ingests PDFs and datasheets
    Product manuals with 86 PDFs ingested
    capabilityBodysame claim 0%same territory 100%Anyone could say itnot a buying criterion
    Unique for now
  • Ingests support tickets
    Support tickets with 9,840 tickets / 480 tickets ingested
    capabilityBodysame claim 0%same territory 100%Anyone could say itnot a buying criterion
    Unique for now
05

How this was calculated

Sameness measures how much your claims overlap with the sites compared. It does not measure message quality or whether buyers prefer you.

AI-analyzed: an AI read each page on its own and grouped the claims that say the same thing. No score here was written by a model — every number is computed from those groupings in our own code, with the weights below.

How the score is built
The six category scores, their weights, and what a high score in each one means
CategoryWeightYoursWhat a high score means
Messaging
Category framing, who it is for, and the outcome promised
30%61The most expensive kind of sameness. A buyer cannot tell what job you do that the others do not.
Claims
Attribute and benefit claims — speed, ease, quality, ROI
30%50Every shared claim is a line already read on another tab. Cut the ones nobody owns and spend the space on something they cannot.
Features
Capabilities and functions the page lists
15%68Expected in a mature category, and the least alarming of the six. Feature parity is normal; leading with it is the mistake.
Proof
The kinds of evidence offered: customer logos, numbers, testimonials, case studies, badges
10%31Same kinds of proof as everyone means the proof stops working as proof. It is scored on the kind of evidence, not on which customers are named.
Structure
Section order, navigation, CTA language and placement
10%50The generic SaaS template — hero, logos, three-feature grid, testimonial, CTA. Familiar is not the same as memorable.
Visual
Palette family, imagery style, layout patterns
5%30Weighted lowest on purpose: buyers rarely decide on this. Worth knowing, rarely worth fixing first.

Each site was read on its own first, with no knowledge of the others, so your page gets no benefit of the doubt a competitor’s does not. A category nothing could be measured for drops out and the rest are re-weighted, rather than counted as zero.

What we compared (3 pages read)
Your page
Kapa.ai
kapa.ai
Competitor
Inkeep
inkeep.com
Competitor
Exa
exa.ai
What it cannot tell you

The index can find where two pages converge. It cannot say whether a buyer would notice, or which of your reasons to buy actually land. A single check also moves several points between runs, so read the band and the ranking, not the last digit.

Your highest-impact changes

  1. 1
    Table stakesYour hero copy says “Connect your unstructured data and give your agents accurate, cited context from your knowledge base via API…”.

    Inkeep and Exa all say it too. Buyers may still need it, but shared ground cannot carry your hero — move it lower and give that space to something only you can say.

  2. 2
    Table stakesYour section copy says “Pre-built connectors for Web crawls, Zendesk, GitHub, Slack, PDFs, and more”.

    Keep the fact, lose the position: inkeep and Exa all say it too, and your section is spending its first impression on the same territory as theirs.

  3. 3
    Table stakesYour section copy says “Trusted by category-defining startups and Fortune 500 enterprises”.

    This is the set's common ground — Inkeep and Exa all say it too. It will not set you apart wherever it sits, and in the section it costs you the one place a distinctive claim would be read.

  4. 4
    Table stakesYour body copy says “Every source's auth, pagination, and rate limits, handled”.

    A buyer with three tabs open reads a version of this on every one of them. Say it further down for the readers who need it; the body should carry a claim they will only find here.

  5. 5
    Table stakesYour body copy says “API reference with 206 endpoints”.

    True of you and true of them: Inkeep and Exa all say it too. That is why it decides nothing, and why the body is the wrong place to spend it.

  6. 6
    Table stakesYour body copy says “GitHub repos with 412 files ingested / 12 repos ingested”.

    In the body: Keep the fact, lose the position: inkeep and Exa all say it too, and your body is spending its first impression on the same territory as theirs.

  7. 7
    Table stakesYour body copy says “Documentation source with 1,284 pages / 2,107 pages ingested”.

    In the body: This is the set's common ground — Inkeep and Exa all say it too. It will not set you apart wherever it sits, and in the body it costs you the one place a distinctive claim would be read.

  8. 8
    Table stakesYour body copy says “Product manuals with 86 PDFs ingested”.

    In the body: Inkeep and Exa all say it too. Buyers may still need it, but shared ground cannot carry your body — move it lower and give that space to something only you can say.

  9. 9
    Surface“Every source's auth, pagination, and rate limits, handled” is yours alone, and buyers weigh it.

    Nobody in the set says this. It sits in body copy, where few readers reach it — worth testing higher up the page; only buyers can tell you whether it lands.

The only way to know if it matters.

This report can tell you where your messaging overlaps. It cannot tell you whether a buyer would care, or which of your reasons to buy actually land. Put the page in front of real B2B buyers in your target market and ask them.

Test it with real buyers
Trusted by
HubSpotRingCentralShopifyCognismPaddleVeeamRipplingMiro