Skip to content

Documents12 companies

12 companies, 12 swept. The median identity score is 84 of 100. Rows are ordered by readiness; rows not yet rated sit last.

12 companies · 12 swept · last sweep 2026-09-16

Identity scores in Documentsswept12 swept

84/100

median

0identity score, 0-100100
12 swept

Ranked by readiness

  • #1

    Postman

    postman.com

    Designs, tests, and manages APIs with integrated documentation, mocking, monitoring, and workflow automation.

    bots in llms.txt

    100/100

    #1 of 12 in Documents
  • #2

    GitBook

    gitbook.com

    Hosts documentation with tools to detect stale content and verify accuracy against product changes.

    bots in llms.txt

    86/100

    #2 of 12 in Documents
  • #3

    Rossum

    rossum.ai

    Extracts data from transactional documents, validates against master data, and routes to downstream systems.

    bots in llms.txt

    84/100

    #3 of 12 in Documents
  • #4

    Instabase

    instabase.com

    Extracts and validates data from complex document packets using multi-model AI agents with cross-document understanding.

    bots in llms.txt

    78/100

    #4 of 12 in Documents
  • #5

    Dbt labs

    getdbt.com

    Transforms raw data into tested, version-controlled SQL models for analytics and AI pipelines.

    bots partly llms.txt

    69/100

    #5 of 12 in Documents
  • #6

    Mindee

    mindee.com

    Extracts structured data from documents via API, supporting OCR, classification, splitting, and cropping across any format or language.

    bots in llms.txt

    69/100

    #6 of 12 in Documents
  • #7

    Nanonets

    nanonets.com

    Extracts data from business documents and automates enterprise workflows like accounts payable, order management, and claims processing.

    bots partly llms.txt

    59/100

    #7 of 12 in Documents
  • #8

    Affinda

    affinda.com

    Extracts and validates data from documents using AI, handling handwriting, tables, and multiple formats.

    bots partly llms.txt

    58/100

    #8 of 12 in Documents
  • #9

    Sensible

    sensible.so

    Extracts structured data from documents using hybrid AI and rule-based methods, returning validated JSON.

    bots in llms.txt

    56/100

    #9 of 12 in Documents
  • #10

    DocuPanda

    docupanda.io

    Extracts structured data from documents of any layout, language, or format using machine learning.

    bots in llms.txt

    54/100

    #10 of 12 in Documents
  • #11

    Swimm

    swimm.io

    Analyzes codebases to extract architecture, dependencies, and business logic for modernization and migration projects.

    bots partly llms.txt

    34/100

    #11 of 12 in Documents
  • #12

    Klarity (check domain)

    klarity.com

    No line written yet.

    bots out llms.txt

    8/100

    #12 of 12 in Documents

How to read a row

swept
Read by a crawler: HTTP and parsing, no model call. Every row, every week.
measured
The full analyzer run, asked of four engines. Shown only where the company agreed to show it.

Can a retrieval engine find, resolve and quote this site? Four parts, each read by the sweep, each shown beside the total.

Entity 25 · Machine-readable 35 · Retrieval access 20 · Footprint 20

A part that was never read is left out and the total is rescaled to what was. It is never counted as zero.

Measure your own site