Skip to content
Good for Bots

Pathnovo

pathnovo.com

AI agents that parse engineering documents into structured data for enterprise systems

Click badge to copy

Changed your site? Run a new scan.

Claim to rescan

Claiming is free. Verify ownership to request a rescan after deploying your changes.

Unclaimed listings are checked automatically every 90 days.

Good for Bots report

Pathnovo scores 84 of 100: 48 of the 50 base points, 36 of the 50 agent-layer points, no penalties.

Earned

48 of the 50 base points above the line, 36 of the 50 agent-layer points below it.

84

3 prompts away

Copy them below ↓

14

Still open

Behind the checks below, with prompts in Pro.

2
Penalties0
How scoring works · current methodology →

Through a bot’s eyes

“Pathnovo is an engineering document intelligence platform that uses AI agents to parse, reconcile, and convert complex engineering documents into structured data for industrial systems.”

Simulated from the response GoodForBotsBot received on , not a live answer from any AI assistant.

What it can read, and what stays hidden →

What a bot can read

A language model reading this input learns that Pathnovo offers AI-led document intelligence and managed services for parsing engineering records. The provided homepage content details the five-stage ingestion and deployment workflow, case studies with metrics from EPC projects, and specific capabilities such as cross-document reconciliation and compliance monitoring. The provided llms.txt file significantly expands on this by listing exact supported document types, enterprise integrations, regulatory standards, and credit-based pricing tiers. The JSON-LD schema adds structured organization details, including founding date, employee count, and ISO certifications.

What is Pathnovo?

Pathnovo is an engineering document intelligence platform and managed service that builds AI agents to parse, reconcile, and deliver structured data from engineering documents. It is designed for manufacturing, energy, and EPC companies working with complex technical records such as P&IDs, BOMs, datasheets, isometrics, and HAZOP studies.

Read more →

The platform automates tag extraction, datasheet reconciliation, BOM validation, and compliance checks, pushing structured outputs into enterprise systems like SAP PM, IBM Maximo, and AVEVA NET.

3 prompts, 14 points

For Claude Code, Cursor or Codex. The score moves after the next scan.

+10score from 84 to 94

Serve Markdown when asked for it

None of the 4 pages read pass. Asking the homepage for markdown still returns text/html.

Read prompt

+2.5score from 94 to 97

Declare AI usage preferences

No usage declaration found in robots.txt or the final successful homepage response header.

Read prompt

+1.3score from 97 to 98

Expose control roles, states and relationships

77 of 82 assessed homepage controls and role declarations meet the static role, state and relationship checks.

Read prompt

Pro

12 more prompts in Pro

Pro shows a prompt for every fix and drafts your llms.txt and robots.txt. It never changes the score.

$9 once, per listingSee Pro

Deployed your fixes? Rescan to update your score →

Select a check to see what we found and how it is graded.

Readable base

47.9 / 50

Passes.

3.3 of 3.3 points.

robots.txt is valid, with 2 user-agent groups.

Passes.

8.3 of 8.3 points.

Sitemap declared in robots.txt, listing at least 200 URLs.

Passes.

8.3 of 8.3 points.

All 4 pages read pass. The homepage describes itself to machines as Organization, ImageObject, QuantitativeValue in valid schema.org JSON-LD.

Passes.

10 of 10 points.

All 4 pages read pass. The homepage marks its content root, its heading outline and its landmarks.

Passes.

8.3 of 8.3 points.

2 of 2 assessed homepage navigation links expose usable destinations.

Passes.

4.2 of 4.2 points.

74 of 74 assessed homepage links and buttons have a nonempty accessible name.

Passes.

2.5 of 2.5 points.

1 of 1 assessed homepage form fields meet the static naming and field-semantics checks.

Half passes.

1.3 of 2.5 points.

77 of 82 assessed homepage controls and role declarations meet the static role, state and relationship checks.

Passes.

1.7 of 1.7 points.

3 of 4 pages read pass; 1 partly. The page at /pricing declares its canonical URL only in a <link> in the body; HTML head processing and Google's guidance expect it in the head.

Agent layer

36.3 / 50

Passes.

11.3 of 11.3 points.

/llms.txt is well formed: 143 links across 13 sections.

Passes.

7.5 of 7.5 points.

/llms.txt is descriptive: a summary plus descriptions on 100% of links.

Passes.

12.5 of 12.5 points.

/llms-full.txt serves the site's content as markdown (42 KB).

Fails.

0 of 10 points.

None of the 4 pages read pass. Asking the homepage for markdown still returns text/html.

Passes.

5 of 5 points.

States a position on 13 AI crawlers, across 3 purposes (training, search, user-fetch).

Fails.

0 of 2.5 points.

No usage declaration found in robots.txt or the final successful homepage response header.

Penalties

− 0

Passes.

Penalty: 0 of 15 points.

Our crawler reaches the page without being challenged.

Passes.

Penalty: 0 of 20 points.

Nothing on the homepage tells a crawler to skip it or to withhold its snippet.

Passes.

Penalty: 0 of 25 points.

The captured homepage HTML contains 2086 content words; completeness is not assessed.

Passes.

Penalty: 0 of 25 points.

robots.txt allows every known search and user-fetch crawler to reach the homepage.

Passes.

Penalty: 0 of 10 points.

The homepage is served over HTTPS and returns HTTP 200.

Passes.

Penalty: 0 of 3 points.

The homepage provides both a non-empty title and meta description.

Passes.

Penalty: 0 of 3 points.

The homepage is reached in 0 redirects.

Passes.

Penalty: 0 of 3 points.

The homepage reached a plain crawler in 100 ms.

Agent capabilities

0 of 10 · no points

Each one found adds +1 to the badge.

Not applicable.

Adds +1 to the badge when found.

No API catalog published.

Not applicable.

Adds +1 to the badge when found.

No /auth.md published.

Not applicable.

Adds +1 to the badge when found.

No AI Catalog published, so no MCP server card can be discovered.

Not applicable.

Adds +1 to the badge when found.

No A2A Agent Card at /.well-known/agent-card.json, and none listed in an AI Catalog.

Not applicable.

Adds +1 to the badge when found.

No usable Agent Skills index was found at either well-known path or in the AI Catalog.

Not applicable.

Adds +1 to the badge when found.

No WebMCP declarations were found in the collected HTML or inline scripts. External scripts were not inspected.

Not applicable.

Adds +1 to the badge when found.

No Web Bot Auth key directory at /.well-known/http-message-signatures-directory.

Not applicable.

Adds +1 to the badge when found.

No HTTP 402 payment challenge on the homepage or robots.txt, and no manifest at /.well-known/x402.

Not applicable.

Adds +1 to the badge when found.

No ACP discovery declaration established at /.well-known/acp.json.

Not applicable.

Adds +1 to the badge when found.

No UCP discovery declaration established at /.well-known/ucp.

Access & usage policy

Who can read this page, and for what purpose?

GoodForBotsBot received HTTP 200 via a direct request. No indexing restriction was observed in the supported homepage directives.

Sampled URL: https://pathnovo.com/

Declared crawling rules

No matching refusal for this URL was found for the listed identities.

Sources and scope: declared crawling rules
  • Declared allow · Scan admission

    • GoodForBotsBot/1.0 (+https://goodforbots.com/bot)

    The recorded robots admission decision allowed our crawler to proceed.

    GoodForBotsBot admission along the homepage request chain

    https://pathnovo.com/

    robots.txt is present and valid

  • Declared allow · robots.txt

    • goodforbotsbot (Site diagnostics)
    • * (Default robots group)
    • claude-searchbot (ai-search)
    • mistralai-index (ai-search)
    • kimi-searchbot (ai-search)
    • amzn-searchbot (ai-search)
    • meta-webindexer (ai-search)
    • exasearchbot (ai-search)
    • aiwebindex (ai-search)
    • shapbot (ai-search)
    • yandexadditional (ai-search)
    • yandexadditionalbot (ai-search)
    • mistralai-training (ai-training)
    • kimibot (ai-training)
    • ai2bot (Unknown)
    • webzio-extended (ai-training)
    • claude-user (user-fetch)
    • kimi-user (user-fetch)
    • amzn-user (user-fetch)
    • diffbot-user (user-fetch)
    • firecrawlagent (extraction)
    • webzio (extraction)
    • meta-externalfetcher (user-fetch, agent)
    • googlebot (search)
    • bingbot (search)

    No matching refusal for this URL in the applicable robots group. This is a declaration, not an access test.

    Acquisition of https://pathnovo.com/

    https://pathnovo.com/robots.txt

    User-agent: *
    allow: /

    robots.txt is present and validStates a position on AI crawlersAllows search and user-fetch crawlers

  • Declared allow · robots.txt

    • oai-searchbot (ai-search)
    • perplexitybot (ai-search)
    • duckassistbot (ai-search)
    • gptbot (ai-training)
    • claudebot (ai-training)
    • applebot-extended (ai-training)
    • chatgpt-user (user-fetch)
    • perplexity-user (user-fetch)
    • mistralai-user (user-fetch)
    • google-extended (ai-training, grounding)
    • meta-externalagent (ai-training, product-improvement)
    • amazonbot (ai-training, product-improvement)
    • ccbot (open-dataset, ai-training)
    • claude-web (Unknown)
    • anthropic-ai (Unknown)
    • bytespider (Unknown)
    • diffbot (search)
    • cohere-ai (Unknown)
    • cohere-training-data-crawler (Unknown)
    • youbot (Unknown)

    No matching refusal for this URL in the applicable robots group. This is a declaration, not an access test.

    Acquisition of https://pathnovo.com/

    https://pathnovo.com/robots.txt

    User-agent: gptbot
    User-agent: oai-searchbot
    User-agent: chatgpt-user
    User-agent: google-extended
    User-agent: claudebot
    User-agent: claude-web
    User-agent: anthropic-ai
    User-agent: perplexitybot
    User-agent: perplexity-user
    User-agent: applebot-extended
    User-agent: bytespider
    User-agent: ccbot
    User-agent: diffbot
    User-agent: meta-externalagent
    User-agent: cohere-ai
    User-agent: cohere-training-data-crawler
    User-agent: amazonbot
    User-agent: duckassistbot
    User-agent: mistralai-user
    User-agent: youbot
    allow: /

    robots.txt is present and validStates a position on AI crawlersAllows search and user-fetch crawlers

Observed access

GoodForBotsBot received HTTP 200 via a direct request.

Sources and scope: observed access
Coverage and interpretation
  • Crawler coverage includes the registered identities and up to 50 user-agent entries from robots.txt. Source excerpts and detail groups are bounded; omissions are labelled.
  • Observed access describes GoodForBotsBot only. Other crawler entries describe declared rules, not tests of those operators' access.
  • Coverage is the sampled homepage response and collected robots.txt files. No indexing or citation is guaranteed. Usage preferences are declarations, not a legal determination of permission.
  • This explanation adds no points or penalties. Any score effects belong to the linked checks.
  • AIPREF and Content Signals keep their own definitions. AIPREF search can include processing used exclusively for search; training preferences are not a blanket ruling on every search process.

In its category

1st of 20 in AI Products & Agents

Whole category →
  1. 1Pathnovo · this reportpathnovo.com
  2. 2Picturesque AIaipicturesque.com
  3. 2SnapPChartsnappchart.app

Category average 58. The percentile appears once the catalogue holds a few hundred sites.

Scanned

Site profile

Login
Not established
API
Not established
Commerce
Not established

It decides which conditional checks apply. “Not established” means no proof in what we read, not proof of absence.

Scan details

Every scan is kept. Pro draws the history as a chart →

Scanned
3 Oct 2026
Rubric
0.9.0
Requests
23
Fetched
3.3 MB
Took
13.4 s
Homepage
100 ms · direct
Pages read
4
robots.txt
allowed
Which pages the scan read

The homepage and up to three pages from the sitemap and homepage links. Structured data, page structure, Markdown negotiation and typed links check each page read and take the median page; other checks read the homepage.

  • Homepagealways read
  • /blog/agentic-document-processing-ai-agentsarticle · from the sitemap · read
  • /pricingpricing · linked from the homepage · read
  • /blogcollection · from the sitemap · read

More like Pathnovo:AI Products & AgentsAI & Machine LearningB2BEnterprise

Also AI & Machine Learning

goodforbots.com

Evaluates website readability for AI crawlers and language models with a scoring rubric

Same category

peakanswer.com

Tracks AI search visibility, finds competitor gaps, and closes them with automated content and outreach

Same category

app.pallix.in

AI visibility and answer engine optimization platform for marketing teams

Same category

snappchart.app

AI chart analysis tool that grades trade setups from uploaded screenshots

Same category

aipicturesque.com

Freemium AI creative studio combining image, video, audio, and motion models in one workspace

Same category

pic2videoai.com

Browser-based image-to-video tool to turn still photos into animated clips