Skip to content
Good for Bots

StartupIdeasDB

startupideasdb.com

AI-powered startup research platform with validated market opportunities and business blueprints

Click badge to copy

Changed your site? Run a new scan.

Claim to rescan

Claiming is free. Verify ownership to request a rescan after deploying your changes.

Unclaimed listings are checked automatically every 90 days.

Good for Bots report

StartupIdeasDB scores 41 of 100: 36 of the 50 base points, 6 of the 50 agent-layer points, no penalties.

Earned

36 of the 50 base points above the line, 6 of the 50 agent-layer points below it.

41

3 prompts away

Copy them below ↓

32

Taken back

Earned, then taken back by penalties.

1

Still open

Behind the checks below, with prompts in Pro.

26
Penalties0
How scoring works · current methodology →

Through a bot’s eyes

“StartupIdeasDB is an AI-powered startup research platform providing over 12,000 validated business ideas, market opportunity scores, execution roadmaps, and problem statements sourced from various online platforms.”

Simulated from the response GoodForBotsBot received on , not a live answer from any AI assistant.

What it can read, and what stays hidden →

What a bot can read

A language model reading this input learns that StartupIdeasDB is a research platform offering 12,000 plus validated startup ideas derived from online communities and review sites. The text outlines specific features, including problem statement databases, pain point collections, an AI builder for whitepapers and technical guides, landing page creation tools, beginner journals, and unicorn case studies. The homepage markdown provides user testimonials, platform metrics, and core feature descriptions. The provided llms.txt file expands significantly on this by detailing specific data sources such as subreddits, target user segments, a comprehensive list of guide topics on business formation and legal structures across multiple countries, pricing plan locations, technical stack details, and AI usage guidelines including attribution rules and restrictions.

What stays hidden

The input does not state the exact pricing numbers for the Beginner and Pro plans, only providing a URL link to the pricing page. It also lacks specific instructions on how to access the database API, as the text notes no public API is currently available and requires direct email contact for partnerships.

What is StartupIdeasDB?

StartupIdeasDB is an AI-powered startup research platform designed for aspiring entrepreneurs, first-time founders, and business students. It offers a database of over 12,000 validated business ideas and real-world problem statements gathered from sources such as Reddit, Product Hunt, Hacker News, Upwork, and user reviews on app stores.

Read more →

The platform provides AI opportunity scores, market size estimates, target audience analysis, revenue models, execution roadmaps, and case studies of successful startups. Additionally, it features tools for generating business whitepapers, technical guides, and landing pages to help users research, validate, and launch products.

3 prompts, 32 points

For Claude Code, Cursor or Codex. The score moves after the next scan.

+12.5score from 41 to 54

Serve Markdown at its own URL

No /llms-full.txt and no markdown versions linked from llms.txt.

Read prompt

+10score from 54 to 64

Serve Markdown when asked for it

None of the 4 pages read pass. Asking the homepage for markdown still returns text/html.

Read prompt

+8.9score from 64 to 73

Describe the site with schema.org JSON-LD

2 of 4 pages read pass. The homepage publishes no structured data at all: no JSON-LD, no microdata, no RDFa.

Read prompt

Pro

16 more prompts in Pro

Pro shows a prompt for every fix and drafts your llms.txt and robots.txt. It never changes the score.

$9 once, per listingSee Pro

Deployed your fixes? Rescan to update your score →

Ready-made files

Pro drafts these for this site.

Pro

A ready-made llms.txt and robots.txt

Pro drafts both files for StartupIdeasDB from its scan, keeps only links and rules our parsers can verify, and hands them over ready to review and publish. It never changes the score.

See Pro

Select a check to see what we found and how it is graded.

Readable base

35.7 / 50

Passes.

3.6 of 3.6 points.

robots.txt is valid, with 1 user-agent group.

Passes.

8.9 of 8.9 points.

Sitemap declared in robots.txt, listing at least 200 URLs.

Fails.

0 of 8.9 points.

2 of 4 pages read pass. The homepage publishes no structured data at all: no JSON-LD, no microdata, no RDFa.

Half passes.

5.4 of 10.7 points.

None of the 4 pages read pass fully; 4 partly. The homepage is partly structured: The outline jumps from h1 to h3, which the standard forbids.

Prompt in Pro →
Passes.

8.9 of 8.9 points.

6 of 6 assessed homepage navigation links expose usable destinations.

Passes.

4.5 of 4.5 points.

74 of 74 assessed homepage links and buttons have a nonempty accessible name.

Not applicable.

0 of 0 points.

No applicable form fields were observed in the homepage HTML.

Passes.

2.7 of 2.7 points.

74 of 74 assessed homepage controls and role declarations meet the static role, state and relationship checks.

Passes.

1.8 of 1.8 points.

All 4 pages read pass. The homepage declares itself as its canonical URL.

Agent layer

5.6 / 50

Half passes.

5.6 of 11.3 points.

/llms.txt has a title but lists no links to follow.

Prompt in Pro →
Fails.

0 of 7.5 points.

/llms.txt lists no links, so there is nothing to describe.

Prompt in Pro →
Fails.

0 of 12.5 points.

No /llms-full.txt and no markdown versions linked from llms.txt.

Fails.

0 of 10 points.

None of the 4 pages read pass. Asking the homepage for markdown still returns text/html.

Fails.

0 of 5 points.

robots.txt names no AI crawlers, so every operator has to guess.

Prompt in Pro →
Fails.

0 of 2.5 points.

No usage declaration found in robots.txt or the final successful homepage response header.

Prompt in Pro →

Penalties

− 0

Passes.

Penalty: 0 of 15 points.

Our crawler reaches the page without being challenged.

Passes.

Penalty: 0 of 20 points.

Nothing on the homepage tells a crawler to skip it or to withhold its snippet.

Passes.

Penalty: 0 of 25 points.

The captured homepage HTML contains 1214 content words; completeness is not assessed.

Passes.

Penalty: 0 of 25 points.

robots.txt allows every known search and user-fetch crawler to reach the homepage.

Passes.

Penalty: 0 of 10 points.

The homepage is served over HTTPS and returns HTTP 200.

Passes.

Penalty: 0 of 3 points.

The homepage provides both a non-empty title and meta description.

Passes.

Penalty: 0 of 3 points.

The homepage is reached in 0 redirects.

Passes.

Penalty: 0 of 3 points.

The homepage reached a plain crawler in 40 ms.

Agent capabilities

0 of 10 · no points

Each one found adds +1 to the badge.

Not applicable.

Adds +1 to the badge when found.

No API catalog published.

Not applicable.

Adds +1 to the badge when found.

No /auth.md published.

Not applicable.

Adds +1 to the badge when found.

No AI Catalog published, so no MCP server card can be discovered.

Not applicable.

Adds +1 to the badge when found.

No A2A Agent Card at /.well-known/agent-card.json, and none listed in an AI Catalog.

Not applicable.

Adds +1 to the badge when found.

No usable Agent Skills index was found at either well-known path or in the AI Catalog.

Not applicable.

Adds +1 to the badge when found.

No WebMCP declarations were found in the collected HTML or inline scripts. External scripts were not inspected.

Not applicable.

Adds +1 to the badge when found.

No Web Bot Auth key directory at /.well-known/http-message-signatures-directory.

Not applicable.

Adds +1 to the badge when found.

No HTTP 402 payment challenge on the homepage or robots.txt, and no manifest at /.well-known/x402.

Not applicable.

Adds +1 to the badge when found.

No ACP discovery declaration established at /.well-known/acp.json.

Not applicable.

Adds +1 to the badge when found.

No UCP discovery declaration established at /.well-known/ucp.

Access & usage policy

Who can read this page, and for what purpose?

GoodForBotsBot received HTTP 200 via a direct request. No indexing restriction was observed in the supported homepage directives.

Sampled URL: https://startupideasdb.com/

Declared crawling rules

No matching refusal for this URL was found for the listed identities.

Sources and scope: declared crawling rules
  • Declared allow · Scan admission

    • GoodForBotsBot/1.0 (+https://goodforbots.com/bot)

    The recorded robots admission decision allowed our crawler to proceed.

    GoodForBotsBot admission along the homepage request chain

    https://startupideasdb.com/

    robots.txt is present and valid

  • Declared allow · robots.txt

    • goodforbotsbot (Site diagnostics)
    • * (Default robots group)
    • oai-searchbot (ai-search)
    • claude-searchbot (ai-search)
    • perplexitybot (ai-search)
    • mistralai-index (ai-search)
    • kimi-searchbot (ai-search)
    • amzn-searchbot (ai-search)
    • meta-webindexer (ai-search)
    • exasearchbot (ai-search)
    • aiwebindex (ai-search)
    • duckassistbot (ai-search)
    • shapbot (ai-search)
    • yandexadditional (ai-search)
    • yandexadditionalbot (ai-search)
    • gptbot (ai-training)
    • claudebot (ai-training)
    • mistralai-training (ai-training)
    • kimibot (ai-training)
    • ai2bot (Unknown)
    • applebot-extended (ai-training)
    • webzio-extended (ai-training)
    • chatgpt-user (user-fetch)
    • claude-user (user-fetch)
    • perplexity-user (user-fetch)
    • mistralai-user (user-fetch)
    • kimi-user (user-fetch)
    • amzn-user (user-fetch)
    • diffbot-user (user-fetch)
    • firecrawlagent (extraction)
    • webzio (extraction)
    • meta-externalfetcher (user-fetch, agent)
    • google-extended (ai-training, grounding)
    • meta-externalagent (ai-training, product-improvement)
    • amazonbot (ai-training, product-improvement)
    • ccbot (open-dataset, ai-training)
    • googlebot (search)
    • bingbot (search)

    No matching refusal for this URL in the applicable robots group. This is a declaration, not an access test.

    Acquisition of https://startupideasdb.com/

    https://startupideasdb.com/robots.txt

    User-agent: *
    allow: /

    robots.txt is present and validStates a position on AI crawlersAllows search and user-fetch crawlers

Observed access

GoodForBotsBot received HTTP 200 via a direct request.

Sources and scope: observed access
Coverage and interpretation
  • Crawler coverage includes the registered identities and up to 50 user-agent entries from robots.txt. Source excerpts and detail groups are bounded; omissions are labelled.
  • Observed access describes GoodForBotsBot only. Other crawler entries describe declared rules, not tests of those operators' access.
  • Coverage is the sampled homepage response and collected robots.txt files. No indexing or citation is guaranteed. Usage preferences are declarations, not a legal determination of permission.
  • This explanation adds no points or penalties. Any score effects belong to the linked checks.
  • AIPREF and Content Signals keep their own definitions. AIPREF search can include processing used exclusively for search; training preferences are not a blanket ruling on every search process.

In its category

17th of 20 in AI Products & Agents

Whole category →
  1. 1Pathnovopathnovo.com
  2. 2Picturesque AIaipicturesque.com
  3. 17StartupIdeasDB · this reportstartupideasdb.com

The prompts above would take StartupIdeasDB to 73, 5th in its category.

To the prompts ↑

Category average 58. The percentile appears once the catalogue holds a few hundred sites.

Scanned

Site profile

Login
Detected
A sign-in or sign-up link
API
Not established
Commerce
Not established

It decides which conditional checks apply. “Not established” means no proof in what we read, not proof of absence.

Scan details

Every scan is kept. Pro draws the history as a chart →

Scanned
3 Oct 2026
Rubric
0.9.0
Requests
22
Fetched
1.5 MB
Took
14 s
Homepage
40 ms · direct
Pages read
4
robots.txt
allowed
Which pages the scan read

The homepage and up to three pages from the sitemap and homepage links. Structured data, page structure, Markdown negotiation and typed links check each page read and take the median page; other checks read the homepage.

  • Homepagealways read
  • /blog/ai-agent-ideasarticle · from the sitemap · read
  • /pricingpricing · linked from the homepage · from the sitemap · read
  • /blogcollection · linked from the homepage · from the sitemap · read

More like StartupIdeasDB:AI Products & AgentsProduct ManagementAI & Machine LearningStartups

Also AI & Machine Learning

goodforbots.com

Evaluates website readability for AI crawlers and language models with a scoring rubric

Same category

feelpair.com

AI-assisted live communication and mediation tool for couples experiencing conflict

Same category

removemate.com

Automated AI background removal for product photos, portraits and pets

Same category

darkharbor.ai

Digital employees for reception, lead follow-up, and administrative tasks

Same category

skipbait.app

AI YouTube video summarizer and chat extension for Chromium browsers

Same category

ai.cosolution.cc

AI customer service operations platform with unified multi-channel inbox and KPI tracking