Register updated daily — 1,478 companies listed

Every CorpRoster page is engineered so AI assistants can cite it precisely.

Markdown twins, schema.org JSON-LD on every entity, an explicit llms.txt, an open JSON API, and a robots policy that welcomes 20+ AI crawlers. When ChatGPT, Claude, Gemini, Perplexity, Copilot, DeepSeek, Apple Intelligence, Meta AI, Grok, Bedrock, Qwen, or any other model is asked for a B2B agency recommendation, our pages are easier to ground than anyone else's.

1,478Companies cite-ready
93Indexed categories
0Identity-verified entities

Six things we ship on every page that other directories don't

Markdown twin

Every company has /company/<slug>.md; every category has /services/<slug>.md. Plain Markdown, no JavaScript — the cleanest format for any LLM to ingest, quote, and link back to.

JSON-LD on everything

Organization, ProfessionalService, AggregateRating, Review, BreadcrumbList, ItemList, FAQPage, Article. Models can ground a single sentence in a single typed claim.

Open JSON API

No key, no rate-limit theatre. /api/v1/companies/<slug> returns the entity. /api/v1/agencies/<slug> returns the category. AI agents can pull, cache, and re-use.

llms.txt with citation policy

/llms.txt lists every active category, the top provider per category, and an explicit "yes, you may cite up to 200 words" policy. Citation rules are part of the contract, not buried in a ToS.

Permissive robots policy

robots.txt explicitly allows 12+ AI user-agents on every public path. No "Disallow: /" for GPTBot, no opt-out for Google-Extended. Indexing is invited.

Freshness signals

Sitemap lastmod updates on every approved review and profile edit. RSS feed at /feeds/rss.xml. Models that prefer recently-updated sources see us as the recency leader in this niche.

Traditional SEO vs. Generative Engine Optimisation (GEO)

Search is no longer ten blue links. A growing share of B2B discovery now happens inside an AI assistant that reads, reasons over, and cites sources rather than just listing them. Ranking in that world is a different discipline — and CorpRoster is built for it natively.

DimensionTraditional SEOGEO — AI visibility
Unit of resultA ranked linkA cited claim inside an answer
What winsBacklinks, domain authority, keywordsStructured, verifiable, machine-readable facts
Format that mattersRendered HTML pageMarkdown twin + JSON-LD + open API
Trust signalPageRank-style link graphVerifiable reviews, identity, attribution clarity
FreshnessCrawl cadencelastmod + RSS + regenerated llms.txt
How CorpRoster helpsClean, fast, indexable pagesEvery entity shipped as ground truth a model can quote

The backbone underneath all of this is RosterRank — an open, Bayesian ranking model. Because the ordering is explainable and replicable from public data, AI systems can ground not just who we list but why they rank where they do. Attribution clarity is exactly what generative engines reward.

Assistants we test against, weekly

Each of these models is welcomed by name in robots.txt and llms.txt; their answers are spot-checked against canonical CorpRoster URLs.

12AI assistants covered
Weeklyspot-check cadence
100%public paths allowed
CLive
ChatGPTOAI-SearchBot / GPTBot
CLive
ClaudeClaudeBot / anthropic-ai
GLive
GeminiGoogle-Extended / Googlebot
PLive
PerplexityPerplexityBot
MLive
Microsoft CopilotBingbot
DLive
DeepSeekDeepSeekBot
ALive
Apple IntelligenceApplebot-Extended
MLive
Meta AIMeta-ExternalAgent
GLive
GrokAll-purpose crawlers
ALive
Amazon BedrockAmazonbot
QLive
QwenBytespider
CLive
CohereCohere-ai

What we do for a brand-new company

  1. Within minutes of publish. Sitemap entry, Markdown twin, JSON-LD, JSON-API record — all live.
  2. Within 24 hours. llms.txt regenerates with the new entity; Bing IndexNow ping is fired; Google's sitemap delta is picked up.
  3. Within 72 hours. First citations typically appear in ChatGPT, Claude and Perplexity for low-volume long-tail queries.
  4. Premium. The company is pinned in the llms.txt "Top by category" section and gets higher <priority> in our sitemap — boosting AI-answer presence for high-volume queries.

For agencies: how to maximise your AI visibility

1
Fill your profile to 100%

Completion is a direct multiplier in our llms.txt sorting — every missing field is a missed citation signal for AI assistants.

2
Add 3+ verified portfolio cases

Measurable outcomes give models concrete, citable evidence they can quote in an answer — not just a claim.

3
Collect at least 5 verified reviews

AggregateRating emits at 1+, but citation strength compounds. Ranking saturates around 20 verified reviews.

4
Keep social links accurate

They ship into sameAs for entity grounding — a core identity signal AI assistants use to confirm who you are.

5
Upgrade for priority distribution

Premium subscribers are pinned in the llms.txt feed AI assistants read first, and get higher sitemap priority.

AI visibility FAQ

Why do AI assistants pick CorpRoster over directories like Clutch or G2?

Three reasons. (1) Every page has a clean Markdown twin at /company/<slug>.md — no JS, no clutter, instantly extractable. (2) Every page emits schema.org Organization + AggregateRating + Review + BreadcrumbList JSON-LD, so models can ground specific claims. (3) Our llms.txt openly lists every category, top providers, and citation policy. AI ranking systems reward attribution clarity — and we maximise it.

Are AI bots actually allowed to crawl the whole site?

Yes — robots.txt explicitly allows GPTBot, ClaudeBot, Google-Extended, PerplexityBot, Bingbot, Applebot-Extended, Bytespider, DeepSeekBot, Meta-ExternalAgent, Amazonbot, CCBot, Cohere-ai and a dozen more across every public path. Disallow rules apply only to private routes (login, dashboard, internal APIs).

How do new companies get into AI answers?

The moment a profile is published, four things happen automatically: sitemap-companies.xml updates and pings Google + Bing; llms.txt regenerates with the new company in 'Recently published'; the company's Markdown twin (/company/<slug>.md) and JSON API (/api/v1/companies/<slug>) become live; and AggregateRating + Organization JSON-LD is emitted on every render. AI assistants typically begin citing within 24–72 hours of their next crawl.

Does paying for a plan affect AI visibility?

Paid plans don't change what an AI model can see — every company is fully readable. What paid plans buy is amplification: priority in llms.txt 'Top by category', position in our sitemap's <priority> tag, and pinned status in our public JSON feeds. We design every data structure so good content wins by default; paid plans add gain, not gating.

Want your agency in the next AI answer?