New · monthly release

Sites like any domain, in one API.

Give any domain and get the sites most like it: competitors, alternatives and look-alike companies, ranked by how many linking sites they share. Found in the Common Crawl web graph and refreshed every month.

Independent tool, not affiliated with Common Crawl or the sites it lists. Public data only.

any domainas the seed
co‑citationshared linking sites
133Mdomains in the graph
monthlyrelease of the web graph
2–13 stypical run
Live data

Why they match.

Two sites are alike when the same websites link to both. Here are the strongest matches for each seed, with the number of linking sites they share. Your run reads the same index.

How a match is found

Three linking sites each link to both your domain and another domain. The more linking sites two domains share, the more similar they are. your site look-alike 3 shared linking sites

Blogs, directories and news sites link to the companies in their field. When the same sites keep linking to two domains, those domains are usually in the same market. We count those shared linkers for every pair in the graph.

What you get

One row per similar site, most similar first.

Each row pairs your seed with one look-alike: how similar they are, how many linking sites they share, and how strong the look-alike is.

Output fields
fieldtypewhat it tells you
domainstringThe seed you asked about.
positionintegerPlace in the list, 1 = most similar.
similarDomainstringA site like the seed.
similaritynumberHow alike the two are, from the share of linking sites they have in common. Higher = closer.
sharedReferrersintegerHow many domains link to both: the evidence behind the match.
similarDomainAuthorityinteger 0–100The look-alike's Open Authority, to tell market leaders from small players.
ccReleasestringThe Common Crawl release the match comes from, e.g. cc-main-2026-jul-aug-sep.
Use cases

Built for people who map markets.

founders & marketers

Competitor discovery

Start from your own domain and find the companies the web groups you with, including the ones you hadn't heard of.

SEO & content teams

"Alternatives to X" research

Build honest alternatives pages and comparison content from the sites the web itself puts next to X.

investors & analysts

Market maps

Seed with one company and get its neighbourhood, with OA to tell leaders from long-tail players.

sales teams

Look-alike lead lists

Take your best customer's domain and find companies like it, ready to enrich and contact.

AI agents

Map a niche from one domain

Over MCP, an agent can expand one domain into a market, then look up authority and links for each player.

How to run

From one domain to its market in seconds.

  1. 01

    Open it on Apify

    Sign in with a free Apify account. No card needed to try.

  2. 02

    Add one or more seeds

    Any domain works. Each seed gets its own ranked list.

  3. 03

    Get your map

    Download JSON, CSV or Excel, call the API, schedule it monthly, or hand it to an AI agent.

{
  "domains": ["apify.com", "booking.com"],
  "maxReferrersPerDomain": 25
}
Reliability

Runs that don't break.

The same factory promises as every CrawlPlant Actor. How the factory works

Monthly release

Every match is computed from the latest Common Crawl web graph.

Answers in seconds

Lookups read a ready index. A typical run takes 2 to 13 seconds.

Evidence on every row

sharedReferrers shows how many linking sites back each match.

Checked every 3 hours

Canary runs repeat known lookups and compare the answers.

Release on every row

ccRelease tells you which month's graph a match comes from.

Pricing

Pricing

$5.00per 1,000 similar sites

You pay for the similar sites you receive. A seed with a 25-site list costs about $0.13.

  • Similarity, shared referrers and OA on every row
  • Many seeds in one run, one dataset
  • No subscription. The free plan is enough to try it
FAQ

About this Actor.

General questions are on the home page. Something else? piotr@crawlplant.com

How does it know two sites are similar?

From who links to them. Sites that the same websites link to are usually in the same market: the same blogs, directories, reviews and news sites cover them together. We count those shared linking sites for every pair in the Common Crawl web graph and rank the matches by it, so every result comes with its evidence in sharedReferrers.

Does it work for huge platforms?

It works best for companies, brands and niche sites, where the sites linking to you say a lot about your market. For a platform the size of facebook.com, almost the whole web links to it, so its matches are other giants rather than a tight market.

Can I use several seeds at once?

Yes. Pass a list of domains and each one gets its own ranked list in the same dataset, with domain on every row so you can group them.

How fresh is it?

Every match comes from the latest monthly release of the Common Crawl web graph, currently cc-main-2026-jul-aug-sep, named on every row in ccRelease.

Can an AI agent use it?

Yes. Add https://mcp.apify.com?tools=crawlplant/similar-sites-finder to any MCP client. Paired with the domain-authority Actor, an agent can map a whole niche from one domain.