Sites like any domain, in one API.
Give any domain and get the sites most like it: competitors, alternatives and look-alike companies, ranked by how many linking sites they share. Found in the Common Crawl web graph and refreshed every month.
Independent tool, not affiliated with Common Crawl or the sites it lists. Public data only.
Why they match.
Two sites are alike when the same websites link to both. Here are the strongest matches for each seed, with the number of linking sites they share. Your run reads the same index.
Top 5 matches for
- shared
How a match is found
Blogs, directories and news sites link to the companies in their field. When the same sites keep linking to two domains, those domains are usually in the same market. We count those shared linkers for every pair in the graph.
One row per similar site, most similar first.
Each row pairs your seed with one look-alike: how similar they are, how many linking sites they share, and how strong the look-alike is.
| field | type | what it tells you |
|---|---|---|
domain | string | The seed you asked about. |
position | integer | Place in the list, 1 = most similar. |
similarDomain | string | A site like the seed. |
similarity | number | How alike the two are, from the share of linking sites they have in common. Higher = closer. |
sharedReferrers | integer | How many domains link to both: the evidence behind the match. |
similarDomainAuthority | integer 0–100 | The look-alike's Open Authority, to tell market leaders from small players. |
ccRelease | string | The Common Crawl release the match comes from, e.g. cc-main-2026-jul-aug-sep. |
Built for people who map markets.
Competitor discovery
Start from your own domain and find the companies the web groups you with, including the ones you hadn't heard of.
"Alternatives to X" research
Build honest alternatives pages and comparison content from the sites the web itself puts next to X.
Market maps
Seed with one company and get its neighbourhood, with OA to tell leaders from long-tail players.
Look-alike lead lists
Take your best customer's domain and find companies like it, ready to enrich and contact.
Map a niche from one domain
Over MCP, an agent can expand one domain into a market, then look up authority and links for each player.
From one domain to its market in seconds.
- 01
Open it on Apify
Sign in with a free Apify account. No card needed to try.
- 02
Add one or more seeds
Any domain works. Each seed gets its own ranked list.
- 03
Get your map
Download JSON, CSV or Excel, call the API, schedule it monthly, or hand it to an AI agent.
{
"domains": ["apify.com", "booking.com"],
"maxReferrersPerDomain": 25
}
# sites most like ahrefs.com
curl -X POST \
"https://api.apify.com/v2/acts/crawlplant~similar-sites-finder/run-sync-get-dataset-items?token=$APIFY_TOKEN" \
-H "Content-Type: application/json" \
-d '{ "domains": ["ahrefs.com"] }'
# 1. Add these MCP tools to your agent
https://mcp.apify.com?tools=crawlplant/similar-sites-finder,crawlplant/commoncrawl-domain-metrics
# 2. Ask in plain language
Map the market around allbirds.com: find similar brands,
then rank them by Open Authority and tell me which are growing.
Runs that don't break.
The same factory promises as every CrawlPlant Actor. How the factory works
Monthly release
Every match is computed from the latest Common Crawl web graph.
Answers in seconds
Lookups read a ready index. A typical run takes 2 to 13 seconds.
Evidence on every row
sharedReferrers shows how many linking sites back each match.
Checked every 3 hours
Canary runs repeat known lookups and compare the answers.
Release on every row
ccRelease tells you which month's graph a match comes from.
Pricing
You pay for the similar sites you receive. A seed with a 25-site list costs about $0.13.
- Similarity, shared referrers and OA on every row
- Many seeds in one run, one dataset
- No subscription. The free plan is enough to try it
How does it know two sites are similar?
From who links to them. Sites that the same websites link to are usually in the same market: the same blogs, directories, reviews and news sites cover them together. We count those shared linking sites for every pair in the Common Crawl web graph and rank the matches by it, so every result comes with its evidence in sharedReferrers.
Does it work for huge platforms?
It works best for companies, brands and niche sites, where the sites linking to you say a lot about your market. For a platform the size of facebook.com, almost the whole web links to it, so its matches are other giants rather than a tight market.
Can I use several seeds at once?
Yes. Pass a list of domains and each one gets its own ranked list in the same dataset, with domain on every row so you can group them.
How fresh is it?
Every match comes from the latest monthly release of the Common Crawl web graph, currently cc-main-2026-jul-aug-sep, named on every row in ccRelease.
Can an AI agent use it?
Yes. Add https://mcp.apify.com?tools=crawlplant/similar-sites-finder to any MCP client. Paired with the domain-authority Actor, an agent can map a whole niche from one domain.