Public web data as an API

Web data that keeps growing.

Retail prices, bank rates and energy plans, refreshed every night. And a monthly index of the web's link graph: 133 million domains and 2.15 billion links for SEO work and AI agents. Every record is validated and kept with its history, served in seconds, and you pay only per result.

  • Refreshed nightly or monthly
  • History built in
  • Pay only per result
Pipeline: schedule, crawl, parse, validate, harvest 01 schedule nightly · monthly 02 crawl paced and polite 03 parse typed fields 04 validate diff vs. last run 05 harvest snapshot + history
Catalog

Find your dataset.

Two families, one factory: nightly retail, finance and energy data, and a monthly index of the web's links. Every Actor works the same way. Pick one to see live numbers, sample output and pricing.

  • Need another site?

    Send the site and the fields you need. Demand decides what we plant next.

    Request a scraper
Why CrawlPlant

Scrapers that don't break.

These are promises of the factory, not features of one scraper. Every Actor we publish keeps them.

Runs don't fail

Your run reads from the latest snapshot (last night's for prices, rates and plans, this month's for the link index), not from a live site under load. If a page can't be reached, you still get every other record, each marked with when it was collected.

01latest snapshotdefault
02live fetchif needed
03last known valuewith timestamp
last crawl
records tracked
sites crawled nightly

partial results, never an empty failed run

History and change feeds

Every change between snapshots is kept: nightly for prices and rates, monthly for authority and rank. Chart any value over time, or ask only for what changed.

field   old    new    at
value   12.40 → 11.90  02:14
status  on    → off    02:14

Clean, typed fields

Consistent names, units and types. Stable ids, so the same record keeps the same key forever.

{ "id": "stable",
  "collected_at": "ISO 8601",
  "source_url": "https://…" }

Answers in seconds

The crawl already happened. A run only reads and filters the harvest: link lookups answer in milliseconds and a typical run takes 2 to 13 seconds.

Monitored around the clock

Canary runs check every Actor every 3 hours and alert us before you notice anything.

every 3 h
Built for AI agents

Give your agent data it can trust.

All 8 Actors work as tools in any MCP client through the Apify MCP server. Your agent gets typed records with timestamps and history, so it can say what changed and how fresh each number is.

https://mcp.apify.com?tools=crawlplant/<actor>

Questions agents ask

  • groceriesIs this week's price on 2 L milk a real special, or just back to normal?
  • financeWhich bank pays the most on a 12-month term deposit today?
  • linksWho links to scrapy.org's competitors but not to scrapy.org?
  • linksHow strong is firecrawl.dev, and is it rising?
MCP config · any MCP client
{
  "mcpServers": {
    "crawlplant": {
      "url": "https://mcp.apify.com?tools=crawlplant/commoncrawl-domain-metrics,crawlplant/link-gap-finder,crawlplant/au-bank-rates",
      "headers": { "Authorization": "Bearer <APIFY_TOKEN>" }
    }
  }
}
agent · example
you

Which sites are most like apify.com? I need five to benchmark against.

calls crawlplant/similar-sites-finder · domain apify.com
agent

These five share the most linking sites with apify.com in the latest release:

#similar site
1brightdata.com
2zyte.com
3scrapingbee.com
4scraperapi.com
5octoparse.com
Pricing

Pay per result. That's it.

Every Actor charges per 1,000 results, at the rates below and on its card. Try any of them on Apify's free plan before you spend anything.

Compare rates in the catalog
  1. 01

    Pay per result

    You're billed for the records you receive. A smaller run costs less.

  2. 02

    No subscription

    No seats, no minimums, no contract. Run it once or every hour.

  3. 03

    Free to try

    The free Apify plan is enough to test any Actor on real data.

Prices per 1,000 results for every Actor
Actoryou pay forper 1,000good to know
Retail, finance & energyrefreshed every night · Australia
Woolworths Australiaproducts$0.30$0.50 on Apify's Free plan
Coles Australiaproducts (store rows included)$0.50$0.80 on Apify's Free plan
Australian Bank Ratesrates$1.00$1.50 on Free · rate-change events priced separately
Australian Energy Plansplans$1.25$2.00 on Free · plan-change events priced separately
SEO & link datamonthly web-graph index · global
Bulk Domain Authority & Referring Domainsdomain profiles$1.00down to $0.70 on Apify's Gold plan
rank history rows$1.50rank history since 2018
link rows$2.00referrers, link gap, similar sites and outgoing-links modes
Referring Domains Checkerreferring domains$2.00strongest first, each with its OA
Link Gap Finderprospects$4.00hubs and infrastructure left out by default
Similar Sites Findersimilar sites$5.00works for any domain
FAQ

Questions, answered.

Anything else? Write to piotr@crawlplant.com.

How fresh is the data?

Retail, finance and energy sources are crawled every night. The link-graph Actors read an index of the Common Crawl web graph that we rebuild every month, with 12 monthly releases of rank history. Every record carries a timestamp or release name, so you always know how fresh a value is.

What if a site changes or blocks requests?

Your run is served from the latest snapshot, so it keeps working. Canary checks every 3 hours tell us about a change long before it reaches you, and we fix it before the next refresh. You get partial results, never an empty failed run.

What is Open Authority (OA)?

A domain's rank in the web's link graph, on a 0–100 log scale. It is computed from the Common Crawl web graph every month, so it is open and repeatable. OA 60+ is roughly the top 150,000 domains; OA 80+ roughly the top 5,000.

Why do referring-domain counts move from month to month?

Each monthly crawl samples a different set of pages, so around 17% of referring domains change between releases even when nothing happened. We label those rows "newly seen" and "no longer seen", and use the 3- and 12-month trends for real direction.

What formats can I get?

Each Actor runs on the Apify platform. Results come as JSON, CSV or Excel from the Console, through the Apify API or client libraries, on a schedule, or as a tool for AI agents via MCP.

Is this legal? What data do you collect?

We collect publicly available data only: no personal data, nothing behind a login. Crawls are paced and run at night to keep load on each site low. How you use the data is up to you, under the source's terms and your local law.

Can you build a scraper for another site?

Yes. Email us the site and the fields you need. We prioritise by demand and let you know when it's live.