Public web data as an API

Web data that keeps growing.

Retail prices, bank rates and energy plans, refreshed every night. A monthly index of the web's link graph: 133 million domains and 2.15 billion links for SEO work and AI agents. And every tennis match since 1990, with odds and Elo, updated every hour. Every record is validated and kept with its history, served in seconds, and you pay only per result.

  • Refreshed nightly or monthly
  • History built in
  • Pay only per result
Pipeline: schedule, crawl, parse, validate, harvest 01 schedule nightly · monthly 02 crawl paced and polite 03 parse typed fields 04 validate diff vs. last run 05 harvest snapshot + history
Catalog

Find your dataset.

Three families, one factory: nightly retail, finance and energy data, a monthly index of the web's links, and tennis data updated every hour. Every Actor works the same way. Pick one to see live numbers, sample output and pricing.

  • Need another site?

    Send the site and the fields you need. Demand decides what we plant next.

    Request a scraper
Sports data

Every tennis match since 1990, as an API.

ATP, WTA, Challenger and ITF, live and back to 1990, updated every hour. Each match comes with odds from about 20 bookmakers, career head-to-head, both players' Elo and their form. Ask it from your code, a spreadsheet or your AI agent.

Sinner v Alcaraz · head-to-head21 meetings
Sinner v Alcaraz, head-to-head by surface: wins of each player
surfaceshare of winsSinnerAlcaraz
hard57
clay24
grass20
exhibition01
all 21912

Last meeting on 2026-04-12, Monte Carlo, clay: Sinner won 7-6(5) 6-3.

  • Who is the favourite?

    Elo and the market, side by side. Tsitsipas v Hijikata on 2026-09-29: Elo gave Tsitsipas 64%, the fair probability from 19 UK bookmakers 73%. He won 6-4 6-3.

    Flashscore Tennis · $1.00 / 1,000 matches
  • What is their head-to-head?

    Every meeting, with the surface and the score. Sinner v Alcaraz: 21 meetings, 9–12. Sabalenka v Swiatek: 15 meetings, 6–9.

    mode h2h · $5.00 / 1,000 head-to-heads
  • Who was No. 1 that week?

    The ranking of any week, ATP from 1973 and WTA from 2001. On 2008-08-18 Nadal led with 6,700 points, ahead of Federer on 5,930.

    mode rankings · $1.00 / 1,000 ranking rows
  • How does a player do against the top 10?

    Record and serve and return statistics by season, surface, tour or opponent's rank. Sinner against the top 10: 74–41, a 64.3% win rate.

    mode playerStats · $5.00 / 1,000 statistics rows
matches since 1990
950k+
players with an Elo rating
13k+
ATP ranking weeks since 1973
2,383
matches with point-by-point
250k+

Elo rates each player's strength right before every match, overall and on each surface. It is built from every match since 1990 and recalculated every night, so you can set it against the bookmakers' fair probability.

Why CrawlPlant

Scrapers that don't break.

These are promises of the factory, not features of one scraper. Every Actor we publish keeps them.

Runs don't fail

Your run reads from the latest snapshot (last night's for prices, rates and plans, this month's for the link index), not from a live site under load. If a page can't be reached, you still get every other record, each marked with when it was collected.

01latest snapshotdefault
02live fetchif needed
03last known valuewith timestamp
last crawl
records tracked
sites crawled nightly

partial results, never an empty failed run

History and change feeds

Every change between snapshots is kept: nightly for prices and rates, monthly for authority and rank. Chart any value over time, or ask only for what changed.

field   old    new    at
value   12.40 → 11.90  02:14
status  on    → off    02:14

Clean, typed fields

Consistent names, units and types. Stable ids, so the same record keeps the same key forever.

{ "id": "stable",
  "collected_at": "ISO 8601",
  "source_url": "https://…" }

Answers in seconds

The crawl already happened. A run only reads and filters the harvest: link lookups answer in milliseconds and a typical run takes 2 to 13 seconds.

Monitored around the clock

Canary runs check every Actor every 3 hours and alert us before you notice anything.

every 3 h
Built for AI agents

Give your agent data it can trust.

All 8 Actors work as tools in any MCP client through the Apify MCP server. Your agent gets typed records with timestamps and history, so it can say what changed and how fresh each number is.

https://mcp.apify.com?tools=crawlplant/<actor>

Questions agents ask

  • groceriesIs this week's price on 2 L milk a real special, or just back to normal?
  • financeWhich bank pays the most on a 12-month term deposit today?
  • linksWho links to scrapy.org's competitors but not to scrapy.org?
  • linksHow strong is firecrawl.dev, and is it rising?
MCP config · any MCP client
{
  "mcpServers": {
    "crawlplant": {
      "url": "https://mcp.apify.com?tools=crawlplant/commoncrawl-domain-metrics,crawlplant/link-gap-finder,crawlplant/au-bank-rates",
      "headers": { "Authorization": "Bearer <APIFY_TOKEN>" }
    }
  }
}
agent · example
you

Which sites are most like apify.com? I need five to benchmark against.

calls crawlplant/similar-sites-finder · domain apify.com
agent

These five share the most linking sites with apify.com in the latest release:

#similar site
1brightdata.com
2zyte.com
3scrapingbee.com
4scraperapi.com
5octoparse.com
Pricing

Pay per result. That's it.

Every Actor charges per 1,000 results, at the rates below and on its card. Try any of them on Apify's free plan before you spend anything.

Compare rates in the catalog
  1. 01

    Pay per result

    You're billed for the records you receive. A smaller run costs less.

  2. 02

    No subscription

    No seats, no minimums, no contract. Run it once or every hour.

  3. 03

    Free to try

    The free Apify plan is enough to test any Actor on real data.

Prices per 1,000 results for every Actor
Actoryou pay forper 1,000good to know
Retail, finance & energyrefreshed every night · Australia
Woolworths Australiaproducts$0.30$0.50 on Apify's Free plan
Coles Australiaproducts (store rows included)$0.50$0.80 on Apify's Free plan
Australian Bank Ratesrates$1.00$1.50 on Free · rate-change events priced separately
Australian Energy Plansplans$1.25$2.00 on Free · plan-change events priced separately
SEO & link datamonthly web-graph index · global
Bulk Domain Authority & Referring Domainsdomain profiles$1.00down to $0.70 on Apify's Gold plan
rank history rows$1.50rank history since 2018
link rows$2.00referrers, link gap, similar sites and outgoing-links modes
Referring Domains Checkerreferring domains$2.00strongest first, each with its OA
Link Gap Finderprospects$4.00hubs and infrastructure left out by default
Similar Sites Findersimilar sites$5.00works for any domain
FAQ

Questions, answered.

Anything else? Write to piotr@crawlplant.com.

How fresh is the data?

Retail, finance and energy sources are crawled every night. The link-graph Actors read an index of the Common Crawl web graph that we rebuild every month, with 12 monthly releases of rank history. Every record carries a timestamp or release name, so you always know how fresh a value is.

What if a site changes or blocks requests?

Your run is served from the latest snapshot, so it keeps working. Canary checks every 3 hours tell us about a change long before it reaches you, and we fix it before the next refresh. You get partial results, never an empty failed run.

What is Open Authority (OA)?

A domain's rank in the web's link graph, on a 0–100 log scale. It is computed from the Common Crawl web graph every month, so it is open and repeatable. OA 60+ is roughly the top 150,000 domains; OA 80+ roughly the top 5,000.

Why do referring-domain counts move from month to month?

Each monthly crawl samples a different set of pages, so around 17% of referring domains change between releases even when nothing happened. We label those rows "newly seen" and "no longer seen", and use the 3- and 12-month trends for real direction.

What formats can I get?

Each Actor runs on the Apify platform. Results come as JSON, CSV or Excel from the Console, through the Apify API or client libraries, on a schedule, or as a tool for AI agents via MCP.

Is this legal? What data do you collect?

We collect publicly available data only: no personal data, nothing behind a login. Crawls are paced and run at night to keep load on each site low. How you use the data is up to you, under the source's terms and your local law.

Can you build a scraper for another site?

Yes. Email us the site and the fields you need. We prioritise by demand and let you know when it's live.