Public web data as an API

Web data that grows every night.

CrawlPlant runs scrapers that don't break. Every source is crawled nightly, validated and kept with its full history, then served to you in seconds. You pay only per result.

  • Fresh every night
  • History built in
  • Pay only per result
Nightly pipeline: schedule, crawl, parse, validate, harvest 01 schedule every night 02 crawl layered proxies 03 parse typed fields 04 validate diff vs. yesterday 05 harvest snapshot + history
Actors live4
records tracked24,927
sites crawled nightly4
last crawltoday
All systems operational live
Catalog

Find your dataset.

Every Actor comes out of the same factory, so they all work the same way. Pick one to see live numbers, sample output and pricing.

  • Need another site?

    Send the site and the fields you need. Demand decides what we plant next.

    Request a scraper
Why CrawlPlant

Scrapers that don't break.

These are promises of the factory, not features of one scraper. Every Actor we publish keeps them.

Runs don't fail

Your run reads from last night's snapshot, not from a live site under load. If a page can't be reached, you still get every other record, each marked with when it was collected.

01nightly snapshotdefault
02live fetchif needed
03last known valuewith timestamp
last crawl
today
records tracked
96,975
sites crawled nightly
4

partial results, never an empty failed run

History and change feeds

Every nightly change is kept. Chart any value over time, or ask only for what changed.

field   old    new    at
value   12.40 → 11.90  02:14
status  on    → off    02:14

Clean, typed fields

Consistent names, units and types. Stable ids, so the same record keeps the same key forever.

{ "…Id": "stable",
  "scrapedAt": "ISO 8601",
  "source": "cache | live" }

Answers in seconds

The crawl already happened overnight. A run only reads and filters the harvest, so results come back while you wait.

Monitored around the clock

Canary runs check every Actor every 3 hours and alert us before you notice anything.

every 3 h
For AI agents

Give your agent data it can trust.

Plug any CrawlPlant Actor into your agent through the Apify MCP server. The agent gets typed records with timestamps and history, so it can say what changed and how fresh each number is.

https://mcp.apify.com?tools=crawlplant/<actor>
  • Works with any MCP clientClaude, Cursor and other agents that speak MCP.
  • Typed records, not page textIds, prices, units and a timestamp on every value.
  • Answers in secondsServed from the nightly harvest, paid per result.
agent · example
you

What changed since last week? Give me the five biggest moves and when each was seen.

calls crawlplant/<actor> · changes since 7 days ago
agent

Five records moved more than 10% this week. The top three, all from last night's harvest:

recordchangeseen
#10482−18.2%Tue 02:14
#20917+15.0%Mon 02:11
#11305−12.7%Sun 02:13
Pricing

Pay per result. That's it.

Each Actor has a rate per 1,000 results, shown on its card and page. Try any of them on Apify's free plan before you spend anything.

Compare rates in the catalog
  1. 01

    Pay per result

    You're billed for the records you receive. A smaller run costs less.

  2. 02

    No subscription

    No seats, no minimums, no contract. Run it once or every hour.

  3. 03

    Free to try

    The free Apify plan is enough to test any Actor on real data.

FAQ

Questions, answered.

Anything else? Write to piotr@crawlplant.com.

How fresh is the data?

Every source is crawled every night, and every record carries a scrapedAt timestamp, so you always know exactly how fresh a value is. History goes back to the first night we crawled it.

What if a site changes or blocks requests?

Your run is served from the nightly snapshot, so it keeps working. Crawls fall back through proxy layers, and canary checks every 3 hours tell us about a change long before it reaches you. You get partial results, never an empty failed run.

What formats can I get?

Each Actor runs on the Apify platform. Results come as JSON, CSV or Excel from the Console, through the Apify API or client libraries, on a schedule, or as a tool for AI agents via MCP.

Is this legal? What data do you collect?

We collect publicly available data only: no personal data, nothing behind a login. Crawls are paced and run at night to keep load on each site low. How you use the data is up to you, under the source's terms and your local law.

Can you build a scraper for another site?

Yes. Email us the site and the fields you need. We prioritise by demand and let you know when it's live.