Live
Woolworths Australia Scraper
Prices, specials, unit prices and price history for the full Woolworths catalog.
CrawlPlant builds and runs reliable scrapers for public web data: prices, products, catalogs. Crawled on a schedule, monitored around the clock, paid per result.
Each Actor is built, crawled and monitored with the same system. We add sites one at a time and keep every one of them running.
Live
Prices, specials, unit prices and price history for the full Woolworths catalog.
In progress
Prices, specials and unit prices from Coles, in the same output schema as Woolworths.
Planned
The same product matched across Australian supermarkets, with prices side by side.
Tell us which site and which fields you need. Demand decides what we build next.
Request a scraperEvery scraper goes through the same pipeline, and keeps going through it for as long as it is live.
We map the site's structure, limits and data model before writing code.
A typed scraper, tested against saved pages and live runs.
Scheduled crawls refresh the cache while traffic is low.
Canary checks every 3 hours, with alerts before users notice.
Fallbacks absorb blocks and layout changes while a fix ships.
New fields and features come from real user requests.
Nightly crawls keep a fresh cache, so most runs return results in seconds.
A run never fails because of one blocked request.
Proxy tiers fall back automatically, and each run logs which layer served it.
Continuous canary runs and alerting catch breakage before it reaches your data.
Stable ids, timestamps and the source of every record, in the same schema every run.
Track how records change over time, not just today's snapshot.
Every CrawlPlant Actor works as a tool through the Apify MCP server. Add one URL to your MCP client and ask in plain language.
https://mcp.apify.com?tools=crawlplant/woolworths-au
Exact prices are on each Actor's Store page.
An Actor is a cloud program on the Apify platform. You run it from the Apify Console, the API, a schedule or an MCP client, and get structured results you can export as JSON, CSV or Excel.
Every supported site is crawled nightly, so cached results are usually less than a day old. When the cache can't answer, the Actor fetches the page live. Every record carries a timestamp.
The run falls back: another proxy layer first, then the last known data, marked with its timestamp. One blocked request never fails the whole run, and each run logs which path served it.
Public data only: prices, product details and catalog structure. No personal data, and nothing behind a login.
Yes. Email us the site and the fields you need. We prioritise by demand and feasibility, and tell you when it goes live.