Features

Everything you need, nothing you do not

Fields suggested for you

We read the page six different ways — its own structured data, its embedded app data, tables, feeds and repeating blocks — and name every column we find.

Images in a ZIP, not just links

Every image is downloaded and packaged with a manifest tying each file back to its row. Most tools hand you URLs and stop there.

See it before you commit

A real preview with sample rows, image thumbnails and a fill rate on every column, so a broken selector is obvious before you spend a credit.

Follows pagination

We find the “next” link or the page number in the URL, and you set how far to go.

Fix anything by hand

Rename a column, change its type, or edit the CSS selector and test it instantly against the live page.

Polite by default

We identify our crawler honestly, honour robots.txt, and pause between pages. Being a good citizen keeps us welcome.

Frequently asked questions

No. Paste a URL, tick the columns you want, and download the file. If you do know CSS you can edit any selector by hand, but you never have to.

One credit is used each time we load a webpage. It does not matter whether we extract one record from it or a thousand — loading the page counts as a single credit.

It is legal in many situations, but it depends on the type of data being collected, how it is used, the website's terms of service, and the laws that apply where you are. You are responsible for making sure your use complies with the law and with the terms of the sites you submit. This is not legal advice — if you are unsure, take your own.

Public pages: online shops, marketplaces, job boards, directories, property sites, news and blog archives. It does not support social media platforms — including LinkedIn, Instagram, Facebook, X and TikTok — nor anything behind a login or a paywall.

Downloaded images remain the copyright of their original owners. Use them for internal analysis, research and cataloguing — not for republication or resale — and check the source site's terms first.

Yes, by default. We also honour the crawl delay a site asks for and pause between pages. Our crawler identifies itself honestly and publishes a contact address.

We read many JavaScript-built sites already, because modern frameworks ship their data inside the page. Where a page genuinely renders nothing without a browser, we tell you so rather than hand back an empty file. Full browser rendering is on the way.

A running scrape stops cleanly and keeps whatever it already collected. Your allowance resets on the same date each month.

That depends on your plan — see the pricing table. Generated files expire sooner and can always be rebuilt from the scrape.

Get in touch and we will add your domain to our blocklist. Our crawler also honours robots.txt, so a disallow rule takes effect on its own.

Ready to pull your first dataset?

Create an account, paste a URL, and see your columns in seconds. The free plan needs no card.

Get started free