Scrape any data off your own or a competitor's site

Scrape any data off your own or a competitor's site

SEO & AEO
Screaming Frog
Configuration, then Custom, then Custom Extraction, and a crawl turns into a spreadsheet of whatever element you pointed it at.
Source
Screaming Frog
Tools
Screaming Frog SEO Spider
Runs
On demand
Cost
Free up to 500 URLs
Last updated
August 7, 2026

What it is.

You open Configuration, then Custom, then Custom Extraction, and add an extractor. You can build it visually by clicking the element on a rendered page, or write the XPath by hand if you know what you want.

Then you crawl, and the Custom Extraction tab holds one column per extractor for every URL. That is how a crawl becomes a dataset: authors, publish dates, prices, schema blocks, Open Graph tags, hreflang.

The published XPath examples cover headings, hreflang, structured data, and social meta tags, which is most of what an audit needs.

What you get.

  • One column per extractor across every crawled URL.
  • Visual selection, so you can point at an element instead of writing XPath.
  • Published XPath examples for headings, hreflang, structured data, and social tags.
  • An export of the extracted data alongside the rest of the crawl.
  • The same capability against a competitor's public pages.
HOW TO USE IT

How to set it up.

1

Open Configuration, then Custom, then Custom Extraction, and add an extractor.

2

Use visual extraction first and click the element you want on a rendered page.

3

Switch to manual XPath when the visual pick grabs too much or too little.

4

Name each extractor for what it holds, because eight columns called Extractor 1 are useless later.

5

Test on a handful of URLs before running the full crawl, since a wrong expression wastes the whole run.

6

Crawl, then read the Custom Extraction tab and export.

Use cases

Pull one field from every page

Set the extractor and get a column across the whole crawl.

Point instead of writing a selector

Use visual selection when nobody wants to write XPath.

Watch a competitor's pricing page

Extract the values you care about and re-crawl later to compare.

Best for

Content audits at scale

Pulling author, date, and word count across a whole blog is a crawl, not a spreadsheet exercise.

Checking schema across a site

Extracting the structured data block per page finds the templates where it is missing.

Competitor research

Anything on a competitor's public pages can be extracted and counted.

Open the original.

Hosted on Screaming Frog, free to open.
Open the Tool

Questions about Scrape any data off your own or a competitor's site

What's the easiest way to build an extractor?
How durable is an extractor?
Why is my crawl so slow?

Questions about SEO & AEO

What's worth extracting from a competitor's site?
Is scraping a competitor safe?
Strengths
  • Visual extraction means you do not need XPath to get started.
  • The examples cover headings, hreflang, structured data, and social tags, so common jobs are documented.
  • The extracted columns sit alongside crawl data, so you can filter by status code or depth at the same time.
Limitations
  • The free version stops at 500 URLs, so a large site needs a license or directory-by-directory crawls.
  • XPath breaks when a site changes its markup, so an extractor is not a permanent integration.
  • JavaScript-rendered elements need rendering turned on, which slows the crawl considerably.
  • Scraping a competitor may breach their terms, so check what your legal team allows.
Skip this if
  • Skip it if you already have a scraper for the fields you need.
Ideal for
  • Technical SEO leads
  • Content strategists
  • Stage: Any stage

Want this running without building it yourself?

TripleDart has scaled 300+ tech companies with expert operators and AI workflows behind every play.