Skip to content
Tabitha

Chrome extension

Page Extractor

Give it a list of addresses, get one row per page.

All nine tools

Where the list tool reads many items on one page, this one reads one item from many pages. Paste or import the addresses, pick what to take from each, and Tabitha visits them in order under the same politeness budget.

Without any selector work you already get the page title, the meta description, the canonical address, the Open Graph fields, the first heading and the visible email addresses. Add your own columns by clicking the fields on a sample page.

Screenshot coming with the tool. Nothing here is a drawing, so this space stays empty until the real capture exists.

How it works

  1. Step 1

    Bring the addresses

    Paste them, import a CSV, or send a column from a table you already extracted. A list of product links from the list tool goes straight in.

  2. Step 2

    Open one page as the sample

    Click the fields you want on that page. Tabitha builds the selectors once and reuses them on the rest of the list.

  3. Step 3

    Set the pace

    Gentle, normal or fast. Gentle is slower than a person reading, normal sits around human speed, and any of them still respects the per host budget and robots.txt.

  4. Step 4

    Run it and watch the log

    Every address gets a row, including the ones that failed, with the reason recorded. You can rerun only the failures.

When to reach for it

  • Enriching a list: you have product or company links and need the detail that only the detail page shows.
  • Auditing your own site or a competitor across a few hundred pages.
  • Any job phrased as one row per URL.
  • Checking whether a list of addresses still resolves, and what changed on them.

What it will not do

  • One row per page. If a page holds a list, use the list tool inside it instead.
  • Pages behind a login work in the browser, not in the cloud.
  • Politeness costs time, so a long list runs long. Thousands of addresses belong in a cloud run.