For the complete documentation index, see llms.txt. This page is also available as Markdown.

Extract Part

What is Extract Part?

This feature lets you extract data selectively by choosing only the areas you want. Try it when you want to improve the structure and accuracy of your extracted data.


How to Use

1

Click the Listly Icon on your target website

Navigate to the web page you want to extract data from in your Listly-installed browser and click the Listly icon.

2

Select 'Extract Part'

Select the [Extract Part] button in the Listly extension.

3

Select the areas you want to scrape

Click to specify the areas you want to extract.

When you click an extraction area with a matching pattern, a pop-up will automatically appear, allowing you to select all similar areas at once.

  • [Group all] : Click to select all matching areas without having to click them individually.

  • [Keep separate] : Click if the areas you wish to extract are distinct individual sections that do not share a pattern.

4

Select extraction options for each area (optional)

If needed, configure extraction options within the individual selection box.

5

Click 'RUN'

Once you've finished selecting extraction areas, click the [RUN] button.

6

Select Tab to extract in results

After reviewing the extracted data in the results page, click the [Excel] or [Google Sheets] button to export.

7

Download your data

Check your downloaded file.


The number of items to extract varies by page. Is group extraction still possible?

Yes, it's possible. For example, as shown in the image above, let's say the first item area on page 1 is replaced by a different element like an advertisement on page 2. (Areas with information you want to extract are marked with green boxes)

Listly's Parts automatically detects data locations with the same structure.

If you extract data on page 1 using the 'automatically select repeating elements (auto-select repeating elements)' method to specify desired areas, then perform group extraction based on that page, data extraction will succeed without failures even if the number of selected areas differs on other web pages.

I want to extract a specific page using the 'individually select desired areas' method, then perform group extraction. However, some pages don't have the element I 'individually selected'. What happens in this case?

As shown in the image above, missing information will be extracted as blank cells.

Let's assume you selected image, title, and description areas separately using Parts on page A, extracted the data, then performed group extraction with pages having the same structure.

If page C has an empty description area that you selected on page A, the data in that cell will be extracted as blank, as shown in the image above.

When selecting repeating areas, unnecessary areas are also selected. How can I solve this?

We plan to add a feature that allows you to remove unnecessary areas by clicking for more precise data area selection. Currently, you can resolve this by manually editing the CSS Selector value of that area.

Last updated

Was this helpful?