> For the complete documentation index, see [llms.txt](https://help.gsctool.com/llms.txt). Markdown versions of documentation pages are available by appending `.md` to page URLs; this page is available as [Markdown](https://help.gsctool.com/getting-started/blog/how-to-extract-200k-urls-from-a-large-xml-sitemap.md).

# How to extract 200k URLs from a Large XML Sitemap

Last week, a customer asked me how to use [Sitemap Extractor](https://sitemapextractor.com) ([https://sitemapextractor.com](https://sitemapextractor.com/)) for extracting large sitemaps.&#x20;

I guided them through the process, and now you can easily do it too by following the steps outlined here.

{% embed url="<https://assets.ytuong.dev/gsctool/extract-200k-urls-from-a-large-xml-sitemap.mp4>" %}
Extract XML Sitemap URLs Tool
{% endembed %}

To extract a large sitemap using [Sitemap Extractor](https://sitemapextractor.com) ([https://sitemapextractor.com](https://sitemapextractor.com/)), follow these steps:

1. **Visit the Website**: Go to [Sitemap Extractor](https://sitemapextractor.com) ([https://sitemapextractor.com](https://sitemapextractor.com/)).
2. **Enter the URL**: Input the URL of the sitemap you want to extract in the designated field.
3. **Start Extraction**: Click on the 'Extract' button to begin the process.
4. **Download**: Once the extraction is complete, download the sitemap file in your preferred format.

**Sub-Sitemap Handling**: The tool also supports downloading sub-sitemaps, which can be found within a large sitemap, providing greater flexibility in managing and analyzing specific sections of your site.

By default, the tool can efficiently handle up to **200,000 URLs** at a time, allowing you to extract and process a significant amount of sitemap data seamlessly.&#x20;

However, if you reach this limit, it’s important to note that it isn’t a constraint of the tool itself, but rather a limitation imposed by web browsers. In such cases, you can clean up the already processed URLs from your list and continue with the download of additional URLs.&#x20;

This approach ensures that you maximize the tool's capabilities while working within the constraints of your browser's performance.

<figure><img src="/files/uW1VqB2EgAkgxDuXNnAh" alt="Sitemap Extractor (https://sitemapextractor.com) is a tool designed to extract and download large sitemaps from websites."><figcaption><p>Sitemap Extractor (<a href="https://sitemapextractor.com/">https://sitemapextractor.com</a>) is a tool designed to extract and download large sitemaps from websites.</p></figcaption></figure>

### FAQs

* **What is Sitemap Extractor?**\
  Sitemap Extractor ([https://sitemapextractor.com](https://sitemapextractor.com/)) is a tool designed to extract and download large sitemaps from websites.
* **How do I use Sitemap Extractor?**\
  To use Sitemap Extractor, visit the website, enter the URL of the sitemap you want to extract, click the 'Extract' button, and then download the extracted sitemap file.
* **What formats can I download the extracted sitemap in?**\
  The tool allows you to download the sitemap in your preferred format, though specific format options are not mentioned in the provided information.
* **Can Sitemap Extractor handle sub-sitemaps?**\
  Yes, Sitemap Extractor supports downloading sub-sitemaps found within a large sitemap, offering flexibility in managing specific sections of your site.
* **How many URLs can Sitemap Extractor process at once?**\
  By default, Sitemap Extractor can efficiently handle up to 200,000 URLs at a time.
* **What should I do if I reach the 200,000 URL limit?**\
  If you reach the 200,000 URL limit, you can clean up the already processed URLs from your list and continue with the download of additional URLs.
* **Is the 200,000 URL limit a constraint of the tool itself?**\
  No, the 200,000 URL limit is not a constraint of the tool but rather a limitation imposed by web browsers.
* **How can I extract more than 200,000 URLs?**\
  To extract more than 200,000 URLs, you can process the first 200,000, clean up the list, and then continue with additional URLs in subsequent extractions.
