Skip to main content

Sitemap to URL Crawler: RAG & AI Data Feeder

Run on Apify Python client

Sitemap to URL Crawler: RAG & AI Data Feeder — a Python client for the Apify Actor. Pull Sitemap to URL data at scale with no API-key hassle: thousands of clean, structured results you can export to JSON, CSV or Excel.

▶️ Run it on Apify: https://apify.com/logiover/sitemap-to-url-crawler

Install

pip install sitemap-to-url-crawler

Usage

from sitemap_to_url_crawler import scrape

items = scrape({}, token="YOUR_APIFY_TOKEN")
print(len(items), "results")
print(items[0])

Get your free Apify token at https://console.apify.com/account/integrations.

Input

Field Type Description
startUrls array Start URLs (Domain or Sitemap)
maxUrls integer Max URLs to Extract
proxyConfiguration object Proxy Configuration

All fields optional — run with empty input {} for a broad default result set.

Output

Each result item includes fields such as: url, lastmod, changefreq, priority, sourceSitemap.

Why use this

  • ⚡ Thousands of results per run, auto-paginated
  • 🔑 No Sitemap to URL login or reverse-engineering — just call the Actor
  • 📦 Export to JSON, CSV, Excel, JSONL, XML
  • ☁️ Runs on Apify cloud — schedule it, add webhooks, wire into Make / Zapier / n8n

FAQ

Do I need an API key?

Only a free Apify token (grab one at https://console.apify.com/account/integrations). No Sitemap to URL login and no scraping setup on your side.

How many results can I get?

Thousands per run. Raise the limit in the input to pull more — the Actor paginates for you.

What export formats are supported?

Results are plain JSON in code; from the Apify dataset you can export CSV, Excel, JSON, JSONL or XML.

Is this an official Sitemap to URL API?

No. It is an unofficial Sitemap to URL data client on the Apify platform — a maintained alternative when there is no official or affordable API.

Links


MIT © 2026 logiover · Client library for the hosted Apify Actor. Not affiliated with Sitemap to URL.

Release files for sitemap-to-url-crawler 1.0.0

For a detailed explanation of source distributions (sdists) and built distributions (wheels), please see the package formats documentation.

Source distribution (sdist)

Source distribution for sitemap-to-url-crawler 1.0.0
File Size Uploaded
sitemap_to_url_crawler-1.0.0.tar.gz 4.0 kB Details

Built distribution (wheel)

Table of built distributions (wheels) for sitemap-to-url-crawler 1.0.0
File Interpreter ABI Platform
sitemap_to_url_crawler-1.0.0-py3-none-any.whl Python 3 none any Details

Total release size: 8.4 kB

Release files / sitemap_to_url_crawler-1.0.0.tar.gz

Download URL sitemap_to_url_crawler-1.0.0.tar.gz
Size 4.0 kB
Tags Source
SHA-256 checksum
How to use checksums
9654b9a35f06f3fda527439204eaa021c4fed54ddb74f5eaa558851cae3b9095
BLAKE2b-256 checksum
How to use checksums
5fbca2076153614fb3f95037aece7276df417f4ff532018d478bd6331812b667
Upload date
Uploaded using Trusted Publishing?
What is trusted publishing?
No
Uploaded via twine/6.2.0 CPython/3.14.6

Release files / sitemap_to_url_crawler-1.0.0-py3-none-any.whl

Download URL sitemap_to_url_crawler-1.0.0-py3-none-any.whl
Size 4.4 kB
Tags Python 3
SHA-256 checksum
How to use checksums
1a6b10bd794bbafe2a820001c54dd54b9fcc64c9f73bd52269e2dc3959401341
BLAKE2b-256 checksum
How to use checksums
d887cc8b8888377f6706f7e28b815c3fbf7b5bdc20025aef72b280cdc7013ee9
Upload date
Uploaded using Trusted Publishing?
What is trusted publishing?
No
Uploaded via twine/6.2.0 CPython/3.14.6

Release history Release notifications | RSS feed

This release

1.0.0 This release

2 release files

Anthropic, PBC Visionary sponsor Bloomberg Visionary sponsor Hudson River Trading Visionary sponsor Meta Visionary sponsor NVIDIA Visionary sponsor Microsoft Sustainability sponsor Depot Continuous Integration AWS Cloud computing and Security Sponsor Datadog Monitoring Fastly CDN Google Download Analytics Sentry Error logging StatusPage Status page