Skip to content
#

web-spidering

Here are 5 public repositories matching this topic...

Language: All
Filter by language

Generate a list of file links you can feed to wget for easy downloading! Mainly used for spidering web folders with lots of files. Can even generate a sitemap.txt or XML file for your website!

  • Updated Jul 18, 2026
  • Shell

Cross-platform Python crawler that finds and verifies downloadable media, documents, and other files, then creates wget-ready URL lists for fast bulk downloading. It also scans sitemap trees and generates validated text or XML sitemaps, with HTTP, HTTPS, FTP, persistent SQLite history, resumable crawls, robots support, and no pip dependencies.

  • Updated Jul 18, 2026
  • Python

Improve this page

Add a description, image, and links to the web-spidering topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the web-spidering topic, visit your repo's landing page and select "manage topics."

Learn more