Skip to content
#

webspider

Here are 79 public repositories matching this topic...

Serritor is an open source web crawler framework built upon Selenium and written in Java. It can be used to crawl dynamic web pages that require JavaScript to render data.

  • Updated Jul 7, 2022
  • Java

Cross-platform Python crawler that finds and verifies downloadable media, documents, and other files, then creates wget-ready URL lists for fast bulk downloading. It also scans sitemap trees and generates validated text or XML sitemaps, with HTTP, HTTPS, FTP, persistent SQLite history, resumable crawls, robots support, and no pip dependencies.

  • Updated Aug 6, 2026
  • Python

Add this topic to your repo

To associate your repository with the webspider topic, visit your repo's landing page and select "manage topics."

Learn more