Skip to content
#

web-scraping-python-projects

Here are 5 public repositories matching this topic...

how-to-parse-xml-in-python

Follow this in-depth technical tutorial to learn how to parse XML data in Python, what libraries you should use, how to handle invalid XML, and more.

  • Updated Sep 25, 2025
  • Python

Cross-platform Python crawler that finds and verifies downloadable media, documents, and other files, then creates wget-ready URL lists for fast bulk downloading. It also scans sitemap trees and generates validated text or XML sitemaps, with HTTP, HTTPS, FTP, persistent SQLite history, resumable crawls, robots support, and no pip dependencies.

  • Updated Jul 18, 2026
  • Python

A curated guide to the best web scraping tools in 2025, comparing leading web scraping software and APIs to help you choose the right solution for reliable and scalable data extraction.

  • Updated Apr 14, 2026

Improve this page

Add a description, image, and links to the web-scraping-python-projects topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the web-scraping-python-projects topic, visit your repo's landing page and select "manage topics."

Learn more