WebSource Harvester is an educational web-source harvester that crawls a site (BFS, depth-controlled), downloads browser-visible assets (HTML, CSS, JS, images, fonts, PDFs), and rewrites paths so pages work offline, including nested routes. It enforces same-origin limits and is designed for learning, offline analysis, and safe portfolio demos.
javascriptcsspythonhtmlpdffontscrawlerofflineimagesweb-scrapingrobotseducationalbfssrcsetstudent-projectsame-origingithub-actionsauthorized-testingsite-mirroringpath-rewriting
-
Updated
Mar 2, 2026 - Python