Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Badzoneyv4n/rlrc-pdf-crawler

Domaine:

digital infrastructure

Type de record:

software
Créateur:
Bad
Hôte:
This is a Tampermonkey-powered automation script that crawls the Rwanda Law Reform Commission (RLRC) website and automatically downloads all active law PDFs skipping only those laws not in force. # 🕷️ RLRC Rwanda Laws PDF Crawler This Tampermonkey-based automation script crawls the Rwanda Law Reform Commission (RLRC) website and **automatically downloads all active law PDFs**. It recursively navigates the site’s folder structure, skips deprecated sections (like "Laws not in force"), and downloads available PDFs one by one. Once complete, it shows a confirmation modal and plays a success sound. > ✅ **Tested and working on Chrome (with Tampermonkey extension)**. > ⚠️ Keep the page in **focus** during the crawl — Chrome may pause background tabs and interrupt downloads. --- ## ✨ Features - 🔁 Fully **recursive folder crawler** - 📥 Auto-downloads every PDF found - 🚫 **Skips ignored folders** like “Laws not in force” - 📍 Uses breadcrumb + internal stack to **track navigation** - ✅ Shows a **completion modal** on finish - 🔔 Plays a **sound notification** when done - 💾 Pure client-side script — no server or API needed --- ## ⚙️ How It Works 1. The script starts on the root RLRC laws page. 2. It: - Downloads all PDFs in the current folder - Stores subfolder links in a **stack** - Visits each folder one by one (depth-first) - Goes back after finishing each 3. When everything is done: - It plays a “ding” sound - Pops a “✅ Done retrieving PDFs” modal in the center of the screen --- ## 🚀 How To Use ### 1. Install Tampermonkey Install Tampermonkey extension for your browser (Chrome recommended). ### 2. Add the Script - Create a new Tampermonkey script - Paste the contents of `rlrc-crawler.user.js` - Save ### 3. Visit the Website Go to: `rlrc.gov.rw` ### 4. Let It Run - It will begin crawling and downloading PDFs - Keep the tab active (don’t switch tabs) - On completion, a modal will appear and a sound will play --- ## 🛠 Configuration You can update this array to skip more folders: ```js const IGNORED_FOLDERS = ["Laws not in force"]; ``` ## 📁 File Structure rlrc-pdf-crawler/ ├── rlrc-crawler.user.js # Tampermonkey sc …

Visit

github.com

Licenses

MIT

Similaires

AdrianoPereira/sasscalweathernet-crawlerRestioson/isixhosa-crawlerttomsin/grio-crawlerSetraC4Ci/Gasy-Corpus-CrawlerThembaGqaza/stats-sa-crawlerNorbert49/webtext-crawler-nlp

AdrianoPereira/sasscalweathernet-crawler

Bot for download data meteorological data from West Coast of the African Continent powered by SASSCA

Restioson/isixhosa-crawler

(Undergrad independent study project) Focused web crawler aiming to discover documents written in is

ttomsin/grio-crawler

The africa data ingestion engine # grio-crawler The data ingestion engine for **Grio** — Nigeria's

SetraC4Ci/Gasy-Corpus-Crawler

python script for scraping malagasy langage websites # GCC (Gasy Corpus Crawler) GCC (Gasy Corpus

ThembaGqaza/stats-sa-crawler

A Python-based web crawler to extract and collect data from the Stats SA (Statistics South Africa) w

Norbert49/webtext-crawler-nlp

High-volume multilingual web scraping and dataset curation pipeline for low-resource languages (Swah