Tool for downloading websites.
Go to file
2020-04-29 20:33:08 +02:00
.github/workflows ci: Split ci in two jobs 2020-04-29 18:54:33 +02:00
src Merge branch 'master' into selective_html_parsing 2020-04-29 16:35:30 +02:00
tests args: Change quiet to verbose since default behavior is quiet 2020-04-18 16:27:21 +02:00
.gitignore gitignore: Ignore tags folder 2020-04-09 11:36:20 +02:00
Cargo.lock scraper: Selective parsing fully working 2020-04-29 11:24:24 +02:00
Cargo.toml scraper: Selective parsing fully working 2020-04-29 11:24:24 +02:00
LICENSE Initial commit 2019-10-29 18:21:52 +01:00
README.md readme: Add contributing links 2020-04-29 20:33:08 +02:00

SuckIT

SuckIT allows you to recursively visit and download a website's content to your disk.

Features

  • Vacuums the entirety of a website recursively
  • Uses multithreading
  • Writes the website's content to your disk
  • Enables offline navigation
  • Saves application state on CTRL-C for later pickup
  • Offers random delays to avoid IP banning

Options

Option Behavior
-h|--help Displays help information
-v|--verbose Activate Verbose output
-d|--depth Specify the level of depth to go to when visiting the website
-j|--jobs Number of threads to use
-o|--output Output directory where the downloaded files are written
-t|--tries Number of times to retry when the downloading of a page fails

Example

A common use case could be the following:

suckit http://books.toscrape.com -j 8 -o /path/to/downloaded/pages/

Want to contribute ? Feel free to open an issue or submit a PR !