Tool for downloading websites.
Go to file
2020-04-29 22:30:05 +02:00
.github/workflows ci: Split ci in two jobs 2020-04-29 18:54:33 +02:00
src misc: Remove dead code 2020-04-29 22:29:24 +02:00
tests args: Change quiet to verbose since default behavior is quiet 2020-04-18 16:27:21 +02:00
.gitignore gitignore: Ignore tags folder 2020-04-09 11:36:20 +02:00
Cargo.lock scraper: Selective parsing fully working 2020-04-29 11:24:24 +02:00
Cargo.toml scraper: Selective parsing fully working 2020-04-29 11:24:24 +02:00
LICENSE Initial commit 2019-10-29 18:21:52 +01:00
README.md Merge branch 'prettify-readme' of github.com:skallwar/suckit into prettify-readme 2020-04-29 20:37:53 +02:00

Build and test

SuckIT

SuckIT allows you to recursively visit and download a website's content to your disk.

Features

  • Vacuums the entirety of a website recursively
  • Uses multithreading
  • Writes the website's content to your disk
  • Enables offline navigation
  • Saves application state on CTRL-C for later pickup
  • Offers random delays to avoid IP banning

Options

Option Behavior
-h, --help Displays help information
-v, --verbose Activate Verbose output
-d, --depth Specify the level of depth to go to when visiting the website
-j, --jobs Number of threads to use
-o, --output Output directory where the downloaded files are written
-t, --tries Number of times to retry when the downloading of a page fails

Example

A common use case could be the following:

suckit http://books.toscrape.com -j 8 -o /path/to/downloaded/pages/

Want to contribute ? Feel free to open an issue or submit a PR !