Archive.org / wayback machine scraper

Status
Not open for further replies.

xiphre

Regular Member
Joined
Jun 9, 2007
Messages
290
Reaction score
87
Hi!

Looking for someone who can make a simple archive.org site scraper in php(yes, php/curl - nothing else please). Needed functions in a nutshell:

1. Form where you enter the url, f.ex:

https://web.archive.org/web/20090228184644/http://www.blackhatworld.com/

Thus we would want all files in the 2009 02 28 cache.

2. The script first returns the amount of files that will be downloaded, hence you may decide whether or not to proceed with the download.

3. Files are downloaded - progress meter "x out of y files downloaded" is displayed.

4. When done the whole site will be presented in a zip file. The zip file should hold an error log of all files that were not successfully scraped.

Payment with Paypal and platonic xmas love. Send me a PM with price and contact information.
 
yo i have been wanting this for a while myself..let me know when you get it
 
Interested in this as well.
OP if you want to share costs of the bot we can JV on this.

Cheers
 
Status
Not open for further replies.
This thread has been auto closed due to the forum's thread age policy. Read more.
Back
Top