One thing that is annoying with Scrapebox is the free proxies, OK I wont moan as they don't cost me anything, but I just prefer to find other proxy sources from the mighty Google! Here is a couple of really easy ways to sourcing free proxies :
First one:
Copy paste the following footprint in Scrapebox
+":8080? +":3128? +":80? filetype:txt
You can also add words like "latest" or "new" to the footprint to attempt to find the newest proxy lists, once done simply make sure the URLS are live and then add them to your proxy source in scrapebox
Second One:
Another good way that is just twisting the above very slightly is by using the Scrapebox Sitemap scraper addon!
What you do is find a free proxy source in Google, I prefer to use free proxy lists .txt as a footprint to find proxy lists. Once you have found a site you like the look of then you need to find the sitemap by simply putting sitemap.xml at the end of the URL.
Once you have found a decent Sitemap you will need to tidy up the list to filter out the main pages of the site and stuff you will not need I.E. crap stuff
Once you have done this you can simply use the URLS from the scrape and put them in the proxy source page(harvest proxies, then add source)
Make sure you dont go to crazy with this though as it will hammer your machine, I tend to just use about 20k and test those
I found this way in the network. I checked and it works but it takes a couple of hours to the selection of sites
If i help you: you know what to do
this is rare Footprint
First one:
Copy paste the following footprint in Scrapebox
+":8080? +":3128? +":80? filetype:txt
You can also add words like "latest" or "new" to the footprint to attempt to find the newest proxy lists, once done simply make sure the URLS are live and then add them to your proxy source in scrapebox
Second One:
Another good way that is just twisting the above very slightly is by using the Scrapebox Sitemap scraper addon!
What you do is find a free proxy source in Google, I prefer to use free proxy lists .txt as a footprint to find proxy lists. Once you have found a site you like the look of then you need to find the sitemap by simply putting sitemap.xml at the end of the URL.
Once you have found a decent Sitemap you will need to tidy up the list to filter out the main pages of the site and stuff you will not need I.E. crap stuff
Once you have done this you can simply use the URLS from the scrape and put them in the proxy source page(harvest proxies, then add source)
Make sure you dont go to crazy with this though as it will hammer your machine, I tend to just use about 20k and test those
I found this way in the network. I checked and it works but it takes a couple of hours to the selection of sites
If i help you: you know what to do
this is rare Footprint
Last edited: