tried scrapebox fresh install update to scrapebox version 2.0.0.85 and update automator to 2.0.0.42 but still stucked at version 2.0.0.37
View attachment 89210
Did you put the fresh version in a new folder? Because it would have to be using a local version because they only keep 1 version on the server and I get the latest version when I download it. Try deleting your old folder all together.
It could be that there are no write permissions, so it can't save the downloaded version, or it could be that security software is stopping the download. So whitelist scrapebox in all security software.
But something is either stopping scrapebox from writing to the folder, or stopping the download.
I do have some questions! I am trying to set up my scrape box. Right now I have 3,500 searches that have about 500-1000 results each.
And I'm trying to figure out the right set up so that I can run as quickly as possible without hitting any roadblocks. I watched your video on safely scraping Google in 2017 and followed your suggestions. I have 30 proxies from Proxy Rack with 50 connections. I use the detailed harvester and have set a 15 second delay. It's currently going at a pace of about 2000 results per hour, which seems slow and would take me quite a long time to finish my scrape. It is only using one thread a time. For every keyword, it does also seem to try multiple proxies before coming to one that works (which I don't understand because it's cycling through the same 30 throughout). I also remove failed proxies.
Is there anything I could do to improve my speed significantly? I'm slowly learning how this all works and what all the setting mean, so sorry of if these are newbie questions.
So proxy rack are back connect proxies yes? With a back connect/reverse etc.. type of proxy the ip changes on the back end, so thats why scrapebox can cycle thru the same ones over and over because the ip is changing on the back end.
As for usage, if they are indeed back connect proxies, lose the delay. Thats only useful if you have shared or private proxies.
So if they are back connect proxies I would try and use the custom harvester. Then go to settings >> connections timeouts and other settings >> more harvester options >> proxy retries - and set this to the max.
Then in that same section go to the connections tab and set it to like 10 or 15.
Try that, if its costing you too much accuracy (Which it might) then go back to the detailed harvester, but don't use a delay. Detailed harvester will always use only 1 connection but will have unlimited proxy retries.
@loopline I also am willing to upgrade to the 100 connections package as well if that means I can get through my search faster (currently @ 50)! I just don't know how to set everything up for maximum results.
I am searching this (and my thousands more of these) on Google:
site:instagram.com + "10..50 posts" + "10.1k followers"
I figured that is an advanced enough search that it would require using the detailed harvester, which only utilizes one thread. Idk. I'm slightly confused!
You don't need more connections, because thats still pulling from the same ip pool on the back end, which is the limiting factor. Also you could try deeperweb and google api, and I think start page, they all are google powered. But first go to settings >> harvester engine configuration >> import >> download default engines.
These other 3 that are google powered get results from google, but have their own ip bans.
Does scrapebox work on mac osx?
It does not work natively on a mac, Yet. A mac version is in development, but may take some time yet to complete. But you can also do parallels, VMware or you can get a VPS, which would work even if you had a mac version already.