Permanently Closed Marketplace Sales Thread

Status
Not open for further replies.
Your support is really fast. I got one issue, i am not able to load proxies in expired domain crawler. Plus its showing already registered domains as expired. I mean when i check domain with godaddy they are already registered plus few of them are alive. So i am not sure what the issue is, as i already update whois list and servers. Is this a proxy issue ?

We released an update around the same time you made your post, so please update to the latest version of the plugin and it should be working.

  • Hey i have around 8 sites and i am buying off page seo services to rank my keywords.. can you please tell me how effective scrapebox can be for me? i was thinking of buying it but i am not sure about the google updates and if the software works properly... can i get backlinks for all my sites?
can anyone please let me know what can the benefits for me buying scrapebox in this times? thanks a lot :D

ScrapeBox is quite versatile, there's an almost unlimited amount of things you can do with it. Best place to start is by checking out all the tutorial video's located here https://www.youtube.com/user/looplinescrapebox/videos

Facing issue in Scrape box License. I did buy it for 97$ for lifetime , now trying to verifying it but it says SOFTWARE WILL TAKE 12 HOURS to verify but nothing updated in software from last 2 days. Please resolve my issue on urgent basis pls! HIT ME back on skype: good_personality

If the license transfer has not been processed, then the issue is almost certainly with the license details being entered incorrectly such as the wrong email, or a typo in the details you entered. Please contact support with your license details and we can see what's wrong http://www.scrapebox.com/contact-us
 
We released an update around the same time you made your post, so please update to the latest version of the plugin and it should be working.



ScrapeBox is quite versatile, there's an almost unlimited amount of things you can do with it. Best place to start is by checking out all the tutorial video's located here https://www.youtube.com/user/looplinescrapebox/videos



If the license transfer has not been processed, then the issue is almost certainly with the license details being entered incorrectly such as the wrong email, or a typo in the details you entered. Please contact support with your license details and we can see what's wrong http://www.scrapebox.com/contact-us

Hi any tutorials on scraping more than 1000 results? I tried scraping shopify stores using the site:myshopify.com footprint, but scrapebox only returned less than 1000 results although there are hundreds of thousands of results on google
 
Hi any tutorials on scraping more than 1000 results? I tried scraping shopify stores using the site:myshopify.com footprint, but scrapebox only returned less than 1000 results although there are hundreds of thousands of results on google

Unfortunately Google doesnt return hundreds of thousands of results, Google along with most other engines like Yahoo, Bing etc will only return a maximum of 1,000 results per query. If you manually click through the results in your browser, you will see the results stop.
 
Hi any tutorials on scraping more than 1000 results? I tried scraping shopify stores using the site:myshopify.com footprint, but scrapebox only returned less than 1000 results although there are hundreds of thousands of results on google
In addtion to what Sweetfunny said you can do stuff like

site:myshopify.com keyword

so like

site:myshopify.com a
site:myshopify.com b
site:myshopify.com c
site:myshopify.com 1
site:myshopify.com 2
site:myshopify.com 3
site:myshopify.com car
site:myshopify.com green

Then just remove duplicate urls when done.

This forces google to return different sets of results from their index. mind you this is an advanced operator so you need to be mindful to not go too fast lest you get your proxies banned.

This video should be helpful


but defintely don't neglect bing as an engine, they are quite lax on bans.

Google soft caps for a lot of queries with advanced operators at 300-600 results and sometimes even less.
 
Hi any tutorials on scraping more than 1000 results? I tried scraping shopify stores using the site:myshopify.com footprint, but scrapebox only returned less than 1000 results although there are hundreds of thousands of results on google
Also another idea I thought of is that if you go to setting >> harvester engine configuration >> import >> download default engines. There are several new engines scrapebox added recently.

So you can use more then 1 engine and thus get more results.
 
Also another idea I thought of is that if you go to setting >> harvester engine configuration >> import >> download default engines. There are several new engines scrapebox added recently.

So you can use more then 1 engine and thus get more results.

Wow thanks a lot. I'll test this.
 
Hey guys, what would you recommend as the best bang for your buck as far as private proxies go that have the most success for scraping Google? I'm targeting to run a couple million searches and trying to scrape a hundred million + URLs in the most efficient way possible! @loopline - do you happen to use Skype or any other platform to get in touch with you outside of BHW? I would absolutely love 5 minutes of your time, you're a legend in my book!
 
hello. i am unable to update my automator plugin to 2.0.0.42 and i am stucked at 2.0.0.37.. any tips? i have already tried to uninstall and install the plugin
 
Wow thanks a lot. I'll test this.

Your welcome. :)

Hey guys, what would you recommend as the best bang for your buck as far as private proxies go that have the most success for scraping Google? I'm targeting to run a couple million searches and trying to scrape a hundred million + URLs in the most efficient way possible! @loopline - do you happen to use Skype or any other platform to get in touch with you outside of BHW? I would absolutely love 5 minutes of your time, you're a legend in my book!

I would try some back connect proxies. I know storm proxies recently released some residential ips, and those sound like they would work good for scraping, in theory anyway.

As for skype etc... no, I don't use it. Just because every time I log on people would always be messaging me. I like helping people but I can spend all day on skype and not actually get any real work done, so I just quit using all that sort of thing.

Did you have any specific questions?

hello. i am unable to update my automator plugin to 2.0.0.42 and i am stucked at 2.0.0.37.. any tips? i have already tried to uninstall and install the plugin

Does it give you an error?

I would try to download an entire new copy of scrapebox and unzip it to a new folder on your desktop or in your documents folder.

http://www.scrapebox.com/payment-received
 
I can't install any add on? My task manager showed some internet connection when I tried to instal add on, but it always failed without any warning. The pop up just disappear, and the add on list stays red, all of them.
 
tried scrapebox fresh install update to scrapebox version 2.0.0.85 and update automator to 2.0.0.42 but still stucked at version 2.0.0.37
error.PNG
 
:)

I would try some back connect proxies. I know storm proxies recently released some residential ips, and those sound like they would work good for scraping, in theory anyway.

As for skype etc... no, I don't use it. Just because every time I log on people would always be messaging me. I like helping people but I can spend all day on skype and not actually get any real work done, so I just quit using all that sort of thing.

Did you have any specific questions?

I do have some questions! I am trying to set up my scrape box. Right now I have 3,500 searches that have about 500-1000 results each.

And I'm trying to figure out the right set up so that I can run as quickly as possible without hitting any roadblocks. I watched your video on safely scraping Google in 2017 and followed your suggestions. I have 30 proxies from Proxy Rack with 50 connections. I use the detailed harvester and have set a 15 second delay. It's currently going at a pace of about 2000 results per hour, which seems slow and would take me quite a long time to finish my scrape. It is only using one thread a time. For every keyword, it does also seem to try multiple proxies before coming to one that works (which I don't understand because it's cycling through the same 30 throughout). I also remove failed proxies.

Is there anything I could do to improve my speed significantly? I'm slowly learning how this all works and what all the setting mean, so sorry of if these are newbie questions.
 
@loopline I also am willing to upgrade to the 100 connections package as well if that means I can get through my search faster (currently @ 50)! I just don't know how to set everything up for maximum results.

I am searching this (and my thousands more of these) on Google:
site:instagram.com + "10..50 posts" + "10.1k followers"

I figured that is an advanced enough search that it would require using the detailed harvester, which only utilizes one thread. Idk. I'm slightly confused!
 
I can't install any add on? My task manager showed some internet connection when I tried to instal add on, but it always failed without any warning. The pop up just disappear, and the add on list stays red, all of them.
Ok, I successfully installed the add on after changing connection. Thanks
 
tried scrapebox fresh install update to scrapebox version 2.0.0.85 and update automator to 2.0.0.42 but still stucked at version 2.0.0.37View attachment 89210

Did you put the fresh version in a new folder? Because it would have to be using a local version because they only keep 1 version on the server and I get the latest version when I download it. Try deleting your old folder all together.

It could be that there are no write permissions, so it can't save the downloaded version, or it could be that security software is stopping the download. So whitelist scrapebox in all security software.

But something is either stopping scrapebox from writing to the folder, or stopping the download.

I do have some questions! I am trying to set up my scrape box. Right now I have 3,500 searches that have about 500-1000 results each.

And I'm trying to figure out the right set up so that I can run as quickly as possible without hitting any roadblocks. I watched your video on safely scraping Google in 2017 and followed your suggestions. I have 30 proxies from Proxy Rack with 50 connections. I use the detailed harvester and have set a 15 second delay. It's currently going at a pace of about 2000 results per hour, which seems slow and would take me quite a long time to finish my scrape. It is only using one thread a time. For every keyword, it does also seem to try multiple proxies before coming to one that works (which I don't understand because it's cycling through the same 30 throughout). I also remove failed proxies.

Is there anything I could do to improve my speed significantly? I'm slowly learning how this all works and what all the setting mean, so sorry of if these are newbie questions.

So proxy rack are back connect proxies yes? With a back connect/reverse etc.. type of proxy the ip changes on the back end, so thats why scrapebox can cycle thru the same ones over and over because the ip is changing on the back end.

As for usage, if they are indeed back connect proxies, lose the delay. Thats only useful if you have shared or private proxies.

So if they are back connect proxies I would try and use the custom harvester. Then go to settings >> connections timeouts and other settings >> more harvester options >> proxy retries - and set this to the max.

Then in that same section go to the connections tab and set it to like 10 or 15.

Try that, if its costing you too much accuracy (Which it might) then go back to the detailed harvester, but don't use a delay. Detailed harvester will always use only 1 connection but will have unlimited proxy retries.

@loopline I also am willing to upgrade to the 100 connections package as well if that means I can get through my search faster (currently @ 50)! I just don't know how to set everything up for maximum results.

I am searching this (and my thousands more of these) on Google:
site:instagram.com + "10..50 posts" + "10.1k followers"

I figured that is an advanced enough search that it would require using the detailed harvester, which only utilizes one thread. Idk. I'm slightly confused!

You don't need more connections, because thats still pulling from the same ip pool on the back end, which is the limiting factor. Also you could try deeperweb and google api, and I think start page, they all are google powered. But first go to settings >> harvester engine configuration >> import >> download default engines.

These other 3 that are google powered get results from google, but have their own ip bans.

Does scrapebox work on mac osx?

It does not work natively on a mac, Yet. A mac version is in development, but may take some time yet to complete. But you can also do parallels, VMware or you can get a VPS, which would work even if you had a mac version already.
 
So proxy rack are back connect proxies yes? With a back connect/reverse etc.. type of proxy the ip changes on the back end, so thats why scrapebox can cycle thru the same ones over and over because the ip is changing on the back end.

As for usage, if they are indeed back connect proxies, lose the delay. Thats only useful if you have shared or private proxies.

So if they are back connect proxies I would try and use the custom harvester. Then go to settings >> connections timeouts and other settings >> more harvester options >> proxy retries - and set this to the max.

Then in that same section go to the connections tab and set it to like 10 or 15.

Try that, if its costing you too much accuracy (Which it might) then go back to the detailed harvester, but don't use a delay. Detailed harvester will always use only 1 connection but will have unlimited proxy retries.



You don't need more connections, because thats still pulling from the same ip pool on the back end, which is the limiting factor. Also you could try deeperweb and google api, and I think start page, they all are google powered. But first go to settings >> harvester engine configuration >> import >> download default engines.

These other 3 that are google powered get results from google, but have their own ip bans.

So I don't think these are back connect proxies. I think they are just private proxies. I attached a picture of the options I have. I tried the custom harvester and it looks like it did sacrifice accuracy quite a bit.. missing 70% of the data (why does this happen by the way?). When I run detailed harvester, I get every single result. So with detailed harvester, is there no way to speed up the process at all?

I'm thinking about buying another set of proxies and running two instances (splitting up the keywords).. but that seems like a hack-y way to speed up the results. Again I'm new to this whole landscape, so apologies if it's evident I don't know what I'm talking about!
 

Attachments

  • Screen Shot 2017-04-06 at 2.46.52 PM.png
    Screen Shot 2017-04-06 at 2.46.52 PM.png
    182.1 KB · Views: 92
@loopline - also I decided to watch it run for a little while and a couple of things I've observed, that I'd love your insight on.

When I first ran.. it was running decently fast. Still slow, but at the same 2-3k results per hour clip. As time has gone on (I'm on result 100k), it's going slower and slower and trying wayyy more proxies for each attempt before a successful harvest... sometimes taking 3-5 minutes+ of trying a different proxy of my same 30 proxy list just to get one successful harvest. Really have no idea what is going on.
 
Status
Not open for further replies.
Back
Top