Permanently Closed Marketplace Sales Thread

Status
Not open for further replies.
I have been having issues with scraping google.

I've tried two 4g proxies with 5 threads. Google flags them.

I got 11 private tried 7 threads. Seemed okay until I ran the email verifier for a few emails. Still passes google proxy test but wont scrap anything.

I've just tested Google and it's working without any issues.

Please go to Settings >> Harvester Engines Configuration >> Select Google, then click "Test Engine" Do you get search results returned here? Then click the Next button do you also get results for the next page?

If so then ScrapeBox and the search engine harvester is working and the issues is with your proxies. But if not then let me know as you can fully modify the search engine harvester, add new engines etc.
 
I've just tested Google and it's working without any issues.

Please go to Settings >> Harvester Engines Configuration >> Select Google, then click "Test Engine" Do you get search results returned here? Then click the Next button do you also get results for the next page?

If so then ScrapeBox and the search engine harvester is working and the issues is with your proxies. But if not then let me know as you can fully modify the search engine harvester, add new engines etc.
How many threads per proxy for the private proxy do you recommend? Also since they don't rotate. How long is a scrape ban?
 
I need to extract the Google cache date for thousands of pages. Can I do this with scrape box? Can I just crawl slowly or are proxies definitely required?
 
I've just tested Google and it's working without any issues.

Please go to Settings >> Harvester Engines Configuration >> Select Google, then click "Test Engine" Do you get search results returned here? Then click the Next button do you also get results for the next page?

If so then ScrapeBox and the search engine harvester is working and the issues is with your proxies. But if not then let me know as you can fully modify the search engine harvester, add new engines etc.
Just purchased Scrapebox due to the great Customer service, I would appreciate the activation at your earliest convinience, regards.
 
Last edited:
How many threads per proxy for the private proxy do you recommend? Also since they don't rotate. How long is a scrape ban?

The proxies do rotate with every request in the search engine scraper. If you have more threads than proxies what happens is each proxy may be doing 2 or more searches at the exact same time. That's not natural behavior except maybe in a home/office with multiple people sharing the one IP address and 2 people may search at the same time, but it's not going to be happening over and over again every few seconds for hours.

So for that reason you should always use less threads than proxies. Please check out this video which goes over scraping limits


I need to extract the Google cache date for thousands of pages. Can I do this with scrape box? Can I just crawl slowly or are proxies definitely required?

Yes ScrapeBox can scrape the cache dates. For thousands of pages proxies are definitely recommended. It's also similar to my comment above with regards to not performing multiple searches at the same time with the same IP this will get the proxy temp blocked.

With Google the slower you perform requests the better. When i need to scrape a lot, i just wind the connections right back, use proxies and just let it run in the background or even overnight.

Just purchased Scrapebox due to the great Customer service, I would appreciate the activation at your earliest convinience, regards.

Welcome aboard! :) All license activations have been processed, so you should be up and running and a welcome email sent.
 
Hey I need help, I just started using scrapebox again after years of not using it. I updated the app to the latest version, but can't seem to scrape anything using proxies. I try to scrape without proxies and it will scrape everything but Google, I'm not sure why Google isn't working, my ip should not be blocked as I haven't done much heavy scraping and haven't used this program in years but want to now! Please help
Anyone?
 

Please go to Settings >> Harvester Engines Configuration >> Select the engine that's not working, then click "Test Engine" Do you get search results returned here? Then click the Next button do you also get results for the next page.

If not then please go to Settings >> Harvester Engines Configuration >> Import >> Download Default Engines From Server.

This will update the engines file to ensure you have the latest version and all are restored to default. You can also completely modify the engines to work however you like such as adding completely new engines and modifying existing engines. There's a video on this here, it's a little old but the method of editing/adding engines is the same.

 
Damn. I went back and re-read page 1 of this thread.
People I haven't seen in a very long time.. Woah.
This forum thread is iconic.

@Sweetfunny, way back in 2008, did you have any idea that people would still be using this software 15 years later?
This is legendary.
 
Hi, how do i reset my license in new laptop? do I have to contact support team or is there any facility available to do it myself?
 
@Sweetfunny,

can you help me with ScrapeBox support, please. I am using YouTube scraper. I would like to scrape all videos within channel. I have tried all posible combinations of URL (using chnnel id, using handle, etc.), but I cannot get it done.

I have contacted support. In order to avoid any issue with YT channel I have given support a channel that belongs to Google.

Support cam back and asked me if I have tested on smaller channel and I have replied that, yes, I have.

However, since than, support is ghosting me.

I would just like to know, I I can scrade all videos within channel in order to further process them. (e.g. download them)

Can somebody help me, please. Or perhaps alert ScrapeBox support.
 
Any idea how to scrape address and contacts of new home or business owners in a country?can this do it?
 
Damn. I went back and re-read page 1 of this thread.
People I haven't seen in a very long time.. Woah.
This forum thread is iconic.

@Sweetfunny, way back in 2008, did you have any idea that people would still be using this software 15 years later?
This is legendary.

Yes it's been a long time that's for sure when folk like HaRRo, Essential Clix, zen19, Yuki etc were around. No i had no idea at the time, many tools then cost a lot and were often gone several months later when methods stopped working or they were shut down by the authorities like all the cookie stuffing scripts, ad click tools etc.

So it wasn't real common for blackhat tools to be around long then. I had a lot of success with SEO and out ranked some of the best like Aaron Wall, Rand Fishkin etc. But the process of scraping URLs, filtering lists, leaving comments, checking pages for your backlinks, catching dropped Godaddy domains etc was a real pain in the neck back then. So i just knew if a tool like ScrapeBox would help me, then it would help countless other people like me too.

Hi, how do i reset my license in new laptop? do I have to contact support team or is there any facility available to do it myself?

No you dont need to contact support, you just need to download the latest version of ScrapeBox and run it then enter your license Name, Email and Transaction ID.

can I use it to scape a web or a blog and their article?
can I get discount to buy it ?:D

Yes ScrapeBox has a basic article scraper addon included when you purchase. We also have a far more advanced article scraper premium plugin which has additional features like article translation, article posting, article spinner which you can see here https://www.scrapebox.com/article-scraper-plugin

Also yes the BHW discount link is https://www.scrapebox.com/bhw

@Sweetfunny,

can you help me with ScrapeBox support, please. I am using YouTube scraper. I would like to scrape all videos within channel. I have tried all posible combinations of URL (using chnnel id, using handle, etc.), but I cannot get it done.

I have contacted support. In order to avoid any issue with YT channel I have given support a channel that belongs to Google.

Support cam back and asked me if I have tested on smaller channel and I have replied that, yes, I have.

However, since than, support is ghosting me.

I would just like to know, I I can scrade all videos within channel in order to further process them. (e.g. download them)

Can somebody help me, please. Or perhaps alert ScrapeBox support.

From memory i recall sending a couple of emails to you about this, the component we use for YouTube (yt-dlp.exe) needs to update due to changes at YouTube. Please go to the Premium Plugins menu and update the YouTube plugin to the latest version which should be working for channel based scraping.

Can scrapebox harvest only urls from paid results (adwords)?

Do you mean from Adwords ads which are shown in the search results pages? If so then yes you should be able to duplicate the Google Engine in the harvester, and give it a new name. Then for the "Just before url" and "Just after url" setting add the HTML just before and just after the URL in the ad rather than what's before and after the link in the organic results.


Any idea how to scrape address and contacts of new home or business owners in a country?can this do it?

That one is difficult to answer. Does your country have some sort of publicly viewable directory of new home and business owners? We have the Yellow Pages Scraper premium plugin. The Yellow Pages is like an online phone book for businesses, but it's not just for new businesses.
 
Do you mean from Adwords ads which are shown in the search results pages? If so then yes you should be able to duplicate the Google Engine in the harvester, and give it a new name. Then for the "Just before url" and "Just after url" setting add the HTML just before and just after the URL in the ad rather than what's before and after the link in the organic results.
Yes, exactly. But after using inspect in browser I cannot see easy way to pick it as most adword embed page like: <span class="BTu2cd">websitename.com</span> And that class is always different.
 
Can scrapebox grab meta from search results? Like grab meta from harvested urls. It would be useful for checking data of indexed links.
 
Yes, exactly. But after using inspect in browser I cannot see easy way to pick it as most adword embed page like: <span class="BTu2cd">websitename.com</span> And that class is always different.

Do you have any plain text before the link such as:

role="text">

or

data-dtld="

Can scrapebox grab meta from search results? Like grab meta from harvested urls. It would be useful for checking data of indexed links.

Yes on the main GUI when you go to Grab/Check > Grab meta info from harvested URL list.
 
Status
Not open for further replies.
Back
Top