Permanently Closed Marketplace Sales Thread

Status
Not open for further replies.
Your connections are 1000 percent too high. Put connections to 1, not 10. You want a ratio of connections vs proxies at 1:10 or 1:15+ depending on your query. So you have 15 proxies so connections should be no more then 1. 20 proxy retries will be fine, since you only have 15 proxies anyway.

proxies are not a problem. i tried to use more but it still only scrapes 3k keywords max.
and if i add a shit load of public ones then it go's up to 10k keywords.

thanks, i will try to add more proxies. :)
-=-
 
  1. loopline im using around 2k google passed proxies, so it should not be an issue
Not sure why its not trying to complete all the keywords,i watched the troubleshoot video but wasnt able to fix it.

When disabling multi-threaded i have some completed status (even with 0 results), some Error 302 (normal) and some scraped results, but the problem is that it isnt going through all the keywords when running.

Right now what i need to do to fully scrape my keywords is exporting not completed and restarting till it complete it all
 
Last edited:
  1. loopline im using around 2k google passed proxies, so it should not be an issue
Not sure why its not trying to complete all the keywords,i watched the troubleshoot video but wasnt able to fix it.

When disabling multi-threaded i have some completed status (even with 0 results), some Error 302 (normal) and some scraped results, but the problem is that it isnt going through all the keywords when running.

Right now what i need to do to fully scrape my keywords is exporting not completed and restarting till it complete it all

You could try the custom harvester and turn the retries up to 99, that would probably get as close as possible with public proxies. I only use private for scraping, and I never have an issue with uncompleted keywords. I can't say that it gets 100% all the time, but if a few are missed here and there I don't care, but accuracy is high anyway.

The ones that are completed with 0 results could be a proxy returning modified data so it seems there are no results or it could be that there simply were no results for that query. Try a few of those queries in a browser, does it have results? Try the same queries as a test with no proxies, do they work?

Public proxies in general aren't your best bet, IMHO, if you are after 100% accuracy. Public proxies with a multi threaded harvester are designed to tear the heart out of the engines and provide as many results as possible as fast as possible with accuracy being less important.

If you need accuracy, set the custom harvester to 1 connection with at least 15 private proxies or more and then set proxy retries to 99 and thats probably as good as it gets.
 
Actually my public proxies are port scanned proxies from a service here in BHW, so its not that public (scraped from sources)

Ill try the custom harverster and see what happens! Thanks


I have 10 semi dedicated proxies but had no luck harvesting with them, maybe pvt proxies would be better, but well now im using these proxies lets see if i can tweak it all and get the most out of scrapebox

regards


edit: im noticing something here. My scraping start at around 100urls/s and after like 1 min it keep dropping till 20urls/s...is it normal?
 
Last edited:
I noticed that my sb is always crashing like a baws? Do others have same experience? in any thing I do, keyword scraping, It always crash lately.
 
I noticed that my sb is always crashing like a baws? Do others have same experience? in any thing I do, keyword scraping, It always crash lately.

The last 2 times I've done some harvesting I crashed...I just assumed it was the proxies going bad but it shouldn't lock up anyways. :(
 
@Sweetfunny

I bought Scrapebox years ago but I wasn't using it anymore.
How can I reactivate my license?
 
@Sweetfunny

I bought Scrapebox years ago but I wasn't using it anymore.
How can I reactivate my license?

Submit for a transfer. Download it, unzip to a folder on your desktop or your documents folder. Then run it, click activate and enter your details, hit submit and wait up to 12 hours for manual activation and your set.

You can download scrapebox from here:
http://www.scrapebox.com/payment-received






Actually my public proxies are port scanned proxies from a service here in BHW, so its not that public (scraped from sources)

Ill try the custom harverster and see what happens! Thanks


I have 10 semi dedicated proxies but had no luck harvesting with them, maybe pvt proxies would be better, but well now im using these proxies lets see if i can tweak it all and get the most out of scrapebox

regards


edit: im noticing something here. My scraping start at around 100urls/s and after like 1 min it keep dropping till 20urls/s...is it normal?

Its possible that at that point proxies are startting to get banned, and scrapebox is having to cycle thru more proxy retries to get results. My Urls a sec seem to start a tad high and level off, but not that much, but with private proxies its going to be fairly consistent. Public proxies, I would think, would be volatile. Personally I don't pay much attentions to urls/sec as I walk away and come back to results so if its a little slower or a little faster, I don't care. Im more worried about consistency and that it actually works.

I noticed that my sb is always crashing like a baws? Do others have same experience? in any thing I do, keyword scraping, It always crash lately.

I have over 30 instances running across about 10 pcs at any given time, and I have no issues. Does it give popup errors? Does it say not responding? Can you be more specific about "always crashing like a baws" ?? Does it ghost white or does it black out?

Do counts not increase but the window is still able to be moved around?

I could go on all day, there are thousands of possibilities, its about equivilant to saying "my car doesn't work". The more detail you give the more accurately I can direct you. :)


The last 2 times I've done some harvesting I crashed...I just assumed it was the proxies going bad but it shouldn't lock up anyways.

Same as above, can you give more details?
 
Last edited:
Bing / Yahoo harvester crashed a few times, it also have given my (only yahoo and bing) results like: Hi / <a href=" / <b>example</b> so no urls :P
Any idea what could be wrong?
 
bing harvester crashes

yahoo harvester is broken...not scraping urls properly
 
Ok, I made a video to get into the details. Sorry for not posting earlier.

youtu . be / -_kB9_c7pY8

please remove the spaces (I can't post links)
 
Bing / Yahoo harvester crashed a few times, it also have given my (only yahoo and bing) results like: Hi / <a href=" / <b>example</b> so no urls :P
Any idea what could be wrong?

Try the custom harvester, it works fine. Settings >> use custom harvester. Yahoo is broken and its a known issue, in fact its already fixed in the next version. But there are 3 harvesters in scrapebox, the most powerful one is the custom harvester, which you can update the engines for at any time and it can be modified easily if something breaks. It works, and works fine and supports over 20 engines. The old multi threaded harvester is whats broken, and perhaps the single threaded harvester.

bing harvester crashes

yahoo harvester is broken...not scraping urls properly

Again yahoo is fixed in the next version, but you should try the custom harvester for both bing and yahoo, its working fine.

Ok, I made a video to get into the details. Sorry for not posting earlier.

youtu . be / -_kB9_c7pY8

please remove the spaces (I can't post links)


In your main scrapebox folder there should be a file call bugreport.txt. Can you please send that file to support, it contains the crash data.

Else if there is not, next time you get a crash, can you please choose the "save report" option or click do not send to server, and then it will save a bugreport and can you send me that please.

scrapeboxhelp (at] gmail [dot} com

or

support {at) scrapebox |dot\ com
 
Try the custom harvester, it works fine. Settings >> use custom harvester. Yahoo is broken and its a known issue, in fact its already fixed in the next version. But there are 3 harvesters in scrapebox, the most powerful one is the custom harvester, which you can update the engines for at any time and it can be modified easily if something breaks. It works, and works fine and supports over 20 engines. The old multi threaded harvester is whats broken, and perhaps the single threaded harvester.



Again yahoo is fixed in the next version, but you should try the custom harvester for both bing and yahoo, its working fine.

Yes I tested yahoo in custom harvester and is not working for me.

But with the custom harvester how do I know which keywords/footprint were processed successfully?

With Multi-Harvester you can "Export all not completed keywords to keyword list" after harvest is complete. With custom harvester I don't even know what keyword has been completed successfully.
 
Yes I tested yahoo in custom harvester and is not working for me.

But with the custom harvester how do I know which keywords/footprint were processed successfully?

With Multi-Harvester you can "Export all not completed keywords to keyword list" after harvest is complete. With custom harvester I don't even know what keyword has been completed successfully.

Yahoo made some changes which broke all harvesters, the RankTracker etc these have all been fixed in the update below. With the custom harvester it can download a fresh engine file from the server, we updated this fairly quickly but if you didn't select to download the latest engines file this would explain why it didn't work.

Also there's no way to export the not completed keywords in the custom harvester, it can use around 30 search engines at once by default. So a keyword that's completed in one engine might not be completed in the other 29 engines because they all harvest at different speeds. It would need to export 30 keyword files, and as you can see things will get messy.

A new update is available.

ScrapeBox v1.16.3:
New: Notification when proxy port blacklist/whitelist enabled
New: Email grabber filter for emails containing
New: Option to remove urls with specific extensions in harvester
New: Proxy Sources
Fix: Yahoo Search Engine in all harvesters.
Fix: Indexification service

Automator:
New: Alexa Rank Checker addon
New: Export and randomize command
New: Export and split command

RankTracker:
New: URL's and Keyword Export option
Fix: Yahoo Search Engine

Article Scraper:
New: Article Directory Added

Addons:
New: Vanity Name Checker Addon
New: Sitemap Scraper Addon numerous options
New: Alexa Addon numerous options
New: Article Scraper Addon Directory Added
Fix: Social Checker Addon Google +1


Also a cool new video by Loopline:

 
Last edited by a moderator:
Awesome update and video. I'll test it and see if I still crash. With the Vanity Name Checker the 2 times I've tried it I get socket errors as most of the results. Any idea what that means?
 
Awesome update and video. I'll test it and see if I still crash. With the Vanity Name Checker the 2 times I've tried it I get socket errors as most of the results. Any idea what that means?

That means something forcibly closed the connection. Does it happen with and without proxies? You need to close down the addon, uncheck the use proxies box and then restart the addon. Then try the same thing but check the use proxies box.

Does it happen in both cases or only with proxies etc..?

Generally its either security software such as a firewall, anti-virus, etc... or its proxy related.
 
I did have better results when checking for weebly accounts but I still got socket errors. I'll try your suggestion tomorrow when I try the add-on again.

I also previously mentioned my harvester was locking up but by process of elimination Bing is causing it. After so many results and just gets stuck and when I try to abort the harvest Google stops fine but bing wont and I have to force shut it down.

Thanks for the help.
 
I did have better results when checking for weebly accounts but I still got socket errors. I'll try your suggestion tomorrow when I try the add-on again.

I also previously mentioned my harvester was locking up but by process of elimination Bing is causing it. After so many results and just gets stuck and when I try to abort the harvest Google stops fine but bing wont and I have to force shut it down.

Thanks for the help.

The harvester lockup happens on some machines in the multi threaded harvester. So I would recommend using the custom harvester, as it doesn't happen in the custom harvester and I think there are 30 engines in there vs 4 in the multi threaded harvester.
 
Status
Not open for further replies.
Back
Top