Permanently Closed Marketplace Sales Thread

Status
Not open for further replies.
An update on my previous issue, SB support have been great as usual and it looks certain the be fixed.
Top Tip: i find being polite and asking specific questions to be more effective than screaming abuse with capslock, just saying!
 
NIXMY
this is just advice, this is sweets sales thread. its one thing
to make comment about what u believe is broken on scrapebox
doesnt seem to effect anyone else.

but once you start throwing in your tools and testing your promoting
your product in another persons sales thread. strictly against bhw
rules, so ide quit doing it. called thread hijacking

dont forget u were warned last week for having a go about
scrapebox up in the proxy section by a mod so ide stop
doing it. just advice .
 
Last edited:
My TDNAM scraper stopped working. Its saying "unable to download new datafile, using local datafile".

I even tried deleting .dat file in addons folder. Now getting "unable to download datafile"

Anyone had such an issue? How to solve it?
 
Hm, still can't seem to access the scrapebox site to download scrapebox again. Loopline any eta you got for when you think the site will be working again?
Main site is up for me and you can download from here:
Code:
[URL]http://www.scrapebox.com/payment-received[/URL]
 
quote_icon.png
Originally Posted by slim_dusty http://www.blackhatworld.com/blackhat-seo/buy-sell-trade/129096-scrapebox-ultimate-serp-scraper-auto-blog-commenter-prstorm-mode-267.html#post4400092
Hi,
I'm getting error 302 (IP blocked) for harvesting results using google passed IPs. Sent email to SB support. Wondering if anyone having similar issue. Was working a few days ago. Using latest version of SB v 1.15.48.


Assuming your using advanced operators? If so:

Hi Loopline,
thanks for getting back to me -
I think my problem was more an academic one rather than a practical problem:

1. I am using public shared proxies for harvesting - not using advanced operators.
2. When I scrape with multi-harvesting, I get good results... no problems.
3. Just recently, I was scraping using the single harvester using google passed proxies (public shared proxies, not scraped from SB). I noticed that 9 out of 10 of the proxy queries were Error 302 (IP blocked) even though I was not using advanced operators and the proxies were all "google passed" by SB checker.

I understand the different proxy bans, but I was surprised to see over 90% of google passed proxies showing as error 302 (IP blocked) even for simple search. As I said, the multi-harvester is working fine.

I'm wondering how accurate the SB google proxy checker is and whether others have the same problem when doing single harvesting. Curious to know what percentage of public proxies are IP blocked for even simple keyword searches.

At the moment this isn't affecting my scraping. Never done scraping with single harvester, so not sure if this has always been the case.

appreciate your feedback loopline.

thanks
 
hay loopline havnt we been saying what u said above for about 3 yrs.
i think most people no by now alot of the proxys returning anon status
are transparent, but it matter not if you only scrape, so its a minor
false positive, but 1 i can live with. doesnt effect me 1 bit or almost
any other user.
 
My TDNAM scraper stopped working. Its saying "unable to download new datafile, using local datafile".

I even tried deleting .dat file in addons folder. Now getting "unable to download datafile"

Anyone had such an issue? How to solve it?

Are yo using the latestest tdnam scraper update? A couple weeks ago godaddy made some changes and there was a new addon released to compensate. Make sure you install the update for the addon.

Hm, still can't seem to access the scrapebox site to download scrapebox again. Loopline any eta you got for when you think the site will be working again?

Its up for me ATM, and it seemed to be working late yesterday, although very slow and I had to retry, was trying to download something as well.

quote_icon.png
Originally Posted by slim_dusty
Hi,
I'm getting error 302 (IP blocked) for harvesting results using google passed IPs. Sent email to SB support. Wondering if anyone having similar issue. Was working a few days ago. Using latest version of SB v 1.15.48.


Assuming your using advanced operators? If so:

Hi Loopline,
thanks for getting back to me -
I think my problem was more an academic one rather than a practical problem:

1. I am using public shared proxies for harvesting - not using advanced operators.
2. When I scrape with multi-harvesting, I get good results... no problems.
3. Just recently, I was scraping using the single harvester using google passed proxies (public shared proxies, not scraped from SB). I noticed that 9 out of 10 of the proxy queries were Error 302 (IP blocked) even though I was not using advanced operators and the proxies were all "google passed" by SB checker.

I understand the different proxy bans, but I was surprised to see over 90% of google passed proxies showing as error 302 (IP blocked) even for simple search. As I said, the multi-harvester is working fine.

I'm wondering how accurate the SB google proxy checker is and whether others have the same problem when doing single harvesting. Curious to know what percentage of public proxies are IP blocked for even simple keyword searches.

At the moment this isn't affecting my scraping. Never done scraping with single harvester, so not sure if this has always been the case.

appreciate your feedback loopline.

thanks

If you not using the advanced operators and you getting it, its because google has recently made some changes that are killing of proxies much quicker. You might notice that you can proxy test and keep google passed and then do another check and a ton of those are dead. Sweetfunny told me they were working very fervently on the issue, but there is no pattern nor rhyme or reason to how google is doing it. I am confident they will come up with a solution. But for now its one of those things were google has changed the rules to the game, and everyone is having to reinvent the strategy by which they play the game.



hay loopline havnt we been saying what u said above for about 3 yrs.
i think most people no by now alot of the proxys returning anon status
are transparent, but it matter not if you only scrape, so its a minor
false positive, but 1 i can live with. doesnt effect me 1 bit or almost
any other user.

Hehe, yeah, I agree with you. I have never once received a complaint when using public proxies in this category. But always someone new coming along with the same thing. Kind of like posting your transaction ID in the thread or asking what setting in scrapebox do you set to automatically create a high page rank, D0f0ll0w, low OBL auto approve list. Its a broken record that is a necessary evil to play. lol
 
Loopline: I use 40 private proxies(8connections) on vps with 2gb ram, single 1.5ghz cpu on 100mbps connection - I am getting from 130 to 160 urls/s

Are proxies limiting me from getting more url/s ? (this is scraping) or is it ram/cpu ?

Another question: When I am scraping big list (over 3k keywords) sometimes it just stops pulling new urls completely and I can't figure out a reason. My private proxies are premium and every time I check them they work and pass google.
edit: I just found out that this problem occurs when I use intitle/inurl. I watched your video loopline and at first I thought proxies are banned from using those lines BUT then I decided to test them individually and they can all still use -inline -inurl ... :s
 
Last edited:
hay don, 40 private proxies is way over kill for scraping
10/20 is enough even though personally i wouldnt scrape
with private proxies. but i wont go into that here. not my
thread. and loopline, how many times do u see some1
swipe a service with a posted ID. they never learn
 
Loopline: I use 40 private proxies(8connections) on vps with 2gb ram, single 1.5ghz cpu on 100mbps connection - I am getting from 130 to 160 urls/s

Are proxies limiting me from getting more url/s ? (this is scraping) or is it ram/cpu ?

Another question: When I am scraping big list (over 3k keywords) sometimes it just stops pulling new urls completely and I can't figure out a reason. My private proxies are premium and every time I check them they work and pass google.
edit: I just found out that this problem occurs when I use intitle/inurl. I watched your video loopline and at first I thought proxies are banned from using those lines BUT then I decided to test them individually and they can all still use -inline -inurl ... :s


130-160 urls sec is pretty good actually. I don't have an "explanation" for it, but somewhere along the way you run into diminishing returns. What I mean by that is, somewhere between 150-300 urls/sec is where it seems to "cap" Meaning if I set connections on the private proxies to 10 or 50 (assuming I have either 50 or 250 private proxies) I still typically see 150-200 urls/sec. So you are doing well. So probably the proxies, but if you get any more I would run them in a separate instance. Scraping doesn't use all that much memory or cpu, but you can watch the performance tab of task manager while its running to get a pretty good gauge.

When it stops pulling urls completely, stop the harvest, then go to settings >> use multi threaded harvester - and uncheck this. Then harvest again. Watch the status column. Is it 302 errors?

If you are testing them individually in a browser it works different. Because Javascript is turned on, and images, and google uses cookies etc... Google can "tell" if your using a proxy more accurately, so they loosen the ban rules. However to do this in scrapebox would require that images load, java script be on etc... and all that translates to a single threaded harvester, no muli threading at all. So that doesn't make sense so scrapebox uses sockets for the scraping, which don't support javascript, but javascript is one way google can detect proxies, and without that the ban rules tighten etc... ( in so many words)

In other words, its all a game, and our job is to learn the unwritten rules and strategies that let us win. :) Often times the only way to do that is to experiment, which it sounds like your doing, but go ahead and push the envelope on your experimenting.

For advanced operators, try setting connections to 10% or 4 instead of 8. Alternatively you can leave that multi threaded harvester unchecked, which will be essentially at 1 connection and you can even turn on the delay for a couple seconds if you want. But "likely" at 40 proxies and 1 connection, you probably wouldn't need a delay.

Or... bet yet, if you wanted to push the envelop and challenge your brain some, try and figure out how to make the footprint you want, without using advanced operators. Quite often it can be done, and then you don't face these issues. The cheese has moved, go find new cheese. Phenomenal book - Who moved my cheese by Spencer Johnson - if you are in business, you should read it.

Anyway I have almost always been able to find a way to get what I want without advanced operators. For example, I won't share how I do it, but I have developed a method for coercing google into giving me all of my competitors backlinks and if I want, their anchor text too. I used to use link:http://www.domain.com Now I don't even use an advanced operator, and I get 1000 times more data then I used to. Its not fast, and it takes work, but the results are very rewarding. Think outside the box, no pun intended.
 
Last edited:
Loopline thanks for such detailed answer :) Btw you were right 1 threaded scraper shows 302(ip blocked).
Time to figure out how to get around this problem.

Cheers
 
I trust that Sweetfunny wouldn't lie about something like this. With that in mind, mr.friend is now on a permanent vacation from BHW. "Wiz"
This is Partiality, may i know why Moderator is favoring Sweetfunny and For what reason my account is banned. I have purchased the software legally from my legit paypal, i have uploaded the payments proof in the forum, I never used any credit card to make the payments, my paypal is verified through a VCC.
 
I have this stupid question

Once you throw in a bunch of free proxies for harvesting keywords and one dies, it gets the other one right ? But is there any "downtime" so to speak ? That it shows your REAL ip while switching ? What if the list of proxies runs out before the job is done (number of requested keywords) ? Are you fried or the job stops ?

Also, i use Proxymultiply to get a list of fresh proxies, so the proxies are already checked as i ask the tool to do so, but when i start to scrape, it asks if i want to test the proxies ? What shall i do ? Test again or it's worthless or scrapebox NEEDS to do so to make sure it doesn't reveal the real IP.
In other words, trust proxymultiply testings only ?

Finally, is it testing the proxies using my own IP (scrapebox that is)?
Thanks
 
testing the proxies doesnt effect you
yes test the the proxies b4 you scrape
as a means to remove the dead ones
 
Anyone of you used scrapebox with hidemyass ?
Or it's big mistake ?

EDITED:
I just tried to grab emails to test the hole thing and i got all failed 414

free proxy entered into scrapebox and using HMY
So i guess it doesn't work ?

HMY was refusing the connection if i understood this right ?
 
Last edited:
I have this stupid question

Once you throw in a bunch of free proxies for harvesting keywords and one dies, it gets the other one right ? But is there any "downtime" so to speak ? That it shows your REAL ip while switching ? What if the list of proxies runs out before the job is done (number of requested keywords) ? Are you fried or the job stops ?

Also, i use Proxymultiply to get a list of fresh proxies, so the proxies are already checked as i ask the tool to do so, but when i start to scrape, it asks if i want to test the proxies ? What shall i do ? Test again or it's worthless or scrapebox NEEDS to do so to make sure it doesn't reveal the real IP.
In other words, trust proxymultiply testings only ?

Finally, is it testing the proxies using my own IP (scrapebox that is)?
Thanks

Yes when one dies it moves on to the next, there is no "downtime" between proxies. Its not like its a solid connection, scrapebox opens a socket to do whatever (scrape,post etc..), if that proxy fails, it then submits a new request via the new proxy, and so on until it works. There are filters built in on scraping, so if a proxy fails it gets added to a list. Each time it fails the fail count goes up for that proxy. At 3 fails its removed. If it succeeds the count goes down. So up and down it goes until a proxy fails too many times and is removed from the array of working proxies for that run (it is not deleted from the list of proxies you load into scrapebox, its just no longer used for that scraping run). If all proxies die, then the scraping stops.

Posting is different, it just posts, if it fails, it skips it and moves on. So you can repost to the failed submissions, but "ideally" you would be using private proxies for posting so this would be less of an issue.

As proxygo said, test the proxies, nuke the bad ones, saves you a lot of time. Scrapebox also works with the proxies in some cases and store session data, so you want to test and then send them immediately back to scrapebox, so you can keep this session data. Scrapebox proxy checks them for anonymity against servers owned by scrapebox. So no issue with your IP if it happens to be transparent. Then it google tests only with the proxies that pass the anonymity test.

Anyone of you used scrapebox with hidemyass ?
Or it's big mistake ?

EDITED:
I just tried to grab emails to test the hole thing and i got all failed 414

free proxy entered into scrapebox and using HMY
So i guess it doesn't work ?

HMY was refusing the connection if i understood this right ?

I have used scrapebox with HMA sure. It works fine. I am curious why you would use HMA AND then turn around and use proxies too? Seems counter productive. The point of using HAM is so you wouldn't have to use proxies.

Your 414 error is likely a proxy issue. That means the URL is too long. Typically the standard is over 2000 characters. You would have to have some insane directories going on or a LOT of encoding to even reach that point. Unless you hit a proxy that has it set way lower, I never had this issue with HMA at all.

If your using HMA, don't use proxies. If you want to use proxies, don't use HMA. For Email scraping/posting etc.. HMA is going to be more solid then public proxies, because its a solid paid service, like wise private proxies will be better then public. However HMA won't get you terribly far when scraping. Although you can set it change IPs every few mins, but you will get failed keywords during the change, so you just have to re scrape the failed ones a bunch of times. A good public proxy list is probably better then HMA for scraping.
 
I see, also i think you're not transparent when it changes proxies. I have a one year deal so might use it but like you said.
HMA for scraping using
Private for posting

By the way, can you elaborate why i wouldn't go far with HMA when scraping ? You mean it will get blacklisted fast ? After a few hundred i suppose ?
I'm still learning the hole thing. By the way, your clips are awesome !
 
I see, also i think you're not transparent when it changes proxies. I have a one year deal so might use it but like you said.
HMA for scraping using
Private for posting

By the way, can you elaborate why i wouldn't go far with HMA when scraping ? You mean it will get blacklisted fast ? After a few hundred i suppose ?
I'm still learning the hole thing. By the way, your clips are awesome !

Glad you find my videos helpful. :)

With HMA, you get 1 IP. So say it scrapes for 4 minuets and then the IP is blocked. So you timed it and figured that out, so you set the HMA software to change the IP every 4 mins. During the time HMA is changing the IP, scrapebox is still running. So its using your real IP, which will quickly get blocked, which is fine cause you can still use google in a browser with a captcha worst case. Once that happens, scrapebox is still hammering away for the 30 seconds to 1 minute while HMA is changing IPs ever 4 mins. So you will burn thru a lot of keywords that will fail while HMA is busy changing IPS. Thats all I meant. Also I have a 1 year deal for HMA and I have posted with it without any issues as well. I use it for other things, but I tested it with scrapebox.
 
Status
Not open for further replies.
Back
Top