Permanently Closed Marketplace Sales Thread

Status
Not open for further replies.
Is anybody using the blog analyzer to clean up the lists before posting? Do you find it accurate enough, or does it skip good pages? Not to mention it takes a lot of time on a 100k list and 10 good proxies.
It's my impression that it marks good pages bad for some reason. Would it be better practice to just post to the list and weed it out that way instead of running the blog analyzer first?

edit: I notice in the current analyzer instance that I'm running that I get several good pages for a certain site, but then many pages marked with error 403 for the same site. I get a bunch good, a bunch bad, one or two good, several bad, then another bunch good, then bad and so on. No particular pattern, like all good at the beginning and then all bad. But the thing is, I checked several bad pages and they all show fine in the browser, with comment form and all. What's the issue here?

ok, another edit: blog analyzer is fast this time around on my 100k list - turns out that so far I still had some free proxies being used by the program and speed was really crappy. It will probably take some 15minutes to go through the entire list at 100 connections.
But still the questions remain: do you use the analyzer on a regular basis and how accurate do you find it to be?
 
Last edited:
Quick question on running Scrapebox...I have a pc that has a TB of hard drive space, 8gb of RAM, is an Intel Core i5 CPU 760 @2.80ghz...would this be sufficient to run Scrapebox without having to purchase a VPS?
 
I cant see exactly when its running but when it hits a proxy that doesn't work - it just skips it , right ?

If I remember correctly, when sbox runs into a bad proxy or a proxy dies while it is in use by sbox, the app selects a different proxy for use on the operation currently under way.
 
Quick question on running Scrapebox...I have a pc that has a TB of hard drive space, 8gb of RAM, is an Intel Core i5 CPU 760 @2.80ghz...would this be sufficient to run Scrapebox without having to purchase a VPS?

Yes those are quite adequate specs to run multiple instances of ScrapeBox (though there are advantages to running it on a VPS or the like), but what you're describing is plenty beefy a machine to run the app.
 
Yes those are quite adequate specs to run multiple instances of ScrapeBox (though there are advantages to running it on a VPS or the like), but what you're describing is plenty beefy a machine to run the app.

Thanks Greg! Also, what are the main advantages to running it on a VPS? I have seen people talk about running on one but not really why.
 
Thanks Greg! Also, what are the main advantages to running it on a VPS? I have seen people talk about running on one but not really why.

There are a few; if your ISP has monthly bandwidth caps or gets "suspicious" about data bursts, a VPS would get you around that (sbox can chew through a lot of data if you use it heavily). For some people, a VPS can also let the app do its magic more quickly in terms of pipeline speed (depends on your local ISP's actual performance).

It also allows you to have the app running as much as you want without tying up your PC (the app is light-weight in terms of resources, but if you really hammer it - like running multiple instances - it can be nice to have it running elsewhere).

And it gives you a bit of isolation from your local machines in terms of anonymity of the source.
 
a bit off topic, why in this thread there are 232 pages, but when I click on it I am taken to page 230 only LOL? :D
 
I already Own Scrapebox and done some commenting a while back ago but just in a low volume... I want to use it to only scraping urls now ....

I can't use Google because all the proxies fail at Google test. I used an alternative proxy finding method but it's not on its potetional.

So I'm using Yahoo now:

My problem is:

I don't get relevant matches. I used intitle, tried allintitle, inanchor, inurl or just throw keywords in with quotes ... and let's say I target "restaurant" keyword and I got websites about law and free games where not a single restaurant phrase is written anywher (this wasn't only with inanchor...)

I'm sure this question was already asked but what can I do to get relevant matches on Yahoo? - and If I remember right I have had similar issues with Google also when I used to do commenting ...

thx
 
Last edited:
This looks great, worth trying out?

Yes of course it is. Did you really need to ask that question? BTW, posts in the BST section (which is this) don't inflate your post count, which is likely what you were doing. :p

I cant see exactly when its running but when it hits a proxy that doesn't work - it just skips it , right ?

i ask cos i was scraping a big directory and they banned my ip even with proxies I'd checked --- so if you extract data and teh proxie is a dud -- it skips it - ?


less important q
how does it know if it works or not ?



ta

Under settings >> adjust multi threaded proxy harvester retries - you can adjust the retires used when scrapebox encountres a bad proxy. I think its 3 at default. So if it finds a bad proxy, it skips it and goes to the next one. It continues until it reaches the setting you have in the above, then it just skips the query all together.

It knows if it doesn't work based on error codes and such. So if the proxy returns a 407 or 403 or 404 it knows its not working. If it times out, or just ignores the request, its also not working etc....


Thanks Loopline, I'm experimenting with connections.

Your welcome. :)

Is anybody using the blog analyzer to clean up the lists before posting? Do you find it accurate enough, or does it skip good pages? Not to mention it takes a lot of time on a 100k list and 10 good proxies.
It's my impression that it marks good pages bad for some reason. Would it be better practice to just post to the list and weed it out that way instead of running the blog analyzer first?

edit: I notice in the current analyzer instance that I'm running that I get several good pages for a certain site, but then many pages marked with error 403 for the same site. I get a bunch good, a bunch bad, one or two good, several bad, then another bunch good, then bad and so on. No particular pattern, like all good at the beginning and then all bad. But the thing is, I checked several bad pages and they all show fine in the browser, with comment form and all. What's the issue here?

ok, another edit: blog analyzer is fast this time around on my 100k list - turns out that so far I still had some free proxies being used by the program and speed was really crappy. It will probably take some 15minutes to go through the entire list at 100 connections.
But still the questions remain: do you use the analyzer on a regular basis and how accurate do you find it to be?

Thats a loaded question. I use blog analyzer only to determine platform type of a url. If you were using slow poster it would be much quicker to use the blog analyzer then to post to the whole list. If its faster poster, its quicker to just post to the list.

Blog analyzer is an addon that is work in progress, the developer has plans to improve it, but its one of those things that fistly blogs are ever changing, and secondly with a constant time crunch for development, developing things like the actual commenter always come first.

Blog analyzer has its faults, but is good over all.

Maybe you are not ready yet to see the good things on page 231 and 232? :D

LOL, nice! I am appearantly unready to see them too, cause its always been this way for me. What kind of wonderful things are you hiding on the pages we can't see? hehe

Heya - Does your software spider and auto reply to vbulletin forums or just blogs?

No, its for blogs only on the commenting side of things. It can crawl the engines for vbulletin footprints, but it doesn't spider sites on the whole, with, I think, the exception of the sitemap addon.

Scrapeboard works with forums, but its not for sale yet.

I already Own Scrapebox and done some commenting a while back ago but just in a low volume... I want to use it to only scraping urls now ....

I can't use Google because all the proxies fail at Google test. I used an alternative proxy finding method but it's not on its potetional.

So I'm using Yahoo now:

My problem is:

I don't get relevant matches. I used intitle, tried allintitle, inanchor, inurl or just throw keywords in with quotes ... and let's say I target "restaurant" keyword and I got websites about law and free games where not a single restaurant phrase is written anywher (this wasn't only with inanchor...)

I'm sure this question was already asked but what can I do to get relevant matches on Yahoo? - and If I remember right I have had similar issues with Google also when I used to do commenting ...

thx

Well you really just need better proxies for google, sources that aren't being used by a billion other people. I have a video on it here:
http://www.youtube.com/watch?v=o_NMkQgK1PY


However if you google up yahoo search operators (hehe, funny pun) you will notice yahoo doesn't support some of the operators you are using, they are google specific operators. So thats reason 1 why your having bad luck. Reason 2 is that some yahoo searches shunt you off to other yahoo products and scrapebox doesn't follow that redirect. Reason 3 is that yahoo and bing together only make up about 1/4 of total search volume. There is a reason for this, their algorithims aren't as "accurate" or "on target" as googles, IMHO.

I personally wouldn't settle for yahoo only, just find better proxies. But if you are going to use yahoo, you have to understand firstly, its not google and it doesn't work like google.
 
Taking forever to get activated on my new windows install... Should I email support or wait it out?
 
I asked for a refund within 7 days after i bought scrape-box because i could not get it to work on my version of parallel desktop. It has now been over 2 weeks since then, and i still have not received any response.
Can anyone help me?
 
Is there any way to recover the last session before it crashed. I was on Fast Poster.

Appreciated
 
How many keywords should i use as my name for commenting?
thank
 
Hi there ... .so I finally took the plunge ...

Your transaction ID for this payment is: 96C063****. ...

Can't wait to start using this ...
 
Last edited:
Anybody scrapebox having problem accessing decaptcha?

I tried directly access to decaptcha website and I can login. But from scrapebox I can't. Unable to check balance. I have tried from 2 different servers from different locations.
 
Hi there ... .so I finally took the plunge ...
Your transaction ID for this payment is: 96C063****. ...
Can't wait to start using this ...

never post Your transaction ID
not unless u want some1 to steal your license
 
Status
Not open for further replies.
Back
Top