Permanently Closed Marketplace Sales Thread

Status
Not open for further replies.
I'm having some serious issues with scraping large amounts of URLs at a time. I hit about 1.5mil this time before it just mysteriously crashed (the ever unhelpful "this application needs to close" message from windows).

Could it be my keyword list making it crash? I'm loading about ~10k keywords per run. I'm on Win7, anyone else experienced issues with these large lists?

Nah mate, Im running a i5 with 8GB RAM and win 7 64 bit and I run with 250K+ keywords and pull over 10 millions urls in a run. Sounds like its your windows setup, how is your RAM when that happens? I would venture to guess that your running out of resources. A guess.
 
Harvesting problem in Google. When trying to scrape longer text strings with multiple operators.

Example:
"powered by wordpress" -"wordpress" site:.com inurl:wordpress <-just a example.

It works fine in Yahoo, but in Google it doesn't (it works sometimes when running major keywordlists).
It works when i do a manual search in google. I've tried multiple combinations and it's not necessarily a Scrapebox issue.

And... some suggestions:
Ability to choose a custom value in "Delay:" in the comment poster, makes it possible to run scrapebox 24/7 safely...

And maybe additional custom fields in "learning mode" which would allow posting to forums, etc... Just suggestions though.
-

Scrapebox is an amazing piece of software for it's price.
 
Nah mate, Im running a i5 with 8GB RAM and win 7 64 bit and I run with 250K+ keywords and pull over 10 millions urls in a run. Sounds like its your windows setup, how is your RAM when that happens? I would venture to guess that your running out of resources. A guess.

Thanks for the suggestion. I thought about that too, but Scrapebox is actually really resource friendly to me. I only have 4gb of ram and am also running it on a Win7 Macbook Pro (late 08), but my computer isn't slowing down and my CPU usage is really low.

I'm really baffled by it, I'm not running anything fancy in the background with it, all I had open earlier was just a text editor. At least now I know it's not a problem with the keyword list.
 
Harvesting problem in Google. When trying to scrape longer text strings with multiple operators.

Example:
"powered by wordpress" -"wordpress" site:.com inurl:wordpress <-just a example.

It works fine in Yahoo, but in Google it doesn't (it works sometimes when running major keywordlists).
It works when i do a manual search in google. I've tried multiple combinations and it's not necessarily a Scrapebox issue.

And... some suggestions:
Ability to choose a custom value in "Delay:" in the comment poster, makes it possible to run scrapebox 24/7 safely...

And maybe additional custom fields in "learning mode" which would allow posting to forums, etc... Just suggestions though.
-

Scrapebox is an amazing piece of software for it's price.

Ok, I'm trying to wrap my head around what your trying to do here. Your example isn't making sense, which is confusing me on what you are asking.

What do you mean by "it doesn't work"?

Does it return no results or just different results then when you do it manually?


Also in your example:

"powered by wordpress" -"wordpress" site:.com inurl:wordpress

You have contradicting elements there, I assume that you aren't acually using that string for your searches? Just checking.
 
Does it always only happen with Yahoo?

No there are different search engines and dif error codes.. Like this time did i set all four off the engines on 75 threads each and I just now realized afew hours later that it froze after 24 min in and 1 million scraped urls.

What does the the different error codes mean. I got a 301 this time...

Well here is my experience. I am running a i5 8GB Ram machine on windows 7 64bit. (these are as of my connection being 24Mbit/2Mbit. I just got my 50/10 recently, but I haven't retested, although I feel it will be the same or similar.)

When harvesting I typically don't push it above 75-125 connections. 75 seems to be optimal for me. Similar in posting/link checking etc... 100 connections.

There seems to be a threshold for ideal performance. I can run 3 instances at 75 connections harvesting and get quicker more stable results then 1 instance at 150. I can run 5 instances posting or link checking at 100 connections with much better results then I can run 1 connection at 200 connections. I get way more timeout errors an other errors.

Point being I turn connections down and increase my instances and get better results. Depends on your system configuration, I just tested different settings and settled on what I use now.

Also it works with sessions when harvesting. So Google for instance sees your IP, but they also see your scrapebox instance as a separate session from other scrapebox instances running on your PC. So each SB instance gets its own treatment in addition to your IP getting treatment rules.

So keeping your connections down will help your particular session/instance will help minimize your issues. (this is from my understanding of it, which is not rock solid, so someone correct me if I am wrong).

But even when unblocking ips via the proxy harvester for a particular SB instance with Google doesn't unblock them for the other instances.


The short version there is that Greg is on point, the first thing I would do is try turning your connections down, like real low and see what happens and then slowly increase them and see where your cut off point is. If that doesn't work then go from there.


I will try to mess around with that try that.... Thanks for the help

AND BTW I pay $30 for my internet connection 100 m/bits down and 10 m/bit up.

Got to love sweden :)
 
Questions here. Is it better to use VPN services rather than proxies in running scrapebox?
 
I have a question on Scapebox. How do you promote keyword using Scrapebox? For example when you comment on a blog, it's your name (not the keyword) that becomes anchor text. How would that help you to rank for the keyword? Google as I know look at the anchor text to rank you.

Should I be using the keyword as name?
 
Scrapebox keeps popping up a login/password box for proxy when using slow commenter. They are private proxies which works great in all other modes. Any way around this?
 
I have a question on Scapebox. How do you promote keyword using Scrapebox? For example when you comment on a blog, it's your name (not the keyword) that becomes anchor text. How would that help you to rank for the keyword? Google as I know look at the anchor text to rank you.

Should I be using the keyword as name?

A good strategy that I've used is to go ahead and make a Scrapebox blast with your keyword as your name. A lot of blogs will reject these comments, but that's okay. Take the ones that weren't successful and blast again (if you want) with a false name [it still counts some]. Most importantly, take the ones that went through as your keyword as save them--they're either autoapprove or the owner just doesn't care that much. You can use them in future blasts. Automated software is largely a game of numbers, so don't be worried too much when a lot of keyword comments get rejected.
 
I would like to purchase the 2nd Scrapebox license. Please advise...
 
hi,

i want to include a few websites in one blast. Can you help me. I remember there was an option like that:

i want scrapebox to choose the correct name (anchor) for the corresponding link.

name1 (link1) name2(link2)

i hope you know what i mean thx
 
hi,

i want to include a few websites in one blast. Can you help me. I remember there was an option like that:

i want scrapebox to choose the correct name (anchor) for the corresponding link.

name1 (link1) name2(link2)

i hope you know what i mean thx

I believe you are referring to Link Lock - there's some info on it in this thread, try an advanced search for sweetfunny's posts with that term in the post.
 
I'd like to request a new feature: removing duplicate keywords in the main window.
 
public proxies = scraping
private ones = posting

u can do both with public ones if you
a light user.
dont bother with the ones in scrapebox
there over saturated
 
I believe it has to do with your Internet Explorer security settings
 
Anybody know why I am getting an "error 400" while running the backlink checker addon?
 
If i am using a hub page as "buffer" site, pointing to my main site, and i blast this "buffer" site is it safe for my adsense account if i have google ads on my main site?


And how much does it cost to buy private proxy?
 
I would like to purchase the 2nd Scrapebox license. Please advise...

Pretty basic. Just buy another one like you did the first one and submit the activation info and your done. Thats what I did for my second license.

Just go back to the first page of this thread, and click the buy link.
 
If i am using a hub page as "buffer" site, pointing to my main site, and i blast this "buffer" site is it safe for my adsense account if i have google ads on my main site?


And how much does it cost to buy private proxy?

It should be safe, generally speaking.

Go to help in Sbox and choose proxies from there or you can search around here on BHW. Price can vary a good bit.
 
Status
Not open for further replies.
Back
Top