Permanently Closed Marketplace Sales Thread

Status
Not open for further replies.
A buddypress account username david20160110bowie created
with the learning system.

Code:
[setup]FriendlyName=BudyPress
Platform=Blog
PageMustContain=/buddy
PageMustNotContain=
LoadUrl=register
Footprint="proudly powered by wordpress and buddypress" inanchor:cheap
Success=<h2>Sign Up Complete!</h2>|activate your account via the email we have just sent to your address.</p>
Failed=<p><strong>ERROR:<
Markup=HTML
UseBlackList=1
UseWhiteList=1
[STEP]
FormMustContain=name="signup_form"*name="field_1"
FormMustNotContain=
PostUrl=%host%%path%register/
signup_username=david20160110bowie
g2_form[fullName]=David Bowei
[email protected]
signup_password=00001111
signup_password_confirm=00001111
field_1=David
field_4_day=
field_4_month=
field_4_year=
field_7_day=
field_7_month=
field_7_year=
field_58=USA
field_47=Phonix
field_14=I am older than 13.
field_5=%rnd-option%
captcha_code=%captcha%
signup_submit=%ignore%
 
Hi,
how can we extract only do follow links from a list of urls?
Any suggestions...
 
Last edited:
Review:

Just writing a short review, not that Scrapebox needs them at this point. :D

Simply put, Scrapebox is one of the best(at least Top 3) pieces of software created for the SEO industry. I have been using it over 5 years and I am truly grateful it was invented. Thanks a million guys!
 
I think the automator has issues with the harvester module. I've pasted in ~100 common word into the keyword field in the automator and set it to merge it with about edit5:"1000" footprints I've got and made from GSA SER.

Now when I launch this job in the automator, the scrapebox instance seems to take forever to merge the 100 keywords I've specified(these are clearly loaded into the keyword fields) with about ~950-1000 footprints from a file I've specified.

I've done the same manually and it only takes ~1 second to the merging. I've waited about 3 minutes for the first time before killing the scrapebox instance.

See the image attached:
http://s12.postimg.org/o49cznn5n/20160111_automator_stuck_with_merging.png

edit: okay, bhw forbids postimg.org, but I won't change that.
edit2: everything is updated to the latest version.
edit3: I've done my best to url decode with meyerweb SER's fucked up encoding before importing into Scrapebox.

Now I can get around to it by merging the footprints with the keywords and just paste them into the keywords field in the automator job creator, but I thought I just let you know.

edit4:
Okay, now the automator job creator crashes when I try to manouver with 90k keywords(keywords and footprints combined).

edit5:
got it working without crashing by having ~1000 footprints in the automator job creator "harvester" module keyword field and have it merge with the 100 keywords. It looks good from testing.

OS:
Windows SRV 2012 R2 Latest build
Dual e5420
16GB ddr2 ram

Ok, Im slightly confused I had trouble following all that. So it is or is not working?

If not and you want to send me your automator job file I can have a look, or if you want to do a video.

Hi,
how can we extract only do follow links from a list of urls?
Any suggestions...

You can do the ******** test addon. You need to know a link already on the page. I am not sure if you have a list of urls with links and you want to know the ******** ones or if your trying to just get an assumption that if you build a link on a page if it will be ******** or not?

video
https://www.youtube.com/watch?v=SBiy0PUFZGc
 
Use the test addon or if you want to know all the ......:) links:
http://beginnersbook.com/2012/11/********-vs-nofollow-backlinks/
http://imgur.com/a/4Coa7
Additional work after data downloaded is needed.
You can scan page first.
 
Last edited:
Yes, I got it working, but I had to do it in a reversed fashion than I originally wanted to, but it works for now.

I also edited my post multiple times hence adding to the confusion.

I had to paste in the footprints(about 1000 for articles from SER) into the keyword field and use the 100 keywords to use to merge with the footprints.

###

Do you know what URLs are in the cloud blacklist? I believe I only saw an encrypted or machine compiled code in the file there. At least notepad ++ can't read it.

Ok, Im slightly confused I had trouble following all that. So it is or is not working?

If not and you want to send me your automator job file I can have a look, or if you want to do a video.



You can do the ******** test addon. You need to know a link already on the page. I am not sure if you have a list of urls with links and you want to know the ******** ones or if your trying to just get an assumption that if you build a link on a page if it will be ******** or not?

video
https://www.youtube.com/watch?v=SBiy0PUFZGc
 
You can do the ******** test addon. You need to know a link already on the page. I am not sure if you have a list of urls with links and you want to know the ******** ones or if your trying to just get an assumption that if you build a link on a page if it will be ******** or not?

video
https://www.youtube.com/watch?v=SBiy0PUFZGc

Thanks for your support Matt,

But actually I'am scraping expired domain.So Now I don't want to extract no follow comment(and such) links but actual in content do follow links.I have large number of URLs from various domains and want to extract only do follow links.
Any suggestions for doing that?
 
Used yepl scraper, 4/5 it gets stuck for some reason. It process 99% of the process and waits for 2-8 threads to finish, which i left the whole day to process. End up force closing the program and end up empty handed.
It scrapes the info extremely fast for 1100 results, in like less than 40s. But took a whole day without success for 2-8 threads to finish?
Im using public proxies. Might be the issue here, what format could i use?
Any idea what could be the problem?
 
Yes, I got it working, but I had to do it in a reversed fashion than I originally wanted to, but it works for now.

I also edited my post multiple times hence adding to the confusion.

I had to paste in the footprints(about 1000 for articles from SER) into the keyword field and use the 100 keywords to use to merge with the footprints.

###

Do you know what URLs are in the cloud blacklist? I believe I only saw an encrypted or machine compiled code in the file there. At least notepad ++ can't read it.

Ok, good deal. I don't know what is in the cloud blacklist. I presume scrapebox.com is and then stuff like matt cutts blog etc... but I don't actually know whats there.

Thanks for your support Matt,

But actually I'am scraping expired domain.So Now I don't want to extract no follow comment(and such) links but actual in content do follow links.I have large number of URLs from various domains and want to extract only do follow links.
Any suggestions for doing that?

I suppose you could try the custom data grabber and build the after as just a link, thus it won't scrape any that have the nofollow tag as it wouldn't match. At least thats the best I can think of, its going to depend on the end sites formatting so I would expect given that some sites don't follow convention and for errors that you will get some false positives and miss some potential links. Video here:

https://www.youtube.com/watch?v=X3Ep-NXg4kY

Used yepl scraper, 4/5 it gets stuck for some reason. It process 99% of the process and waits for 2-8 threads to finish, which i left the whole day to process. End up force closing the program and end up empty handed.
It scrapes the info extremely fast for 1100 results, in like less than 40s. But took a whole day without success for 2-8 threads to finish?
Im using public proxies. Might be the issue here, what format could i use?
Any idea what could be the problem?

Well the threads got locked. I can't say for sure what caused it, could be security software like anti-virus, could be some 3rd party program on your pc, could just be a bad connection with the end site. Does it always do this or only this once?

Make sure the plugin is white listed in all security software and try closing down any unneeded programs and see if that does it.
 
Well the threads got locked. I can't say for sure what caused it, could be security software like anti-virus, could be some 3rd party program on your pc, could just be a bad connection with the end site. Does it always do this or only this once?

Make sure the plugin is white listed in all security software and try closing down any unneeded programs and see if that does it.

No antivirus everything whitelisted, no programs open. Still the same problem. Used public proxies that SB liked (SB proxy checker).Any other solutions?
 
No antivirus everything whitelisted, no programs open. Still the same problem. Used public proxies that SB liked (SB proxy checker).Any other solutions?

Give your computer a swift kick. :D hehe

So if you load the same keywords in does it do it over and over or is it random? Meaning can you intentionally reproduce the issue with a given set of data?

Because if so then support can find the problem specifically if its on the urls its self etc..

None the less Ill mail them about it to see if anything can be done without specific data.

If you can reproduce it can you post/share the data and settings (screenshot is great) here or you can just mail it to support.

scrapeboxhelp (at] gmail [dot) com
 
"You can have multiple [Step] configured for multi-step forms that may require you to fill out info on 2 or more pages."
This is an example, registration for a free forum.
Code:
[setup]FriendlyName=G
Platform=Image
PageMustContain=<html
PageMustNotContain=Registration is disabled
Footprint=


Success=%header-phpbb3%|>Your first forum<
Failed=Registration is disabled




loadurlfromanchor=register
loadurl=register
Markup=html
UseBlackList=1
UseWhiteList=1
[step]
FormMustContain=terms|/register.html
FormMustNotContain=
PostUrl=register.html


terms_agree=on
submit=%ignore%
DoStepIf=terms_agree
[step1]
FormMustContain=register.html
FormMustNotContain=
PostUrl==register.html


forumname=board online                   
forum_dir={%rnd-option%|3}                             
title=Test Board            
desc=Poster Best                   
admin_login=007Admin                   
admin_pass1 =00001111 
*mail*[email protected]                    
submit=%ignore%
[step2]
FormMustContain=register.html
FormMustNotContain=
posturl=register.html


admin_pass2=00001111  
submit=%ignore%
 
Just know about 50-70% the whole logic of learning system of .ini file, also I am still a newbie so the example above umm, is just an "example".
 
Last edited:
That's your email. You were logged into that gmail in one of your older videos.

Give your computer a swift kick. :D hehe

So if you load the same keywords in does it do it over and over or is it random? Meaning can you intentionally reproduce the issue with a given set of data?

Because if so then support can find the problem specifically if its on the urls its self etc..

None the less Ill mail them about it to see if anything can be done without specific data.

If you can reproduce it can you post/share the data and settings (screenshot is great) here or you can just mail it to support.

scrapeboxhelp (at] gmail [dot) com
 
Adding new Platforms to ScrapeBox


Open all .ini definition files, remove page must contain content. Save all.
9WOVr7K

Add a new .ini file;
EJDV364


http://imgur.com/gbcLjOX

It's FriendlyName shows in the Select Platforms window.

4pF6Hh0


Start Poster:
lolzV9F


acadamia.edu registration

yeLgaEe


G4X67WR


Code:
[setup]



FriendlyName=General
Platform=Image Comment


PageMustContain=<title
PageMustNotContain=
Footprint=whatever
Success=Confirmation email|login_token|/confirm/sent
Failed=


loadurlfromanchor=Sign Up
loadurl=/login


Markup=html
UseBlackList=1
UseWhiteList=1
[step]
DoStepIf=>Sign Up
loadurlfromanchor=Sign Up


FormMustContain=terms|/register.html|action="/user/registration.html"
FormMustNotContain=xxx


PostUrl=/registrations|register.html|regist|registration


username=depphotmail
[email protected]
password=00001111 
*first_name*=Depp
*last_name*=John
*mail*[email protected]
*pass*=00001111 
terms_agree=on
submit=%ignore%
DoStepIf=>Terms|value="Sign Up"|terms_agree|registrations|action="/user/registration.html"




[step1]
FormMustContain=terms|/register.html|action="/user/registration.html"
FormMustNotContain=xxx
PostUrl=register.html|regist|registration
[email protected] 
password=00001111 
username=blacktion
*mail*[email protected] 
*pass*=00001111 
terms_agree=on
submit=%ignore%
DoStepIf=terms_agree|registrations|action="/user/registration.html"
[step2]
FormMustContain=register.html
FormMustNotContain=
PostUrl==register.html


forumname=Olgar                   
forum_dir={%rnd-option%|3}                             
title=T Board            
desc=Olgar Oak                  
admin_login=007Admin                   
admin_pass1 =00001111 
*mail*[email protected]                    
submit=%ignore%
DoStepIf=xxxx
[step3]
FormMustContain=register.html
FormMustNotContain=
posturl=register.html


admin_pass2=00001111  
submit=%ignore%
DoStepIf=xxxx
 
Hi, I've been reading up on scrapebox now and I'm convinced I must buy it! I come across peoples post that state it's $57 but I think the price has gone up? Or is there a page I'm supposed to go to?
 
Hi, I've been reading up on scrapebox now and I'm convinced I must buy it! I come across peoples post that state it's $57 but I think the price has gone up? Or is there a page I'm supposed to go to?

Full price is $97 but the BHW price is currently $67.

http://www.scrapebox.com/bhw

Or you can get the program plus the 3 premium plugins for $99 instead of buying it for $67 plus each plugin at $20 times 3. So for $2 more then the normal price of just the program you get the program plus $60 worth of plugins. Its an old halloween promotion.

http://www.scrapebox.com/halloween
 
Is there a way to specify a folder's files in the automator to be processed in an alphabetical order?

I've got over 100 million links that I want to extract outbound links from and these are over in 1000 files. Making 1000 automator steps for each would take days.

Full price is $97 but the BHW price is currently $67.

http://www.scrapebox.com/bhw

Or you can get the program plus the 3 premium plugins for $99 instead of buying it for $67 plus each plugin at $20 times 3. So for $2 more then the normal price of just the program you get the program plus $60 worth of plugins. Its an old halloween promotion.

http://www.scrapebox.com/halloween
 
Status
Not open for further replies.
Back
Top