qlithe
Supreme Member
- Feb 14, 2012
- 1,252
- 288
Are you using the latest version of the page authority addon? Does it give you an error message?
Yes
No errors
It seems to work just fine, it's just that all the fields except backlinks is blank
Are you using the latest version of the page authority addon? Does it give you an error message?
Hi, and thanks for your link, i already check your youtub page, this is lot of job...
I have some probleme when i finish to start the Harvested, i already check on "usage guid" but i dint find anything...
i do a screenshot becaus my english is so bad this picture going to speak better than me ^^: View attachment 45741
and when i close this window i have no url's harvested : View attachment 45742
did i do somthing wrong ?
thanks per advance.
Yes
No errors
It seems to work just fine, it's just that all the fields except backlinks is blank
Scrapebox link extractor addon closes after some internal link extraction
Im scraping something with search operator "site:".When the scrape started the avg urls/sec was 60 something.Now it droped to 4.Why there was such an segnificant drop?I use 20PP and 2 connections.
I use 20 private proxies.IT could be because your proxies are getting blocked and scrapebox is spending its time cycling thru bad proxies rather then scraping results, which makes the average url slower. Its also possible there are less results for those sites later on so it would slow it down.
Id place my money 10:1 on proxies getting blocked. Are you using no more then 1 connection for every 15 proxies that you have? I would probably do more like 1 connection for every 20 proxies, but 15 would be a start. How many proxies do you have and how many connections are you using and which engines?
I use 20 private proxies.
I was using 10:1 like you mentioned before.And yesterday i setup to 20:1 and it still has dropped to 5 urls/s and for both i used only the google SE. I doubt there is lack of results.
Its hard to say for sure. I mean it could still be proxies being blocked. Like if you are using advanced operators etc... It could still be proxies performing slowly over time.
It could be that later in the mix the results are low, so the number of urls per page are low which slows it down.
Ultimately I don't really actually pay attention to urls/sec. I don't honestly care. If it works it works and if it doesn't work it doesn't work. If you leave it alone there is no guarantee that 5 mins after you stop watching it its not going to go back to 60 urls a sec.
You would have to do a bunch of split testing.
Try different keywords. Try 1 connection for every 30 or 50 proxies.
Try different proxies all together.
It will be process of elimination and lots of split testing.
If it simply doesn't work, its easy enough to nail down, but yours works, its just not as fast as you want. So lots of testing is the only way to find out, then once you understand it all you can tweak the bottle neck and speed it up.
Purchased yesterday and got the license email today but when I start scrapebox. It keeps giving "unable to access any scrapebox server at the moment. please try again in a moment"
Does having a dynamic IP pose a problem?
NVM my antivirus/firewall had crashed making connecting to server impossible. Works now.
Also I purchased the article scraper plugin. How do I go about downloading and activating?
Update: Thats done as well now I get libeay32.dll missing error. how to fix that.
Site: IS an advanced operator. using site: means that you get blocked quicker then if you used just likeNah, i only use "site:google.com mykeyword" this format (but not that domain doooooh).That's nothing special, doubt this is that "advanced" footprint.
Well i also don't pay attention on the urls/sec but the amount of results it scrapes.It used to scrape like 1m results a day now only 100k, i assume something went wrong here.
But thanks for the info, will do that.
Have a beer for me, you are just amazing for what you do.
Site: IS an advanced operator. using site: means that you get blocked quicker then if you used just like
"car"
its not as bad as linke
inurl:car intitle:house inanchor:red
etc...
but 1:10 ratio won't cut it any longer. Go at least 1:15, probably 1:20, possibly 1:30
Ha, I don't like beer, but thanks.![]()
Also if your using just the site: then you can use the link extractor. If you have the automator, you can have the link extractor integrate and have it load in, pull internal links, then export those, filter them to remove anything you don't want like
#
etc...
Then reimport them into the link extractor and extract, export and filter, then reimport etc...
Do that 3 -5 times and you will have a vast majority of the sites internal pages.
Also you could do a
site:
on the domains and then use the link extractor. I find that I always get way more links using the link extractor then just by using site:
You could let the automator do it all, start to finish, and your set.
Are you running the latest version 2.0.0.7? If not please update to the latest version. Also the addon now removes duplicates in real time, plus it saves the urls in real time to the "LinkExtractorData" folder this way if there's problems the urls are not lost. Also if you are extracting all the urls from a 500,000 list, i'm not sure if you realize you could easily end up with 100 million urls if each page has just 200 links.
ScrapeBox Do This?
I need a software/bot that i would be able to send hundreds of comments to a website (blogger).
e.g. Web/URL: (URL to send the comments) http://example.com/music-jaguar-otym.html?m=1 (It should preferably use the mobile site)
Name: (James|Jason|John|Notre|Harry|Barry|Fred|Anonymous )
Comment: (i love this song, its totally awesome|drake has delivered again|This is really good, i cant stop playing it|this guys can't stop making hits|WOW!!!! just WOW|don't like the song, too much autotune|chai, the video just cant wait for the video|every time, wizzy delivers)
can scrapebox do this, i intend using it on blogspot
Thanks again!Well than, have a coffee and think on this post haha (joke).
I am considering buying the automator to be hands free.
But regarding this method, if im only looking for subdomains from that domain, will link extractor do that?
Doesn't it only get some random urls?
About proxies & harvesting, i have a question.
- Testing my proxies say gg 100%
- harvesting turn to error, not harvested
- Manually doing the same give gg capcha on proxy.
I would expect scrapebox to use my capche solver, but not...
I would expect scrapebox to report proxy as dead is capcha requested, but not...
I am stuck on this.
what is best practice on that ?
Yes, i'm running the most recent version. I get an error that says my license isn't valid or something to that effect. I have a valid license and emails that prove I purchased Scrapebox. What should I do? I need this up and running. Can you PM me or tell me who I need to talk to? I've been trying to get this working for 2 weeks now.
support (at} scrapebox {dot) com
mail them directly, that is who you need to talk to about licensing.
No offense, but don't you think i've tried that already...like 10 times already? Are you part of their support team? Is this the kind of support I can expect now from them? I have paid for two licenses for Scrapebox and neither work. I'll try one more time to contact support.
No offense, but don't you think i've tried that already...like 10 times already? Are you part of their support team? Is this the kind of support I can expect now from them? I have paid for two licenses for Scrapebox and neither work. I'll try one more time to contact support.