Permanently Closed Marketplace Sales Thread

Status
Not open for further replies.
hi,

"Mass URL Shorten" plugin have three failed result,i got: dl4.pl, hud.sn,ck2.it .anyone meet this?

and I hope OP can add more shorten service ,like goo.gl ,bit.ly and so on.
 
hi,

"Mass URL Shorten" plugin have three failed result,i got: dl4.pl, hud.sn,ck2.it .anyone meet this?

and I hope OP can add more shorten service ,like goo.gl ,bit.ly and so on.
You can custom it.

++
Email me help beat cheap books. Or I stop posting here from now on.
 
hi,

"Mass URL Shorten" plugin have three failed result,i got: dl4.pl, hud.sn,ck2.it .anyone meet this?

and I hope OP can add more shorten service ,like goo.gl ,bit.ly and so on.

As noted above you can customize and edit the engines and add your own. I did look at goo.gl not long ago helping someone out and it couldn't be added because of how it was formatted for no javascript pages. Thats been a month ago or so so you could check again.
 
Hi,

I am clueless about scrapebox but I have a few qns for it. So currently, my main focus of using SB is to find expired domains from Big G for certain kws and in my country TLD.

1) as I heard SB have a proxy harvester, so do I still need to buy proxies to scrape for expired domains?

2) can I scrape expired domains through a specific directory or local news website?

3) can I harvest all local emails from specific local directory?

Thank you.
 
If you are a heavy user (not me) prepare some proxies (not only for the accounts but also) prevent bans from some sites.
 
Ok, loopline, sorry for long answering. Im talking about proxy harvester, i just harvesting from default resources and mine ones. Like 1 week ago from all proxies i got like 1000+ Google tested proxies, now is only 10-30. I always using 300 threads, but now i changed it to 100 and i still got up to 50 proxies, still with GSA proxies tester.

I got one more question, if i add such a long keyword list, how i can scrape result only per year? Because this this is enabled only in manual searching in google but i cant find it in scrapebox.
 
Hi,

I am clueless about scrapebox but I have a few qns for it. So currently, my main focus of using SB is to find expired domains from Big G for certain kws and in my country TLD.

1) as I heard SB have a proxy harvester, so do I still need to buy proxies to scrape for expired domains?

2) can I scrape expired domains through a specific directory or local news website?

3) can I harvest all local emails from specific local directory?

Thank you.

1) Well if you were going to use proxies for finding expired domains then you probably want private proxies. But it depends on how your going about doing it. If your using google as part of the process or some other engine then public proxies can be fine as long as you find ones that work with the engine your using. If your using proxies for some other part of the process then you would want private.

However the first question I would ask is do I need proxies at all? And the answer is most likely not. Ive seen a fairly wide variety of methods to get expired domains with Scrapebox, as it is the swiss army knife of seo after all, but for most methods you wouldn't even need proxies. Unless your using high connections on a given domain etc...

2) Sure, yes. You can use the link extractor addon or the grab links by crawling a site.

3) Probably. You can use the grab emails by crawling a site function, but if the mails are produced via javascript or require you to be logged into to see them then its not going to work.

Ok, loopline, sorry for long answering. Im talking about proxy harvester, i just harvesting from default resources and mine ones. Like 1 week ago from all proxies i got like 1000+ Google tested proxies, now is only 10-30. I always using 300 threads, but now i changed it to 100 and i still got up to 50 proxies, still with GSA proxies tester.

I got one more question, if i add such a long keyword list, how i can scrape result only per year? Because this this is enabled only in manual searching in google but i cant find it in scrapebox.


GSAs tester works different, its not better or worse, but its different. So it may pick up a couple that scrapebox misses and scrapebox might pick up a couple that gsa misses. However scrapebox is using its own proxy judges, you can specify what to use in GSA and the results may be off due to the judge. So ultimately regardless of what a tester says, what works? I mean if you load the proxies in from GSA do they work? Thats the real test.

As for the custom googles I cover how to do it in this video:

https://www.youtube.com/watch?v=72bC56R_4-M
 
Hi is anyone here who using harvester proxy from public proxy get a low google test proxies? most of them is failed because connection timed out? because since 3 days ago i got this problem. i have whitelist scrapebox on my VPS. can you share your result when you harvest public proxy? thanks

View attachment 76055
Same happens to me
 
Same happens to me

Turn your connections way down and your timeouts way up. Does that change it?

If so great, if not make sure you whitelist scrapebox in all security software.

If it just started happening like a switch flipped, what changed on that day/at that time? If you installed new software or an anti-virus udpated etc.. thats likely the culprit.
 
Can someone help with this please:

How can I scrape google and get as close as possible to the same results I'm seeing in a browser?

I'm ONLY scraping first 10 results per keyword yet I'm seeing some significant differences. I'm not logged into any account or anything when browsing too. Straight up clean search. I get same results in firefox or chrome, yet harvester pulls different. I've checked the harverster engine config for google. Definitely using google.com. I've tried it for just one keyword at a time too without using a proxy or anything and I still see real differences? I'm using the data to ultimately check my competition so it's pretty important for it to be accurate.

Can anyone recommend a different query string that might be a little more accurate?

I'm no scrapebox guru so maybe I'm missing something here? I know google gives different results from different datacenters but this should absolutely be the same when I'm requesting it via the same IP.

Thanks for your help.
 
Your transaction ID for this payment is: xxxxxxxxxxxxxxx.

You should edit that transaction ID out, its part of your license info.

Can someone help with this please:

How can I scrape google and get as close as possible to the same results I'm seeing in a browser?

I'm ONLY scraping first 10 results per keyword yet I'm seeing some significant differences. I'm not logged into any account or anything when browsing too. Straight up clean search. I get same results in firefox or chrome, yet harvester pulls different. I've checked the harverster engine config for google. Definitely using google.com. I've tried it for just one keyword at a time too without using a proxy or anything and I still see real differences? I'm using the data to ultimately check my competition so it's pretty important for it to be accurate.

Can anyone recommend a different query string that might be a little more accurate?

I'm no scrapebox guru so maybe I'm missing something here? I know google gives different results from different datacenters but this should absolutely be the same when I'm requesting it via the same IP.

Thanks for your help.

Setup scrapebox like your browser is if you want to see the same thing. Scrapebox is pulling 100 results per page and only showing you the first 10, however your probably looking at 10 results per page in a browser. 100 vs 10 results per page can show significant different results.

In the screen where you select your engines, right before you start harvesting, near the bottom of the list is a 10 results per page option. Try that.

Im guessing, but if you dont' use proxies, that should get you close, else you can edit that engine and put in your firefox/chrome user agent to get the exact same results.

Bearing in mind that while you say its important to see the same thing, scrapebox is giving you a clue, google displays all sorts of different sets of results. Your customers/users may be seeing what scrapebox sees and not what you see, so that would make your browser data pointless. Im not saying they are, but you could take 50 people and set them up with different browsers and ips and they could get 50 different sets of results, some varying wildly. Just sayin.
 
Whats the best way to have multiple instance instals on the same server?

Currently I installed it once then copied it over and over. It works but some times I get random little bugs. Is this the best way or should I instal scrapebox over and over again from scratch if that makes sense?
 
Hey my scrapebox is getting this error
"error hooking api "loading stringa
dumping first 32 bytes"
 
Thanks for the reply and info.

I already was using 10 results per page so no issue there. It looks like it's ALL about the user agent. Definitely different results based on that. Took me awhile but I understand it now and how to make scrapebox recognize what's "before" and "after" the link, etc...

Messed around with it for a bit and did get it working with one user agent, but then had more issues. Not really worth it I think. I was not aware how google would serve up that many different results based on the user agent. Learn something every day! Ultimately I think it will all come out in the wash. I may indeed over or underestimate a KWD based on the results I scrape vs. results of browser BUT I don't think it's going to be a too big of a difference as long it's all from "google.com".

Thank you for your input. Much appreciated!

You should edit that transaction ID out, its part of your license info.



Setup scrapebox like your browser is if you want to see the same thing. Scrapebox is pulling 100 results per page and only showing you the first 10, however your probably looking at 10 results per page in a browser. 100 vs 10 results per page can show significant different results.

In the screen where you select your engines, right before you start harvesting, near the bottom of the list is a 10 results per page option. Try that.

Im guessing, but if you dont' use proxies, that should get you close, else you can edit that engine and put in your firefox/chrome user agent to get the exact same results.

Bearing in mind that while you say its important to see the same thing, scrapebox is giving you a clue, google displays all sorts of different sets of results. Your customers/users may be seeing what scrapebox sees and not what you see, so that would make your browser data pointless. Im not saying they are, but you could take 50 people and set them up with different browsers and ips and they could get 50 different sets of results, some varying wildly. Just sayin.
 
Whats the best way to have multiple instance instals on the same server?

Currently I installed it once then copied it over and over. It works but some times I get random little bugs. Is this the best way or should I instal scrapebox over and over again from scratch if that makes sense?

Just copy the folders, or you can download a fresh copy. I have a video on this:
https://www.youtube.com/watch?v=aZzdE6ybu38

Hey my scrapebox is getting this error
"error hooking api "loading stringa
dumping first 32 bytes"

Contact support

http://www.scrapebox.com/contact-us

Thanks for the reply and info.

I already was using 10 results per page so no issue there. It looks like it's ALL about the user agent. Definitely different results based on that. Took me awhile but I understand it now and how to make scrapebox recognize what's "before" and "after" the link, etc...

Messed around with it for a bit and did get it working with one user agent, but then had more issues. Not really worth it I think. I was not aware how google would serve up that many different results based on the user agent. Learn something every day! Ultimately I think it will all come out in the wash. I may indeed over or underestimate a KWD based on the results I scrape vs. results of browser BUT I don't think it's going to be a too big of a difference as long it's all from "google.com".

Thank you for your input. Much appreciated!

Your welcome. Yes google started this a couple of years ago I think and its just more prevalent now. Either they did a lot of testing to find out its better or this is the giant test, lol. You don't have to rebuild the whole engine, just edit the existing google engine and change the user agent.

But yes you are correct, it will come out in the wash. If nothing else you may even want to build 2 or 3 different copies of the google engine to do research with, with different user agents and setups so you can see the broad spectrum.



~~~~~~~~~~~~~~~~~

Scrape Mails From Craigs List


 
Last edited by a moderator:
I am in need of some guru assistance if possible!

Currently running scrapes using Ypscraper and the entries begin to flow. However as the search works through keywords/locations, the scraped entries then dissapear? I get to the end of the search and there are no entries to export even though I watched many appear during the search.

Any ideas guys? Thanks.
 
Status
Not open for further replies.
Back
Top