Gscraper Taking Years To Scrape Targets!!!

Did he take a look?

No Mudbutt. I refused. I have no need of his VPS. For the record, I have a shared VPS in Arizona, a server in Chicago, and a blade server in Czechoslovakia. I also have enough bandwidth at home to run any SEO program you can name if I run it in a limited fashion - meaning I am not also streaming a NetFlicks at the same time.

I have licenses to any program that I use, that is not an issue. The reason that I the scraped 1000 key words I mentioned earlier is that I am using for a comparison to what hrefer will scrape while making a buy decision for Xrumer. The reason I kept the threads down with GScraper and used 500 proxies was to keep from burning the proxies up.

I use GScraper all the time. The difference is that I am aware of what has been going on. I also use ScrapeBox. The two compliment each other. However, the only time I use GScraper to post is if I do not care if the URL is passed on.

Around version update 1.1.7.0 I started keeping copies of the old versions of GScraper because there was an update that had a bunch of problems and I needed a way to roll back against future problems.

gsversions_zpsbec91953.png


The version that I have shown the code for was 1.2.0.5. The first version that I have a record of reverse compiling was 1.1.8.8, but I had reversed compiled two versions before that and just overwrote the files because I was simply tracking changes. Long before I reverse compiled the program, I had noticed the problems I mentioned here. This prompted a buy decision for Scrape Box.

I will also answer a question that others may have. I am obviously a programmer, why would I be into SEO. It is not for the money. I can sit in front of the TV doing nothing and I would not have to worry about anything. That income is secure, and it is judgement proof. I have spent more time on the shitter in prison than many of the posters here have been alive, and I damn near got caught up in Operation Sun Devil (You can look it up on Wiki). During my last bid, the government offered to release me from prison if I would go to work for them. I elected to do my 15 years. About 2006 - 7 one of my friends notified me that I was all over the web on mugshot sites. I contacted some sites and requested that they take the images down. They of course refused unless I paid them. I refused, and in turn began working against them on a political level and in other ways. During the process, the operators of those sites smeared me across every consumer complaint site on the web. I used their tactics against them, and created a few of my own. I had thought that my job was finished after the Google update in November when mugshots were demoted and I could rest. Apparently, the job is not done. Also, the smear sites like ROR and the like need to be addressed. To this end, I am working on a program to automate the process. When it will be done, I have no idea. I have my own testers. I also do reputation management/repair on a limited basis.

While performing the above, I noticed that many of the SEO programs were bunk. I have tested many. I have even reversed a few to see why. Many are scams. GScraper does a job, the job that you paid for it to do, but it does another job in the background. You can take it for what it is worth and use it to your advantage, or you can ignore it and it just might bite you. I am not against Gscraper. I simply want the problems fixed. If they are not, I will do what I will do. There is no need on a paid program to send every URL and every proxy to another country. Telemetric bug tracking is one thing, but this is well beyond that. One thing I will make absolutely clear, I have not seen, and do not believe that GScraper is used for any type of spying purpose. It is my belief that Gscraper collects the scraped information to provide the proxys and the auto approve lists that are used in the subscription service.

Now, Let us see what Hawke has to say.
 
I think it's insane that people use someone else's tool to scrape google when you can easily build a scraper yourself with a spare afternoon and some persistence.

To help you guys out, here:

- look up php curl
- open firefox and turn off javascript and cookies
- visit google and do a search
- grab the url
- write your php curl script to use private proxies and visit that url
- use simple html dom to grab the links

voila, you now have a scraper.

Why do that when Scrapebox is a measly $57 for a lifetime license?
 
php curl is good, ruby is good, python can do even better with less effort, only time and energy are matters.
 
Can not be sure, that may be a strict way to block out crack version, just don't know, but if you can not trust the supplier, you'd better stay away.
 
@keith88

I can explain why your scraper was taken so long time. First I should explain to you the scrape's principle, your thread just have said you have 30000 footprints, not include how many keywords, and assume there have 10000 keywords. And your "Maximum results per search" of GScraper is 500:
222.jpg

One footprint and one keyword need request Google 5 times, that's you know. And how many times your scrape TASK that need request Google now? Need 30000*10000*5=1500000000, 1.5 billion times. Fine, we all know GScraper is Multi-threading, you set the highest 1500 threads, Assuming your bandwidth is large enough (such as Dedicated100M), and every threads can work fine. So each thread need request Google 1000000 times. We assuming request and response from Google 1 time need 10s on average, so GScraper will need 1000000*10/60/60/24=116 days to finish your TASK. This is the most ideal state, not even think faster than this!

I suggest to you, streamline your Footprints and keywords, and set the appropriate thread number according to your network, that can allow GScraper to be higher efficiency.
 
Last edited:
@JustUs

Why do you think GScraper will send the clients URLs and proxies to our server? We estimated conservatively we have 3000 clients(much more than that), and have 1000 clients will online to use GScraper. The data amount of use GScraper to scrape and post count in GB every day, we assuming the amount of date is 1GB every day. That according to your logic, GScraper's Server will receive 1000G(1TB) data from our clients. What do you think of which server can withstand so large amount of data?

Of course, GScraper does have communication with our server, but only limited to the following:
1\ Login to GScraper, we will sent the use key to our server in order to verify user is or not legal. This process includes presented and the returned, the data is encrypted.
2\Get GScraper unlimited proxies, this will follow with whole scrape and post process. Because GScraper need to continue to obtain proxies from our server. The proxies from our server return to GScraper is encryption, then GScraper to decrypt the proxies. So, if you not enabled this function, GScraper will not communication with the server.
3\Click the such as "Start scrape", "Start post" button, will triggers the server authentication.

So All our communication is order to valid. We are welcome somebody to use Wireshark or Fiddler to sniff our date, you will discover this is small data, and the necessary authentication request. This test is so easy, Look at how much date are send to our server, and how much date to your local.

In addition, we use encryption and decryption just for authentication and obtain proxies, we don't encryption the clients links and proxies send to our server. I thinks no one SEO tool to do this. Because it is too unrealistic, we should use the hundreds of times valued server clusters, which can withstand so large amount of data. The most IMPORTANT, we provide the soft and provide the services, products by heart are our mission. Stolen customer data is dishonest, this program will no one to use finally, to Business Company, is completely self destruction. So the important is we not to do this , and ever!

GScraper depletion the server's bandwidth resources have occurred sometimes, but its means GScraper are working fine, it's uninterrupted to request Google and get data, post website and leave your link. That's our clients need, and GScraper can give. If you think GScraper bring the high traffic which let your confused. We apologize to you, we apologize for GScraper's high efficiency and high performance work.
 
BTW, our email support is always userful , if you have question ,welcome to email us. our email is [email protected].

Because of Time difference, we will reply you in Chinese work time AQAP.
 
I am waiting for a Final Answer to these accusations as they seem very serious.
 
@JustUs

Why do you think GScraper will send the clients URLs and proxies to our server? We estimated conservatively we have 3000 clients(much more than that), and have 1000 clients will online to use GScraper. The data amount of use GScraper to scrape and post count in GB every day, we assuming the amount of date is 1GB every day. That according to your logic, GScraper?s Server will receive 1000G(1TB) data from our clients. What do you think of which server can withstand so large amount of data?

Of course, GScraper does have communication with our server, but only limited to the following:
1\ Login to GScraper, we will sent the use key to our server in order to verify user is or not legal. This process includes presented and the returned, the data is encrypted.
2\Get GScraper unlimited proxies, this will follow with whole scrape and post process. Because GScraper need to continue to obtain proxies from our server. The proxies from our server return to GScraper is encryption, then GScraper to decrypt the proxies. So, if you not enabled this function, GScraper will not communication with the server.
3\Click the such as ?Start scrape?, ?Start post? button, will triggers the server authentication.

So All our communication is order to valid. We are welcome somebody to use Wireshark or Fiddler to sniff our date, you will discover this is small data, and the necessary authentication request. This test is so easy, Look at how much date are send to our server, and how much date to your local.

In addition, we use encryption and decryption just for authentication and obtain proxies, we don?t encryption the clients links and proxies send to our server. I thinks no one SEO tool to do this. Because it is too unrealistic, we should use the hundreds of times valued server clusters, which can withstand so large amount of data. The most IMPORTANT, we provide the soft and provide the services, products by heart are our mission. Stolen customer data is dishonest, this program will no one to use finally, to Business Company, is completely self destruction. So the important is we not to do this , and ever!

GScraper depletion the server?s bandwidth resources have occurred sometimes, but its means GScraper are working fine, it?s uninterrupted to request Google and get data, post website and leave your link. That?s our clients need, and GScraper can give. If you think GScraper bring the high traffic which let your confused. We apologize to you, we apologize for GScraper?s high efficiency and high performance work.

Hawke, check your PM as I will send a link of the source code; you will have to pardon the fact that it is in C# versus the native VB, that I have played with it, that the symbols are different from what you use, and that I have some sections disabled or completely removed. From there we can discuss the matter privately.

To address what is here publicly. With the subscription service you are sending the proxies both ways. The first is to deliver the proxy to the customer, and the second is to *notify* your server that the proxy has been used. Those proxies you were, and maybe still are storing in the registry under the key "HKCU/Software/Jitesi." In the older versions, you also stored non subscriptions in the same location, whereas now they are in there own user defined file. Each proxy is encrypted and sent to the GScraper server. The same also occurs with each URL. I had long ago determined that part of this is due to validation permission. Each of these is wrapped in a 512 byte encrypted wrapper. Those wrappers are sent to GScrapers servers. I cannot speak of why they are encrypted, but it may simply be that you distrust your government even less than I trust mine. However, for validation there is no reason, good or bad, to send a users non subscription proxies, keywords, or URL's to your servers.

As to your bandwidth, you maintain that you scrape your own proxies. This in turn raises doubts as to your claims of bandwidth limitations. There are other issues on this that I would raise privately.

Because I am working on a program, and because you will know for certain that I have your source, as a courtesy, if you will offer me the same guarantee of confidentiality that I am giving you, then I will send you the source of what I am working on so that you can be assured that the code is different and proceeds along a different methodology than Gscraper.
 
I said again, we have not sent customers proxy or url or keywords or anything to our Server, this is for sure. We just get the users authentication information(include user key, machine code and random chars), and this is only apply in authentication. If you dont agree it, I have nothing to say. And If anyone want get GScraper to test it which sniff the network when GScraper are working, we can provide the Software and key, we are welcome!


Hawke, check your PM as I will send a link of the source code; you will have to pardon the fact that it is in C# versus the native VB, that I have played with it, that the symbols are different from what you use, and that I have some sections disabled or completely removed. From there we can discuss the matter privately.

To address what is here publicly. With the subscription service you are sending the proxies both ways. The first is to deliver the proxy to the customer, and the second is to *notify* your server that the proxy has been used. Those proxies you were, and maybe still are storing in the registry under the key "HKCU/Software/Jitesi." In the older versions, you also stored non subscriptions in the same location, whereas now they are in there own user defined file. Each proxy is encrypted and sent to the GScraper server. The same also occurs with each URL. I had long ago determined that part of this is due to validation permission. Each of these is wrapped in a 512 byte encrypted wrapper. Those wrappers are sent to GScrapers servers. I cannot speak of why they are encrypted, but it may simply be that you distrust your government even less than I trust mine. However, for validation there is no reason, good or bad, to send a users non subscription proxies, keywords, or URL's to your servers.

As to your bandwidth, you maintain that you scrape your own proxies. This in turn raises doubts as to your claims of bandwidth limitations. There are other issues on this that I would raise privately.

Because I am working on a program, and because you will know for certain that I have your source, as a courtesy, if you will offer me the same guarantee of confidentiality that I am giving you, then I will send you the source of what I am working on so that you can be assured that the code is different and proceeds along a different methodology than Gscraper.
 
The code is written by me, we know we not do that like you say. If you have doubts about the code, we can explain to you.

Hawke, check your PM as I will send a link of the source code; you will have to pardon the fact that it is in C# versus the native VB, that I have played with it, that the symbols are different from what you use, and that I have some sections disabled or completely removed. From there we can discuss the matter privately.

To address what is here publicly. With the subscription service you are sending the proxies both ways. The first is to deliver the proxy to the customer, and the second is to *notify* your server that the proxy has been used. Those proxies you were, and maybe still are storing in the registry under the key "HKCU/Software/Jitesi." In the older versions, you also stored non subscriptions in the same location, whereas now they are in there own user defined file. Each proxy is encrypted and sent to the GScraper server. The same also occurs with each URL. I had long ago determined that part of this is due to validation permission. Each of these is wrapped in a 512 byte encrypted wrapper. Those wrappers are sent to GScrapers servers. I cannot speak of why they are encrypted, but it may simply be that you distrust your government even less than I trust mine. However, for validation there is no reason, good or bad, to send a users non subscription proxies, keywords, or URL's to your servers.

As to your bandwidth, you maintain that you scrape your own proxies. This in turn raises doubts as to your claims of bandwidth limitations. There are other issues on this that I would raise privately.

Because I am working on a program, and because you will know for certain that I have your source, as a courtesy, if you will offer me the same guarantee of confidentiality that I am giving you, then I will send you the source of what I am working on so that you can be assured that the code is different and proceeds along a different methodology than Gscraper.
 
The code is written by me, we know we not do that like you say. If you have doubts about the code, we can explain to you.

I am currently under the influence of a migraine and cannot go into the matter tonight - publicly or privately - however the download link has now been deleted along with the material that was on the server.
 
GScraper Team will welcome feedback any doubt or question about our software to us. First, we guarantee not steal any information of the customer absolutely. Then we can explain and solve the customers misunderstanding or problem.

Early rest and when you have good spirits, welcome email us or reply here.

I am currently under the influence of a migraine and cannot go into the matter tonight - publicly or privately - however the download link has now been deleted along with the material that was on the server.
 
Yea sometime i had faced same issues.. It just work too slow and take lot time to scrape.
 
Back
Top