Scrapebox Index Checker Error

Radarkitai

BANNED
Joined
Nov 2, 2017
Messages
738
Reaction score
232
Guys,

I am getting this error In index checker - It is working good with other search engine but it shows error in google.

How to fix this issue ? does anyone have faced I am newbie :"( Scrapebox support is taking so I thought I should ask here

XZFaAx.jpg
 
AS noted above, its ip bans. In late 2017 google got really strict on advanced operator queries, and stuff like

info:domain.com

which is what scrapebox uses - gets banned fast. you can build a custom test for the above format to test your proxies against the index checker:

 
Big G doesn't like your proxies? Did you G-test them?
No I didn test mate :(
AS noted above, its ip bans. In late 2017 google got really strict on advanced operator queries, and stuff like

info:domain.com

which is what scrapebox uses - gets banned fast. you can build a custom test for the above format to test your proxies against the index checker:


Ok Let me watch the video I hope it fixes the issue
 
Some resolvent:
1. check your proxy
2. check your sitemap
3. check your robot.txt file
4. check your public folder on your server
 
If you're running the latest version of Scrapebox, it's probably about proxies. Test them first for Google. Also, it's recommended to set your connections to 1 and maybe use a little delay, too.
 
Some resolvent:
1. check your proxy
2. check your sitemap
3. check your robot.txt file
4. check your public folder on your server
How does checking 2,3,4 help him scrape on SB?

Also OP: just use Bing, proxies are cheaper and you can get really good results. Once you have made a seed list with bing you want to do some internal/external link finding with that. You can scrape bing and get 400K links, then once you've scrapped internal/external you can times this by 20
 
Thank you <3

The problem was with Proxies.

Thread can be closed <3
 
AS noted above, its ip bans. In late 2017 google got really strict on advanced operator queries, and stuff like

info:domain.com

which is what scrapebox uses - gets banned fast. you can build a custom test for the above format to test your proxies against the index checker:

I'm having the same problem with Fresh proxies, tested them and they all work perfect.
Still, some of the urls showing ERROR on index check even though everything fine with them (check manually).
 
I'm having the same problem with Fresh proxies, tested them and they all work perfect.
Still, some of the urls showing ERROR on index check even though everything fine with them (check manually).
Its the exact query type thats banned. You can build a custom test for

info:domain.com style and test your proxies against that.

 
AS noted above, its ip bans. In late 2017 google got really strict on advanced operator queries, and stuff like

info:domain.com

which is what scrapebox uses - gets banned fast. you can build a custom test for the above format to test your proxies against the index checker:


hey,

why does sb uses info:domain.com ?
its not reliable.
I just made a quick test and site:https://www.xxx.com/buy-backlinks/ shows that the url is indexed but info:https://www.xxx.com/buy-backlinks/
does not give a result...

whats the best way to check if url is indexed? is there a php curl script, if so, how many queries per ip to avoid ban?
How many queries per id per hour/day ?
Is there a service paid or free with an API?

Thanks
 
hey,

why does sb uses info:domain.com ?
its not reliable.
I just made a quick test and site:https://www.xxx.com/buy-backlinks/ shows that the url is indexed but info:https://www.xxx.com/buy-backlinks/
does not give a result...

whats the best way to check if url is indexed? is there a php curl script, if so, how many queries per ip to avoid ban?
How many queries per id per hour/day ?
Is there a service paid or free with an API?

Thanks
info is the "higher" level result. Im calling it higher just so I can explain it. So this is totally made up but it helps to undestand it.

If you could think of google as Tiers.

Tier 0 is a url google knows about but its not worth indexing
Tier 1 is a url that shows up sometimes in some results or google engines, or if its directly searched but may not show up for site:
Tier 2 is a url is worthy of showing up for site consistently
Tier 3 is a url that shows up for info.

If it shows up for info at Tier 3, its of a high enough quality its going to show up everywhere.

So its not that info: is "not" reliable, its that info is the "most" reliable, which is why its used.

Sort of more or less.
 
info is the "higher" level result. Im calling it higher just so I can explain it. So this is totally made up but it helps to undestand it.

If you could think of google as Tiers.

Tier 0 is a url google knows about but its not worth indexing
Tier 1 is a url that shows up sometimes in some results or google engines, or if its directly searched but may not show up for site:
Tier 2 is a url is worthy of showing up for site consistently
Tier 3 is a url that shows up for info.

If it shows up for info at Tier 3, its of a high enough quality its going to show up everywhere.

So its not that info: is "not" reliable, its that info is the "most" reliable, which is why its used.

Sort of more or less.
This guy knows about SB more than SB support team including owner.
 
I'm starting to get the feeling scrapebox is a crappy tool. It may have worked in the past but right now, it doesn't work for what I've used it. The index checker always FAILS. I have 50 private proxies which work great with RANKERX and GSA SER but Scrapebox either freezes all together or gives false positives. I have 1000 urls to check and it tells me none of them are indexed when I know most of them are. Also, the article builder is bad. I need a tool that I can give it keywords and other specs like number of words and it returns rows. Good luck with scrapebox. Scrapebox is a major fail in my book.

Big G doesn't like your proxies? Did you G-test them?
 
I have 1000 urls to check and it tells me none of them are indexed when I know most of them are.

Can you post just 1 single URL here that's returning a false positive so we can see?

Also, the article builder is bad. I need a tool that I can give it keywords and other specs like number of words and it returns rows. Good luck with scrapebox.

The article scraper has never had this feature, so if you bought it (or you buy anything for that matter) and expected it to do things which were never advertised then your always going to have issues. It will certainly scrape articles, but it wont cut them off at a specific amount of words.
 
I'm starting to get the feeling scrapebox is a crappy tool. It may have worked in the past but right now, it doesn't work for what I've used it. The index checker always FAILS. I have 50 private proxies which work great with RANKERX and GSA SER but Scrapebox either freezes all together or gives false positives. I have 1000 urls to check and it tells me none of them are indexed when I know most of them are. Also, the article builder is bad. I need a tool that I can give it keywords and other specs like number of words and it returns rows. Good luck with scrapebox. Scrapebox is a major fail in my book.
Id say this is a lack of understanding about how google indexing things not that scrapebox is crappy.

Scrapebox reports to you exactly what google gives it, the question is do you understand how google works and are you giving scrapebox what it needs to get what you want? The answer is most likely no on at least 1 of those counts or more.

different googles index things differently. Your url may be indexed and on page 1 in google.com and on page 10 in google UK and not indexed at all in Google China.

Google redirects your requests based on the geographic location of your ips.

So if you load in proxies from China, Italy and Kenya and you index check the same url you could get it as being idexed or not indexed for any of those locations. Its not that scrapebox is crappy its that google is redirecting your request to the google that is associated with the geo of the ip your using. Then scrapebox accurately reports whatever google gives it.

It goes deeper then this, but you need to understand that there is no such thing as indexed or not indexed in "google". There is such a thing as indexed or not indexed in google.com and indexed or not indexed in google UK and indexed or not indexed in google Italy, ASSUMING all things are equal.

There is also such a thing as indexed in google italy for firefox and not indexed in google italy for chrome. Or indexed in google UK for 10 results per page and not indexed in google UK for 100 results per page.

User agent, geo of the ip, etc... all play a part. You could take 10 machines and search the same thing with USA IPs and get 10 different sets of results, sometimes missing some urls and urls ordered differently.

So indexed is subjective to the same thing, google returns all kinds of different results for all kinds of different things, plus if your urls are weak when it comes to links, then that just makes it even more variance. Higher authority domains will be "less" subject to this, but still subject to geo and relevane and user agent and 10 results vs 100 and more.

So really you need to decide on your target, like my customers are from the USA, and then index check with only USA ips.

Or you could just do what I do and say is index checking even worth doing in the first place or should I just put that time and effort into actually building as many quality links as I can?

But do as you want, but before you blame a tool for being crappy, you should examine if perhaps the knowledge you have of how google works is lacking or if the entire idea of index checking is even worth it?

Definitely there is a lot to understand about google and when something doesn't work as I expect I ask why and dig rather then just blame the tool. Because I know 1 very important thing to be completely true.

The Tool is Only as Good as the Master. So I always examine and endevour to make sure I am as good as possible so I can get the tool to give me what I want.

My 2 cents.

Also as a side note, its generally frowned upon to resurrect old threads like you did, BHW mods prefer you just start a new relevant thread or you could post in the scrapebox sales thread.
 
Back
Top