Permanently Closed Marketplace Sales Thread

Status
Not open for further replies.
I was using allintitle:keyword from Google Competition Finder addon.

I was using 50 dedicated proxy, 5 connection and a list of 17,000 keywords. But the result is showing 0 for all of them.

I have checked manually few of them and results are not 0.

How can I fix this issue or use "allintitle" footprint?
Try a sample test with no proxies, does it work?

It could well be and likely is a proxy issue.

also make sure to choose broad and not exact. Exact wraps the entire query in quotes and that will skew it in this case.
 
Try a sample test with no proxies, does it work?

It could well be and likely is a proxy issue.

also make sure to choose broad and not exact. Exact wraps the entire query in quotes and that will skew it in this case.
Yeah, I was using exact, it is working fine when I choose broad.

Thanks man.
 
I want to get all indexed pages of a website. How would I go about doing that?
I entered site:<domain> in the keyword harvester then clicked start harvesting in the url harvester and I'm not getting any results...what am I doing wrong?
 
Yeah, I was using exact, it is working fine when I choose broad.

Thanks man.

glad its working :)

I want to get all indexed pages of a website. How would I go about doing that?
I entered site:<domain> in the keyword harvester then clicked start harvesting in the url harvester and I'm not getting any results...what am I doing wrong?

Yeah, site:domain.com is probably the best option. If your getting no results, its probably an ip block. Go to help >> show error log >> harvester - what are the errors? 503 or 302 for google is an ip ban.
 
There seems to be some issue with Link Extractor and Windows 10. On this same exact PC (no hardware changes), only 1 day ago on Windows 7, I was running 200 threads at only ~50-60% CPU max.

Long story short, I had to upgrade to Win10 (using LTSB 2016, 1607). Now the same 200 threads runs at 99-100% CPU after only a few minutes of link extracting.

Edit: I am running some more tests now. Trying only 100 threads, also trying some other stuff... Will report back.
 
Last edited:
I thought it was going to end up being GSA Proxy Scraper using as Internal Server... so I exported some proxies, re-tested in SB, and tested 300 threads without GSA Proxy Scraper open at all. Within 15 minutes, the CPU is at 100%. What could cause such high CPU usage? I ran this same machine on Windows 7 x64 for weeks at 200-300 threads at the highest I ever saw was 50-60%. What happens after ~15 minutes that causes the CPU usage to skyrocket?
 
In Process Explorer I'm comparing the threads between the Windows 10 machine and my 3rd SB license on Windows 7. Granted the Win7 machine is 6core/12thread, and the Win10 is 6core/6thread.

On the Win10 machine, I see certain "threads" (we're talking about threads that inside of ScrapeBox Link Extractor, looking at the details of them) that aren't canceled out properly. They hang at High CPU for way way too long. Usually it's 1-2 "threads" (NOT CPU threads, internal SB threads) that hang at 15-20% CPU. On the Win7 machine, I'm running 1000 threads right now at only 30-40% CPU. Granted, the Win7 CPU has more CPU-threads, but still, when analyzing the threads inside of ScrapeBox, I see it is much, much more efficient at ending certain threads, and evenly distributing the CPU power across the threads .

For example, the top 5 threads on Win7 are all using 3-5%, followed by hundreds with only .02-1% CPU. Opposed to Win10, where the top 2 threads may be using 20% each, and they never end. I've seen my Win10 CPU be at 100% usage for 10+ minutes straight (being caused by a few threads that refuse to close). On the exact same hardware, on Win7, at the exact same threads, using the exact same proxies, the highest I ever saw was 50-60%.

I would've loved to keep Win7 on this machine, but the motherboard is an H370, and there are no Intel USB drivers because they've started using their own on-board PCH instead of ASMedia. So basically, I had to use an external GPU and external PCIe controller for USB ports. Even worse, the motherboard only has 1 PCIe slot. Using Win10 was clearly a better option... or so I thought? After seeing this I've really tempted to re-install Win7 and live with the consequences. (Mainly, every time troubleshooting, I have to remove GPU, add PCIe USB controller... huge pain, but better than running at 100% CPU with only 100 threads).

xoj9B8U
 
Last edited:
There seems to be some issue with Link Extractor and Windows 10. On this same exact PC (no hardware changes), only 1 day ago on Windows 7, I was running 200 threads at only ~50-60% CPU max.

Long story short, I had to upgrade to Win10 (using LTSB 2016, 1607). Now the same 200 threads runs at 99-100% CPU after only a few minutes of link extracting.

Edit: I am running some more tests now. Trying only 100 threads, also trying some other stuff... Will report back.

There's a very large Windows 10 userbase for ScrapeBox, more than Windows 7 and ScrapeBox is also coded and developed on Windows 10 and there's no other reports of this happening. So it could be a security software scanning the threads and causing the slowdown or something to do with your list for example if you are scanning an auto approve list with thousands of comments the HTML can be huge.

If you want to, you can send the list to support and we can run it on any OS and compare.
 
Hi!
I'm doing Link Checking for 2000k links with 40 shared proxies and 200 threads. Lately link checker after checking all links is stuck with some threads, usually +- 100 and even when i stop it, checker is still stuck with those threads and nothing happening. To stop it i need to stop scrapebox from task list and this means I'm losing all checked links and I need to do job from 0.
This problem started around 2 months ago, and if i'm lower threads to 50, link checker will check my 2mil. links for ages. and I noticed that link checker always stuck with some active treads at the end always but if it stuck with lower count of threads then I can stop it and everything is fine.
Why this happening and is there some other way to stop Link checker when it is stuck with more that 100 treads without ending scrapebox program itself.

Thanks!
 
There's a very large Windows 10 userbase for ScrapeBox, more than Windows 7 and ScrapeBox is also coded and developed on Windows 10 and there's no other reports of this happening. So it could be a security software scanning the threads and causing the slowdown or something to do with your list for example if you are scanning an auto approve list with thousands of comments the HTML can be huge.

If you want to, you can send the list to support and we can run it on any OS and compare.

I'll just use the Win10 machine for GSA stuff only as it doesn't require much CPU usage (seems likely that the problem is that the Win10 machine doesn't have hyperthreading and SB now takes full advantage of all threads, which I love). Tweaking my router TCP timeouts, it's about to get pretty wild at my house. (Unlimited bandwidth Internet)

Few feature requests though:
Why is there no option to, for example, "remove indexed -> yes or no" the "yes" would remove all URLs marked as indexed "Yes," whereas the No would remove all "No" URLs. This seems like such a simple feature but would save countless amounts of minutes in the long-run instead of having to manually select the URLs that are indexed already. Just add it into the "Remove more" dialog "Remove Indexed URLs" "Remove non-Indexed URLs". I've been waiting for this option since like 2012.

On that note, why not a "Export Non-Indexed URLs to Clipboard" and "Export Indexed URLs to Clipboard". Seems this would be a very useful feature.

Allow us to use custom hotkeys. The current hotkeys in SB are to only basic functions of the application. Some instances for example I only use Link Extractor, but I cannot make a hotkey for this addon. Again, those seconds saved from having to click it every time add up to minutes over time.

Lastly, a way to remove the dialog from clearing stuff, like clearing proxies "Are you sure?" I'd like to be able to turn that off.

Thanks a lot!
 
Last edited:
Hi!
I'm doing Link Checking for 2000k links with 40 shared proxies and 200 threads. Lately link checker after checking all links is stuck with some threads, usually +- 100 and even when i stop it, checker is still stuck with those threads and nothing happening. To stop it i need to stop scrapebox from task list and this means I'm losing all checked links and I need to do job from 0.
This problem started around 2 months ago, and if i'm lower threads to 50, link checker will check my 2mil. links for ages. and I noticed that link checker always stuck with some active treads at the end always but if it stuck with lower count of threads then I can stop it and everything is fine.
Why this happening and is there some other way to stop Link checker when it is stuck with more that 100 treads without ending scrapebox program itself.

Thanks!

Its locked threads. If you go to settings >> connection timeouts and other settings >> other - there is a link checker min threads. You can set that to like 100 or whatever and when it reaches that number the link checker will try to auto kill the remaining threads.

But the cause can be security software such as anti-virus, malware checkers and firewalls. So you should whitelist scrapebox in all security software and then you can whitelist the entire scrapebox folder as well.



Further any program that accesses the internet can lock threads, things like skype, utorrent etc… So you can try closing down any unneeded programs. Then if its working you can turn programs back on 1 by 1 to find the culprit.



Further pc optimization software can lock threads so you can shut any such software down.



Take note that disabling security software (such as anti-virus, malware checkers and firewalls) often only stops new rules form forming, but allows existing rules to still fire. So you have to fully whitelist in the security software or uninstall the security software(as a test).



Further some security softwar requires you to whitelist in more then one place before it takes effect.



Also note that disabling a router firewall, does actually fully disable it.





Basically you have to sort out what is locking the threads, because scrapebox is forced to wait until all threads are released. On occasion it can be windows that does it, so you can try restarting your machine and/or lowering total connections.




I'll just use the Win10 machine for GSA stuff only as it doesn't require much CPU usage (seems likely that the problem is that the Win10 machine doesn't have hyperthreading and SB now takes full advantage of all threads, which I love). Tweaking my router TCP timeouts, it's about to get pretty wild at my house. (Unlimited bandwidth Internet)

Few feature requests though:
Why is there no option to, for example, "remove indexed -> yes or no" the "yes" would remove all URLs marked as indexed "Yes," whereas the No would remove all "No" URLs. This seems like such a simple feature but would save countless amounts of minutes in the long-run instead of having to manually select the URLs that are indexed already. Just add it into the "Remove more" dialog "Remove Indexed URLs" "Remove non-Indexed URLs". I've been waiting for this option since like 2012.

On that note, why not a "Export Non-Indexed URLs to Clipboard" and "Export Indexed URLs to Clipboard". Seems this would be a very useful feature.

Allow us to use custom hotkeys. The current hotkeys in SB are to only basic functions of the application. Some instances for example I only use Link Extractor, but I cannot make a hotkey for this addon. Again, those seconds saved from having to click it every time add up to minutes over time.

Lastly, a way to remove the dialog from clearing stuff, like clearing proxies "Are you sure?" I'd like to be able to turn that off.

Thanks a lot!

On the indexed you can just go to import/export metrics >> export as excel. This will export the grid and you can sort them. It will say PR at the top of the column but its indexed, its just exporting the grid.
 
Its locked threads. If you go to settings >> connection timeouts and other settings >> other - there is a link checker min threads. You can set that to like 100 or whatever and when it reaches that number the link checker will try to auto kill the remaining threads.

But the cause can be security software such as anti-virus, malware checkers and firewalls. So you should whitelist scrapebox in all security software and then you can whitelist the entire scrapebox folder as well.



Further any program that accesses the internet can lock threads, things like skype, utorrent etc… So you can try closing down any unneeded programs. Then if its working you can turn programs back on 1 by 1 to find the culprit.



Further pc optimization software can lock threads so you can shut any such software down.



Take note that disabling security software (such as anti-virus, malware checkers and firewalls) often only stops new rules form forming, but allows existing rules to still fire. So you have to fully whitelist in the security software or uninstall the security software(as a test).



Further some security softwar requires you to whitelist in more then one place before it takes effect.



Also note that disabling a router firewall, does actually fully disable it.





Basically you have to sort out what is locking the threads, because scrapebox is forced to wait until all threads are released. On occasion it can be windows that does it, so you can try restarting your machine and/or lowering total connections.






On the indexed you can just go to import/export metrics >> export as excel. This will export the grid and you can sort them. It will say PR at the top of the column but its indexed, its just exporting the grid.

Exporting them would take even longer than sorting and highlighting / copying them to a notepad file........................................ (infinite amounts of periods here). The function I propose would be instantaneous, without even having to sort at all. It would save A LOT of time for everyone that uses index checking on Scrapebox.
 
Can we scrape searchblogspot.com ?
I tried customizing the search, but with no success.
 
Can we scrape searchblogspot.com ?
I tried customizing the search, but with no success.


I had a quick look, and it seems that it requires javascript and scrapebox uses raw sockets and threads - these do not support javascript.

That said, you could just search google for

site:blogspot.com keyword

and probably get the same results.
 
I had a quick look, and it seems that it requires javascript and scrapebox uses raw sockets and threads - these do not support javascript.

That said, you could just search google for

site:blogspot.com keyword

and probably get the same results.
The searchblogspot.com would yield only blogspot results.

But yeah, google it is.
 
when I want to launch article scraper addon, I get an error message, saying this plugin is singed by gunter kramer.... and nothing happens when I hit ok.

is this plugin not available anymore?
 
The searchblogspot.com would yield only blogspot results.

But yeah, google it is.

So does
site:blogspot.com car

It only returns blogspot blogs. thats what the site: operator does.

so just replace "car" with your keyword or put the site: in the footprints box at the top.

Unless Im just totally missing the point of what your trying to do, but try it in a browser and see if you get what you want.

when I want to launch article scraper addon, I get an error message, saying this plugin is singed by gunter kramer.... and nothing happens when I hit ok.

is this plugin not available anymore?

Make sure your scrapebox and article scraper addon are both up to date using the latest versions. If it doesn't work try uninstalling the addon and reinstalling, but you will need the latest version of scrapebox as well for it to work.
 
Is there a way to scrape websites those contain a specific url in html?
 
So does
site:blogspot.com car

It only returns blogspot blogs. thats what the site: operator does.

so just replace "car" with your keyword or put the site: in the footprints box at the top.

Unless Im just totally missing the point of what your trying to do, but try it in a browser and see if you get what you want.



Make sure your scrapebox and article scraper addon are both up to date using the latest versions. If it doesn't work try uninstalling the addon and reinstalling, but you will need the latest version of scrapebox as well for it to work.

thanks - upgrading to latest version solved the problem
 
@Sweetfunny
Support gets really annoying. I am trying to transfer my license. Here is the first mail support sent me regarding transfer date.

Records show you have utilized your monthly free transfer.

ScrapeBox is a single PC license and can only be transferred once per calendar month for free to facilitate obtaining a new PC, or system reformat.

Your next free transfer is available 2nd July 2018.


After I tried to transfer today, I got the second email saying.

Records show you have utilized your monthly free transfer.

ScrapeBox is a single PC license and can only be transferred once per calendar month for free to facilitate obtaining a new PC, or system reformat.

Your next free transfer is available 3rd July 2018.


Wtf is going on? You say license can only be transferred once per month. But we on 2nd of July and you don't allow to transfer anyway.
 
Status
Not open for further replies.
Back
Top