Hello scrapebox team, thanks for awesome software!
I noticed a bug at page scanner plugin, it appeared 2 times in a row to me.
After I imported list of 1m urls to page scanner and run it, it go fine until the end, at the end some threads stopped at urls forever, lets say 5 of them, they didn't timed out, usually I click abort, so they all stop and I can export results.
But 2 days ago after scan finished and I click abort because of some dead threads, one remain after that and I couldn't export anything, I had to kill process and start again. Yesterday I tried again, it was scanning all day, but in the end there was 1 threat that I couldn't abort, so again, all day wasted.
Short version: sometimes it is impossible to abort page scanner plugin and export results.
That means that something has locked 1 or more of the threads. This can be security software such as anti-virus, malware checkers and firewalls. So you should whitelist scrapebox in all security software and then you can whitelist the entire scrapebox folder as well.
Further any program that accesses the internet can lock threads, things like skype, utorrent etc… So you can try closing down any unneeded programs. Then if its working you can turn programs back on 1 by 1 to find the culprit.
Further pc optimization software can lock threads so you can shut any such software down.
Take note that disabling security software (such as anti-virus, malware checkers and firewalls) often only stops new rules form forming, but allows existing rules to still fire. So you have to fully whitelist in the security software or uninstall the security software(as a test).
Further some security softwar requires you to whitelist in more then one place before it takes effect.
Also note that disabling a router firewall, does actually fully disable it.
Basically you have to sort out what is locking the threads, because scrapebox is forced to wait until all threads are released. On occasion it can be windows that does it, so you can try restarting your machine and/or lowering total connections.
Hi matt
For some reason i need to list Search result from the last page , and scrapebox scrape result seems random (not based the position in serp)
any trick to find out which result is in position let say 80 or 99? from scrape log maybe?
Well you could use detailed harvester as it keeps them in order, but custom harvester is all about mass and speed so it doesn't care about order, it just goes as quick as possible and fills results.
Or if you have the rank tracker plugin, that might be a convenient use of that plugin.
loopline!!
I am having problems "grabbing emails" from URLs that I harvested
1. I entered keyword
2. added my 50 proxies. 'use proxies' cehceked
3. i tried proxy results 1000, 44, 55. I still have errors with grabbing emails (see #6
4. I clicked Start Harvesting. tried with google / yahoo/bing. i still get errors (see #6
5. I click "grab e-mails from crawling sites ( harvested URLs)
6. Useragent: blank, also tried clicking random one
box checked for 'use harvested urls' deptth: 1.... delay=none. fixed 0, random between 0 and 0, proxy retries 0. I click START. .....then it says processing but it doesn't move?????????? it just stays stuck at processing
edit: i tried it again with a lower proxy result and it worked. Does this mean I need more proxies to get more results?
I replied your other threads, but you should use either no proxies or only private proxies. But really just try 2 connections and no proxies and it will probably hum along nicely.