I found a few bug:
1. Lately there was update for multi-thread harvesting and I think it still doesn't work properly. When I use proxy for harvest url with the multi-thread option checked, it will stop harvesting when the proxy fails. It should switch to another proxy instead of ending the task. When I turn off multi-thread it works properly.
Sometimes in multithread it also harvest only 100 / 200 / 300 etc. URLs and then stopped instead of harvesting to 1000 URLs. Maybe it's switching to another proxy which fails but I'm not sure. When I turn off multi-thread it harvest properly till 1000 results.
Yes it works, check it in a packet monitor. If any proxy is blocked or 404 ScrapeBox will retry with another proxy however many times you have specified in the Settings menu. Are you using free proxies? If so then the feature won't suddenly give them private proxy reliability and allow you to multi-thread on an unstable connection with 100% accuracy.
The feature most definitely helps with free proxies, i done a lot of split testing on it but it won't perform miracles unfortunately. As i've mentioned before, if accuracy of results and the number returned is important don't use the multi-threaded harvester with free proxies.
If you want a shed load of URL's in the shortest time possible, multi-thread it.
2. There is addon - Alive Check - we have option "Recheck Failed" when we use it, it will check all the URLs not only the one which failed.
It's only rechecking the failed ones for me, perhaps you have a really old version of that addon? Check to see if there's any updates for it.
3. Addon - Link Extractor - now we can choose if we want to check "internal links' but it will be nice if we can decide if we want to check "external links" (now it always checks external links). There are situations when we need only to check internal links. It will be fine if you can add such an option to switch off checking external links
So you mean you want to harvest internal links, besides doing site:domain.com and harvesting them or using the Sitemap ripper addon? We will see.
It is a bug I can always first see PR of URL and then look for Domain PR of same url and it should work....
and it will be better if we let Sweetfunny decide that it's a bug or not
Yes Run-on-Top is right that's by design, if a PR value is populated it will not check it again it will only check the URL's with the "---" not checked flag. If you are checking a 200k list and your proxies burn out 3/4 the way through, you most definitely don't want to get new proxies and check the lot from scratch you want to continue where you left off.
Ok.. why not a preview,
ScrapeBox Learning:
An unknown platform, in this case Joomla:
http://i40.tinypic.com/34hh11c.png
Enter Learning Mode and Teach SB:
http://i39.tinypic.com/rsslth.png
Now SB knows the form type:
http://i41.tinypic.com/j7ti0j.png
Another example, this one Posterous.com
http://i40.tinypic.com/15qyihx.png
Should open up some possibilities, i even just taught SB a Wordpress email contact form and it was able to populate and submit the same form on other sites.
