- Jul 13, 2008
- 2,219
- 5,431
Hi
Is there a way to extract in bulk URL Links from local files (text,html ,etc..) using Scrapebox or any other tool ?
Thanks in advance
We dont, not for URL's. We have the option to extract emails from local files but not URL's. You could upload the files to your hosting, and then extract the data from them as they are now online. However we will take a look at such a feature for local files.
Hello
I'm facing a problem regarding the avg speed while scrapping. I am using backconnect rotating proxies, the 80 Threads Package. I'm using the 2 main proxies for scraping and the average speed it's only 18 urls per second. Anyway it takes around 3 hours to scrape 200000 urls with my internet connection of 65 Mbps. I already check the option "use custom harvester" and in connections settings I use the recommended maxium number of 20 threads in Harvester, with 60 seconds timeout.
I don't know what I'm doing wrong.
Someone help me find a solution to increase the avg scraping speed?
When using 2 proxies and 60 seconds delay, if you strike a lot of proxies that are timing out that can consume a lot of time. Also how many Google passed proxies in the pool will have a huge impact, if there's not many then the time to scrape results can really add it with ScrapeBox struggling to find any of your proxies that are working.
You could try lowering the timeout to 10 seconds, and also experiment between 10 and 20 connections at that timeout setting to test if this works better for your situation. I can be just a matter of tweaking, measuring the results and repeating to find the perfect amount of proxies, connections, keywords, timeouts etc for your PC, network and proxies.