Permanently Closed Marketplace Sales Thread

Status
Not open for further replies.
Hi
Is there a way to extract in bulk URL Links from local files (text,html ,etc..) using Scrapebox or any other tool ?
Thanks in advance

We dont, not for URL's. We have the option to extract emails from local files but not URL's. You could upload the files to your hosting, and then extract the data from them as they are now online. However we will take a look at such a feature for local files.

Hello
I'm facing a problem regarding the avg speed while scrapping. I am using backconnect rotating proxies, the 80 Threads Package. I'm using the 2 main proxies for scraping and the average speed it's only 18 urls per second. Anyway it takes around 3 hours to scrape 200000 urls with my internet connection of 65 Mbps. I already check the option "use custom harvester" and in connections settings I use the recommended maxium number of 20 threads in Harvester, with 60 seconds timeout.
I don't know what I'm doing wrong.
Someone help me find a solution to increase the avg scraping speed?

When using 2 proxies and 60 seconds delay, if you strike a lot of proxies that are timing out that can consume a lot of time. Also how many Google passed proxies in the pool will have a huge impact, if there's not many then the time to scrape results can really add it with ScrapeBox struggling to find any of your proxies that are working.

You could try lowering the timeout to 10 seconds, and also experiment between 10 and 20 connections at that timeout setting to test if this works better for your situation. I can be just a matter of tweaking, measuring the results and repeating to find the perfect amount of proxies, connections, keywords, timeouts etc for your PC, network and proxies.
 
Can ScrapeBox scrape more emails than D7? I’m currently using D7 to scrape.
 
Hi,

Can I use Scrapebox with Mac OS Big Sur? I bought the license yesterday and upgraded to Big Sur today and now I can´t open Scrapebox anymore. It just closes after few seconds.

Thanks
 
how much the Email premium plugin costs??
thanks

Sorry for the delay, i didnt have a notification from BHW of your post.

ScrapeBox does come with an included basic email scraper, you can view details of this here: http://www.scrapebox.com/email-scraper

It's available on the main GUI of ScrapeBox under the "Grab/Check" button.

There also is a Premium Plugin with additional features as can be seen here which is $47 for this plugin on top of the cost of ScrapeBox: http://www.scrapebox.com/email-scraper-plugin

The basic email scraper only collects emails, with the option of saving the URL it was scraped from along with the email. The premium plugin also has the ability to train it to scrape other data as well as many other features which can be seen in the video and our website.

Can ScrapeBox scrape more emails than D7? I’m currently using D7 to scrape.

I'm not really familiar with D7 so i cannot comment on which would scrape more emails. However i notice it's about $100/month for 100 daily searches and 30,000 daily leads. The ScrapeBox premium plugin is only $47 for a lifetime license with no limits on the amount of use so ScrapeBox is significantly cheaper.

Hi,

Can I use Scrapebox with Mac OS Big Sur? I bought the license yesterday and upgraded to Big Sur today and now I can´t open Scrapebox anymore. It just closes after few seconds.

Thanks

At the moment ScrapeBox cannot run on Big Sur due to issues between the socket code, HTTPS connections and using proxies multi-threaded with HTTPS. We also rely on the code compiler and some third party components to bring in compatibility too. But we are working on a solution.
 
Do you have any plans on adding plugins for other domain auctions than godaddy?
 
Does the harvester google in the scrapebox work for you? Checking the cache in google through scrapebox with the same proxy works ok. Google harvesting with the same proxy via hrefref (from Xrumer) works normally. Only in scrapebox does google harvesting not work at all recently.
 
Do you have any plans on adding plugins for other domain auctions than godaddy?

We dont have any immediate plans for this, unfortunately with large changes to the latest Mac Big Sur release it's going to take quite a lot of development time getting this sorted which doesn't have any benefit to 95% of users. As good as it would be just adding new features, unfortunately there's a stack of back end compatibility and housekeeping for the Mac which has to be done.

Does the harvester google in the scrapebox work for you? Checking the cache in google through scrapebox with the same proxy works ok. Google harvesting with the same proxy via hrefref (from Xrumer) works normally. Only in scrapebox does google harvesting not work at all recently.

Have you tried updating to the latest version of ScrapeBox, plus downloaded the latest search engine definition file?

If not then please go to Settings >> Harvester Engines Configuration >> Import >> Download Default Engines From Server.

This will update the engines file to ensure you have the latest version and all are restored to default. The latest Google engine definition is working. You are also free to update or change any engines if they are not working how you like. Also if you go to Help >> Show error log >> Harvester.log you can see what errors are being returned when an engine doesnt work.
 
how to find outbound links on any doomain. In link extractor i need to put all urls of domain
 
how to find outbound links on any doomain. In link extractor i need to put all urls of domain

You could do this in 2 steps, first you could scrape all internal URL's of a domain. If they have a sitemap you can use the SiteMap Scraper addon which is the fastest and simplest way to get all the sites URL's. Otherwise you can use the Grab/Check > Grab Links by Crawling a Site option on the harvester to crawl the site for all it's internal URL's.

Then load all the internal URL's in to the Link Extractor to scrape all the outbound links.
 
I have the Article Scraper/Spinner premium addon. Is there any chance of you adding the ability to import from .csv file to the spinner in the future?
 
I don't seem to be able to update any of the Harvest engine settings, I have to save as a new engine instead. Testing engine doesn't seem to enable me to update.
 
vyPB3iz.png

Sweetfunny

Can I make a pause button? PLs )
How can I update proxies with new ones if parsing takes a long time? During this time, proxies stop working and the parsing speed drops significantly ?
 
@Sweetfunny
In the field of keywords -> Is it possible to implement data substitution macros?
macros

{az:a:z} - substitution of all characters from a to z (a, b, c, ..., x, z)
{az:aaa:zzz} - substitution of all characters from aaa to zzz (aaa, aab, aac, ..., zzx, zzz)
{az:a:zz} - substitution of all characters from a to zz (a, b, c, ... aa, ab, ..., zx, zz)
{az:00:99} - substitution of all numbers from 00 to 99(00, 01, 02, ..., 98, 99)
{az:а:яяя} - substitution of all Cyrillic characters from а to яяя (а, ..., аа, аб, ..., яяю, яяя)
https://en.a-parser.com/wiki/query-format/
 
I have the Article Scraper/Spinner premium addon. Is there any chance of you adding the ability to import from .csv file to the spinner in the future?

Hello, we have added it to the feature suggestion list.

I don't seem to be able to update any of the Harvest engine settings, I have to save as a new engine instead. Testing engine doesn't seem to enable me to update.

From the latest version the default engines are not editable, this way we can keep them updated without wiping your customizations every time. Previously if Google made a change and we fixed it, we would be answering hundreds of support emails for months telling people to update the engines file.

So now if you want to make changes to an engine, you can either make the changes to an existing engines then save it as a new engine or create a new engine from scratch.

View attachment 152785
Sweetfunny

Can I make a pause button? PLs )
How can I update proxies with new ones if parsing takes a long time? During this time, proxies stop working and the parsing speed drops significantly ?

The "Proxies" button in your screenshot has the option to auto load proxies from a file while harvesting and you can also set the harvester to load the proxies every X minutes or when it runs out of working proxies.

@Sweetfunny
In the field of keywords -> Is it possible to implement data substitution macros?
macros

{az:a:z} - substitution of all characters from a to z (a, b, c, ..., x, z)
{az:aaa:zzz} - substitution of all characters from aaa to zzz (aaa, aab, aac, ..., zzx, zzz)
{az:a:zz} - substitution of all characters from a to zz (a, b, c, ... aa, ab, ..., zx, zz)
{az:00:99} - substitution of all numbers from 00 to 99(00, 01, 02, ..., 98, 99)
{az:а:яяя} - substitution of all Cyrillic characters from а to яяя (а, ..., аа, аб, ..., яяю, яяя)
https://en.a-parser.com/wiki/query-format/

No sorry, there's no data substitution macro option in ScrapeBox sorry.

@Sweetfunny
Please make it possible to save several companies at once for parsing )
In reality, it will be very convenient.

ScrapeBox has a lot of features, which one did you mean? When you say several companies, did you mean the Yellow Pages Scraper?
 
I am trying to find expired .es domains to register but all of them get marked as "No, domains seems to be taken", even they are not. Also, the premium plugin is not able to check .es domains.

Is this a Scrapebox issue or I am doing something wrong?
 
Hey, I'm trying to post a batch of articles to a WP site through the Article Scraper plugin. But after a few articles i get error 415. Any idea what causes it?

Thanks
 
Status
Not open for further replies.
Back
Top