Permanently Closed Marketplace Sales Thread

Status
Not open for further replies.
I have doubt hope you can let me know whether it's possible do it through scrapebox

I have a list of domains with extensions and is it possible to check whether it's possible to register them ?

Sure, just use the domain availability checker under grab/check.


So you can try going to settings >> connections timeouts and other settings >> more harvester opttions - and turn the proxy retries all the way up. Because whats happening is many of your rotating proxies are no doubt blocked and when the retries are reached the keyword is skipped.

You could alternatilvey try the detailed harvester, which is built for accuracy. It will take longer but likely yield more results. But a lot of your proxies are blocked which is why its happening.

do you have a discount Op?

Yes its in the very first post. :)

http://www.scrapebox.com/bhw

$67
 
Hello,

Yesterday I've purchased ScrapeBox because it has a feature to scrape from the "Dogpile" search engine.

Now that I got Scrapebox I went to the "costume harvester configuration" window to check the Dogpile engine out, but when I click "Test engine" it says "No links could be retrieved"....

I've tried making it work by changing the engine's settings but with no success...

Please help me out with this, I need the dogpile engine...

Cheers
 
Hello,

Yesterday I've purchased ScrapeBox because it has a feature to scrape from the "Dogpile" search engine.

Now that I got Scrapebox I went to the "costume harvester configuration" window to check the Dogpile engine out, but when I click "Test engine" it says "No links could be retrieved"....

I've tried making it work by changing the engine's settings but with no success...

Please help me out with this, I need the dogpile engine...

Cheers


I've just updated the Dogpile engine, so if you restart ScrapeBox it will automatically download the latest engine definitions and it should work.
 
So you can try going to settings >> connections timeouts and other settings >> more harvester opttions - and turn the proxy retries all the way up. Because whats happening is many of your rotating proxies are no doubt blocked and when the retries are reached the keyword is skipped.

You could alternatilvey try the detailed harvester, which is built for accuracy. It will take longer but likely yield more results. But a lot of your proxies are blocked which is why its happening.
The proxy retries was on max
 
thanks for the solution much appreciated I got one more doubt

is there a way to compare two domain list and save the only difference ?
 
Hey, I've got a question about proxy usage...

is there a feature to wait between search engine queries for each proxy?
for example proxy1 has to wait 60 seconds between each search it does.

This is really impotent to me so my proxies wont get banned.

Cheers
 
You need Viagra, lol

Delay's default time is 0 (each search engine, in Harvest Engine Configuration);
Increase it/them. (?)
 
Last edited:
The proxy retries was on max

ok then your proxies are just blocked. I mean you have to bear in mind that many other users of the rotating proxies service are using proxies with google and getting them blocked. So Scrapebox is retrying and retrying and when it reaches the retry max it skips the keyword.

So maybe you should try deeperweb and google api, they are google powered but have their own ip bans and/or you can try bing, which is super lenient on bans.

thanks for the solution much appreciated I got one more doubt

is there a way to compare two domain list and save the only difference ?

Yes you can use the compare feature.

Import >> import and compare url list

or import and compare on the domain level.

So you can compare the exact url list or based of of domains.

Hey, I've got a question about proxy usage...

is there a feature to wait between search engine queries for each proxy?
for example proxy1 has to wait 60 seconds between each search it does.

This is really impotent to me so my proxies wont get banned.

Cheers

Yes, you can use the detailed harvester and use the delay option. That is a delay per keyword not per proxy. You can go into the settings >> harvester engine configuration - and then set a delay for each engine. This delay is per query.

Both those work in the detailed harvester and they do combine, so if you set a delay per query and a delay per keyword you will wind up with both stacking.

You can also go to settings >> connections timeouts and other settings >> more harvester options >> proxy change interval - and set this. This is how often you force a proxy change. So you could force a proxy change after every request if you want. Mind you if you use the custom harvester the delay option isn't there, but forcing when to chnage a proxy may achieve what you want in the end as well.
 
Yes, you can use the detailed harvester and use the delay option. That is a delay per keyword not per proxy. You can go into the settings >> harvester engine configuration - and then set a delay for each engine. This delay is per query.

Both those work in the detailed harvester and they do combine, so if you set a delay per query and a delay per keyword you will wind up with both stacking.

You can also go to settings >> connections timeouts and other settings >> more harvester options >> proxy change interval - and set this. This is how often you force a proxy change. So you could force a proxy change after every request if you want. Mind you if you use the custom harvester the delay option isn't there, but forcing when to chnage a proxy may achieve what you want in the end as well.



The proxy change interval is well... I can't call it reliable because I don't really know how many queries the proxy does in a set amount of time.

so really there's isn't such feature for fast scraping because the details harvester only uses 1 thread

Making a feature that creates search engine queries timers for each proxy could really save them from getting banned, no matter how many threads you use.

For example:

you have 10 fast private proxies and you really don't want them to get banned, you could make a safe bet and set that each proxy would search every 60 seconds.
That means 10 search engine queries every 60 seconds.
and if you're willing to risk it to have faster scraping you could lower the query delay for each proxy.
and of course you can get more proxies to make it scrape faster with the proxy delay.

This is just a safe bet feature that the proxies won't get banned, resulting in faster scraping in the long term.
could you guys impalement this feature?
I think this feature would greatly benefit scrapebox and its users.
 
I read several times still a little confused.
This will lower the speed of scraping significantly unnecessary.
 
I read several times still a little confused.
This will lower the speed of scraping significantly unnecessary.

it may lower the speed of scraping if you use it, but your proxies wont get banned in the middle of scraping, so the scraping will be fully finished with unbanned proxies.

Gscraper and gsa search engine ranker has that feature, and its GREAT.
 
The proxy change interval is well... I can't call it reliable because I don't really know how many queries the proxy does in a set amount of time.

so really there's isn't such feature for fast scraping because the details harvester only uses 1 thread

Making a feature that creates search engine queries timers for each proxy could really save them from getting banned, no matter how many threads you use.

For example:

you have 10 fast private proxies and you really don't want them to get banned, you could make a safe bet and set that each proxy would search every 60 seconds.
That means 10 search engine queries every 60 seconds.
and if you're willing to risk it to have faster scraping you could lower the query delay for each proxy.
and of course you can get more proxies to make it scrape faster with the proxy delay.

This is just a safe bet feature that the proxies won't get banned, resulting in faster scraping in the long term.
could you guys impalement this feature?
I think this feature would greatly benefit scrapebox and its users.

Its not exactly linear but more or less if you have 10 proxies and your going to set a delay of 60 then just use 1 connection and set a delay of 6. I mean thats going to get your ips banned any way you slice it because its too fast, but you get the point. If you reach the point where you have enough proxies you don't need a delay then just let it run endlessly at 1 connection.

Every engine in the custom harvester gets its own array for keywords completed and proxies being used. So if you then tack on timers per proxy and 1 person selects 20 engines then you have 20 engines running with 20 arrays tracking 20 sets of keywords.

Then you have them all tracking proxies too as they will auto refresh for each engine as needed and not refresh if not needed.

Then if you add timers per proxy per engine, your looking at a nightmare. Aside from the fact that you could get into all sorts of potential crashes its just going to suck up more memory and slow down harvest.

Basically why worry about setting 10 connections for 10 ips and a delay per ip when you can just 1 connection and an appropriate delay and your rotating thru proxies slow enough your not getting them blocked. You wind up with the same outcome without having to worry about a ton of potential headache in the custom harvester.

I have a 20+ minute video dedicated just to this.


Thats my 2 cents, I don't get to make the decision if its implemented or not, but they would just wind up doing a lot of work to achieve the same end as the detailed harvester already does and does well.
 
The main thing is that every searchengine has its own algo where sets up some limits/signals and when you jump over them you are going to be banned or flagged. There are many details which are crucial to mark your actions but there are still some ratios under which you can quite safely do your job and avoid bans for proxies. These ratios won't let you scrape with insane speed but you will be able to manage your goals but you'll have to be more accurate in choosing keywords that will allow you to minimize the number of queries sent, and thus save proxies and set a safe ratio without causing excessive elongation your harvest.

In simple words you have to be more careful and patient.

Of course, you can always also use more search engines at one time.
 
YT's tech is amazing!
Code:
https://www.youtube.com/watch?v=NW9Bx7IG8AM&t=4s
Pr641az
Pr641az

WEzj61X

However this rapping can't be "transcript-ed";
(I don't know what he's singing about.)
Code:
https://www.youtube.com/watch?v=UZOe90aQQ3s
5dUCG2g
 
Hello

I want to transfer my SB to VPS and have problems doing it - "We have not been able to submit your activation request".

Any ideas?
 
Can scrapebox scrape blog posts from wordpress based domains? Could be a good addition for article scrape if we could put a custom domain to scrape from.
 
Sweet i tried contacting support they are very unhelpful i cannot get my scrapebox activated, i've been a supporter of your software i even bought the doomed and unfinished scrapeboard. now i cant get my scrapebox reactivated.. please help. I've sent several email to support and they could not help. i always refer people to you and your software but its causing me a lot of anguish to get the software I bought reactivated. Especiially since i have bought all your software thus far and did not ask for a refund on scrapboard. I trust you will help.. thanks
 
Status
Not open for further replies.
Back
Top