Permanently Closed Marketplace Sales Thread

Status
Not open for further replies.
Im using back connect proxies, but I pulled all 2200 ish footprints from GSA, added them on 24 hour google (which netted 200K results) and scraped google. 100 connections and also 25 connections. When it gets near the end the threads do decrease but nothing like your saying, it reaches the point where all keywords are active in a thread and decreases from there till its done. Its quite fast for me.

Ill keep testing, but I can't reproduce what your talking about, grant it Im working on lower connections, but as you say it should be worse. Although you are running 5000 connections right? I mean at a point you are going to have 5000 active queries and its going to reach the end of the list of keywords and the last 5000 are going to have to finish, which may take a bit, but I would assume not hours like you are experiencing.

That's right. I'm using 5000 threads (now experimenting with 7k and 8k). I guess it's a combination between using public proxies and large amounts of threads. My guess for this is: when the keywords get less than the max threads set some of those public proxies go bad and scrapebox is waiting for the timeout until it declares it bad and switches to another proxy. So from 5000 public proxies probably a lot go bad and scrapebox has to wait for each one some time until it can switch to the next one. And if it switches to another bad proxy it will still have to wait until it gets to the next one. And the more bad proxies + the more threads = the more the waiting. This has to be the cause for this.

I'm getting like 2k urls/s average until the point where the threads start to decrease. From there on until the end the average speed goes to around 50 urls/s. I really have to do something about this, but unfortunately this process is completely in scrapebox control. I cannot add any external script or anything really to stop the process before it starts decreasing the threads. I'm still thinking how can I get around this.

I really can't see how to solve this. It represents a real issue for public proxy users. Maybe there should be different modes of scraping in scrapebox when using public vs private proxies. Private proxies are a bit expensive for what I'm trying to accomplish and many users are still using public proxies for scraping. Can anything be done to improve public proxies scraping with massive thread number?
 
Last edited:
Hi

Sorry to interrupt the discussion, but any tips on improving the efficiency of indexing links with rapid indexer? I know how to index links, as I've done it for myself, but if I wanted to ensure that links are really gettings indexed what would I need to do without passing my links to external services? Thanks for any tips.

What do external link index service providers do?
How do they index your links?
What system do they use?
 
That's right. I'm using 5000 threads (now experimenting with 7k and 8k). I guess it's a combination between using public proxies and large amounts of threads. My guess for this is: when the keywords get less than the max threads set some of those public proxies go bad and scrapebox is waiting for the timeout until it declares it bad and switches to another proxy. So from 5000 public proxies probably a lot go bad and scrapebox has to wait for each one some time until it can switch to the next one. And if it switches to another bad proxy it will still have to wait until it gets to the next one. And the more bad proxies + the more threads = the more the waiting. This has to be the cause for this.

I'm getting like 2k urls/s average until the point where the threads start to decrease. From there on until the end the average speed goes to around 50 urls/s. I really have to do something about this, but unfortunately this process is completely in scrapebox control. I cannot add any external script or anything really to stop the process before it starts decreasing the threads. I'm still thinking how can I get around this.

I really can't see how to solve this. It represents a real issue for public proxy users. Maybe there should be different modes of scraping in scrapebox when using public vs private proxies. Private proxies are a bit expensive for what I'm trying to accomplish and many users are still using public proxies for scraping. Can anything be done to improve public proxies scraping with massive thread number?

There comes a threshold where too many threads is slower then less threads. All this passes thru windows messaging and its just not built to handle this. I mean you would be way faster to run 4 machines at 2,000 threads then 1 at 8,000. Even google uses lots of smaller machines.

Try a loop at 7,000 threads and then try the same exact loop at 200 threads. How much of an actual time differences is there and speed difference? If you want to mail me your exact footprints Ill do some apples to apples comparisons.

service (at} mattborden {dot) com

Hi

Sorry to interrupt the discussion, but any tips on improving the efficiency of indexing links with rapid indexer? I know how to index links, as I've done it for myself, but if I wanted to ensure that links are really gettings indexed what would I need to do without passing my links to external services? Thanks for any tips.

What do external link index service providers do?
How do they index your links?
What system do they use?

Well I won't discuss your last 3 questions, as its beyond the scope of this thread.

As for Scrapebox, just build links to them. Comments, Guestbook posts, Image comments, Trackbacks etc.. and they will get indexed. No need to pass to an external service.
 
Asking again, is there any way you guys could implement sorting the columns in the SB 2.0 rank tracker plugin? Currently there's no way to order the columns (for example, lowest rankings to high) and in the 1.0 version you just click the heading to sort.
 
So, my outbound link checker will freeze up. It'll almost finish then get stuck at "1" connection, even when stopping. It'll never fully stop and I'm unable to do anything with it.
 
So, my outbound link checker will freeze up. It'll almost finish then get stuck at "1" connection, even when stopping. It'll never fully stop and I'm unable to do anything with it.

Which version, the one for ScrapeBox V1.x or V2.x?
 
Submitted for license transfer few days ago with no confirmation or any reply when i emailed you guys.

Can you please look into it? Transfers take a long time these days too i noticed.
 
Scrapebox v2.0.0.31 Beta Released

  • Added minimize to Proxy Manager and Proxy Harvester
  • Added to save proxies to file during harvesting proxies
  • Added "Remove urls with more than xxx character" to Automator
  • Added "Text File Tool" to tools menu
  • Fixed a bug in linkchecker where anchor link adn anchor text was not saved to csv
  • Keyword area/remove containing/not containing can now have multiple words
  • Added to email grabber filter "Keep emails containing"
  • Fixed a bug in proxy manager where sorting and filtering by speed failed
  • Fixed a bug in classify proxy sources sorting function


Submitted for license transfer few days ago with no confirmation or any reply when i emailed you guys.

Can you please look into it? Transfers take a long time these days too i noticed.

If you didn't get any confirmation, then it's likely you entered an invalid email or a typo and can't receive our reply. Please ensure your details are right, the average time transfers take hasn't changed in 5 years.
 
I've finally found the issue. It's the "enable auto load proxies from file" not doing it's work.

May I ask EXACTLY how does the "enable auto load proxies from file" work and what EXACTLY it is supposed to do? Because I have no idea what it does, but it definitely does not work as intended. I've been struggling for months with this issue and in the end it turns out a basic feature doesn't work right.

I did the following test:

1. Selected "enable auto load proxies from file" as the only source for proxies.
2. Set the "Load after x minutes" to 5 minutes
3. Set the "Select auto load proxy file" to proxies.txt
4. Intentionally inserted bad proxies into proxies.txt
5. Set SB to 150 threads
6. Started the harvest
7. A bunch of errors started coming up and the avg urls showed some low amount like 5-6
8. I replaced the proxies in proxies.txt with 30 private proxies in the mean time while scraping
9. After 5 minutes has passed, the harvesting status changed to "Refreshing proxies"
10. The refreshing proxies completed, BUT THE ERRORS WERE STILL PRESENT AND THE SPEED NEVER CHANGED! WHICH MEANS THE PROXIES WERE NEVER UPDATED FROM THE TXT FILE AND APPARENTLY THAT DOESN'T EVEN WORK.

I've been banging my head for more than a month now with this and it turns out my proxies were never updated during the harvesting sessions.


EDIT: It seems like this has been fixed in 2.0.0.31. It looks like it updates the proxies from file on the fly now whenever the file has been altered even before the delay set. Nice job!
 
Last edited:
Suggestion for Scrapebox V2:

Maybe I missed it, but it seems that it is only possible to apply 1 mask/keyword at a time when filtering urls in the harvester for certain terms that you want to remove such as

porn
blogspot
wiki

etc.

It would be very useful to be able to apply a .txt file containing all words we want filtered out, as opposed to adding one at a time.

maybe this is already possible and I missed the functionality?

Thanks.
 
So, my outbound link checker will freeze up. It'll almost finish then get stuck at "1" connection, even when stopping. It'll never fully stop and I'm unable to do anything with it.

I replied to your pm with a bunch of info, check that.

Suggestion for Scrapebox V2:

Maybe I missed it, but it seems that it is only possible to apply 1 mask/keyword at a time when filtering urls in the harvester for certain terms that you want to remove such as

porn
blogspot
wiki

etc.

It would be very useful to be able to apply a .txt file containing all words we want filtered out, as opposed to adding one at a time.

maybe this is already possible and I missed the functionality?

Thanks.

Its under remove/filter >> remove urls containing entries from and then select the text file.
 
If you didn't get any confirmation, then it's likely you entered an invalid email or a typo and can't receive our reply. Please ensure your details are right, the average time transfers take hasn't changed in 5 years.
My details are correct as there was no problem on previous transfers. I have used another email address to contact you guys, and even submitted a ticket on your support page.

It's been over 12hrs now and i've not gotten any reply, could you or anyone on support send me a pm? I can't send you a pm though.
 
Hey there,

Is anyone getting incorrect results from the competition-finder addon with v2.0-beta? I entered a list of about 10K keywords to get their "exact" competition. Using 25-back-connect proxy ports, with a delay of 3 seconds.... .some of the results come back as "error -1" ... I wonder what that is, but most importantly, I manually checked some of the results that seemed weird and I've seen huge discrepancies between what SB reports and what I see by entering the keyword manually in my browser (with quotes) ..

Everytime I use the comp-finder I only go as high as 10 connections. Some of the errors are 503 but I understand what those are ...

Any ideas about those incorrect results? Should I lower the number of connections using 25 back-connect proxies that rotate every 10 mins? Should I get more ports? Is there a setting that will provide more accurate results?

Sorry too many questions.. but u get the idea..
 
Last edited:
i'm newbie in link building
As i researched on BWH, everyone says that SB and GSA is best combination for link building
i want to buy both... but before purchasing them can someone answer my questions? please

1. so, as far as i know scrapebox can scrape urls and are these urls good to sumbit these urls by GSA ?
2. what does scrapebox have unique ? what are its best features ?
3. as i see scrapebox 2 is already released and can i buy this second version for only 57 USD ?
4. do i need proxy for running Scrapebox ? and as i know scrapebox scrapes proxies too, so can i use these scraped proxies for running GSA, are they reliable ?
 
When using scrapebox premium article scraper to post to Blogger blog, error is returned when trying to post single article or batch. I've made sure to use correct login and password, etc. Is there something special about the formatting of the blog URL that I should? How can I resolve the error? Thx
 
Status
Not open for further replies.
Back
Top