Permanently Closed Marketplace Sales Thread

Status
Not open for further replies.
Hi loopline
would i be able to get data from Yellowpages with Custom Data Grabber or any other directory site?

Yes in theory, I haven't tested it on yellow pages, but as long as the data can be matched with regex or even if its just consistent, then you don't need regex. Also of course it can't be in javascript etc... as the custom data grabber uses sockets, which don't support javascript

So probably, yes.

Thank you for the help. I understand it now. I normally check the raw domain lists or www as I go through different metrics. I tried it after the update and those two are not working. I will add the http:// part myself now as I understand the logic.

But could you please inform me one small feature if it is available in scrapebox and if not, could you please add in a future update. It would be great if you could add a button to prefix http:// in a whole list and another button to prefix www. in a whole list in the harvester. The reason I am saying about 2 buttons is because both of them have different use. In case of expired domains for example, the www and non-www has different metrics for Moz etc. So, at times both need to be checked. These three would work for checking root, subdomain and page metrics as those are different in many cases. Just a thought. If not, I would do that manually. :)

As noted by soft touch you can already prefix things and really you can add any prefix you want and just do a find replace to replace it with whatever you want.
 
Now It Will only do, slow commenting!

ScrapeBox still has the fast poster, it's selectable on the comment poster. http://www.scrapebox.com/learning-poster-v2

You can harvest and post to the following:

  • 4image
  • Advanced Guestbook
  • AkoBook
  • Ard Guestbook
  • Aska BBS
  • ASP Blog
  • Basti Guestbook
  • BeepWorld
  • Bella Guestbook
  • Blogengine
  • Burning Book
  • Chinese Blog
  • cms2day
  • CoderWorld
  • Coppermine
  • DedeIms
  • DRB Guestbook
  • e107 Forum
  • EasyBook Reloaded
  • GA Guestbook
  • Gallery2 Image
  • Icybook Guestbook
  • Jambook Guestbook
  • Jax Guestbook
  • Joomla Comment
  • K2 Blog
  • Pixelpost
  • Plogger
  • Serendipity Blog
  • Sitebuilder Guestbook
  • TextCube Guestbook
  • WordPress
  • WPTrackback

The platform files are in plain text, so you can also tweak existing platforms, or train new platforms if you like.
 
As noted by soft touch you can already prefix things and really you can add any prefix you want and just do a find replace to replace it with whatever you want.

Sorry my friend, I never noticed the prefix urls in harvester grid with http:// inside the more list tools. Not sure how I overlooked that option all these years. Thank you for the help. :)
 
ScrapeBox v2.0.0.48 Update Available

  • Fixed a bug in custom harvester related to loading proxies from file
  • Fixed a bug in META grabber
  • Added export as .csv to meta grabber export menu
  • Changed the way how the delay works in detailed harvester
  • Added antigate.com to the list of available captcha solver services
  • Addon: Google Competition Finder fixed
 
Thanks for your continued work on scrapebox and I'm glad you're better now. Any idea on why scrapebox v2 32bit goes up to 100% cpu usage after version 2.00.39? Never had this issue on my old rusty laptop running scrapebox 2.0039 and earlier. Not complaining, just really want to know what I can optimise or not. Thank you. My present version is 2.0047 with all free/2 paid plugins installed(autom, article scraper). All up to date.
ScrapeBox v2.0.0.48 Update Available
  • Fixed a bug in custom harvester related to loading proxies from file
  • Fixed a bug in META grabber
  • Added export as .csv to meta grabber export menu
  • Changed the way how the delay works in detailed harvester
  • Added antigate.com to the list of available captcha solver services
  • Addon: Google Competition Finder fixed
 
Takes already longer then 12 hours to active. I have send an email to technical support as well.
 
Thanks for your continued work on scrapebox and I'm glad you're better now. Any idea on why scrapebox v2 32bit goes up to 100% cpu usage after version 2.00.39? Never had this issue on my old rusty laptop running scrapebox 2.0039 and earlier. Not complaining, just really want to know what I can optimise or not. Thank you. My present version is 2.0047 with all free/2 paid plugins installed(autom, article scraper). All up to date.

What engines are you using and with how many connections? I don't have this issue, mind you I mainly only use google, and google is fairly optimized.

Takes already longer then 12 hours to active. I have send an email to technical support as well.

Make sure you sent in the correct info, and that you haven't already transferred this month.
 
Hi Matt I mostly use google.com, bing and yahoo. I also built custom googles based on google.com but replacing the country tld-s. They work the same. It's currently installed on my laptop with stock/default settings for the harvester. When I had a windows vps at the time of v2.0039 and earlier versions, scrapebox harvesting didn't use a lot cpu. It was very light. It mostly used cpu on link extracting. The vps I had was this: https://www.shape.host/windows-vps/, plan 3 but it had 8 gb ram at the time. I've never seen v2*39 and earliers use so much cpu on either that vps or on my home acer aspire 5736z 32bit win7 ultimate laptop. Many others have noted the increase of cpu usage after v40 and laters. v39 release comment: http://www.blackhatworld.com/blackhat-seo/seo-link-building/129096-scrapebox-ultimate-serp-scraper-auto-blog-commenter-prstorm-mode-post7873950.html#post7873950 v41 release comment: http://www.blackhatworld.com/blackhat-seo/seo-link-building/129096-scrapebox-ultimate-serp-scraper-auto-blog-commenter-prstorm-mode-post7909208.html#post7909208 Complaints begin about cpu usage: http://www.blackhatworld.com/blackhat-seo/seo-link-building/129096-scrapebox-ultimate-serp-scraper-auto-blog-commenter-prstorm-mode-post7912879.html#post7912879 Your reply to that one: http://www.blackhatworld.com/blackhat-seo/seo-link-building/129096-scrapebox-ultimate-serp-scraper-auto-blog-commenter-prstorm-mode-post7914941.html#post7914941 There's satyr85: http://www.blackhatworld.com/blackhat-seo/seo-link-building/129096-scrapebox-ultimate-serp-scraper-auto-blog-commenter-prstorm-mode-post8039438.html#post8039438 v39 album http://imgur.com/a/9l7Z6 v42 album http://imgur.com/a/LCgLa Another one: http://www.blackhatworld.com/blackhat-seo/seo-link-building/129096-scrapebox-ultimate-serp-scraper-auto-blog-commenter-prstorm-mode-post8042386.html#post8042386 I've done some testing right now on v47, and these major engines kill my cpu: Google.com Yahoo blocked me Bing.com Stock minor engines don't really worked // no cpu death Google nationalised domains I don't know what change since v39, perhaps it got better in performance and may require more resources. Perhaps these search engines got more resource "needy". Thanks regardless. Fuck bhw formatting, really tired of it.
What engines are you using and with how many connections? I don't have this issue, mind you I mainly only use google, and google is fairly optimized.
 
Last edited:
Hi,
Which proxy provider do you recommend ?

The providers SB recommends are here:
http://www.scrapebox.com/proxy

Hi Matt I mostly use google.com, bing and yahoo. I also built custom googles based on google.com but replacing the country tld-s. They work the same. It's currently installed on my laptop with stock/default settings for the harvester. When I had a windows vps at the time of v2.0039 and earlier versions, scrapebox harvesting didn't use a lot cpu. It was very light. It mostly used cpu on link extracting. The vps I had was this: https://www.shape.host/windows-vps/, plan 3 but it had 8 gb ram at the time. I've never seen v2*39 and earliers use so much cpu on either that vps or on my home acer aspire 5736z 32bit win7 ultimate laptop. Many others have noted the increase of cpu usage after v40 and laters. v39 release comment: http://www.blackhatworld.com/blackh...ter-prstorm-mode-post7873950.html#post7873950 v41 release comment: http://www.blackhatworld.com/blackh...ter-prstorm-mode-post7909208.html#post7909208 Complaints begin about cpu usage: http://www.blackhatworld.com/blackh...ter-prstorm-mode-post7912879.html#post7912879 Your reply to that one: http://www.blackhatworld.com/blackh...ter-prstorm-mode-post7914941.html#post7914941 There's satyr85: http://www.blackhatworld.com/blackh...ter-prstorm-mode-post8039438.html#post8039438 v39 album http://imgur.com/a/9l7Z6 v42 album http://imgur.com/a/LCgLa Another one: http://www.blackhatworld.com/blackh...ter-prstorm-mode-post8042386.html#post8042386 I've done some testing right now on v47, and these major engines kill my cpu: Google.com Yahoo blocked me Bing.com Stock minor engines don't really worked // no cpu death Google nationalised domains I don't know what change since v39, perhaps it got better in performance and may require more resources. Perhaps these search engines got more resource "needy". Thanks regardless. Fuck bhw formatting, really tired of it.

Ha, yes the BHW formatting isn't so good.

I don't have issues, but I can duplicate it to some degree, if I really crank my threads way way up and especially with Bing. That said it doesn't totally max out my CPU or anything and I run optimized budget servers. That said I know they had to add some extra code for thread monitoring because someone on here was wanting to use SB a certain way and it was causing issues. Perhaps this had something to do with it. Perhaps its as you noted that its a combination of SE getting more needy and Scrapebox getting more efficient, thus using more resources.

I do know there are multiple people that noted this, for me its hard because on servers that are way less powerful then what others state and I can't duplicate it to the degree they can. I know I talked to support and on a budget vps they couldn't duplicate it either.

My guess is its something to do with a handful of peoples setup, some 3rd party software, something and support can't isolate it so its hard to compensate for it. I don't know exactly, I know across nearly a dozen machines I can't duplciate any issues nearly as large as what anyone else is stating, and Im talking even on a machine that costs $75 and is split into 2 vps so getting half of a $75 machines worth of resources, and support has a $30 vps they couldn't get it to duplciate on.

Did you have any common 3rd party software installed on the vps and the laptop?

i wait for 12hour, but still not get active, please help me

You will need to contact support directly at

support (at} scrapebox (dot] com
 
Last edited:
Q: Any common software? No, not really. Other than your typical desktop needs, like dropbox and firefox. My default browser has been firefox in the last 8 years I think, but that shouldn't change the in-built "browser" or whatever scrapebox uses to harvest or should it? I don't think my default browser setting has anything to do with resource usage on scrapebox. Other common software at that time was GSA Search Engine Ranker and Captcha breaker. SER, CB and multiple copies scrapebox copies run along merrily on that server I linked to above. But something changed and CPU usage became insane after v41. I understand and appreciate your honesty about not being able to reproduce it. My server ran on win srv 08 and I've used/still use win 7 at home. I assume server 08 and win 7 share the same kernel. Perhaps it's something to the with Windows' kernel from 2008 era? Honestly no idea. I forgot to add that yesterday when I wrote the previous comment of mine, scrapebox went apeshit on cpu on both 10 threads and stock 50 threads. CPU usage goes mad about 1 to 3 seconds after running the custom harvester. Detailed harvester doesn't seem to go mad, uses random 5 to 30% CPU on this crap laptop.
... Did you have any common 3rd party software installed on the vps and the laptop? support (at} scrapebox (dot] com
 
Thanks for your continued work on scrapebox and I'm glad you're better now. Any idea on why scrapebox v2 32bit goes up to 100% cpu usage after version 2.00.39? Never had this issue on my old rusty laptop running scrapebox 2.0039 and earlier.

This bug was reported long time ago (its 2+ months old bug), but SB support dont care about this bug. Maybe this bug affect 1% customers but its still a bug and it should be fixed.

Edit:
I know I talked to support and on a budget vps they couldn't duplicate it either.
I talked to support and i offered full access to dedicated server where they could reproduce bug. Support told me its too much work to move all stuff to my server or something like this. Support had enough proof, and help from people like me to track this bug and fix it. Sad to say but from my personal poitn of view, they dont want to fix it.
 
Last edited:
Hello, i want to try scrapebox for adult site, is there any way to disable blacklist and harversting blacklist site ?

thanks
 
This bug was reported long time ago (its 2+ months old bug), but SB support dont care about this bug. Maybe this bug affect 1% customers but its still a bug and it should be fixed.

Edit:

I talked to support and i offered full access to dedicated server where they could reproduce bug. Support told me its too much work to move all stuff to my server or something like this. Support had enough proof, and help from people like me to track this bug and fix it. Sad to say but from my personal poitn of view, they dont want to fix it.

To clarify, nobody said that its too much work to move things around.

Here is the part of the email I sent you regarding that:
"If the CPU would be really at 100% (physical cpu), then you would not really be able to do anything else on the vps, but that?s not the case, I can start other programs, move windows around just fine on your vps, even it shows 100% usage.
As I mentioned, I cannot run the IDE debugger, it would involve installing the whole development software."


Running a CPU at 100% does not necessarily mean its a bug, but could mean that the software is able of fully utilize your CPU.
You might see 100% CPU due to the way the virtualization is implemented on your vps.
As I wrote, even when SB shows 100% CPU, I still can run other programs on your vps, move smoothly windows around without a problem.
If a physical CPU would run at 100%, that would almost not be possible.

So what is your problem with the CPU? It seems not to cause any issue on your vps.

And nobody said we don't care about this, the problem is effecting just a couple of people from the whole user base but both myself, sweetfunny and loopline have spent more time investigating this then any other problem since V2's release.
 
Q: Any common software? No, not really. Other than your typical desktop needs, like dropbox and firefox. My default browser has been firefox in the last 8 years I think, but that shouldn't change the in-built "browser" or whatever scrapebox uses to harvest or should it? I don't think my default browser setting has anything to do with resource usage on scrapebox. Other common software at that time was GSA Search Engine Ranker and Captcha breaker. SER, CB and multiple copies scrapebox copies run along merrily on that server I linked to above. But something changed and CPU usage became insane after v41. I understand and appreciate your honesty about not being able to reproduce it. My server ran on win srv 08 and I've used/still use win 7 at home. I assume server 08 and win 7 share the same kernel. Perhaps it's something to the with Windows' kernel from 2008 era? Honestly no idea. I forgot to add that yesterday when I wrote the previous comment of mine, scrapebox went apeshit on cpu on both 10 threads and stock 50 threads. CPU usage goes mad about 1 to 3 seconds after running the custom harvester. Detailed harvester doesn't seem to go mad, uses random 5 to 30% CPU on this crap laptop.

I run windows 2008 on most of my machines because I like it better then 12, but I do have 2012 on a few servers, and I run windows 8 at home and office. So I don't think its the OS. Its a good a excuse for you to buy a new laptop, hehe

Truely though I don't know and I can't reproduce it.

Hello, i want to try scrapebox for adult site, is there any way to disable blacklist and harversting blacklist site ?

thanks

The blacklist isn't to do with adult sites, its to do with sites you wouldn't want to post on no matter what. Like matt cutts blog for instance, thats like asking to get penalized. So its a non issue. Scrapebox will work fine with adult sites.

To clarify, nobody said that its too much work to move things around.

Here is the part of the email I sent you regarding that:
"If the CPU would be really at 100% (physical cpu), then you would not really be able to do anything else on the vps, but that?s not the case, I can start other programs, move windows around just fine on your vps, even it shows 100% usage.
As I mentioned, I cannot run the IDE debugger, it would involve installing the whole development software."


Running a CPU at 100% does not necessarily mean its a bug, but could mean that the software is able of fully utilize your CPU.
You might see 100% CPU due to the way the virtualization is implemented on your vps.
As I wrote, even when SB shows 100% CPU, I still can run other programs on your vps, move smoothly windows around without a problem.
If a physical CPU would run at 100%, that would almost not be possible.

So what is your problem with the CPU? It seems not to cause any issue on your vps.

And nobody said we don't care about this, the problem is effecting just a couple of people from the whole user base but both myself, sweetfunny and loopline have spent more time investigating this then any other problem since V2's release.

Thats interesting on the still being able to use it. I know on my servers when it hits 100% CPU due to whatever all I have running its 100%, its unresponsive. I can't do anything, windows won't load, it takes a fortnight to run a program it seems, and I often have to disconnect from the server and reconnect several time as its responsive only for a brief period when I first log in.

Cottonwolf, when it happens to you can you still use the laptop at all, or does it more or less freeze? Just curious.
 
nobody said that its too much work to move things around.
True, but you didnt want to move debuger to my dedi (its probably easiest way to find bug) although i gave full access to server for unlimited time. On other hand i understand that moving sensible data to not your server can be not comfortable but probably its only way to find this bug.

You might see 100% CPU due to the way the virtualization is implemented on your vps.
I mentioned it many times in screenshots and emails I dont use VPS, I u use beefy dedicated servers (dual cpu, 12 cores, 24 threads) and SB is able to kill it (two physical CPUs, each CPU with 6 physical cores).

As you can see here SB is using 100% CPU (dual 6 core, 24 threads) dedicated server (standard dedicated servers like Xeon e3 have 4 cores 8 threads) - task manager is showing 24 threads available:


uj1G17c




Another screenshot - newest version, windows 2012R2 OS, other dedicated server - Xeon E5 1650 - SB is killing full physical CPU:

klZziMz




Same server, version .39 - 3 times better lpm with 8-9 times lower cpu usage:

QcY6tgW



Its not possible for scraper to kill such CPUs, but SB is doing it. I can give you full access to this server so you can install debuger there and debug everything for as much time as you need. On this server you will be able to reproduce bug, so there should be no problems with tracking bug.

Im willing to do everything i can to help you fix this bug.
 
Status
Not open for further replies.
Back
Top