Permanently Closed Marketplace Sales Thread

Status
Not open for further replies.
I've been having this issue a lot lately: I harvest proxies, I check them all and SB gets stuck on 1-10 threads for hours and in the end I'm just forced to close ScrapeBox and lose all my work (just lost 1M urls with PageRank). Any idea what's going on?
 
Appreciate the Support Loop!

One more last Question for Twitter:

Is it possible to scrape all "Tweets" from a very Big Twitter Account and harvest all Posts on the Wall and save them into txt File ?

In addition, would be possible to scrape only the Comments without the URL in the Tweet ?

Everything is possible depend how much interesting you want to do.
Without url or urls is pretty basic.
 
I've been having this issue a lot lately: I harvest proxies, I check them all and SB gets stuck on 1-10 threads for hours and in the end I'm just forced to close ScrapeBox and lose all my work (just lost 1M urls with PageRank). Any idea what's going on?

Are you checking proxies or checking page rank? Please express yourself clearly so we can help you not be stuck again.
 
Are you checking proxies or checking page rank? Please express yourself clearly so we can help you not be stuck again.

I'm checking proxies. Is there any way to close the proxy mananger without losing the 1M urls I checked? The Proxy Manager has been stuck like this for hours now.

SB Proxies.PNG
 
Guys do you know what is Google search query to find blogs that accept link in comments? whether ******** or nofollow both is ok, I just want to know what is the search query for that. Been searching for 4-5hours now without satisfying result....
 
Awesome, thanks very much for beta testing the Expired Domain Finder and it's great to hear you found a bunch of available domains. :)

Hi Sweetfunny, do you have an estimated release date for the expired domain finder?
 
I purchased SB a few weeks ago and so far it's been working great. I am still in the learning phase though, watching loopline's videos whenever I can :) I still have a long way to go ... :)
 
Guys do you know what is Google search query to find blogs that accept link in comments? whether ******** or nofollow both is ok, I just want to know what is the search query for that. Been searching for 4-5hours now without satisfying result....
Blogs which ScrapeBox supported accept links in comments?

"powered by wordpress" "(http|www)" "(http|www)" thanks share

Or Custom Grabber etc.
Search For Comments Allow Links.PNG
The link above is an anchor link.
There must be many other methods
to find comments which allow links.
 
I'm checking proxies. Is there any way to close the proxy mananger without losing the 1M urls I checked? The Proxy Manager has been stuck like this for hours now.

View attachment 73564
"without losing the 1M urls"--without losing the 1M proxies
and
10624 proxies--I think it's 0.1M. :)
There are some fancy function in Proxy Manager. Like save to a file every xxx minutes etc.
Or you can create an automator file,
split all proxy sources into several parts and test them separatedly.
Loopine's Youtube Channel has tutorial on this.
 
Appreciate the Support Loop!

One more last Question for Twitter:

Is it possible to scrape all "Tweets" from a very Big Twitter Account and harvest all Posts on the Wall and save them into txt File ?

In addition, would be possible to scrape only the Comments without the URL in the Tweet ?

I think that due to twitters script that loads the page as you scroll that you couldn't get more tweets then load on the initial page load. You could mess around with regex and see if you could avoid the url in the tweets, but Im a beginner at regex and I don't know for sure if it would or would not do it.

I've been having this issue a lot lately: I harvest proxies, I check them all and SB gets stuck on 1-10 threads for hours and in the end I'm just forced to close ScrapeBox and lose all my work (just lost 1M urls with PageRank). Any idea what's going on?


You can't kill just the proxy harvester as its part of the main Scrapebox, but you could do the proxy harvesting in a 2nd instance of Scrapebox and avoid the issue, and then just have it auto save good proxies to file every X mins. Thats the work around, but its better to solve the root issue.

The root issue being something on your machine is most likely locking 1 or more of the threads and not allowing it to finish. Make sure you whitelist Scrapebox in all security software, anti-virus, malware checker, firewall etc... Also you can whitelist it in your router and try closing down any unneeded programs.



Guys do you know what is Google search query to find blogs that accept link in comments? whether ******** or nofollow both is ok, I just want to know what is the search query for that. Been searching for 4-5hours now without satisfying result....

Are you asking for a query for auto approve blogs - if so there is no such thing. You just have to test post and then link check to find live links. If your link is live the site is auto approve. I have a video on it - https://www.youtube.com/watch?v=jYLvarYF6hI

Are you asking for a query to find blogs that have a comment form regardless of whether or not its auto approve or not - then you can use the built in footpints in Scrapebox or just search around here on BHW.

Are you asking for a query to find blogs that accept comments and accept links in he content body - I don't think this is possible.
 
Why Vanity checker doesn't show me the accurate result. Out of 500 scrape blogs, it is showing 90% are available but the truth is, it is not available, already taken. You have to fix this issue.
 
Why Vanity checker doesn't show me the accurate result. Out of 500 scrape blogs, it is showing 90% are available but the truth is, it is not available, already taken. You have to fix this issue.

Post some examples
 
Hi guys,

I've having issues with the harvester for local results.

I have edited the Google engine to be .co.uk, however when searching for a business name, I receive 8 yell.com listings, which I presume means that it is still searching as if it was a .com query.

Thing is, on non local queries it shows .co.uk results, which I why I'm confused!
 
You can't kill just the proxy harvester as its part of the main Scrapebox, but you could do the proxy harvesting in a 2nd instance of Scrapebox and avoid the issue, and then just have it auto save good proxies to file every X mins. Thats the work around, but its better to solve the root issue.

The root issue being something on your machine is most likely locking 1 or more of the threads and not allowing it to finish. Make sure you whitelist Scrapebox in all security software, anti-virus, malware checker, firewall etc... Also you can whitelist it in your router and try closing down any unneeded programs.

Does Stop ProxyManager when running threads are less or equal 10 threads. 0 to turn off. -- help avoid be locked? I set stop Manager less than 50~10 threads and result is good. I have no anti-virus installed.
threads.PNG
Even have recorded a video. Gosh I can't upload it here.
Proxy.PNG
 
Hi guys,
Thing is, on non local queries it shows couk results, which I why I'm confused!
Check result by RankChecker of couk and compare the results. Or capture the ht-t-p sessions of RankTracker, or Google Global extension and see if the edited Google engine is alright.
 
You can't kill just the proxy harvester as its part of the main Scrapebox, but you could do the proxy harvesting in a 2nd instance of Scrapebox and avoid the issue, and then just have it auto save good proxies to file every X mins. Thats the work around, but its better to solve the root issue.

The root issue being something on your machine is most likely locking 1 or more of the threads and not allowing it to finish. Make sure you whitelist Scrapebox in all security software, anti-virus, malware checker, firewall etc... Also you can whitelist it in your router and try closing down any unneeded programs.

Hey Loopline, thanks for the reply. I know I can't kill just the Proxy Manager and that's why I lost all my work the other day when I was forced to shut down ScrapeBox. I don't have any antivirus, malware, etc. software on my VPS, any idea what could be triggering it then?
 
Why Vanity checker doesn't show me the accurate result. Out of 500 scrape blogs, it is showing 90% are available but the truth is, it is not available, already taken. You have to fix this issue.

As soft touch said you can post some examples, that way someone can reproduce the issue.

But you also have to bear in mind that what the vanity checker does is to check for markers in the page that shows that a page is not live. As a general rule with such properties if a page isn't live it can be registered, and thats the assumption that the vanity checker makes.

However some properties, tumblr for example, will ban a domain and when they do that it can't ever be reregistered. So it will show up just like one that is available and you won't know its not available until you try and register it. Since the vanity checker is built for speed and doesn't support authentication etc... it does not attempt to register the properties it only checks to see if they match a given set of markers. So the vanity checker should more be considered a "highly accurate assumption based on probable data" then it should be a "this is 100% available to register".

I hear people about every week say this about tumblr, because it such a popular target, that it has had so many subdomains regsitered and then banned and so many people are trying to register subdomains on it that already have links that the vast majority of anything that meets common criteria is already taken or banned. Its kind of like trying to find a google passed proxy, but worse, everyone wants them and so they are few and far between. Same goes for tumblr, so many get banned and so many people are looking for ones to register that what is left is all the ones that are permanantly banned. So you might be correct, in your sample set 90% of them might not be available, but thats not a fault of the vanity checker, its doing its job correctly in that case, its a fault of high demand and low supply and the nature of the game.

Anyway, you can post some examples, but just bear in mind for some properties its like trying to get the Black Friday deal where the awesome TV is cheap and they have a very small quantity of them and everyone wants them. If your not from the US that won't make sense, but if you are then you get it. So if the vanity checker is saying its available and you load the page and someone already has it registered, then the vanity checker needs to be "fixed", but if its saying its not available and you load the page in a browser and it shows a 404 or not found/doesn't exist, then the vanity checker is working as it should, making the assumption that since it doesn't exist it should be availalbe, and if you then go to try and register it and its not able to be registered, then its just been permanently banned. In this case there is nothing wrong with the vanity checker and for that matter nothing that could even be fixed, its just the nature of low supply, high demand and the way the game is played.

So then you have 2 choices. Churn thru massive amounts of them to try and find a few you can register. Pick a different platform thats not so heavily used, or build in your own platform. (I have a video on this - https://www.youtube.com/watch?v=-F2nr_ltCRo )

Hi guys,

I've having issues with the harvester for local results.

I have edited the Google engine to be .co.uk, however when searching for a business name, I receive 8 yell.com listings, which I presume means that it is still searching as if it was a .com query.

Thing is, on non local queries it shows .co.uk results, which I why I'm confused!

What do you get in a browser at google.co.uk for the same terms?

Hey Loopline, thanks for the reply. I know I can't kill just the Proxy Manager and that's why I lost all my work the other day when I was forced to shut down ScrapeBox. I don't have any antivirus, malware, etc. software on my VPS, any idea what could be triggering it then?

You could turn off real time monitoring for windows defender. Also it could be a firewall in a router, so you could try to disable that or bypass the router as test. Until you sort it I would just run things in 2 separate instances. Also I assume you are using V 2.0.0.58 ?
 
Hey there mate...

How is the expired domain finder addon going? That's something I'd definitely like to purchase or BETA test.

Also, could you please add options to export to clipboard or back to Scrapebox for all of the windows and tools? Some addons only allow you to export as a file and I now find "file.xlsx" "testtesttest.txt" and "blahdiblah.csv" littered all over my desktop when all I want to do is move the URLs to the Scrapebox domain list.

Probably the best purchase I've made online by the way, along with Long Tail Pro.
 
Expired.PNG
When environment is normal my observation shows that Expired Domain Finder crawls about 100 urls per 2~3 seconds.
A conservative design?

Everyday I am busy testing it and keep sending problems to ScrapeBox team.
After the Domain Finder trending towards stable, plan to only run Finder and almost nothing else.
But where to host these has-been? Just 301 or sell them?
Often be immersed in deeply, general reflections and not yet get an answer now.
 
Status
Not open for further replies.
Back
Top