i need help on scrapebox,
how can i generate urls so that i can check if the site is alive using the add-ons. see sample below:
http://sample.com/?site=1
http://sample.com/?site=2
i need to generate millions of numbers to check the website.. is it possible with scrapebox?
thanks
You can use random.org to generate the numbers or some other number generator and then use the merge function
http://www.random.org/integers/
merge info
http://scrapeboxfaq.com/how-do-i-use-tokens-with-the-m-merge-option
but you do like
http://sample.com/?site=%kw%
Then just put your numbers in the keywords box and put the above in a text file, save it and hit the M merge button.
Bear in mind the current 32bit version of SB will crash if you put millions of numbers in there and hit merge. Ive tested it in 2.0 64bit version and can do millions with no problem, although I did manage to use 6GB+ of memory just on messing with keywords. (2.0 is not yet publicly released)
Then just save them in the keyword section and open them in an addon. Again importing too many millions of urls in the alive checker will cause it to run out of memory and crash, also not an issue with 64bit 2.0 alive checker addon, but thats not yet out.
Been using SB for a couple of years now, and it still is the most handy tool I have.
I have 2 questions about the proxy checker though.
1: Is there any way to quickly add all the 404/503 etc proxies to the blacklist?
2: Can I somehow remove duplicate proxies even if their ports are different?
Thanks!
What do you mean by add the 404/503 to the blacklist. When you filter the proxies when you are done, they will be removed. Are you talking about the proxy sources that return that or...?
There is no way to remove duplicate with a different port, as they are seen as a unique proxy. Although if you only want certain ports you can do that.
hello looplines..
i have some question ,
1.how do i make scrapebox to scrape target site based site meta tags for example this tags
2. how to fix the 403 eror on premium plugin article scraper when i try to submit an article to a blogspot blog
thank you
Scrapebox can only scrape with parameters that the engines support and google etc... do not allow you to scrape based on meta data, or html tags/data. Only based on page "content". So no you can't do that. You can scrape with a footprint that gets you close and use the page scanner to qualify pages based on the meta data though.
403 would be coming from either your proxies, or blogger blocking your IP, or some security software on your pc. Try it without proxies, does that work? Can you login to the blog in a browser.
Sweetfunny does not answer pms, you should contact them directly
http://www.scrapebox.com/contact-us