Expired Tumbler Scraping Issue in Scrapebox

SiddharthW

Power Member
Joined
Jul 29, 2017
Messages
734
Reaction score
384
I recently started Scraping some Tumbler but Scrapebox giving 404 error on Sensitive Media. Any solution?
 
I recently started Scraping some Tumbler but Scrapebox giving 404 error on Sensitive Media. Any solution?
Can you give me some examples? Im not sure I even understand the issue. Where are you getting the 404 when scraping in the main harvester or when running the vanity name checker addon or ?

Perhaps you can attach a screenshot.
 
Can you give me some examples? I'm not sure I even understand the issue. Where are you getting the 404 when scraping in the main harvester or when running the vanity name checker addon or?

Perhaps you can attach a screenshot.
I'm getting 404 code in vanity name checker addon. But those Aren't actually expired domain. I've configured the Scarpebox in that way Alive = 404 error. But When I started checking vanity name checker addon giving 404 error on this type of site https://raravis-carlqvist.tumblr.com/
 

Attachments

  • Screenshot (445).png
    Screenshot (445).png
    168.6 KB · Views: 9
I'm getting 404 code in vanity name checker addon. But those Aren't actually expired domain. I've configured the Scarpebox in that way Alive = 404 error. But When I started checking vanity name checker addon giving 404 error on this type of site https://raravis-carlqvist.tumblr.com/
I mean by Alive Checker. I'm not using Vanity name checker.
 
I'm getting 404 code in vanity name checker addon. But those Aren't actually expired domain. I've configured the Scarpebox in that way Alive = 404 error. But When I started checking vanity name checker addon giving 404 error on this type of site https://raravis-carlqvist.tumblr.com/
Thats a bit of an anomaly. When I run it, it completes fine, so its entirely possible its your security software etc..

When I load that in a browser it redirects to a safemod on page. That safemode on page doesn't contain data that makes scrapebox think its available, nor does it contain data that tells scrapebox its taken. So the vanity checker, for me, just displays completed.

You could edit the tumblr definition file to show that if it gets to safe mode it shows as taken.

But because its safe mode your security software may be intercepting the request and thus returning a 404. Or your ISP or your router or all sorts of things. So I doubt its your ISP, unless you just have really strict rules in your country.

Else whitelist scrapebox in all security software, then whitelist in your router, if possible. You will also want to whitelist the entire scrapebox folder and then disable any real time scanning.

Make sure you whitelist and not just disable security software.
 
Thats a bit of an anomaly. When I run it, it completes fine, so its entirely possible its your security software etc..

When I load that in a browser it redirects to a safemod on page. That safemode on page doesn't contain data that makes scrapebox think its available, nor does it contain data that tells scrapebox its taken. So the vanity checker, for me, just displays completed.

You could edit the tumblr definition file to show that if it gets to safe mode it shows as taken.

But because its safe mode your security software may be intercepting the request and thus returning a 404. Or your ISP or your router or all sorts of things. So I doubt its your ISP, unless you just have really strict rules in your country.

Else whitelist scrapebox in all security software, then whitelist in your router, if possible. You will also want to whitelist the entire scrapebox folder and then disable any real time scanning.

Make sure you whitelist and not just disable security software.
Can you check this with ALive Check V2.0.0.12?
 
I've observed Vanity Name Checker Is working Fine with Sensitive Site but Alive check giving 404 error in this situation.
 
Just Checked I can't use rotating proxy in Vanity Name Checker, Why? Private Proxy is working but no rotating Proxy
 
Scrapebox giving 404 error on Sensitive Media

Is it possible to spoof your useragent to Googlebot when using Scrapebox?

Interesting fact about Tumblr: Normally you can only see sensitive media when logged into a Tumblr account. But since that would prevent Google from crawling those pages correctly, they show the pages normally to anyone whose useragent is set to / spoofed as Googlebot. If you can set that up to work with Scrapebox, there is a good chance that it would solve your problem.
 
Back
Top