Scrapebox reporting error 404 during harvesting

ljuba973

Newbie
Joined
Jul 23, 2011
Messages
15
Reaction score
0
Hi all,

I am using ScrapeBox together with Auto Hide IP and that combination worked perfectly until yesterday when I tried to harvest some blogs from Google and Yahoo and Yahoo reported error 404. I tried different country's IP and run again just Yahoo and same error I got. I left it until now and still after 24 hours I am getting same error - as on attached image.

I couldn't find what exactly Error 404 means - source unavailable, blocked, ...? How I can fix this issue and be able to harvest Yahoo again?
 

Attachments

  • YahooError.jpg
    YahooError.jpg
    31.3 KB · Views: 37
As far as i know Scrapebox is not harvesting links from yahoo since last few days. Don't know the issue is with proxies or yahoo's end.
 
Thanks a lot ... so 404 IS "source unavailable". Probably Yahoo changed structure of pages. Hopefully new update of SB will be issued with new harvesting algorithm
 
Yes i think Sweetfunny is releasing a new update for this.
 
It's nothing to do with SB. Yahoo is doing an API update or something and should be up very soon.
 
Not sure what the latest is, but it still aint working properly.

What was once an easy way to harvest over 2-3 million URLs now gets you around 3-4,000 URLs before you have to run it again.
 
Not sure what the latest is, but it still aint working properly.

What was once an easy way to harvest over 2-3 million URLs now gets you around 3-4,000 URLs before you have to run it again.

How can you harvest 2/3 million urls? Scrapebox seems to limit to 1 million.:06:
 
How can you harvest 2/3 million urls? Scrapebox seems to limit to 1 million.:06:

1 million is only the amount it can store in the harvested results section.

If you go to the harvested sessions inside the Scrapebox folder it stores them all there.

I'm getting "error 502" when I try harvesting with Yahoo and a quick bit of Googling says that's a Yahoo API gateway error. So it sounds like Yahoo have API problems, or they have made it so that Scrapebox won't function correctly with it.
 
Last edited:
1 million is only the amount it can store in the harvested results section.

If you go to the harvested sessions inside the Scrapebox folder it stores them all there.

I'm getting "error 502" when I try harvesting with Yahoo and a quick bit of Googling says that's a Yahoo API gateway error. So it sounds like Yahoo have API problems, or they have made it so that Scrapebox won't function correctly with it.

Ty for the millions :)

Well, I have the same problem for yahoo since this night.
Edit : Now Yahoo is good for me
 
Last edited:
Hi all,

I am using ScrapeBox together with Auto Hide IP and that combination worked perfectly until

I just tested auto hide ip, and it changes the IE proxy settings, does scrapebox use the IE setting if you dont set proxy in it?
how have you done that?
 
Back
Top