how to scrape expired articles?

Do you have coding experience? If so I can help you greatly. It's actually very easy, take few codes from GitHub, combine them for more features and you have the ultimate tool.
 
I don't know if I understood your wishes well, so in any case, my suggestion would be to copy the article with expired domains and then insert it into the spinner.
In my opinion, it is better this way because you choose the text you want and not depend on some bot.
If you need few hundred articles from same site its way faster with bot than manually
 
https://github.com/hartator/wayback-machine-downloader

This will download the complete archive for any domain, personally I would just export the complete expired domain list for the relevant majestic niche then run all sites through the downloader.
i tried to use it and it works
it is work but it is download the whole html,
sime paged load very slow when i load them
there is some way to get from there only the text of every page?

I don't know if I understood your wishes well, so in any case, my suggestion would be to copy the article with expired domains and then insert it into the spinner.
In my opinion, it is better this way because you choose the text you want and not depend on some bot.
 
I wrote argo-content's team regarding this thing and they answered the following - 'It mainly scrapes from article directories, but you can supply an own list to be scraped via google so theoretically also from sites like waybackmachine, but not directly'. So this program can't scrape WBM directly (
 
To properly scrape expired articles you should code your own script. But if you don't have coding experience you can use content grabber. The steps are:
Step 1: Finding domains that have expired articles by going to expireddomains.net, search for keywords that are in your nice.
Step 2: Filter the keywords with high ACR, it means those sites have more content that archived.
Step 3: Go to archive.org type in the domain and check those sites to see how much content you can take, these steps you have to do manually.
Step 4: Use content grabber or web harvy, scrapebox... to save content to an excel file.
Step 4: Use plugin WP all import to upload these content to your website
 
To properly scrape expired articles you should code your own script. But if you don't have coding experience you can use content grabber. The steps are:
Step 1: Finding domains that have expired articles by going to expireddomains.net, search for keywords that are in your nice.
Step 2: Filter the keywords with high ACR, it means those sites have more content that archived.
Step 3: Go to archive.org type in the domain and check those sites to see how much content you can take, these steps you have to do manually.
Step 4: Use content grabber or web harvy, scrapebox... to save content to an excel file.
Step 4: Use plugin WP all import to upload these content to your website

so afther i find a domain with good content, i can use scrapebox to save only the content from that domain?
do you have guide for it?

and can you suggest me please some WP plugin for automaticly upload content?

thank you very much for your answer!
 
so afther i find a domain with good content, i can use scrapebox to save only the content from that domain?
do you have guide for it?

and can you suggest me please some WP plugin for automaticly upload content?

thank you very much for your answer!
here you go: Scrape Articles from Any Site in Any Language - New Scrapebox Article Scraper Plugin-
 
+ 1 for Scrapebox.
it can crape articles from just about any website, and upload it to WordPress. it also has a bunch of rewriting tools and batch uploading.
 
so afther i find a domain with good content, i can use scrapebox to save only the content from that domain?
do you have guide for it?
Well, I think I'm correct if I tell you that Scrapebox Article Scraper can ONLY scrape articles from online live websites, and not from Archive.org expired websites.

Is it right @loopline ?

I have bought that scrapebox plugin and I used it for to scrape artcles, but I can't find any option to download expired articles.
 
Well, I think I'm correct if I tell you that Scrapebox Article Scraper can ONLY scrape articles from online live websites, and not from Archive.org expired websites.

Is it right @loopline ?

I have bought that scrapebox plugin and I used it for to scrape artcles, but I can't find any option to download expired articles.
Correct, the article scraper does not work with archive.org

The expired domain finder can download all the files for the domain though, so that you can upload it back after you purchase it, if you want to recreate the domain.
 
Back
Top