Scrapebox and manual commenting - How to check duplicates?

retroslice

Registered Member
Joined
Mar 15, 2012
Messages
82
Reaction score
6
Hey guys,

I am planning on using scrapebox to manually comment on high pr blogs to backlink to my money site.

So far I have got a list of 100 quality blogs to comment on that are do follow, and I'll start commenting on them slowly as my site is new.

My question is, is there a method that allows me to check which sites/domains I have already commented on, so if scrapebox harvests the same url in the future it will automatically know and remove it? (if I start a brand new harvest with similar keywords).

I am thinking if I got 100 urls I would save them as a text file.
Then the next day I harvest another 100 urls, I wan't to check them against the ones I already commented on from the text file from the previous day so I can remove them.

Hope this makes sense
 
retroslice,

There are two ways you can go about this. The first would be to import your existing lists into the Harvester section, then use Import URL List > Import URL Lists to compare (on domain level). This will remove any duplicate URL's or Domains from the main list.

Another way you can accomplish this automatically is to add the domains to Scrapebox's Blacklist. You can find the blacklist inside a folder in your Scrapebox directory. However, the last method is only viable for small amounts of domains (a few hundred) otherwise it will begin to severely decrease your harvesting speed (as it needs to check each URL harvested against the list).

Hope this helps,
GG
 
That's brilliant, method 1 was exactly what I was looking for.

Thank you!
 
Without starting a new thread, could you tell me how to find edu links that allow/have comments?

I just harvested 2000 .edu links using the built in footprint "site:.edu", and clicking through them all, the majority do not have comments enabled!
 
Without starting a new thread, could you tell me how to find edu links that allow/have comments?

I just harvested 2000 .edu links using the built in footprint "site:.edu", and clicking through them all, the majority do not have comments enabled!

Try using any of these footprints:

site:.edu "Leave a Comment"
site:.edu "Leave a Response"
site:.edu "Leave a Reply"
site:.edu "Add Comment"
site:.edu "Add Response"
site:.edu "Add Reply"
site:.edu "Post a Comment"
site:.edu "Post a Response"
site:.edu "Post a Reply"
 
I am thinking if I got 100 urls I would save them as a text file.
Then the next day I harvest another 100 urls, I wan't to check them against the ones I already commented on from the text file from the previous day so I can remove them.
You could do this by using Scrape box's Remove/Filter to remove the URL's not required.

Lets say you save the previous day's commented URL's in commented.txt, and the newly harvested URL's are in Scrape box's Harvester. You could use Remove/Filter > Remove Url's Containing Entries From: <Browse to commented.txt>
 
You can also take all of your successfully posted URL's and store them in excel which also has a duplicate checker.
 
Back
Top