Nutch has these protocols implemented : http, https, ftp, file. As long as you get links to your documents in those schemes, Nutch would do the crawl.
Thanks, Tejas On Tue, Jan 28, 2014 at 10:07 PM, rashmi maheshwari < [email protected]> wrote: > I could crawl internet webpage and local directory folder to some extent. > > How to implement email and inranet blogs crawling? > > -- > Rashmi > Be the change that you want to see in this world! >

