[
https://issues.apache.org/jira/browse/NUTCH-1341?page=com.atlassian.jira.plugin.system.issuetabpanels:comment-tabpanel&focusedCommentId=13419115#comment-13419115
]
Lewis John McGibbney commented on NUTCH-1341:
---------------------------------------------
Initially I share your concern about exactly where this should be set Markus
however having looked at the usage of prevModifiedTime throughout the class (~4
occurrences) and the knock on effect it looks fine.
> NotModified time set to now but page not modified
> -------------------------------------------------
>
> Key: NUTCH-1341
> URL: https://issues.apache.org/jira/browse/NUTCH-1341
> Project: Nutch
> Issue Type: Bug
> Affects Versions: 1.5
> Reporter: Markus Jelsma
> Assignee: Markus Jelsma
> Fix For: 1.6
>
> Attachments: NUTCH-1341-1.6-1.patch
>
>
> Servers tend to respond with incorrect or no value for LastModified. By
> comparing signatures or when (fetch.getStatus() ==
> CrawlDatum.STATUS_FETCH_NOTMODIFIED) the reducer correctly sets the
> db_notmodified status for the CrawlDatum. The modifiedTime value, however, is
> not set accordingly.
--
This message is automatically generated by JIRA.
If you think it was sent incorrectly, please contact your JIRA administrators:
https://issues.apache.org/jira/secure/ContactAdministrators!default.jspa
For more information on JIRA, see: http://www.atlassian.com/software/jira