[
https://issues.apache.org/jira/browse/NUTCH-1341?page=com.atlassian.jira.plugin.system.issuetabpanels:comment-tabpanel&focusedCommentId=13482341#comment-13482341
]
Hudson commented on NUTCH-1341:
-------------------------------
Integrated in nutch-trunk-maven #465 (See
[https://builds.apache.org/job/nutch-trunk-maven/465/])
NUTCH-1341 NotModified time set to now but page not modified (Revision
1401288)
Result = SUCCESS
markus :
Files :
* /nutch/trunk/CHANGES.txt
* /nutch/trunk/src/java/org/apache/nutch/crawl/CrawlDbReducer.java
> NotModified time set to now but page not modified
> -------------------------------------------------
>
> Key: NUTCH-1341
> URL: https://issues.apache.org/jira/browse/NUTCH-1341
> Project: Nutch
> Issue Type: Bug
> Affects Versions: 1.5
> Reporter: Markus Jelsma
> Assignee: Markus Jelsma
> Fix For: 1.6
>
> Attachments: NUTCH-1341-1.6-1.patch
>
>
> Servers tend to respond with incorrect or no value for LastModified. By
> comparing signatures or when (fetch.getStatus() ==
> CrawlDatum.STATUS_FETCH_NOTMODIFIED) the reducer correctly sets the
> db_notmodified status for the CrawlDatum. The modifiedTime value, however, is
> not set accordingly.
--
This message is automatically generated by JIRA.
If you think it was sent incorrectly, please contact your JIRA administrators
For more information on JIRA, see: http://www.atlassian.com/software/jira