[jira] [Commented] (HDFS-14624) When decommissioning a node, log remaining blocks to replicate periodically

Stephen O'Donnell (JIRA) Tue, 02 Jul 2019 14:11:05 -0700


    [ 
https://issues.apache.org/jira/browse/HDFS-14624?page=com.atlassian.jira.plugin.system.issuetabpanels:comment-tabpanel&focusedCommentId=16877303#comment-16877303
 ]


Stephen O'Donnell commented on HDFS-14624:
------------------------------------------

This request is driven from requests by our support team and customers, who 
state that tracking decommission progress is hard.

The same information is there in JMX, and the namenode WebUI shows the current 
snapshot nicely. In support cases we tend to get logs rather than JMX samples, 
and it can be useful to see when decommission got stuck, and track the progress 
over time to see if it is getting faster or slower etc.

We can do this by adjusting the log level of the DatanodeAdminManager 'on the 
fly', but that is often only after the problem occurs, so it would be nice to 
have this logged as normal info messages provided they don't create too much 
spam.

> When decommissioning a node, log remaining blocks to replicate periodically
> ---------------------------------------------------------------------------
>
>                 Key: HDFS-14624
>                 URL: https://issues.apache.org/jira/browse/HDFS-14624
>             Project: Hadoop HDFS
>          Issue Type: Improvement
>          Components: namenode
>    Affects Versions: 3.3.0
>            Reporter: Stephen O'Donnell
>            Assignee: Stephen O'Donnell
>            Priority: Major
>         Attachments: HDFS-14624.001.patch
>
>
> When a node is marked for decommission, there is a monitor thread which runs 
> every 30 seconds by default, and checks if the node still has pending blocks 
> to be replicated before the node can complete replication.
> There are two existing debug level messages logged in the monitor thread, 
> DatanodeAdminManager$Monitor.check(), which log the correct information 
> already, first as the pending blocks are replicated:
> {code:java}
> LOG.debug("Node {} still has {} blocks to replicate "
>     + "before it is a candidate to finish {}.",
>     dn, blocks.size(), dn.getAdminState());{code}
> And then after the initial set of blocks has completed and a rescan happens:
> {code:java}
> LOG.debug("Node {} {} healthy."
>     + " It needs to replicate {} more blocks."
>     + " {} is still in progress.", dn,
>     isHealthy ? "is": "isn't", blocks.size(), dn.getAdminState());{code}
> I would like to propose moving these messages to INFO level so it is easier 
> to monitor decommission progress over time from the Namenode log.
> Based on the default settings, this would result in at most 1 log message per 
> node being decommissioned every 30 seconds. The reason this is at the most, 
> is because the monitor thread stops after checking after 500K blocks and 
> therefore in practice it could be as little as 1 log message per 30 seconds, 
> even if many DNs are being decommissioned at the same time.
> Note that the namenode webUI does display the above information, but having 
> this in the NN logs would allow progress to be tracked more easily.



--
This message was sent by Atlassian JIRA
(v7.6.3#76005)

---------------------------------------------------------------------
To unsubscribe, e-mail: [email protected]
For additional commands, e-mail: [email protected]

[jira] [Commented] (HDFS-14624) When decommissioning a node, log remaining blocks to replicate periodically

Reply via email to