[
https://issues.apache.org/jira/browse/HDFS-12345?page=com.atlassian.jira.plugin.system.issuetabpanels:comment-tabpanel&focusedCommentId=16364655#comment-16364655
]
Íñigo Goiri edited comment on HDFS-12345 at 2/14/18 7:32 PM:
-------------------------------------------------------------
I'm personally interested on pushing HDFS as a YARN job into Hadoop itself.
I've gone through that part of the code extensively and I'd like a couple
things tweaked here and there but I like the setup.
We already have tools like GridMix so I think it makes sense to add the
workload replayer too.
I internally have a couple MapReduce jobs that do something pretty similar so I
could also converge into this.
To summarize, my opinion is that we should go over on how to split the work but
I think this most of Dynanometer (if not all) should go into Hadoop.
was (Author: elgoiri):
I'm personally interested on pushing HDFS as a MapReduce job into Hadoop itself.
I've gone through that part of the code extensively and I'd like a couple
things tweaked here and there but I like the setup.
We already have tools like GridMix so I think it makes sense to add the
workload replayer too.
I internally have a couple MapReduce jobs that do something pretty similar so I
could also converge into this.
To summarize, my opinion is that we should go over on how to split the work but
I think this most of Dynanometer (if not all) should go into Hadoop.
> Scale testing HDFS NameNode with real metadata and workloads
> ------------------------------------------------------------
>
> Key: HDFS-12345
> URL: https://issues.apache.org/jira/browse/HDFS-12345
> Project: Hadoop HDFS
> Issue Type: New Feature
> Components: namenode, test
> Reporter: Zhe Zhang
> Assignee: Erik Krogen
> Priority: Major
>
> Dynamometer has now been open sourced on our [GitHub
> page|https://github.com/linkedin/dynamometer]. Read more at our [recent blog
> post|https://engineering.linkedin.com/blog/2018/02/dynamometer--scale-testing-hdfs-on-minimal-hardware-with-maximum].
> To encourage getting the tool into the open for others to use as quickly as
> possible, we went through our standard open sourcing process of releasing on
> GitHub. However we are interested in the possibility of donating this to
> Apache as part of Hadoop itself and would appreciate feedback on whether or
> not this is something that would be supported by the community.
> Also of note, previous [discussions on the dev mail
> lists|http://mail-archives.apache.org/mod_mbox/hadoop-hdfs-dev/201707.mbox/%[email protected]%3e]
--
This message was sent by Atlassian JIRA
(v7.6.3#76005)
---------------------------------------------------------------------
To unsubscribe, e-mail: [email protected]
For additional commands, e-mail: [email protected]