[ 
https://issues.apache.org/jira/browse/FLINK-838?page=com.atlassian.jira.plugin.system.issuetabpanels:comment-tabpanel&focusedCommentId=14087437#comment-14087437
 ] 

Fabian Hueske edited comment on FLINK-838 at 8/6/14 9:34 AM:
-------------------------------------------------------------

I'm not really an expert for classloaders. I would suggest to put this issue 
aside for now. Maybe someone has  a clever idea.

The more important issue is to get the job configuration right. 

Regarding the custom Partitioner/Sorter/Grouper it is not as easy as it sounded 
in my previous comment. I will try to get a prototype of a custom runtime 
operator for Hadoop Reduce which uses the custom sorter and grouper instead of 
the Flink provided sorters and gropers. This requires changes through the full 
job preparation process, i.e., job building, optimization, translation, and 
execution. Let's see if that works...

What do you mean exactly by "support for sorting (custom {{Comparators}})"? Are 
you talking about the custom comparators for the Reduce sorter? How did you do 
that?

Yeah, time flies by. Only 1.5 weeks left by now. Let's try hard to get your 
code into the master branch next week!



was (Author: fhueske):
I'm not really an expert for classloaders. I would suggest to put this issue 
aside for now. Maybe someone has  a clever idea.

The more important issue is to get the job configuration right. 

Regarding the custom Partitioner/Sorter/Grouper it is not as easy as it sounded 
in my previous comment. I will try to get a prototype of a custom runtime 
operator for Hadoop Reduce which uses the custom sorter and grouper instead of 
the Flink provided sorters and gropers. This requires changes through the full 
job preparation process, i.e., job building, optimization, translation, and 
execution. Let's see if that works...

What do you mean exactly by "support for sorting (custom {Comparators})"? Are 
you talking about the custom comparators for the Reduce sorter? How did you do 
that?

Yeah, time flies by. Only 1.5 weeks left by now. Let's try hard to get your 
code into the master branch next week!


> GSoC Summer Project: Implement full Hadoop Compatibility Layer for 
> Stratosphere
> -------------------------------------------------------------------------------
>
>                 Key: FLINK-838
>                 URL: https://issues.apache.org/jira/browse/FLINK-838
>             Project: Flink
>          Issue Type: Improvement
>            Reporter: GitHub Import
>              Labels: github-import
>             Fix For: pre-apache
>
>
> This is a meta issue for tracking @atsikiridis progress with implementing a 
> full Hadoop Compatibliltiy Layer for Stratosphere.
> Some documentation can be found in the Wiki: 
> https://github.com/stratosphere/stratosphere/wiki/%5BGSoC-14%5D-A-Hadoop-abstraction-layer-for-Stratosphere-(Project-Map-and-Notes)
> As well as the project proposal: 
> https://github.com/stratosphere/stratosphere/wiki/GSoC-2014-Project-Proposal-Draft-by-Artem-Tsikiridis
> Most importantly, there is the following **schedule**:
> *19 May - 27 June (Midterm)*
> 1) Work on the Hadoop tasks, their Context and the mapping of Hadoop's 
> Configuration to the one of Stratosphere. By successfully bridging the Hadoop 
> tasks with Stratosphere, we already cover the most basic Hadoop Jobs. This 
> can be determined by running some popular Hadoop examples on Stratosphere 
> (e.g. WordCount, k-means, join) (4 - 5 weeks)
> 2) Understand how the running of these jobs works (e.g. command line 
> interface) for the wrapper. Implement how will the user run them. (1 - 2 
> weeks).
> *27 June - 11 August*
> 1) Continue wrapping more "advanced" Hadoop Interfaces (Comparators, 
> Partitioners, Distributed Cache etc.) There are quite a few interfaces and it 
> will be a challenge to support all of them. (5 full weeks)
> 2) Profiling of the application and optimizations (if applicable)
> *11 August - 18 August*
> Write documentation on code, write a README with care and add more 
> unit-tests. (1 week)
> ---------------- Imported from GitHub ----------------
> Url: https://github.com/stratosphere/stratosphere/issues/838
> Created by: [rmetzger|https://github.com/rmetzger]
> Labels: core, enhancement, parent-for-major-feature, 
> Milestone: Release 0.7 (unplanned)
> Created at: Tue May 20 10:11:34 CEST 2014
> State: open



--
This message was sent by Atlassian JIRA
(v6.2#6252)

Reply via email to