[
https://issues.apache.org/jira/browse/FLINK-2250?page=com.atlassian.jira.plugin.system.issuetabpanels:all-tabpanel
]
Fabian Hueske updated FLINK-2250:
---------------------------------
Component/s: (was: Java API)
(was: Scala API)
DataSet API
> Backtracking of intermediate results
> ------------------------------------
>
> Key: FLINK-2250
> URL: https://issues.apache.org/jira/browse/FLINK-2250
> Project: Flink
> Issue Type: New Feature
> Components: DataSet API, Distributed Runtime
> Reporter: Maximilian Michels
> Assignee: Maximilian Michels
>
> With intermediate results available in the distributed runtime as of
> FLINK-986, we could now incrementally resume failed jobs if we cached the
> results FLINK-1404. Moreover, Flink users could build incremental Flink jobs
> using count/collect/print and, ultimately, also continue from old job results
> in an interactive shell environment like the scala-shell.
> The following tasks need to be completed for that to happen:
> - Cache the results
> - Keep the ExecutionGraph in the JobManager
> - Change the scheduling mechanism to track back the results from the sinks
> - Implement a session management to eventually discard old results and
> ExecutionGraphs from the TaskManagers/JobManager
--
This message was sent by Atlassian JIRA
(v6.3.4#6332)