[ https://issues.apache.org/jira/browse/FLINK-2379?page=com.atlassian.jira.plugin.system.issuetabpanels:all-tabpanel ]
Chesnay Schepler closed FLINK-2379. ----------------------------------- Resolution: Won't Do Closing since flink-ml is effectively frozen. > Add methods to evaluate field wise statistics over DataSet of vectors. > ---------------------------------------------------------------------- > > Key: FLINK-2379 > URL: https://issues.apache.org/jira/browse/FLINK-2379 > Project: Flink > Issue Type: New Feature > Components: Library / Machine Learning > Reporter: Sachin Goel > Assignee: Sachin Goel > Priority: Major > Labels: pull-request-available > Time Spent: 10m > Remaining Estimate: 0h > > Design methods to evaluate statistics over dataset of vectors. > For continuous fields, Minimum, maximum, mean, variance. > For discrete fields, Class counts, Entropy, Gini Impurity. > Further statistical measures can also be supported. For example, correlation > between two series, computing the covariance matrix, etc. > [These are currently the things Spark supports.] -- This message was sent by Atlassian JIRA (v7.6.3#76005)