You are viewing a plain text version of this content. The canonical link for it is here.
Posted to issues@flink.apache.org by "ASF GitHub Bot (JIRA)" <ji...@apache.org> on 2015/09/18 11:47:04 UTC

[jira] [Commented] (FLINK-2312) Random Splits

    [ https://issues.apache.org/jira/browse/FLINK-2312?page=com.atlassian.jira.plugin.system.issuetabpanels:comment-tabpanel&focusedCommentId=14805322#comment-14805322 ] 

ASF GitHub Bot commented on FLINK-2312:
---------------------------------------

Github user sachingoel0101 commented on the pull request:

    https://github.com/apache/flink/pull/921#issuecomment-141406447
  
    Right now, there is no way to achieve this. After support for persisted results is added, this can be re-visited again. Closing for now.


> Random Splits
> -------------
>
>                 Key: FLINK-2312
>                 URL: https://issues.apache.org/jira/browse/FLINK-2312
>             Project: Flink
>          Issue Type: Wish
>          Components: Machine Learning Library
>            Reporter: Maximilian Alber
>            Assignee: pietro pinoli
>            Priority: Minor
>
> In machine learning applications it is common to split data sets into f.e. training and testing set.
> To the best of my knowledge there is at the moment no nice way in Flink to split a data set randomly into several partitions according to some ratio.
> The wished semantic would be the same as of Sparks RDD randomSplit.



--
This message was sent by Atlassian JIRA
(v6.3.4#6332)