Skip to content

"Split out" validation dataset #61

Description

@zeppelinche-cmd

It would be good if we could split out a tiny part (e.g. 0.1%) of all of our source datasets early on in processing to serve as a validation dataset that we don't train on.

Relevant for post-flag efforts.

Metadata

Metadata

Assignees

No one assigned

    Labels

    ideaIdea of thing to do, if time permit

    Type

    No type

    Projects

    Status
    Todo

    Milestone

    No milestone

    Relationships

    None yet

    Development

    No branches or pull requests

    Issue actions