Skip to content

Approach #1

Description

@Lkruitwagen
  • ML test-bench style
  • streaming data? or fixed corpus? (transfer costs excessive on Azure - probably want a fixed corpus)
  • two corpuses: [DL S2 DLSR, S1] and [L2A, EES1]
  • pytorch
  • pytorch parallel generator: https://stanford.edu/~shervine/blog/pytorch-how-to-generate-data-parallel
  • run self-supervised experiments, with supervised finishing on a joint header
  • "cross validate on the same distribution as test"

Metadata

Metadata

Assignees

No one assigned

    Labels

    No labels
    No labels

    Projects

    No projects

      Milestone

      No milestone

      Relationships

      None yet

      Development

      No branches or pull requests

      Issue actions