Skip to content
This repository was archived by the owner on Jan 26, 2021. It is now read-only.
This repository was archived by the owner on Jan 26, 2021. It is now read-only.

data prepare #81

Description

@vision-zhao

there is transcript.txt file in my hand, and i want use it to train a lda model, which is fed by two files like 'docword.nytimes.txt' and 'vocab.nytimes.txt'. so i wonder how to handle this transcript to be the exactly format?(i have wrote a script to handle my file, but it didn't work) i will be very grateful if you can tell me the exactly format I need when you are convenient!

Activity

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Metadata

Metadata

Assignees

No one assigned

    Labels

    No labels
    No labels

    Type

    No type

    Projects

    No projects

      Milestone

      No milestone

      Relationships

      None yet

      Development

      No branches or pull requests

      Issue actions