Skip to content

It seems no extra caption included in the code. #19

Description

@Johann356

In the paper, the method uses class labels or image caption generated by clipcap etc. to composite a text to form text embedding while the actual way in the code doesn't include only hard-coded text and no extra caption or imagenet label. I wonder whether the code is the actual training code used to produce your pretrained model. By the way, the performance is great. I guess maybe the extra training data from Danbooru2021 takes effect but I'm not sure. Thanks for answering.

Metadata

Metadata

Assignees

No one assigned

    Labels

    No labels
    No labels

    Projects

    No projects

    Milestone

    No milestone

    Relationships

    None yet

    Development

    No branches or pull requests

    Issue actions