Skip to content

Request for Training Code for Reproducing VSR (Lip Reading) #4

Description

@HarukiK2

Hello,

Thank you very much for your great work and for releasing CoGenAV — I believe it is a very promising project.

I'm currently working on reproducing visual speech recognition (VSR), specifically lip reading, and I’m very interested in CoGenAV's performance on the LRS2 dataset.

I’ve checked the repository and found the inference scripts (e.g., infer_avse_avss.py, infer_vsr_avsr.py), but I couldn’t find any training scripts or instructions.

Would it be possible for you to share the training code as well?

In particular, I’d appreciate any scripts or configuration files related to training on LRS2 for the VSR task.

I understand that preparing and releasing the training code may take time, but any guidance or partial resources would be greatly appreciated.

Thanks again for your valuable contribution to the community!

Activity

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Metadata

Metadata

Assignees

No one assigned

    Labels

    No labels
    No labels

    Type

    No type

    Projects

    No projects

      Milestone

      No milestone

      Relationships

      None yet

      Development

      No branches or pull requests

      Issue actions