alvanli/canto-audio-llm

5

stars

18

commits

Python

primary language

Jan 10, 2025

updated

README

Canto-DiVA

  • this is an attempt in replicating the DiVA paper and finetune it to Cantonese audio data
    • DiVA provides an easy way to create Audio models without generating a bunch of tasks for audio data, and without instruction data
    • the original code was written in Levanter, so I wanted to turn it into PyTorch
  • The scripts are run on local Ubuntu machine with 2 x 4090s
  • The model has been training since Dec 19, 2024. When completed, results will be uploaded

Contributors

alvanli

18 commits

alvanli/canto-audio-llm

5

stars

18

commits

Python

primary language

Jan 10, 2025

updated

README

Canto-DiVA

  • this is an attempt in replicating the DiVA paper and finetune it to Cantonese audio data
    • DiVA provides an easy way to create Audio models without generating a bunch of tasks for audio data, and without instruction data
    • the original code was written in Levanter, so I wanted to turn it into PyTorch
  • The scripts are run on local Ubuntu machine with 2 x 4090s
  • The model has been training since Dec 19, 2024. When completed, results will be uploaded

Contributors

alvanli

18 commits

Languages

Python

99.2%