CantoCaptions Cantonese ASR

Details

This model is designed for high-quality written Cantonese transcription using the CantoCaptions standards, which mostly follow acceptable usages and variants documented in the words.hk dictionary. Sentence final particles (SFPs) such as 啦 laa1 / 喇 laa3 / 嗱 laa4 are separated out by tone as described on the CantoCaptions website.

The current model was trained using LoRa rank=128 and trained over a single epoch on ~75h of training audio, with an additional 4h dev and 4h test audio derived from the same dataset. At the time of writing, the model is in an early stage of development and may be updated.

Downloads last month
1,652
Safetensors
Model size
2B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for rookes/cantocaptions-cantonese-asr

Adapter
(12)
this model