Skip to content

Releases: openvpi/TIFA

v1.0.0: pretrained model

Choose a tag to compare

@yqzhishen yqzhishen released this 29 Sep 14:02

This is TIFA 1.0 pretrained model.

Datasets

Voice datasets

Natural noise & accompaniments datasets

Reverb dataset

MB-RIRs

Supported languages and special tags

Language Code Accepted forms Quality of coverage
Chinese Mandarin zh Hanzi, pinyin Excellent
English en Words Good
Japanese ja Kanji, kana, romaji Good
Yue (Cantonese) yue Hanzi, jyutping Moderate
Tag Meaning
AP Aspiration
EP Expiration
GS Glottal stop

Specifications

  • #Params: ~40M
  • Max context length: ~1 minute (longer inputs may result in unstable alignment)

Acknowledgements

Data labels are provided by:

License

The model files apply CC BY-NC-SA 4.0 license. This license only constraints the model file itself, not the data that flows through it.