CoolFace
Modelpublic

OmniAICreator/Galgame-Llasa-1B-v3

sourceHugging Facecc-by-nc-4.0updated 1y agoView on Hugging Face
2likes37downloads
Model Card

Galgame-Llasa-1B-v3

Overview

This is the version 3 of the Galgame-Llasa-1B, a Text-to-Speech (TTS) model fine-tuned for Japanese. This model is based on HKUSTAudio/Llasa-1B-Multilingual.

What's New in v3?

The primary improvement in v3 is the modification of the text normalization process during training.

This update leads to more consistent and accurate speech synthesis, further improving upon the advances made in v2.

What's New in v2 (from v1)?

Version 2 was trained on a larger and more diverse dataset, including the original Galgame dataset and other sources.

As a result, v2 offered several key improvements over the original version:

  • —Improved Kanji Reading: The model handled the reading of Kanji characters more accurately.
  • —Enhanced Prosody: The generated speech had more natural intonation and expressiveness.
  • —Greater Voice Diversity: The model could produce a wider range of voice styles than the previous version.

License

This model is licensed under the CC-BY-NC-4.0.