Irodori-TTS-v4.1-Anime
phasefield-audio's anime-voice fine-tune of Irodori-TTS v4.1 for expressive Japanese text-to-speech.
Base model
Model Description
A Japanese text-to-speech model fine-tuned from Aratako/Irodori-TTS-v4.1-Small using anime-style speech data.
The base model's annotation pipeline is not publicly documented, so the fine-tuning data was annotated independently. Consequently, caption conditioning and emoji controls may behave differently from the base model.
Checkpoints
The full-precision checkpoint is available at the repository root.
Quantized variants are provided in the following subdirectories:
int8-weight-onlyint8-dynamicint4-weight-onlyfloat8-weight-onlyfloat8-dynamic
For inference and installation instructions, see the original Irodori-TTS repository.
License
This model follows the same MIT License and ethical restrictions as the base model.
Sign up to read complete case studies, access detailed metrics, and unlock all use cases.
Sign up to read complete case studies, access detailed metrics, and unlock all use cases.