PA

Irodori-TTS-v4.1-Anime

Audioby phasefield-audio·Model page

phasefield-audio's anime-voice fine-tune of Irodori-TTS v4.1 for expressive Japanese text-to-speech.

Share:

Base model

Aratako/Irodori-TTS-v4.1-Small

Model Description

A Japanese text-to-speech model fine-tuned from Aratako/Irodori-TTS-v4.1-Small using anime-style speech data.

The base model's annotation pipeline is not publicly documented, so the fine-tuning data was annotated independently. Consequently, caption conditioning and emoji controls may behave differently from the base model.

Checkpoints

The full-precision checkpoint is available at the repository root.

Quantized variants are provided in the following subdirectories:

  • int8-weight-only
  • int8-dynamic
  • int4-weight-only
  • float8-weight-only
  • float8-dynamic

For inference and installation instructions, see the original Irodori-TTS repository.

License

This model follows the same MIT License and ethical restrictions as the base model.

Author
PA
phasefield-audio
User
phasefield-audio
Details
Downloads0
Likes82
AccessOpen Source
Tasktext-to-speech
Parameters766M
Trending81
Licensemit
CreatedSep 4, 2026
UpdatedSep 4, 2026
View on Hugging Face
Get the full context.

Sign up to read complete case studies, access detailed metrics, and unlock all use cases.

Irodori-TTS-v4.1-Anime — AI Model Details | Applied