Irodori-TTS-v4.1-Anime

A Japanese text-to-speech model fine-tuned from Aratako/Irodori-TTS-v4.1-Small using anime-style speech data.

The base model's annotation pipeline is not publicly documented, so the fine-tuning data was annotated independently. Consequently, caption conditioning and emoji controls may behave differently from the base model.

Checkpoints

The full-precision checkpoint is available at the repository root.

Quantized variants are provided in the following subdirectories:

  • int8-weight-only
  • int8-dynamic
  • int4-weight-only
  • float8-weight-only
  • float8-dynamic

For inference and installation instructions, see the original Irodori-TTS repository.

License

This model follows the same MIT License and ethical restrictions as the base model.

Downloads last month

-

Downloads are not tracked for this model. How to track
Safetensors
Model size
0.8B params
Tensor type
F32
ยท
Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for phasefield-audio/Irodori-TTS-v4.1-Anime

Space using phasefield-audio/Irodori-TTS-v4.1-Anime 1