Instructions to use akhbar/chatterbox-tts-norwegian with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Chatterbox
How to use akhbar/chatterbox-tts-norwegian with Chatterbox:
# pip install chatterbox-tts import torchaudio as ta from chatterbox.tts import ChatterboxTTS model = ChatterboxTTS.from_pretrained(device="cuda") text = "Ezreal and Jinx teamed up with Ahri, Yasuo, and Teemo to take down the enemy's Nexus in an epic late-game pentakill." wav = model.generate(text) ta.save("test-1.wav", wav, model.sr) # If you want to synthesize with a different voice, specify the audio prompt AUDIO_PROMPT_PATH="YOUR_FILE.wav" wav = model.generate(text, audio_prompt_path=AUDIO_PROMPT_PATH) ta.save("test-2.wav", wav, model.sr) - Notebooks
- Google Colab
- Kaggle
Hei Alex!
#5
by mortelil - opened
Hei.
Ja, jeg er Norsk. Hva kan jeg hjelpe deg med?
...
Hei akhbar. Testet denne finetunen nå og den hørtes veldig bra ut! Kan du dele noe om hvilken pipeline du bruker, og hvor mye data du har brukt for å komme lage den?
I'll answer in english in case non-Norwegians stumble in here.
I trained the model using one of the finetuning repositories available on github. If I remember correctly, there are two: One with "standard" finetuning, and one that uses GRPO. I used the one without GRPO.
The model was trained on roughly 6500 hours of data for one epoch.
Thanks for the quick answer!