Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

Are there different voices? Or only [s1] and [s2] in the examples?


We just clarified in the README, sorry for the confusion ;(

Note that the model was not fine-tuned on a specific voice. Hence, you will get different voices every time you run the model. You can keep speaker consistency by either adding an audio prompt (a guide coming VERY soon - try it with the second example on Gradio or HF Space for now), or fixing the seed.




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: