A flow-matching text-to-speech model by SWivid offering high-quality voice synthesis with natural expressiveness.
No resource links recorded.