Hey! I was looking into Fish speech and it looks like it supports having multiple references with audio/text https://docs.fish.audio/api-reference/sdk/javascript/api-reference#ttsrequest
it would be neat if it was possible to pass multiple --voice-ref and --reference-text, I don't believe all backends support this so maybe just pass through the first entry only for those?
Hey! I was looking into Fish speech and it looks like it supports having multiple references with audio/text https://docs.fish.audio/api-reference/sdk/javascript/api-reference#ttsrequest
it would be neat if it was possible to pass multiple
--voice-refand--reference-text, I don't believe all backends support this so maybe just pass through the first entry only for those?