-
@weare_so_back stts2 is using yl's libritts model, doing one-shot voice cloning, the cloning quality isnt great, so the actual voice timbre is changed using RVC. and as usual, RVC models can be trained with like 5 minutes of data. This one is Arona from Blue Archive.
AuroraNemoia’s Twitter Archive—№ 1,442