A single voice actor couldn't produce enough lines to fully train an AI model...
The model is trained on a massive corpus of existing data and then fine tuned to match the target voice actor. Using less than ~30s of reference audio you can get a pretty decent fine tuning the main issue is that it currently isn't on par with the quality and consistency of an in studio voice actor, especially over long time domains.