I'm looking for the best current option for a pretty specific workflow.
I want to record a short video of myself on my phone, then use a reference image of a photorealistic AI character to replace me in the video.
The important part is that I want to preserve as much of my original performance as possible — facial expressions, lip movements, head tilts, hand gestures, body movement and timing — while making the person in the final video consistently look like the same AI character.
I'm not looking for a basic face swap or a talking avatar from a still image. I want my real recorded performance to drive the AI character.
Ideally I'm looking for something cloud/web based rather than a complicated local ComfyUI setup.
For anyone actually doing this currently: what model/platform are you getting the best results with?
I'm especially interested in:
•photorealism
•character consistency between clips
•facial expression/lip movement preservation
•how many rerolls it usually takes to get a usable result
•actual cost per usable clip
Recent experiences/examples would be really appreciated since these models are changing so quickly.