
Turn a single image into a realistic talking avatar using audio-driven animation with WAN LongCatAvatar.
Create talking avatar videos from a single image using audio or text input.
Generate natural head movement, lip-sync, and facial expressions while keeping the face identity consistent.
Animate human, anime, or stylized characters for shorts, explainers, or AI presenters.
Combine LongCatAvatar with voice models, TTS, or audio-driven workflows inside ComfyUI.
Run the full avatar generation locally, giving better control, privacy, and customization.
Read-only preview β Use the expand button for a larger view
Ready to use this workflow? Download the JSON file below β free, no signup needed.
Join the discussion
Sign in to leave a comment or reply
No comments yet
Be the first to share your thoughts!