Andrew Adams · Co-Founder & Operations at Wireflow · AI Talking Photo
Most talking photo tools hide the model behind one button.
Wireflow shows the graph: a Script node feeds an ElevenLabs voice, and a HeyGen Avatar4 node animates the Portrait Photo into a short clip with the lips synced to that voice. Free to build, pay per generation.
Free to build · no credit card

Why animate a photo on a canvas
Most talking photo tools are a single upload box wired to a hidden model: drop a face in, press a button, and accept whatever clip comes back. When the motion is stiff or the framing is wrong, there is nowhere to turn, no view of how the video was made, and no way to reuse the recipe on the next portrait. The upload box is the whole interface, and its ceiling arrives fast.
Wireflow takes a different shape. A talking photo is a small workflow you own on a node canvas: a Portrait Photo node holds the still, a Script node holds the words, an ElevenLabs node turns them into a voice, and a HeyGen Avatar4 node animates the photo to that voice. It runs on hosted compute in the browser with nothing to install, and the same canvas drives the rest of your video work, from an AI avatar generator to full multi-clip pipelines.
What the talking photo flow gives you
Portrait input
An image node takes your front-facing still and passes it to the video model as the start frame.
Script
A text node holds what the photo says, in plain words, and an ElevenLabs node turns it into a voice.
Video render
The HeyGen Avatar4 node reads the photo and the voice and returns a short talking-head clip at 1:1.
Swappable models
Drop Kling AI Avatar or VEED Fabric into the same slot when a look calls for it.
Change the voice
Pick another ElevenLabs voice, or swap in a cloned voice, and the mouth stays synced to the new audio.
Share by link
The workflow is versioned and shareable, so a teammate runs the exact same talking photo graph.
How the talking head graph runs
The workflow behind this page's button is deliberately small: four nodes and three wires.
- Portrait Photo holds the still. One image node takes your front-facing photo and feeds it into the video model as the face that speaks.
- Script holds the words. A block of text says what the photo should say, and it wires into the voice node.
- ElevenLabs voices the script. The voice node turns the words into a speech track and passes it to the video model.
- HeyGen Avatar4 animates it. The video node reads the photo and the speech track and returns a short clip at 1:1.
The motion in this flow is driven by the audio track, so the lips sync to the specific words in your script. To narrate in your own voice, swap the stock voice for a cloned one on the same canvas. And because the graph lives among the platform's hosted models, the clip can roll straight into more work, like an animated still image sequence, without leaving the browser.
When another tool is the better call
If you want a one button phone app that turns a selfie into a canned talking clip with fixed styles and no canvas to touch, a consumer app is built for exactly that, and this page will not pretend otherwise. Wireflow gives you the graph instead, which is more control than a casual one off clip needs.
Two honest limits matter here. First, this flow speaks in a stock ElevenLabs voice, so narration in your own voice needs a cloned voice in place of the stock one. Second, because it animates a real uploaded photo, only use portraits you have the rights to, and never make a talking video of a real person without their consent. To generate a synthetic presenter from scratch instead, start with an AI face generator flow.
More Than Just AI Talking Photo
See the whole pipeline
The flow sits in the open: a Portrait Photo and an ElevenLabs voice made from your script feed the HeyGen Avatar4 node. Rewire it like any AI video generator graph.

One photo, many clips
Reuse a single headshot across versions by editing the script and rerunning. Chain it into an image to video pipeline.

Swap the video model
HeyGen Avatar4 animates by default; drop Kling AI Avatar or VEED Fabric into the same slot. Reuse it inside an AI avatar generator flow.

Add a real voice
This flow already speaks your script in a stock ElevenLabs voice. For narration in your own voice, start from AI voice cloning. If the clip comes from Seedance, our walkthrough on adding your real voice to a Seedance 2.5 video covers the audio chain.

Pay per generation, build free
Building the graph on the canvas costs nothing; you pay only when a clip renders. Iterate like an AI image animator with no monthly seat.

Build Any AI Workflow
AI Models Integrated
Full Commercial License
FAQs
It is a short video made from a single still portrait where the subject appears to move and speak. On Wireflow it is a node workflow: an ElevenLabs node voices your script, and a HeyGen Avatar4 node animates the Portrait Photo to that voice.

Written by
Andrew Adams · Co-Founder & Operations at Wireflow
Runs client operations and content strategy at Wireflow. Works directly with creative teams and agencies to build production AI workflows.
Make a portrait talk, on a graph you control
Open the flow, upload a portrait, write the script, and run it: a short talking-head clip with synced lips from a graph you can rerun and reshape. Building is free; you pay per generation, not per month.