

Hedra Character 3 turns a portrait plus audio into a talking or singing video up to ten minutes long. Its current API route lists 540p at $0.025 per second, 720p at $0.05, and 1080p at $0.0625. Hedra Omnia is a separate newer model for more dynamic character scenes; it does not make Character 3's long-form portrait workflow obsolete. Prepare a clear portrait and clean audio, run a short sample, then verify the complete output and charge before producing a series. Facts and prices checked September 13, 2026.
Upload an image and audio, then review the animated result in Magic Hour.
Try AI Talking PhotoCharacter 3 is Hedra's audio-driven character-video model. Its current model page lists a start image and required audio, with 540p, 720p, and 1080p output options. The older statement that Character 3 is limited to 720p is no longer consistent with that page.
Use it for a speaking or singing character when the source is a still image. Confirm the settings available in your account rather than assuming that every model on the platform has the same input, duration, or output limits.
Character 3 is Hedra's long-form audio-driven portrait model. It requires a start frame and audio, supports 540p, 720p and 1080p, and accepts up to 600 seconds of audio on the current API route. Start here when the deliverable is a presenter, singer, podcast clip or other speech-driven character performance.
Hedra Omnia launched in February 2026 as Hedra's more advanced model for character dialogue in dynamic environments with camera control. Evaluate it when the scene needs more than a largely portrait-led performance. Hedra currently presents Omnia and Character 3 as separate choices, so record the selected model instead of calling every Hedra output “Character 3.”
Plan | Monthly USD price | Included monthly credits |
|---|---|---|
Basic | $15 | 1,500 |
Creator | $30 | 5,400 |
Professional | $75 | 14,400 |
Enterprise | Custom | Custom |
Hedra's pricing page currently lists Basic at $15 monthly with 1,500 credits, Creator at $30 with 5,400, and Professional or Teams at $75 with 14,400. The page lists commercial use on paid plans. Hedra's current video-generator terms say free-plan outputs may be watermarked and are non-commercial. Check the selected model's quoted credit use and your account conditions before production.
Do not turn plan credits into a promised number of videos without the selected model's current quote. Duration, output settings, and repeated attempts affect the amount of work your balance can cover.
The current Character 3 API page lists $0.025 per second at 540p, $0.05 at 720p, and $0.0625 at 1080p. A 60-second attempt is therefore $1.50, $3.00, or $3.75 before retries. A ten-minute 1080p attempt is $37.50 before retries. These are API rates for this model, not a conversion from Hedra subscription credits.
Budget by accepted output. If three one-minute 720p attempts are needed to approve one clip, the generation cost is $9 before any voice, editing, storage or review work. Save the model, resolution, audio duration, attempt count and actual charge with each result.
Sign in through Hedra, open Creative Studio, and choose the talking-character workflow and model. Interface labels can change; the required assets are the useful anchor: a starting image and an approved audio track.
Use a well-lit face with unobstructed eyes and mouth. Leave room for movement around the head and chin. If you create a character image first, save the approved image instead of repeatedly prompting for an approximate replacement.
Record your line or generate speech, then listen to it by itself. Correct pronunciation, pauses, and the message before using it as an input. Hedra's photo-to-video guide describes this image-and-audio workflow.
Choose settings appropriate to the deliverable and review the quoted credit use. Begin with a short segment that contains the expressions and pacing you need. A low-cost preview can screen the inputs, but it does not establish the quality of a different final mode.
Watch the whole output at normal speed. Inspect the mouth at pauses, the eyes during movement, and the face shape through the sentence. Confirm the downloaded file's dimensions and export conditions before making a series.
Use a licensed or original presenter image and a short line such as: “Here's how the cap opens. Turn it once, then lift.” Time the actual recording and make the instruction match the real product.
Pair the presenter with footage of the action. The talking character introduces the demonstration; it should not invent evidence that the product works. Do not present a generated actor as a real customer describing a purchase they never made.
Keep these three assets together: the approved image, the final voice recording, and the visible demonstration. When the instructions change, update the relevant asset and review the combined video. For longer structures, use the product video script templates.
Suppose you need ten clips and each attempt receives a quote of Q credits. One attempt per clip costs 10Q; two attempts per clip cost 20Q. If only eight of those clips are usable, the cost per usable clip is the total spent divided by eight.
This is planning arithmetic, not a Hedra rate or an observed success rate. Record actual charges and usable outputs in your project. A subscription price alone does not show the cost of a finished campaign.
Problem | First revision to try |
|---|---|
Unclear mouth or facial distortion | A sharper, less obstructed portrait |
Awkward pronunciation | Correct the voice recording before animating again |
Too much head movement | A calmer delivery and simpler starting pose; review available motion controls |
The face changes between clips | Reuse the exact approved source image and compare every output |
Missing final words | Check the chosen audio segment and output duration |
The result looks fine but cannot be delivered | Recheck the export resolution, watermark, and account entitlement |
Do not keep repeating an unchanged setup after the same failure. Isolate the image, audio, or setting responsible, or choose a different workflow.
Your task | Useful starting comparison |
|---|---|
Animate a portrait from audio | Hedra Character 3 and Magic Hour Talking Photo both fit this input |
Replace speech in an existing face video | Start with Magic Hour Lip Sync or a dedicated video-lip-sync workflow |
Create a presenter and translated versions | |
Generate a scene without an existing photo or video | Compare the selected text-to-video models, not the talking-photo tools |
Choose on your actual input, export, and workflow requirements. This guide does not establish that one provider has the most accurate lips, fastest processing, or lowest cost for every project. For a broader shortlist, see best AI talking-photo tools.
Hedra offers a free starting route. Its current video-generator page says free outputs may include a watermark and are limited to non-commercial use. The public pricing page does not promise one fixed recurring free-credit allowance, so check the balance, eligible model, charge and download shown in your account.
Yes, the current Character 3 model page lists 1080p alongside 540p and 720p. Check availability for the selected workflow and account; this does not establish the same limit for every Hedra model.
No. A rendered clip is a finished video. An interactive avatar also needs a streaming, voice, and response system. Confirm the current developer product and its billing separately; do not use an old streaming rate to budget ordinary video generations.
Yes. Magic Hour Talking Photo takes an image and audio directly. Magic Hour also offers text-to-video, image-to-video, and editing workflows. It is inaccurate to describe the platform as only transforming existing footage.
Evaluate the available sample first. Choose a paid tier when the result fits your project and you need its credits or export conditions. Confirm that your image, voice, and script can be used for the intended purpose.
