Hedra AI: Character 3 tutorial, pricing & Omnia comparison

Handmade collage showing a portrait becoming an animated character video with audio

Quick answer

Hedra Character 3 turns a portrait plus audio into a talking or singing video up to ten minutes long. Its current API route lists 540p at $0.025 per second, 720p at $0.05, and 1080p at $0.0625. Hedra Omnia is a separate newer model for more dynamic character scenes; it does not make Character 3's long-form portrait workflow obsolete. Prepare a clear portrait and clean audio, run a short sample, then verify the complete output and charge before producing a series. Facts and prices checked September 13, 2026.

Create a talking video

Upload an image and audio, then review the animated result in Magic Hour.

Try AI Talking Photo

Magic Hour publishes this guide and offers a separate talking-photo workflow. Hedra product, model, duration, resolution, and API claims are based on the linked official Hedra sources checked September 13, 2026; this is a documentation review rather than a retained output-quality benchmark.

Magic Hour publishes this guide and includes its own product in the comparison. Treat the recommendations as editorial guidance from a vendor, and verify the linked first-party product details and your own output requirements before choosing a tool.

What is Hedra Character 3?

Character 3 is Hedra's audio-driven character-video model. Its current model page lists a start image and required audio, with 540p, 720p, and 1080p output options. The older statement that Character 3 is limited to 720p is no longer consistent with that page.

Use it for a speaking or singing character when the source is a still image. Confirm the settings available in your account rather than assuming that every model on the platform has the same input, duration, or output limits.

Character 3 versus Hedra Omnia

Character 3 is Hedra's long-form audio-driven portrait model. It requires a start frame and audio, supports 540p, 720p and 1080p, and accepts up to 600 seconds of audio on the current API route. Start here when the deliverable is a presenter, singer, podcast clip or other speech-driven character performance.

Hedra Omnia launched in February 2026 as Hedra's more advanced model for character dialogue in dynamic environments with camera control. Evaluate it when the scene needs more than a largely portrait-led performance. Hedra currently presents Omnia and Character 3 as separate choices, so record the selected model instead of calling every Hedra output “Character 3.”

Hedra pricing: current monthly plans

Plan

Monthly USD price

Included monthly credits

Basic

$15

1,500

Creator

$30

5,400

Professional

$75

14,400

Enterprise

Custom

Custom

Hedra's pricing page currently lists Basic at $15 monthly with 1,500 credits, Creator at $30 with 5,400, and Professional or Teams at $75 with 14,400. The page lists commercial use on paid plans. Hedra's current video-generator terms say free-plan outputs may be watermarked and are non-commercial. Check the selected model's quoted credit use and your account conditions before production.

Do not turn plan credits into a promised number of videos without the selected model's current quote. Duration, output settings, and repeated attempts affect the amount of work your balance can cover.

Character 3 API cost per finished attempt

The current Character 3 API page lists $0.025 per second at 540p, $0.05 at 720p, and $0.0625 at 1080p. A 60-second attempt is therefore $1.50, $3.00, or $3.75 before retries. A ten-minute 1080p attempt is $37.50 before retries. These are API rates for this model, not a conversion from Hedra subscription credits.

Budget by accepted output. If three one-minute 720p attempts are needed to approve one clip, the generation cost is $9 before any voice, editing, storage or review work. Save the model, resolution, audio duration, attempt count and actual charge with each result.

How to create the first talking video

1. Open Hedra and check the selected workflow

Sign in through Hedra, open Creative Studio, and choose the talking-character workflow and model. Interface labels can change; the required assets are the useful anchor: a starting image and an approved audio track.

2. Prepare one clear portrait

Use a well-lit face with unobstructed eyes and mouth. Leave room for movement around the head and chin. If you create a character image first, save the approved image instead of repeatedly prompting for an approximate replacement.

3. Approve the audio before animation

Record your line or generate speech, then listen to it by itself. Correct pronunciation, pauses, and the message before using it as an input. Hedra's photo-to-video guide describes this image-and-audio workflow.

4. Check the model, resolution, and charge

Choose settings appropriate to the deliverable and review the quoted credit use. Begin with a short segment that contains the expressions and pacing you need. A low-cost preview can screen the inputs, but it does not establish the quality of a different final mode.

5. Generate, review, and download

Watch the whole output at normal speed. Inspect the mouth at pauses, the eyes during movement, and the face shape through the sentence. Confirm the downloaded file's dimensions and export conditions before making a series.

A sample brief for a product explainer

Use a licensed or original presenter image and a short line such as: “Here's how the cap opens. Turn it once, then lift.” Time the actual recording and make the instruction match the real product.

Pair the presenter with footage of the action. The talking character introduces the demonstration; it should not invent evidence that the product works. Do not present a generated actor as a real customer describing a purchase they never made.

Keep these three assets together: the approved image, the final voice recording, and the visible demonstration. When the instructions change, update the relevant asset and review the combined video. For longer structures, use the product video script templates.

Estimate the cost of a series

Suppose you need ten clips and each attempt receives a quote of Q credits. One attempt per clip costs 10Q; two attempts per clip cost 20Q. If only eight of those clips are usable, the cost per usable clip is the total spent divided by eight.

This is planning arithmetic, not a Hedra rate or an observed success rate. Record actual charges and usable outputs in your project. A subscription price alone does not show the cost of a finished campaign.

Troubleshoot before repeating a generation

Problem

First revision to try

Unclear mouth or facial distortion

A sharper, less obstructed portrait

Awkward pronunciation

Correct the voice recording before animating again

Too much head movement

A calmer delivery and simpler starting pose; review available motion controls

The face changes between clips

Reuse the exact approved source image and compare every output

Missing final words

Check the chosen audio segment and output duration

The result looks fine but cannot be delivered

Recheck the export resolution, watermark, and account entitlement

Do not keep repeating an unchanged setup after the same failure. Isolate the image, audio, or setting responsible, or choose a different workflow.

Hedra vs Magic Hour vs HeyGen

Your task

Useful starting comparison

Animate a portrait from audio

Hedra Character 3 and Magic Hour Talking Photo both fit this input

Replace speech in an existing face video

Start with Magic Hour Lip Sync or a dedicated video-lip-sync workflow

Create a presenter and translated versions

Compare HeyGen's documented avatar and translation features

Generate a scene without an existing photo or video

Compare the selected text-to-video models, not the talking-photo tools

Choose on your actual input, export, and workflow requirements. This guide does not establish that one provider has the most accurate lips, fastest processing, or lowest cost for every project. For a broader shortlist, see best AI talking-photo tools.

Choose Hedra for the character job it solves

Direct answer: Hedra belongs on the shortlist when audio-driven character performance is central. Compare it with broader talking-photo and video platforms when you also need generation, editing, multiple models or production handoffs.

Test. Use one portrait and 15-second script with pauses, emphasis and a mild head turn. Record model version, voice source, resolution, duration, queue and retries.

Review. Check identity, lips, teeth, blinking, head motion, framing and whether performance matches the audio. Longer clips reveal drift that a short showcase can hide.

Decision rule. Choose a specialist when character performance is the whole job. Choose a broader workspace when talking-photo generation is one step in a larger image, video and audio workflow.

Try the workflow: Open the matching Magic Hour tool. Use the same representative input and acceptance criteria before comparing results.

Frequently asked questions

Hedra offers a free starting route. Its current video-generator page says free outputs may include a watermark and are limited to non-commercial use. The public pricing page does not promise one fixed recurring free-credit allowance, so check the balance, eligible model, charge and download shown in your account.

Yes, the current Character 3 model page lists 1080p alongside 540p and 720p. Check availability for the selected workflow and account; this does not establish the same limit for every Hedra model.

No. A rendered clip is a finished video. An interactive avatar also needs a streaming, voice, and response system. Confirm the current developer product and its billing separately; do not use an old streaming rate to budget ordinary video generations.

Yes. Magic Hour Talking Photo takes an image and audio directly. Magic Hour also offers text-to-video, image-to-video, and editing workflows. It is inaccurate to describe the platform as only transforming existing footage.

Evaluate the available sample first. Choose a paid tier when the result fits your project and you need its credits or export conditions. Confirm that your image, voice, and script can be used for the intended purpose.

Aastha Kochar - author at MagicHour (SaaS MarTech Content Writer)
Aastha Kochar
Content Manager
Aastha Kochar has spent 5+ years creating content for B2B and B2C SaaS brands in the AI and MarTech space. She is well-versed with AI-powered content tools and offers deep comparisons after trying and testing every tool. Her work has helped companies increase organic traffic, earn AI citations, and most importantly — turn readers into users. With a bachelor's and master's degree in Journalism and Mass Communication, she brings strong research skills, authentic storytelling, and a deep understanding of what makes audiences actually care about what they're reading.
View author →

Continue Reading

Magic Hour editorial collage comparing AI lip sync with phoneme drawings and a speech waveform
7 best AI lip sync video tools (2026 comparison)
AI Talking Photo Tools (2026)
Best AI talking photo tools (2026): 5 portrait workflows
6 Best Free AI Lip Sync Tools
Best free AI lip sync tools (2026): limits & watermarks
How to Lip Sync a Video With AI
How to lip sync a video with Magic Hour: 5 steps
Top 6 Best Talking Photo APIs
5 best talking-photo APIs and endpoints for developers
Product Video Script Templates
Product video script templates: 12 formats you can adapt