Hedra
  • Developers
  • Inference
  • Studio
  • Enterprise
Log inSign Up
Open Hedra
Your account
    Explore
  • Home
  • Developers
  • Inference
  • Studio
  • Enterprise
    Log inSign Up
    Open Hedra

Move from avatars to full production

HeyGen makes avatar videos. Hedra Agent and Workspaces produce the whole piece with your own presenter, voice, and footage, on avatar and character models served on Hedra's own infrastructure, HeyGen Photo Avatar 4 among them.

Get an API keyRead the docs
POST /v3/models/omnihuman-15
export HEDRA_API_KEY="<key_id>:<secret>"

curl -X POST \
  "https://api.hedra.com/v3/models/omnihuman-15" \
  -H "Authorization: Key $HEDRA_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{ "input": {
      "prompt": "A presenter speaks to camera in a bright kitchen",
      "aspect_ratio": "16:9",
      "resolution": "1080p",
      "start_image": "https://example.com/presenter.png",
      "audio": "https://example.com/script.mp3"
  } }'
1080p · 60s · 16¢/second · $9.60 per clip

A more modern way to work

Skip the model aggregators. Hedra Agent and Workspaces give your team an autonomous agent that supercharges your workflows with a superintelligent understanding of the latest models: brief it in a shared space and it plans, generates, and edits across video, image, speech, and music, with your team and your agents working alongside it.

Hedra Workspaces home: a composer that says start with an idea, research and generation templates, and a row of shared spaces
Built onSPACES · AGENTS · VIDEO · IMAGE · SPEECH · MUSIC · EDIT VIDEO

GPU infrastructure, an inference engine, and post-training research

Hedra operates GPU clusters and the inference engine that runs diffusion and autoregressive media models on them 8x faster. The same research team post-trains models for customers: supervised fine-tuning, LoRA training, and GRPO, on your data.

An abstract microchip architecture animation
Runs onNVIDIA CUDA · AMD ROCm · GOOGLE TPU · AWS TRAINIUM · CUSTOM ASICS

Hedra and HeyGen: focus at a glance

HeyGen ships avatar products. Hedra serves and post-trains avatar models, HeyGen's included, for teams building their own.

Comparison of Hedra and HeyGen by focus area
DimensionHedraHeyGen
What it isHedraGPU infrastructure, inference and post-training researchHeyGenAvatar video products for marketing and localization
Core technologyHedraHedra's inference engine, its own GPU clusters, and post-training (SFT, LoRA, GRPO)HeyGenStock avatars, photo avatars and voice inside HeyGen's products
ModelsHedraOpen-weight models on Hedra's engine, models post-trained for you, and partner models through a gatewayHeyGenIts own avatar models; Photo Avatar 4 is also on Hedra
How you connectHedraAPI, SDK, CLI, MCP, and Hedra WorkspacesHeyGenHeyGen app and API
Where inference runsHedraHedra's GPU clusters, or your ownHeyGenHeyGen's cloud
Best fitHedraTeams building visual products who need inference, capacity and custom modelsHeyGenMarketing and localization teams making avatar videos

Avatar, character and voice models on Hedra infrastructure, HeyGen's included

Video rates are per second of generated performance; voice rates are per 1,000 characters. Open-weight models run on Hedra's engine; partner models are reached through the gateway.

Models available through Hedra and their rates
ProviderModel IDRate
HeyGenModel IDheygen-photo-avatar-4Rate10¢/s · up to 1080p
ByteDanceModel IDomnihuman-15Rate16¢/s · 1080p
KlingModel IDkling-ai-avatar-v2Rate11.5¢/s · 720p pro
ElevenLabsModel IDelevenlabs-v3Rate12.6¢ / 1,000 characters
MiniMaxModel IDminimax-speech-25-hd-previewRate12.86¢ / 1,000 characters

FAQs

Is HeyGen Photo Avatar 4 on Hedra the same model HeyGen ships?
Yes. HeyGen Photo Avatar 4 is HeyGen's model, reached through Hedra's gateway at the rate above with the same request shape as every other model in the catalog. Open-weight avatar models run on Hedra's engine beside it.
Can Hedra post-train a model for us?
Yes. Hedra runs supervised fine-tuning, LoRA training, and GRPO on models you own or license, on your data and against your evals, then serves the result on the same engine. The weights are yours.
Can this run on our own hardware?
Yes. Hedra Inference deploys on your Kubernetes or bare metal, including air-gapped environments, across NVIDIA, AMD, AWS Trainium, Google TPU, and custom accelerators. Weights, data, and the control plane stay on infrastructure you operate.
Kling V3 Standard: Woman Presenting in a Kitchen

Infrastructure for visual intelligence.

Get an API keyRead the docs
Hedra
Hedra

Product

DevelopersStudioEnterpriseInferencePricing

Resources

Agent documentationDeveloper documentationBlogUse CasesModelsFeedbackChangelogStatus

Company

AboutCareersContactBuilder programSupportAlternatives

Legal

Privacy PolicyTerms of useAcceptable useCookie PolicyBiometric data policy
LinkedinInstagramDiscord
support@hedra.comHedra 2026 — All rights reserved