Move from avatars to full production
HeyGen makes avatar videos. Hedra Agent and Workspaces produce the whole piece with your own presenter, voice, and footage, on avatar and character models served on Hedra's own infrastructure, HeyGen Photo Avatar 4 among them.
export HEDRA_API_KEY="<key_id>:<secret>"
curl -X POST \
"https://api.hedra.com/v3/models/omnihuman-15" \
-H "Authorization: Key $HEDRA_API_KEY" \
-H "Content-Type: application/json" \
-d '{ "input": {
"prompt": "A presenter speaks to camera in a bright kitchen",
"aspect_ratio": "16:9",
"resolution": "1080p",
"start_image": "https://example.com/presenter.png",
"audio": "https://example.com/script.mp3"
} }'A more modern way to work
Skip the model aggregators. Hedra Agent and Workspaces give your team an autonomous agent that supercharges your workflows with a superintelligent understanding of the latest models: brief it in a shared space and it plans, generates, and edits across video, image, speech, and music, with your team and your agents working alongside it.

GPU infrastructure, an inference engine, and post-training research
Hedra operates GPU clusters and the inference engine that runs diffusion and autoregressive media models on them 8x faster. The same research team post-trains models for customers: supervised fine-tuning, LoRA training, and GRPO, on your data.
Hedra and HeyGen: focus at a glance
HeyGen ships avatar products. Hedra serves and post-trains avatar models, HeyGen's included, for teams building their own.
| Dimension | Hedra | HeyGen |
|---|---|---|
| What it is | GPU infrastructure, inference and post-training research | Avatar video products for marketing and localization |
| Core technology | Hedra's inference engine, its own GPU clusters, and post-training (SFT, LoRA, GRPO) | Stock avatars, photo avatars and voice inside HeyGen's products |
| Models | Open-weight models on Hedra's engine, models post-trained for you, and partner models through a gateway | Its own avatar models; Photo Avatar 4 is also on Hedra |
| How you connect | API, SDK, CLI, MCP, and Hedra Workspaces | HeyGen app and API |
| Where inference runs | Hedra's GPU clusters, or your own | HeyGen's cloud |
| Best fit | Teams building visual products who need inference, capacity and custom models | Marketing and localization teams making avatar videos |
Avatar, character and voice models on Hedra infrastructure, HeyGen's included
Video rates are per second of generated performance; voice rates are per 1,000 characters. Open-weight models run on Hedra's engine; partner models are reached through the gateway.
| Provider | Model ID | Rate |
|---|---|---|
| HeyGen | heygen-photo-avatar-4 | 10¢/s · up to 1080p |
| ByteDance | omnihuman-15 | 16¢/s · 1080p |
| Kling | kling-ai-avatar-v2 | 11.5¢/s · 720p pro |
| ElevenLabs | elevenlabs-v3 | 12.6¢ / 1,000 characters |
| MiniMax | minimax-speech-25-hd-preview | 12.86¢ / 1,000 characters |
FAQs
- Is HeyGen Photo Avatar 4 on Hedra the same model HeyGen ships?
- Yes. HeyGen Photo Avatar 4 is HeyGen's model, reached through Hedra's gateway at the rate above with the same request shape as every other model in the catalog. Open-weight avatar models run on Hedra's engine beside it.
- Can Hedra post-train a model for us?
- Yes. Hedra runs supervised fine-tuning, LoRA training, and GRPO on models you own or license, on your data and against your evals, then serves the result on the same engine. The weights are yours.
- Can this run on our own hardware?
- Yes. Hedra Inference deploys on your Kubernetes or bare metal, including air-gapped environments, across NVIDIA, AMD, AWS Trainium, Google TPU, and custom accelerators. Weights, data, and the control plane stay on infrastructure you operate.