Hedra
  • Developers
  • Studio
  • Enterprise
  • Blog
Log inSign Up
Open Hedra
Your account
    Explore
  • Home
  • Developers
  • Studio
  • Enterprise
  • Blog
    Log inSign Up
    Open Hedra
VEED

VEED Fabric 1.0

All video models
Video modelVEED

Talking video with natural lip-sync and expressive animation.

OverviewText + image + audio to video

Overview

VEED Fabric 1.0 is an image-to-video model by VEED for generating realistic talking avatars. By combining a static image and an audio track, it produces precise lip-sync videos where facial expressions and head movements follow the speech rhythm. It is well-suited for personalized marketing content, educational videos, and social ads. Creators often pair it with image generators like Nano Banana Pro or Seedream 4.0 to design custom characters before animating them.

Pricing

480p 10.71¢/second, 720p 21.43¢/second.

VEED Fabric 1.0 All Inputs

Text + image + audio to video — generates video.

Specifications

Input mode
Text + image + audio to video
Accepts
start frame, audio (required)
Aspect ratios
16:9, 9:16, 4:3, 3:4, 1:1
Resolutions
480p, 720p
Native audio
Audio-driven

Build with this model: VEED Fabric 1.0 API on the Hedra Developer Platform.

Best of VEED Fabric 1.0

A close-up shot of a brown horse with its mouth slightly open, standing in a grassy green pasture with a soft-focus wooden fence behind it. The scene is illuminated by warm golden hour sunlight. This 1312x736 resolution video was generated using the VEED Fabric 1.0 Fast model.Majestic Horse in Sunny Pasture — VEED Fabric 1.0 FastA close-up video frame of a small metallic robot with glowing orange eyes on the lunar surface. The robot arranges glowing golden crystals in a grid pattern on the gray soil. In the dark space background, planet Earth is visible. This video has a 1312x736 resolution, generated using VEED Fabric 1.0.Robot on the Moon, by VEED Fabric 1.0A portrait video frame of an elderly Buddhist monk with a long grey beard and orange robes, gesturing as he speaks inside a temple. A golden Buddha statue and lit candles are visible in the background. Generated using VEED Fabric 1.0 Fast at 736x1312 resolution.Buddhist Monk Gesturing in Temple — VEED Fabric 1.0 FastA vertical video frame showing an elderly Buddhist monk with a long white beard and bald head, wearing saffron robes. He is gesturing with his hands inside a temple, with a golden Buddha statue and candles in the background. Generated using VEED Fabric 1.0 Fast at 736x1312 resolution.Buddhist Monk Speaking in Temple — VEED Fabric 1.0 Fast

What is VEED Fabric 1.0 best used for?

VEED Fabric 1.0 is an image-to-video model specialized in creating lip-synced talking avatars. It excels at turning a single static image—whether a photorealistic human, 3D mascot, clay figure, or illustration—and an audio file into a dynamic video. The model synchronizes mouth movements, head gestures, and body language to match the speech rhythm. It is widely used for marketing videos, educational content, and social media ads where you need realistic speech animation without filming.

Who created Fabric 1.0 and are there other versions?

Fabric 1.0 was developed by the video editing platform VEED and launched via API in mid-2025. Powered by a Diffusion Transformer (DiT) architecture, it processes visual and audio data simultaneously. VEED also offers a speed-optimized variant called VEED Fabric 1.0 Fast, which trades a small amount of visual fidelity for significantly faster generation times and lower inference costs.

How can I get the most realistic or expressive results from this model?

For the best lip-sync quality, ensure your input audio is clear and free of background noise. A popular community workflow involves generating a custom base character using image models like Nano Banana 2 or Nano Banana Pro before animating it with Fabric. Additionally, if you are using VEED's text-to-speech features, you can insert bracketed audio tags—such as [excited], [whisper], or [sigh]—directly into your script to control the avatar's emotional delivery and pacing at the sentence level.

Similar models

Kling AI Avatar v2 ProKlingKling AI Avatar v2 StandardKlingVEED Fabric 1.0 FastVEEDHedra AvatarHedraHedra Character 3HedraHedra OmniaHedra

Prompt tips

  • Optimize the Source Image: Provide a high-resolution, forward-facing portrait with a neutral expression. Avoid images where hands or objects obscure the mouth and jawline.
  • Clean Audio is Crucial: The model relies on clear phoneme detection. Ensure your audio track is free of background noise or heavy echo to prevent mouth jitter.
  • Leverage Emotion Tags: If using text-to-speech, insert bracketed tags like [excited], [whisper], or [confident] in your script to drive dynamic, sentence-level facial expressions.
  • Build Custom Spokespeople: Generate a unique character using an image model like Flux 1.1 Pro or Nano Banana, then use Fabric 1.0 to bring them to life with a cloned voice.

What Will You Create?

Open Creative StudioGet the API
Hedra
Hedra

Product

Developer PlatformStudioEnterpriseSovereignPricing

Resources

Agent documentationDeveloper documentationBlogUse CasesModelsFeedbackChangelogStatus

Company

AboutCareersContactAffiliatesSupportAlternatives

Legal

Privacy PolicyTerms of useAcceptable useCookie PolicyBiometric data policy
LinkedinInstagramDiscord
support@hedra.comHedra 2026 — All rights reserved