Hedra
  • Developers
  • Studio
  • Enterprise
  • Blog
Log inSign Up
Open Hedra
Your account
    Explore
  • Home
  • Developers
  • Studio
  • Enterprise
  • Blog
    Log inSign Up
    Open Hedra
A close-up shot of a brown horse with its mouth slightly open, standing in a grassy green pasture with a soft-focus wooden fence behind it. The scene is illuminated by warm golden hour sunlight. This 1312x736 resolution video was generated using the VEED Fabric 1.0 Fast model.

VEED Fabric 1.0

Talking video with natural lip-sync and expressive animation.

Get an API keyOpen Creative Studio
All models

Overview

VEED Fabric 1.0 is an image-to-video model by VEED for generating realistic talking avatars. By combining a static image and an audio track, it produces precise lip-sync videos where facial expressions and head movements follow the speech rhythm. It is well-suited for personalized marketing content, educational videos, and social ads. Creators often pair it with image generators like Nano Banana Pro or Seedream 4.0 to design custom characters before animating them.

Specifications · Text + image + audio to video

Input mode
Text + image + audio to video
Accepts
start frame, audio (required)
Aspect ratios
16:9, 9:16, 4:3, 3:4, 1:1
Resolutions
480p, 720p
Native audio
Audio-driven

Pricing

VEED
480p10.71¢/second
720p21.43¢/second

Build with this model: the VEED Fabric 1.0 API on the Hedra Developer Platform.

VEED Fabric 1.0 API →

Real output · generated with VEED Fabric 1.0

Best of VEED Fabric 1.0

A close-up shot of a brown horse with its mouth slightly open, standing in a grassy green pasture with a soft-focus wooden fence behind it. The scene is illuminated by warm golden hour sunlight. This 1312x736 resolution video was generated using the VEED Fabric 1.0 Fast model.Majestic Horse in Sunny Pasture — VEED Fabric 1.0 FastA close-up video frame of a small metallic robot with glowing orange eyes on the lunar surface. The robot arranges glowing golden crystals in a grid pattern on the gray soil. In the dark space background, planet Earth is visible. This video has a 1312x736 resolution, generated using VEED Fabric 1.0.Robot on the Moon, by VEED Fabric 1.0A portrait video frame of an elderly Buddhist monk with a long grey beard and orange robes, gesturing as he speaks inside a temple. A golden Buddha statue and lit candles are visible in the background. Generated using VEED Fabric 1.0 Fast at 736x1312 resolution.Buddhist Monk Gesturing in Temple — VEED Fabric 1.0 FastA vertical video frame showing an elderly Buddhist monk with a long white beard and bald head, wearing saffron robes. He is gesturing with his hands inside a temple, with a golden Buddha statue and candles in the background. Generated using VEED Fabric 1.0 Fast at 736x1312 resolution.Buddhist Monk Speaking in Temple — VEED Fabric 1.0 Fast

Prompting

Prompt tips

  • Optimize the Source Image: Provide a high-resolution, forward-facing portrait with a neutral expression. Avoid images where hands or objects obscure the mouth and jawline.
  • Clean Audio is Crucial: The model relies on clear phoneme detection. Ensure your audio track is free of background noise or heavy echo to prevent mouth jitter.
  • Leverage Emotion Tags: If using text-to-speech, insert bracketed tags like [excited], [whisper], or [confident] in your script to drive dynamic, sentence-level facial expressions.
  • Build Custom Spokespeople: Generate a unique character using an image model like Flux 1.1 Pro or Nano Banana, then use Fabric 1.0 to bring them to life with a cloned voice.

About the model

Questions, answered

What is VEED Fabric 1.0 best used for?

VEED Fabric 1.0 is an image-to-video model specialized in creating lip-synced talking avatars. It excels at turning a single static image—whether a photorealistic human, 3D mascot, clay figure, or illustration—and an audio file into a dynamic video. The model synchronizes mouth movements, head gestures, and body language to match the speech rhythm. It is widely used for marketing videos, educational content, and social media ads where you need realistic speech animation without filming.

Who created Fabric 1.0 and are there other versions?

Fabric 1.0 was developed by the video editing platform VEED and launched via API in mid-2025. Powered by a Diffusion Transformer (DiT) architecture, it processes visual and audio data simultaneously. VEED also offers a speed-optimized variant called VEED Fabric 1.0 Fast, which trades a small amount of visual fidelity for significantly faster generation times and lower inference costs.

How can I get the most realistic or expressive results from this model?

For the best lip-sync quality, ensure your input audio is clear and free of background noise. A popular community workflow involves generating a custom base character using image models like Nano Banana 2 or Nano Banana Pro before animating it with Fabric. Additionally, if you are using VEED's text-to-speech features, you can insert bracketed audio tags—such as [excited], [whisper], or [sigh]—directly into your script to control the avatar's emotional delivery and pacing at the sentence level.

Same API · one key

Similar models

A medium shot of a brown bay horse with a dark mane standing in a green meadow filled with wildflowers. In the background, soft-focus mountains rise under a blue sky with light clouds. This video is generated by Kling AI Avatar v2 Pro at 1920x1072 resolution.Kling AI Avatar v2 ProKlingA man in a cream linen suit, patterned tie, and dark sunglasses stands on a stone balcony overlooking a Mediterranean coastal town with turquoise water. The shot is a medium close-up, generated using Kling AI Avatar v2 Standard at 1280x720 resolution. He is gesturing with his right hand.Kling AI Avatar v2 StandardKlingA close-up shot of a brown horse with its mouth slightly open, standing in a grassy green pasture with a soft-focus wooden fence behind it. The scene is illuminated by warm golden hour sunlight. This 1312x736 resolution video was generated using the VEED Fabric 1.0 Fast model.VEED Fabric 1.0 FastVEEDA vertical 768x1152 talking-animal video of a huge muscular humanoid cheetah bodybuilder in a gym, delivering an intense coaching line to camera. Generated from a candid-style still and an audio track using Hedra Avatar.Hedra AvatarHedraA widescreen 1080x720 seated talking-animal video of a large bear in a work shirt, delivering a gruff, matter-of-fact roofing estimate to camera. Generated from a still image and an audio track using Hedra Character 3.Hedra Character 3HedraHedra OmniaHedra

What will you create?

Get an API keyOpen Creative Studio
Hedra
Hedra

Product

Developer PlatformStudioEnterpriseSovereignPricing

Resources

Agent documentationDeveloper documentationBlogUse CasesModelsFeedbackChangelogStatus

Company

AboutCareersContactBuilder programSupportAlternatives

Legal

Privacy PolicyTerms of useAcceptable useCookie PolicyBiometric data policy
LinkedinInstagramDiscord
support@hedra.comHedra 2026 — All rights reserved