VEED Fabric 1.0 is VEED's talking-head & lipsync video model on Versely. This page is its structured prompting reference: the 3 parameters its schema actually exposes, the talking-video technique that applies to it, copy-ready templates.
Everything here is grounded in the same sources Versely's agent reads — the model's input schema. Where a line is general craft advice rather than a documented fact about VEED Fabric 1.0, the page says so.
What VEED Fabric 1.0 wants
The exact input surface, from the same schema the Versely agent fetches with get_model_input_schema before every generation.
| Parameter | What it does | Values |
|---|---|---|
image_urlreq | Source character image (URL maxLength 2083) | string (max 2083) |
audio_urlreq | Driving audio (URL maxLength 2083) | string (max 2083) |
resolutionreq | Resolution — REQUIRED, no default per fal docs | 480p · 720p |
Technique that applies here
Lipsync/avatar/audio-driven: text or audio input, delivery and framing controls
- Three different input shapes share this family, and figuring out which one your model uses tells you what 'the prompt' even means: text-is-the-script (HeyGen Avatar V3, VEED Avatars — your text becomes the spoken words), audio-drives-it (Kling Avatar Pro, LTX 2.3 Audio to Video, Wan 2.2 Speech to Video — the words come from your uploaded audio_url; any prompt text only shapes the visual scene/motion around it), or pure re-timing with no prompt field at all (Sync Lipsync 2.0, VEED Lipsync, VEED Fabric 1.0, Sync React 1).
- Kling Avatar Pro's input.prompt is required (max 5000 chars), but the schema describes it as a 'Motion/scene description,' not dialogue — writing your script into it does nothing, since input.audio_url is what the avatar actually lip-syncs to. Wan 2.2 Speech to Video has the same audio-drives-it shape but makes its equivalent prompt field optional.
- Sync React 1 replaces a text prompt with enum dials: emotion (required, one word only, from happy/angry/sad/neutral/disgusted/surprised), model_mode (lips/face/head — how much of the frame reacts, default face), temperature (0–1 — how exaggerated the reaction is, default 0.5), and lipsync_mode (default bounce). Sync Lipsync 2.0 exposes that same length-mismatch choice as sync_mode (cut_off/loop/bounce/silence/remap), defaulting to cut_off instead.
Copy-ready templates
Replace the bracketed slots; each template says when it's the right shape.
voice.prompt: "[SCRIPT — what the avatar says, written as natural spoken sentences]" · character.avatar: [AVATAR ID], voice.voice: [VOICE ID], resolution: [720p/1080p], output_language: [LANGUAGE CODE if dubbing]
Use when: Scripting a HeyGen Avatar V3 talking head from text — no separate audio file needed, HeyGen generates the voice.
Source [video/image]: [description of the face and framing] · Driving audio: [description and length] · [sync_mode or resolution setting]: [VALUE]
Use when: Briefing a Sync Lipsync 2.0, VEED Lipsync, or VEED Fabric 1.0 job — none of them have a creative prompt field, so your only real levers are which source and audio you pair, plus the length-mismatch or resolution setting.
How the Versely agent does this automatically
You can use this page by hand, or let the agent apply the same knowledge. Four real mechanisms — no more, no less:
get_model_input_schema— before generating, the agent looks up VEED Fabric 1.0's exact input fields, required fields, allowed values, defaults, and min/max bounds. The parameter table above is that same surface.- The prompt enhancer's family rules — 12 per-family rewrite rules (this model's family isn't one of the 12, so only general enhancement applies) shape how a rough prompt gets rewritten.
- The per-provider speech guide — for TTS scripts, the agent follows a provider-specific tag scheme — not relevant to this model, but it's why voiceover scripts come out marked up correctly.
expand_movie_scene— in movie flows, brief scene ideas are rewritten into detailed cinematic descriptions before generation.
Mistakes that waste generations
- Typing your script into Kling Avatar Pro's or Wan 2.2 Speech to Video's prompt field and expecting the avatar to say it — both prompts are scene/motion descriptions; the spoken words come only from audio_url.
- Writing a phrase like 'a little sad but trying to smile' into Sync React 1's emotion field — the schema requires exactly one word from a fixed 6-value enum (happy/angry/sad/neutral/disgusted/surprised); anything else is invalid.
- Leaving VEED Fabric 1.0's resolution unset — unlike most other models in this family, it has no default; the schema marks it required (480p or 720p only), so an empty value fails instead of falling back to a default.
This guide also covers
These siblings share VEED Fabric 1.0's prompting-relevant input surface, so their prompting URLs resolve here — tier and pricing differences live on their own model pages:
Frequently asked questions
Does VEED Fabric 1.0 support negative prompts?+
No — VEED Fabric 1.0's published schema has no negative_prompt parameter. Exclusions have to be phrased positively inside the main prompt, or dropped.
How does the Versely agent know VEED Fabric 1.0's parameters?+
Before generating, the agent calls its get_model_input_schema tool, which looks up the exact input fields, required fields, allowed values, defaults, and min/max bounds for the model. Nothing on this page is guessed — it is the same schema surface those tools read.
Does this guide also cover VEED Fabric 1.0 Text?+
Yes. VEED Fabric 1.0 Text share the same prompting-relevant input surface as VEED Fabric 1.0, so their prompting URLs redirect here instead of duplicating this page. Tier and pricing differences live on each model's own /models page.
Related prompting guides
Generate with VEED Fabric 1.0
VEED Fabric 1.0 is live in Versely — paste a template above, or just describe what you want and let the agent map it onto the schema for you.