Faces and human motion
kling_v3from $1.33















Generate video, images, voice and finished cuts straight from your prompts — in Claude, ChatGPT, Cursor or your editor. Your agent installs it in one paste.
You’ll paste this in the next step.
This opens the Add custom connector dialog — paste the URL and click Add. Free plans allow one.
Kling, Veo, Seedance, FLUX 3, Sora, ElevenLabs, HeyGen and the rest behind one interface, with one request shape. Swapping models is a one-line change.
Faces and human motion
kling_v3from $1.33Cinematic, native audio
veo_3_1from $1.68Long shots, native audio
seedance_2_5from $0.70Frontier, synchronized audio
flux_3from $0.90Physical realism
sora_2$1.05Fast and cheap
ltx_2_19b_distilled$0.53Upload a real screen recording. The AI generates the actor and the scene around it, and the cut is composited against your actual UI — not an approximation of it.
No Topaz, no CapCut, no separate lipsync tool, and no shuffling files between them. The agent runs the chain in one sequence and hands back a finished file.
generate_videotext, image or video ingenerate_speechcast by descriptiontranscriptionword-level timingrender_videoclips, overlays, audiovideo · upscalefrom $0.16videotext-to-video, image-to-video, video-to-video, lipsync, upscaleimagegenerate, edit, upscalespeechtext-to-speech, voice search by descriptionmusicgenerate from a prompttranscriptionword-level timestamps for captionsffmpegtrim, resize, slice, probe, arbitrary commandsrenderTSX composition into a single mp4pipelinemulti-step workflowsVideo takes minutes, and a client that blocks on it times out, retries, and bills you twice. Here a call returns a job id immediately — run them in parallel, poll separately, kill the bad ones before they finish burning credits.
Every generation is priced from a published rule before anything runs. The agent can quote it, check the balance, and wait for a yes — instead of surfacing the number on your invoice.
quoted from the live rulecheck_balance()An illustration of the estimate step. Real figures come from estimate_cost, reading the same pricing rules published on /v2/models.
Every file remembers how it was made — the job, the tool, the model and the exact input. Paste a URL back and the agent recovers the recipe, changes one thing, and re-runs it.
get_lineage(url)prompt changedprompt changedaspect_ratio changedreference changedClaude bills Marketing while Cursor bills Product — the setting is per connection and it persists. Output lands in that team’s library, in whichever folder you point it at.
Start from a published template instead of a blank file, cast a voice by describing it rather than scrolling a list, and get word-level timings back for captions that actually line up.
Search published templates by tag, render one as-is, or pull its TSX source, edit it in the chat, and render your own version.
Describe the read you want and get ranked suggestions, or filter on gender, accent and language. ElevenLabs under the hood.
Score a cut from a prompt, and burn in captions from a transcription that carries a start and end timestamp for every word.