Picsart API Platform
← All models

ByteDance OmniHuman

ByteDance · Video · Image → Video

Animate a portrait with realistic body movement driven by audio.

Model ID: bytedance-omnihuman-v1.5Workflow: bytedance/omnihuman/v1.5
Image InputAudio Input1080p
Try on Playground ↗

Overview

Alongside the unified model APIs, we expose compatibility APIs that take each vendor's original parameters exactly as the vendor defines them — nothing renamed, nothing reshaped on the way through.

That makes them the shortest path onto Picsart if you already work with the vendor directly: the request bodies you've already written keep working as they are. You keep their parameter names and defaults, and you get the vendor's full parameter set rather than the subset that is shared across every model.

The cost is that a request is written for one vendor — switching models later means rewriting it, and results come back in the vendor's own shape. When you'd rather write once and change models freely, use the unified API.

Make a request

Call the workflow with ai.apis.run() in TypeScript, or hit /workflows/{workflow}/execute directly — params is passed through untouched either way.

These endpoints are addressed by workflow name rather than model id: the one in this model's header, since one model runs as one workflow. Authentication is unchanged — your Picsart API key as a bearer token (see Authentication).

ts
import { createClient, ApiRunMode } from '@picsart/ai-sdk';

const ai = createClient({
  apiKey: process.env.PICSART_API_KEY,
  apiUrl: 'https://api.green-salad-4efd.toolsminati.workers.dev',
});

// Calls the 'bytedance/omnihuman/v1.5' workflow directly — params are sent as-is.
const { result, usage } = await ai.apis.run('bytedance/omnihuman/v1.5', {
  image_url: "https://cdn.example.com/input.jpg",
  audio_url: "https://cdn.example.com/input.mp3",
  prompt: "A serene mountain lake at golden hour, ultra detailed"
}, {
  mode: ApiRunMode.SYNC,
});

console.log(result); // workflow-specific output
console.log(usage?.credits); // credits charged

Async (submit & poll)

This model can run longer than the sync limit (~20s). In TypeScript, ai.apis.run(…, { mode: ApiRunMode.ASYNC }) polls for you; over HTTP, submit and poll yourself.

ts
import { createClient, ApiRunMode } from '@picsart/ai-sdk';

const ai = createClient({
  apiKey: process.env.PICSART_API_KEY,
  apiUrl: 'https://api.green-salad-4efd.toolsminati.workers.dev',
});

// mode: ASYNC submits the job and polls under the hood — you just await.
const { result } = await ai.apis.run('bytedance/omnihuman/v1.5', {
  image_url: "https://cdn.example.com/input.jpg",
  audio_url: "https://cdn.example.com/input.mp3",
  prompt: "A serene mountain lake at golden hour, ultra detailed"
}, {
  mode: ApiRunMode.ASYNC,
});

console.log(result);

Parameters

8 parameters, sent inside params. These are the vendor's own names, so they line up one-for-one with the vendor's documentation. Required ones must be supplied; the rest fall back to their defaults.

ParameterTypeRequiredDefaultDetails
image_url
Publicly accessible HTTP(S) URL of the source image. Formats: JPG (preferred), PNG, JFIF. Must be under 5 MB and under 4096x4096. The subject can be a person, a pet or an animated character; the sharper the image, the better the result.
stringyes
audio_url
Publicly accessible HTTP(S) URL of the driving audio. Expression, lip-sync and motion are derived from it, so no prompt is needed for emotion. Must be under 60 seconds; its duration defines the output video length and the billed amount.
stringyes
prompt
Optional prompt steering the scene, movements and camera work — it does NOT control lip-sync (the audio does). Supported languages: Chinese, English, Japanese, Korean, Mexican Spanish, Indonesian.
stringno
resolution
Output video resolution. Affects billing and generation time (the vendor quotes a real-time factor of 27 for 1080p and 23 for 720p).
stringno1080p
720p1080p
turbo_mode
Speed up generation at the cost of some quality. Defaults to false.
booleanno
mask_url
Mask image URL(s) selecting which subject in the image speaks — only needed for multi-subject images. Accepts a single URL or a list (several subjects speaking at once). Masks come from the vendor subject-detection step; an ordinary mask image whose white area covers the speaker also works.
string | arrayno
seed
Random seed. -1 (default) generates a random one; the same positive seed with identical inputs reproduces the same result.
numberno-1
options
Options controlling safety checks and drive integration
objectno
safety_checks
Safety check settings
objectno
enabled
Whether to run content moderation. Defaults to true.
booleanno
drive
Save result to Picsart Drive
objectno
name
File name in Picsart Drive
stringyes
attributes
Custom attributes to attach to the file
objectno
folder
Target folder in Picsart Drive
objectno
inputs_transformation
Input transformation settings
objectno
downscale_oversized_images
Whether to downscale oversized input images. Defaults to false.
booleanno

Response

Over HTTP the output arrives inside a status envelope, at response.result. ai.apis.run() unwraps that envelope for you and resolves to { result, usage } instead. Either way the resultitself is the vendor's own shape.

json
{
  "result": {
    "url": "https://cdn.green-salad-4efd.toolsminati.workers.dev/…/result.mp4",
    "video": {
      "url": "https://cdn.green-salad-4efd.toolsminati.workers.dev/…/result.mp4",
      "content_type": "…"
    },
    "duration": 0,
    "resolution": "720p",
    "mimeType": "video/mp4",
    "driveFile": {}
  },
  "usage": {
    "credits": 5
  }
}