digital-humanDigital Human turns a video of a person plus a driving audio clip into a lip-synced talking video: the person keeps their original look, motion and background while the mouth is re-animated to speak the audio. If the audio runs longer than the video, the footage is extended automatically — the output always follows the audio length. Send two URLs (video_url: mp4/mov of the person; audio_url: mp3/wav/m4a, up to 30 minutes) and get an mp4 back at the source video's resolution. Billing is $0.01 per second of the driving audio — a 60-second voiceover costs $0.60, half the per-second price of image-based lip-sync. Use AI Lip-Sync instead when you only have a still portrait image. Typical turnaround is about 2 minutes for short clips. Only successful generations are charged; results are kept for 7 days.
$0.01 per second of audio vs $0.02/s for the image-based ai-lipsync — when you already have footage of the person, this is the cheaper path
Lighting, motion, background and camera work come from your footage; only the mouth is re-animated to match the new audio
Audio longer than the video? The footage is extended automatically — verified with a 6.2s voiceover over a 5s clip
Output matches the source video resolution (tested at 1440×2560) — no forced downscale
POST /api/v1/video/generations with model digital-human, video_url and audio_url — poll the task id for the mp4
$0.01 per second of audio · if the audio outlasts the video, the footage is extended automatically
Your talking video will appear here
Digital Human — это API категории «генерация видео» от APIMODELS. Digital Human turns a video of a person plus a driving audio clip into a lip-synced talking video: the person keeps their original look, motion and background while the mouth is re-animated to speak the audio. If the audio runs longer than the video, the footage is extended automatically — the output always follows the audio length. Send two URLs (video_url: mp4/mov of the person; audio_url: mp3/wav/m4a, up to 30 minutes) and get an mp4 back at the source video's resolution. Billing is $0.01 per second of the driving audio — a 60-second voiceover costs $0.60, half the per-second price of image-based lip-sync. Use AI Lip-Sync instead when you only have a still portrait image. Typical turnaround is about 2 minutes for short clips. Only successful generations are charged; results are kept for 7 days. Через платформу APIMODELS доступ к этой модели идёт по единому API с прозрачной оплатой по факту использования. Текущая цена: per second of audio: $0.01.
Quickly generate brand promotion videos for ad campaigns and social media marketing.
Create compelling short-form video content for platforms like TikTok, Instagram, and YouTube.
Generate product feature demonstrations and tutorials to improve user conversion.
Produce course explanations, knowledge explainers, and training videos at low cost.
Digital Human доступна через APIMODELS по цене: per second of audio: $0.01. Оплата идёт по факту использования — вы платите только за то, что сгенерировали.
Зарегистрируйтесь на APIMODELS, получите API-ключ и обращайтесь к нашему единому эндпоинту. Есть подробная документация с примерами на cURL, Python и Node.js.
APIMODELS отдаёт ту же самую модель Digital Human через свою агрегирующую платформу. Мы предоставляем единый интерфейс API, поэтому отдельные аккаунты у каждого поставщика не нужны — один ключ открывает доступ ко всем моделям.
В APIMODELS Digital Human соседствует более чем с 60 моделями под одним API-ключом и общим балансом, поэтому выбор — это вопрос соответствия задаче, а не привязки к поставщику. Модель поддерживает Video + Audio, Lip-Sync, Keeps Original Motion, Audio up to 30min, Per-Second Pricing; вы можете сравнить её по цене и возможностям с другими моделями категории «генерация видео» и переключиться, изменив одну строку с именем модели — без нового аккаунта и без переписывания интеграции. Все варианты категории «генерация видео» с актуальными ценами — на apimodels.app/models.
Digital Human поддерживает: Video + Audio, Lip-Sync, Keeps Original Motion, Audio up to 30min, Per-Second Pricing. Полный список параметров и примеры вызовов — в документации APIMODELS.
Да. APIMODELS отдаёт Digital Human через единый API и один ключ — отдельные аккаунты у каждого поставщика не нужны, и вам не приходится самостоятельно решать вопросы регионального сетевого доступа к каждому из них.
Принимаем Stripe (Visa, Mastercard и другие международные карты) и Alipay. Баланс становится доступен сразу после оплаты.