digital-humanDigital Human turns a video of a person plus a driving audio clip into a lip-synced talking video: the person keeps their original look, motion and background while the mouth is re-animated to speak the audio. If the audio runs longer than the video, the footage is extended automatically — the output always follows the audio length. Send two URLs (video_url: mp4/mov of the person; audio_url: mp3/wav/m4a, up to 30 minutes) and get an mp4 back at the source video's resolution. Billing is $0.01 per second of the driving audio — a 60-second voiceover costs $0.60, half the per-second price of image-based lip-sync. Use AI Lip-Sync instead when you only have a still portrait image. Typical turnaround is about 2 minutes for short clips. Only successful generations are charged; results are kept for 7 days.
$0.01 per second of audio · if the audio outlasts the video, the footage is extended automatically
Your talking video will appear here
$0.01 per second of audio vs $0.02/s for the image-based ai-lipsync — when you already have footage of the person, this is the cheaper path
Lighting, motion, background and camera work come from your footage; only the mouth is re-animated to match the new audio
Audio longer than the video? The footage is extended automatically — verified with a 6.2s voiceover over a 5s clip
Output matches the source video resolution (tested at 1440×2560) — no forced downscale
POST /api/v1/video/generations with model digital-human, video_url and audio_url — poll the task id for the mp4
Digital Human — это API категории «генерация видео» от APIMODELS. Digital Human turns a video of a person plus a driving audio clip into a lip-synced talking video: the person keeps their original look, motion and background while the mouth is re-animated to speak the audio. If the audio runs longer than the video, the footage is extended automatically — the output always follows the audio length. Send two URLs (video_url: mp4/mov of the person; audio_url: mp3/wav/m4a, up to 30 minutes) and get an mp4 back at the source video's resolution. Billing is $0.01 per second of the driving audio — a 60-second voiceover costs $0.60, half the per-second price of image-based lip-sync. Use AI Lip-Sync instead when you only have a still portrait image. Typical turnaround is about 2 minutes for short clips. Only successful generations are charged; results are kept for 7 days. Через платформу APIMODELS доступ к этой модели идёт по единому API с прозрачной оплатой по факту использования. Текущая цена: per second of audio: $0.01.
Быстро собрать промо-видео бренда под рекламные кампании и продвижение в соцсетях.
Короткие вертикальные ролики для TikTok, Instagram и YouTube.
Показать возможности продукта и снять обучающие ролики, которые повышают конверсию.
Объяснения к курсам, разборы тем и обучающие видео — недорого и на потоке.
Digital Human доступна через APIMODELS по цене: per second of audio: $0.01. Оплата идёт по факту использования — вы платите только за то, что сгенерировали.
Зарегистрируйтесь на APIMODELS, получите API-ключ и обращайтесь к нашему единому эндпоинту. Есть подробная документация с примерами на cURL, Python и Node.js.
APIMODELS отдаёт ту же самую модель Digital Human через свою агрегирующую платформу. Мы предоставляем единый интерфейс API, поэтому отдельные аккаунты у каждого поставщика не нужны — один ключ открывает доступ ко всем моделям.
В APIMODELS Digital Human соседствует более чем с 60 моделями под одним API-ключом и общим балансом, поэтому выбор — это вопрос соответствия задаче, а не привязки к поставщику. Модель поддерживает Video + Audio, Lip-Sync, Keeps Original Motion, Audio up to 30min, Per-Second Pricing; вы можете сравнить её по цене и возможностям с другими моделями категории «генерация видео» и переключиться, изменив одну строку с именем модели — без нового аккаунта и без переписывания интеграции. Все варианты категории «генерация видео» с актуальными ценами — на apimodels.app/models.
Digital Human поддерживает: Video + Audio, Lip-Sync, Keeps Original Motion, Audio up to 30min, Per-Second Pricing. Полный список параметров и примеры вызовов — в документации APIMODELS.
Да. APIMODELS отдаёт Digital Human через единый API и один ключ — отдельные аккаунты у каждого поставщика не нужны, и вам не приходится самостоятельно решать вопросы регионального сетевого доступа к каждому из них.
Принимаем Stripe (Visa, Mastercard и другие международные карты) и Alipay. Баланс становится доступен сразу после оплаты.