digital-humanDigital Human turns a video of a person plus a driving audio clip into a lip-synced talking video: the person keeps their original look, motion and background while the mouth is re-animated to speak the audio. If the audio runs longer than the video, the footage is extended automatically — the output always follows the audio length. Send two URLs (video_url: mp4/mov of the person; audio_url: mp3/wav/m4a, up to 30 minutes) and get an mp4 back at the source video's resolution. Billing is $0.01 per second of the driving audio — a 60-second voiceover costs $0.60, half the per-second price of image-based lip-sync. Use AI Lip-Sync instead when you only have a still portrait image. Typical turnaround is about 2 minutes for short clips. Only successful generations are charged; results are kept for 7 days.
$0.01 per second of audio vs $0.02/s for the image-based ai-lipsync — when you already have footage of the person, this is the cheaper path
Lighting, motion, background and camera work come from your footage; only the mouth is re-animated to match the new audio
Audio longer than the video? The footage is extended automatically — verified with a 6.2s voiceover over a 5s clip
Output matches the source video resolution (tested at 1440×2560) — no forced downscale
POST /api/v1/video/generations with model digital-human, video_url and audio_url — poll the task id for the mp4
$0.01 per second of audio · if the audio outlasts the video, the footage is extended automatically
Your talking video will appear here
Digital Human é uma API de geração de vídeo da APIMODELS. Digital Human turns a video of a person plus a driving audio clip into a lip-synced talking video: the person keeps their original look, motion and background while the mouth is re-animated to speak the audio. If the audio runs longer than the video, the footage is extended automatically — the output always follows the audio length. Send two URLs (video_url: mp4/mov of the person; audio_url: mp3/wav/m4a, up to 30 minutes) and get an mp4 back at the source video's resolution. Billing is $0.01 per second of the driving audio — a 60-second voiceover costs $0.60, half the per-second price of image-based lip-sync. Use AI Lip-Sync instead when you only have a still portrait image. Typical turnaround is about 2 minutes for short clips. Only successful generations are charged; results are kept for 7 days. Pela plataforma APIMODELS, você acessa este modelo por uma API unificada com preços transparentes de pagamento por uso. Preço atual: per second of audio: $0.01.
Gere rapidamente vídeos promocionais da marca para campanhas e redes sociais.
Crie vídeos curtos atraentes para TikTok, Instagram e YouTube.
Gere demonstrações de recursos e tutoriais para melhorar a conversão.
Produza explicações de cursos, divulgação e vídeos de treinamento a baixo custo.
O Digital Human está disponível pela APIMODELS a: per second of audio: $0.01. A cobrança é por uso — você paga apenas pelo que gera.
Cadastre-se na APIMODELS, obtenha sua chave de API e chame nosso endpoint unificado. Oferecemos documentação detalhada com exemplos em cURL, Python e Node.js.
A APIMODELS oferece o mesmo modelo Digital Human pela nossa plataforma de agregação. Fornecemos uma interface de API unificada, então você não precisa de contas separadas por provedor: uma única chave para acessar todos os modelos.
Na APIMODELS, o Digital Human roda junto com mais de 60 modelos usando uma única chave de API e um único saldo, então escolher é questão de adequação, não de dependência. Ele suporta Video + Audio, Lip-Sync, Keeps Original Motion, Audio up to 30min, Per-Second Pricing e você pode avaliá-lo em preço e capacidade frente a outros modelos de geração de vídeo, trocando ao alterar uma única string com o nome do modelo — sem nova conta ou integração. Veja todas as opções de geração de vídeo com preços ao vivo em apimodels.app/models.
O Digital Human suporta: Video + Audio, Lip-Sync, Keeps Original Motion, Audio up to 30min, Per-Second Pricing. Consulte a documentação da APIMODELS para todos os parâmetros e exemplos de chamada.
Sim. A APIMODELS expõe o Digital Human por uma única API unificada e uma só chave, sem contas separadas por provedor e sem precisar lidar com o acesso de rede regional de cada provedor.
Aceitamos Stripe (Visa, Mastercard e outros cartões internacionais) e Alipay. O saldo fica disponível imediatamente após o pagamento.