2026年7月1日

Davinci Magihuman

Open-source AI generating lip-synced talking videos from a single photo and audio/text.
AI WritingContent GenerationResearchEmail WritingSummarizationRewritingAcademic Research瀏覽器擴充功能

總覽
Davinci Magihuman 是什麼?

daVinci-MagiHuman is an advanced, open-source 15B-parameter AI model developed by Sand.ai and GAIR Lab at Shanghai Jiao Tong University. It is designed to generate high-quality, lip-synced talking videos from a single portrait image and a script or audio file. Unlike traditional methods that combine separate text-to-speech and video pipelines, daVinci-MagiHuman utilizes a unified single-stream Transformer to jointly denoise video and audio tokens simultaneously. Released under the Apache 2.0 license, it allows users to inspect weights, run inference locally, and use the technology for commercial purposes. It is optimized for speed, capable of generating short clips in just seconds on professional-grade hardware like the NVIDIA H100.

Open-source AI generating lip-synced talking videos from a single photo and audio/text.

核心功能
Unified Audio + Video generation in a single model pass
Reference photo input allows talking head creation from one image
Multilingual support for broad lip-sync coverage
Open-source Apache 2.0 license for commercial and local use
Fast inference with ~2s generation time for short clips on H100 GPUs
State-of-the-art quality with low Word Error Rates (WER)
熱門使用情境
  • Creating AI-powered marketing avatars from static portraits
  • Developing multilingual educational content with synchronized lip motion
  • Generating low-latency digital humans for interactive applications
  • Prototyping realistic talking head animations for social media
如何使用
  • To use daVinci-MagiHuman
  • upload a clear
  • front-facing portrait photo and provide a script or audio file. Select your desired output resolution (e.g.
  • 256p
  • 720p
  • or 1080p) and start the generation process. Once the AI completes the job
  • you can download your talking video. For local deployment
  • users can download the model checkpoints from Hugging Face and follow the provided CLI instructions.
定價
Davinci Magihuman 採用 Freemium 定價模式,價格與功能可能隨時間調整。
Basic
$19.90/month
1,990 credits (approx. 16 standard or 9 HD video generations/month)
Pro
$31.92/month
3,990 credits, priority processing, batch background removal, and 20% off discount
Max
$47.92/month
5,990 credits, highest priority, dedicated support, and lifetime usage rights
Pay-as-you-go
$1 per 100 credits
Credits never expire, used for one-time top-ups
優惠 / 優惠碼
目前沒有優惠碼。
產品時間線
待核實
驗證資訊
工具狀態
待核實
定價已核驗
待核實
創辦人已認領
否 / 待核實
來源
官網 / 社群提交
相關標籤
AI WritingContent GenerationResearchEmail WritingSummarizationRewritingAcademic Research瀏覽器擴充功能Freemium
你是這個工具的官方團隊嗎?
認領這個資料頁後,你可以更新產品資訊、定價與官方回覆。