Back to Tools Directory
daVinci-MagiHuman is an open-source audio-video generative foundation model for creating human-centric videos with synchronized speech, facial expression, and body motion. It uses a single-stream transformer architecture to deliver multilingual generation and fast inference from reference images and text.
Categories & Topics
multimodal aivideo generationfoundation model
Repository Stats
2.1k
GitHub Stars
213
Forks
2.1k
Watchers
25
Open Issues
Repository
GAIR-NLP/daVinci-MagiHumanLast Updated
Today
August 6, 2026