
open-source-ai• 15 min
daVinci-MagiHuman: A 15B-parameter single-stream Transformer (Apache 2.0) that jointly generates synchronized video and speech from text in a single pass
Jointly generating synchronized video and speech from text in a single pass — no cross-attention or multi-stream branches required.
#davinci#magihuman