ByteDance has recently introduced its state-of-the-art video generation model, OmniHuman-1.5, via its Github page. This innovative model boasts the capability to create lifelike character animations using just a single image and an accompanying speech track. Remarkably, it supports the production of videos exceeding one minute in duration, effortlessly handling intricate multi-character interactions.
