Commit Graph
34 Commits
Author SHA1 Message Date
Bubbliiiing 9da6b79ece Update Prompt && Update Readme && Fix control bug (#131)
* Update Readme

* Update prompt tips

* Fix bug in ui

* Fix control bug and Update readme

* Update prompt again
2024-11-11 13:48:53 +08:00
Bubbliiiing 62de94e2f1 Update V5 (#128)
Update V5
2024-11-08 19:46:50 +08:00
Bubbliiiing d3b8bbbd14 Fix bug in v3 training (#98) 2024-08-22 16:03:14 +08:00
hkz 5ea1bf2450 Update Video Caption (#93)
* Add VILA1.5 in video caption

* update VILA1.5 in video_caption

* add get_video_path_list back & update the accelerate version & auto load the vila quant_path & download models in the main process

* add exception handling & add force_ascii=False

* fix the NCCL timeout & use logger

* update utils

* Update video splitting

* Update pre-filtering

* Update video caption

* Add caption_rewrite.py

* Add VideoCLIPXL

* Add beautiful prompt demo

* fix test

* Update Dockerfile.ds

* update VideoCLIPXL & add filter_meta_train.py

* update stage3

* update README.md

* update README_zh-CN.md

* update README

* update requirements

* fix stage_3_video_recaptioning.sh

* update Beautiful Prompt

* update README

* update vila_video_recaptioning.py & fix the empty gather result
2024-08-21 16:40:54 +08:00
Bubbliiiing f2f0cbccc9 Bug fix/encode prompt (#96)
* update readme

* fix bug in encode_prompt
2024-08-19 16:53:05 +08:00
e4f1a4fe97 Update V4 version (#92)
* update train_lora && update deepspeed && update training with max token length

* fix bug in train.py

* fix bug in training_with_video_token_length

* update v2v && update v2v api

* add rope2d embedding precomputation; move text encoder to dataloader to reduce gpu memory consumpution

* add cuda multi-stream to speedup vae encode

* update new vae && new comfyui

* fix some bug in training code

* Add lcm lora (#89)

Co-authored-by: xuanyuan.lb <xuanyuan.lb@alibaba-inc.com>

* Update Training Code and fix bug in low vram mode

* fix bug in low vram mode

* update report

* update cfg

* actual text clip

---------

Co-authored-by: mengli.cml <mengli.cml@alibaba-inc.com>
Co-authored-by: liubo0902 <38622806+liubo0902@users.noreply.github.com>
Co-authored-by: xuanyuan.lb <xuanyuan.lb@alibaba-inc.com>
2024-08-19 11:22:17 +08:00
bubbliiiing 039d67acf1 support float16 && add reference && fix bug in training 2024-07-24 13:19:03 +08:00
bubbliiiing 1a5bab2234 fix bug in no inpaint model 2024-07-18 21:09:10 +08:00
bubbliiiing 2eef3f9f78 update v4 2024-07-18 16:33:56 +08:00
Bubbliiiing 616a35425d Fix the issue where Lora training cannot load state, fix the issue where training cannot eval (#66)
* rename the comfyui files and new readme

* fix bug in eval and load lora state dict
2024-07-18 15:47:51 +08:00
Bubbliiiing 8b7722463e rename the comfyui files and new readme (#53) 2024-07-13 14:57:13 +08:00
Bubbliiiing 883c0a20a8 update readme (#52) 2024-07-12 18:03:48 +08:00
Bubbliiiing fbbfc818ea Update pipeline (#50) 2024-07-12 17:47:47 +08:00
bubbliiiing 8fd346a9af Merge branch 'main' into comfyui 2024-07-12 16:26:34 +08:00
bubbliiiing c1d7e503d3 update comfyui 2024-07-12 16:24:46 +08:00
Wang Qiang d57c779217 Added exception handling for video read bucket_sampler.py 2024-07-10 15:42:30 +08:00
Bubbliiiingandchenyunkuo.cyk f9eeabe231 Updated to v3 version, supports image generated videos, with a maximum support of 960x960x144 video generation. (#40)
* update v3

* Fix frame start_idx bug, see issue #41.

* update readme and fix bug in training

* update requirements

* update ui

* fix bug in inpaint

* update new ui

* fix bug in auto resize

* fix bug in auto resize

* fix bug in modelscope and eas

* update low gpu memory mode

---------

Co-authored-by: chenyunkuo.cyk <chenyunkuo.cyk@alibaba-inc.com>
2024-07-05 20:27:24 +08:00
bubbliiiing 59bd5de1ad update inpaint model 2024-06-08 09:38:48 +08:00
bubbliiiing 60d205111b update inpaint model and new ui in demo 2024-06-07 10:48:13 +08:00
Bubbliiiing d196f64d93 Add huggingface link and new UI (#22) 2024-06-04 21:33:15 +08:00
zouxinyi0625 dc9305b0fe update ui chinese (#20)
* update ui chinese
2024-06-04 16:24:56 +08:00
Bubbliiiing 58d793659d Update 768x768 Link and new gallery (#15)
update readme and model link
2024-06-04 10:47:57 +08:00
hkunzhe 9e2fba7d40 fix the conflict between autogptq and sglang 2024-05-31 14:27:56 +08:00
Bubbliiiingandzouxinyi0625 fb0916f4da Update EasyAnimateV2 (#5)
* update EasyAnimateV2

* update datasets loader

* update datasets loader

* update fast api

* complete data preprocess pipeline.

* update lots of readme

* update readme bans

* fix bug in validation while training

* provide example for video cut

* update text box

* add arxiv

* delete IDDPM

* update gallery

* update arxiv

* update readme

* update readme

* link fix

* update vae readme

---------

Co-authored-by: zouxinyi0625 <zouxinyi.zxy@alibaba-inc.com>
2024-05-31 10:33:14 +08:00
hkunzhe a8cfe4ec7e fix test 2024-05-28 12:00:49 +08:00
hkunzhe 6efc7de3f2 add the dataset preprocessing pipeline 2024-05-27 22:26:03 +08:00
hkunzhe ac97760d94 update requirements.txt for video caption 2024-04-25 17:23:44 +08:00
hkunzhe 728749039f revert Dockerfile.ds & add sglang runtime shutdown 2024-04-23 19:24:02 +08:00
hkunzhe a2bf7165f8 update docker 2024-04-22 20:11:29 +08:00
huangkunzhe.hkz a3588f838d fix typos 2024-04-19 10:21:27 +08:00
huangkunzhe.hkz 0b02c45b08 fix ref path 2024-04-15 16:29:17 +08:00
huangkunzhe.hkz a4008a6127 add video caption 2024-04-15 16:25:11 +08:00
bubbliiiing 1a1720cfa5 update readme and default params 2024-04-13 15:02:01 +08:00
bubbliiiing ebe1a64a1c Create Code 2024-04-12 17:42:52 +08:00