diff --git a/README.md b/README.md index 16977e6..623bfa0 100755 --- a/README.md +++ b/README.md @@ -331,7 +331,6 @@ The table below summarizes currently supported model families and weights. Video | Model Family | Modality | Supported Tasks | Weight / Download / Description | |--|--|--|--| - | Wan2.2-Fun | Video | Series trained by this project on Wan2.2, covering T2V, I2V, first/last frame, controlled generation, and camera control |
Wan2.2-Fun-A14B-InP🤗🤖Wan2.2-Fun-14B text-to-video generation weights, trained at multiple resolutions, supports start-end image prediction.
Wan2.2-Fun-A14B-Control🤗🤖Wan2.2-Fun-14B video control weights, supporting various control conditions such as Canny, Depth, Pose, MLSD, etc., and trajectory control. Supports multi-resolution (512, 768, 1024) video prediction at 81 frames, trained at 16 frames per second, with multilingual prediction support.
Wan2.2-Fun-A14B-Control-Camera🤗🤖Wan2.2-Fun-14B camera lens control weights. Supports multi-resolution (512, 768, 1024) video prediction, trained with 81 frames at 16 FPS, supports multilingual prediction.
Wan2.2-Fun-5B-InP🤗🤖Wan2.2-Fun-5B text-to-video weights trained at 121 frames, 24 FPS, supporting first/last frame prediction.
Wan2.2-Fun-5B-Control🤗🤖Wan2.2-Fun-5B video control weights, supporting control conditions like Canny, Depth, Pose, MLSD, and trajectory control. Trained at 121 frames, 24 FPS, with multilingual prediction support.
Wan2.2-Fun-5B-Control-Camera🤗🤖Wan2.2-Fun-5B camera lens control weights. Trained at 121 frames, 24 FPS, with multilingual prediction support.
Wan2.2-Fun-Reward-LoRAs🤗🤖Reward LoRAs that optimize Wan2.2-Fun generated videos via reward backpropagation
| | Wan2.2-VACE-Fun | Video | Series trained by this project with the VACE scheme, covering controlled generation and subject reference |
Wan2.2-VACE-Fun-A14B🤗🤖Control weights for Wan2.2 trained using the VACE scheme (based on the base model Wan2.2-T2V-A14B), supporting various control conditions such as Canny, Depth, Pose, MLSD, trajectory control, etc. It supports video generation by specifying the subject. It supports multi-resolution (512, 768, 1024) video prediction, and is trained with 81 frames at 16 FPS. It also supports multi-language prediction.
| | Wan2.2 | Video | Official Wan weights covering T2V, I2V, audio-driven, and character animation; can be used as training baseline for Wan2.2-Fun |
Wan2.2-TI2V-5B🤗🤖Wan2.2-5B text/image-to-video weights
Wan2.2-T2V-A14B🤗🤖Wan2.2-14B text-to-video weights
Wan2.2-I2V-A14B🤗🤖Wan2.2-14B image-to-video weights
Wan2.2-S2V-14B🤗🤖Wan2.2-14B audio-to-video weights, speaker-driven digital human
Wan2.2-Animate-14B🤗🤖Wan2.2-14B character replacement and motion transfer weights; repo contains multiple precision files
| @@ -410,6 +409,7 @@ The table below summarizes currently supported model families and weights. Video ### Wan2.1-Fun-V1.1-14B-Control && Wan2.1-Fun-V1.1-1.3B-Control Generic Control Video + Reference Image: + + - +
@@ -424,6 +424,7 @@ Generic Control Video + Reference Image: Wan2.1-Fun-V1.1-1.3B-Control
@@ -437,11 +438,12 @@ Generic Control Video + Reference Image:
Generic Control Video (Canny, Pose, Depth, etc.) and Trajectory Control: + - +
@@ -453,7 +455,7 @@ Generic Control Video (Canny, Pose, Depth, etc.) and Trajectory Control:
@@ -467,6 +469,7 @@ Generic Control Video (Canny, Pose, Depth, etc.) and Trajectory Control: + + + + +
@@ -493,6 +496,7 @@ Generic Control Video (Canny, Pose, Depth, etc.) and Trajectory Control: Pan Right
@@ -503,6 +507,7 @@ Generic Control Video (Canny, Pose, Depth, etc.) and Trajectory Control:
Pan Down @@ -513,6 +518,7 @@ Generic Control Video (Canny, Pose, Depth, etc.) and Trajectory Control: Pan Up + Pan Right
@@ -599,6 +605,7 @@ Resolution-512
A young woman with beautiful clear eyes and blonde hair, wearing white clothes and twisting her body, with the camera focused on her face. High quality, masterpiece, best quality, high resolution, ultra-fine, dreamlike. diff --git a/README_ja-JP.md b/README_ja-JP.md index b88f54d..4d88eab 100755 --- a/README_ja-JP.md +++ b/README_ja-JP.md @@ -331,7 +331,6 @@ sh scripts/{model_name}/train.sh | モデル系列 | モダリティ | サポートタスク | 重み / ダウンロード / 説明 | |--|--|--|--| - | Wan2.2-Fun | ビデオ | 本プロジェクトがWan2.2で訓練した系列。テキスト/画像から動画、首尾画像、制御生成、カメラ制御をカバー |
Wan2.2-Fun-A14B-InP🤗🤖Wan2.2-Fun-14Bのテキスト・画像から動画を生成するモデルの重み。複数の解像度で学習されており、動画の最初と最後のフレームの予測をサポートしています。
Wan2.2-Fun-A14B-Control🤗🤖Wan2.2-Fun-14Bの動画制御用重み。Canny、Depth、Pose、MLSDなどのさまざまな制御条件に対応しており、軌跡制御もサポートしています。512、768、1024の複数解像度での動画生成が可能で、81フレーム、16fpsで学習されています。多言語対応の予測もサポートしています。
Wan2.2-Fun-A14B-Control-Camera🤗🤖14B Controlにカメラモーション制御を追加
Wan2.2-Fun-5B-InP🤗🤖Wan2.2-Fun-5B テキストから動画生成用の重み。121フレーム、24 FPSで学習され、先頭/末尾フレーム予測をサポート。
Wan2.2-Fun-5B-Control🤗🤖Wan2.2-Fun-5B 動画制御用重み。Canny、Depth、Pose、MLSDなどの制御条件や軌道制御をサポート。121フレーム、24 FPSで学習され、多言語予測に対応。
Wan2.2-Fun-5B-Control-Camera🤗🤖Wan2.2-Fun-5B カメラレンズ制御用重み。121フレーム、24 FPSで学習され、多言語予測に対応。
Wan2.2-Fun-Reward-LoRAs🤗🤖Wan2.2-Fun生成動画を報酬逆伝播で最適化するReward LoRA集合
| | Wan2.2-VACE-Fun | ビデオ | 本プロジェクトがVACE方式で訓練した系列。制御生成と主題参照をカバー |
Wan2.2-VACE-Fun-A14B🤗🤖VACE方式でトレーニングされたWan2.2の制御ウェイト(ベースモデルはWan2.2-T2V-A14B)。Canny、Depth、Pose、MLSD、軌道制御などの異なる制御条件をサポートします。対象を指定して動画生成が可能です。多解像度(512、768、1024)の動画予測をサポートし、81フレームで16FPSでトレーニングされています。多言語予測にも対応しています。
| | Wan2.2 | ビデオ | Wan公式重み。テキスト/画像から動画、音声駆動、キャラクターアニメーションをカバー。Wan2.2-Fun系列の訓練基線としても使用可能 |
Wan2.2-TI2V-5B🤗🤖Wan2.2-5B テキスト/画像から動画生成重み
Wan2.2-T2V-A14B🤗🤖Wan2.2-14B テキストから動画生成重み
Wan2.2-I2V-A14B🤗🤖Wan2.2-14B 画像から動画生成重み
Wan2.2-S2V-14B🤗🤖Wan2.2-14B 音声から動画生成重み、話者駆動デジタルヒューマン
Wan2.2-Animate-14B🤗🤖Wan2.2-14B キャラクター置換・モーション転移重み。リポジトリに複数精度ファイルを含む
| @@ -409,14 +408,15 @@ sh scripts/{model_name}/train.sh ### Wan2.1-Fun-V1.1-14B-Control && Wan2.1-Fun-V1.1-1.3B-Control -Generic Control Video + Reference Image: +汎用制御動画 + 参照画像: + + - +
- Reference Image + 参照画像 - Control Video + 制御動画 Wan2.1-Fun-V1.1-14B-Control @@ -424,6 +424,7 @@ Generic Control Video + Reference Image: Wan2.1-Fun-V1.1-1.3B-Control
@@ -437,11 +438,12 @@ Generic Control Video + Reference Image:
-Generic Control Video (Canny, Pose, Depth, etc.) and Trajectory Control: +汎用制御動画(Canny、Pose、Depth など)と軌跡制御: + - +
@@ -453,7 +455,7 @@ Generic Control Video (Canny, Pose, Depth, etc.) and Trajectory Control:
@@ -467,6 +469,7 @@ Generic Control Video (Canny, Pose, Depth, etc.) and Trajectory Control: + + + + +
@@ -493,6 +496,7 @@ Generic Control Video (Canny, Pose, Depth, etc.) and Trajectory Control: Pan Right
@@ -503,6 +507,7 @@ Generic Control Video (Canny, Pose, Depth, etc.) and Trajectory Control:
Pan Down @@ -513,6 +518,7 @@ Generic Control Video (Canny, Pose, Depth, etc.) and Trajectory Control: Pan Up + Pan Right
@@ -599,6 +605,7 @@ Generic Control Video (Canny, Pose, Depth, etc.) and Trajectory Control:
美しい澄んだ目と金髪の若い女性が白い服を着て体をひねり、カメラは彼女の顔に焦点を合わせています。高品質、傑作、最高品質、高解像度、超微細、夢のような。 diff --git a/README_zh-CN.md b/README_zh-CN.md index 2ebbab5..a650030 100755 --- a/README_zh-CN.md +++ b/README_zh-CN.md @@ -349,7 +349,7 @@ sh scripts/{model_name}/train.sh # 四、视频作品 -Image to Video: +图生视频: @@ -369,14 +369,15 @@ Image to Video:
-Generic Control Video + Reference Image: +通用控制视频 + 参考图像: + + - +
- Reference Image + 参考图像 - Control Video + 控制视频 Wan2.1-Fun-V1.1-14B-Control @@ -384,6 +385,7 @@ Generic Control Video + Reference Image: Wan2.1-Fun-V1.1-1.3B-Control
@@ -397,11 +399,12 @@ Generic Control Video + Reference Image:
-Generic Control Video (Canny, Pose, Depth, etc.) and Trajectory Control: +通用控制视频(Canny、Pose、Depth 等)与轨迹控制: + - +
@@ -413,7 +416,7 @@ Generic Control Video (Canny, Pose, Depth, etc.) and Trajectory Control:
@@ -427,6 +430,7 @@ Generic Control Video (Canny, Pose, Depth, etc.) and Trajectory Control: +