sayakpaul
4c05ea6669
updates
2024-11-18 15:30:23 +05:30
sayakpaul
a40ccd2237
updates
2024-11-18 15:28:37 +05:30
sayakpaul
9852c3d994
updates.
2024-11-18 15:26:52 +05:30
sayakpaul
5ba510e061
updates
2024-11-17 10:18:20 +05:30
sayakpaul
a9adf22386
dataprep.
2024-11-17 08:31:21 +05:30
zhipuch
ac5f417555
sft with multigpu
2024-11-12 15:20:41 +08:00
Aryan
d63a826f37
I2V multiresolution finetuning by removing learned PEs ( #31 )
...
* i2v finetuning without learned pe
* update
* update
* update
* update readme
* refactor
* refactor
2024-11-08 07:59:12 +05:30
Yuxuan.Zhang
cbbac06112
add some script of lora test ( #66 )
...
* multi resolutions support
* full chinese readme
* Update README.md
* Update README.md
Co-authored-by: Sayak Paul <spsayakpaul@gmail.com >
* reformat from pycharm
* dataset.md
* Update README.md
* for merge
* mergeing
* torch update for use
* add test lora script
* Update test_lora_inference.py
* Update requirements.txt
* Update requirements.txt
---------
Co-authored-by: Aryan <aryan@huggingface.co >
Co-authored-by: Sayak Paul <spsayakpaul@gmail.com >
2024-10-30 16:25:14 +05:30
Aryan
bd4062760c
Update README.md ( #73 )
2024-10-28 12:51:25 +05:30
Yuancheng Xu
0affacb229
Update README.md ( #59 )
...
Install diffusers from source to address https://github.com/a-r-r-o-w/cogvideox-factory/issues/46
2024-10-21 04:01:21 +05:30
Johan Mellin
7a6dff2d9f
Fixed optimizers parsing error in bash scripts ( #61 )
...
* replaced lambda statements with local functions to work with pickle in dataset.py and moved collate function outside of main and added its own class in cogvideox_image_to_video_lora.py
* fixed merge issue
* fixed row deletion merge
* fixed row deletion merge
* fixed comma-separated param issue in bash
2024-10-21 04:00:51 +05:30
Aryan
db9b295408
Windows support for T2V scripts ( #48 )
2024-10-21 02:40:05 +05:30
Sayak Paul
6c00cf094b
fix: resuming from a checkpoint when using deepspeed. ( #38 )
...
* fix: resuming from a checkpoint when using deepspeed.
* remove changes to prepare_dataset.py
* propagate to others.
* tackle gradnorm.
---------
Co-authored-by: Aryan <aryan@huggingface.co >
2024-10-20 02:56:32 +05:30
Dhanush
f0d9908bbd
fix: correct type in .py files ( #52 )
2024-10-20 01:22:07 +05:30
Aryan
8da9638896
more dataset fixes from stashed changes ( #49 )
2024-10-19 06:30:26 +05:30
Aryan
f2a1626de2
improve dataset preparation ( #43 )
...
* update
* update
* update
* update
2024-10-18 02:43:22 +05:30
Aryan
8c12f34d4e
update ( #42 )
2024-10-17 17:29:38 +05:30
Aryan
feb2e26547
Improve dataset preparation support + multiresolution prep ( #39 )
...
* update
* make style
* renormalize correctly
* apply suggestions from review
* apply suggestions from review
* update
2024-10-17 16:48:52 +05:30
Farookh Zaheer Siddiqui
a6c246c29d
[Docs] : Update README.md ( #35 )
...
* Update README.md
* Update README.md
---------
Co-authored-by: Aryan <contact.aryanvs@gmail.com >
2024-10-15 19:12:57 +05:30
Johan Mellin
4f2744ee59
Update for windows compability ( #32 )
...
* replaced lambda statements with local functions to work with pickle in dataset.py and moved collate function outside of main and added its own class in cogvideox_image_to_video_lora.py
* Update cogvideox_image_to_video_lora.py
Remove the bug-fix related to image encoding since it's already present in #31
* Update cogvideox_image_to_video_lora.py
Revert to last commit (60c46827404cd4780ee0e7d3c9c56a953e2a04ca)
2024-10-15 19:10:50 +05:30
Aryan
3754d20eac
Lower requirements versions ( #27 )
...
* loosen requirements
* Update requirements.txt
* Update requirements.txt
* Update requirements.txt
2024-10-14 15:35:04 +05:30
Yuxuan.Zhang
1323aeb5b1
Update README and Contribution guide ( #20 )
...
* multi resolutions support
* full chinese readme
* Update README.md
* Update README.md
Co-authored-by: Sayak Paul <spsayakpaul@gmail.com >
* reformat from pycharm
* dataset.md
* Update README.md
---------
Co-authored-by: Aryan <aryan@huggingface.co >
Co-authored-by: Sayak Paul <spsayakpaul@gmail.com >
2024-10-12 02:37:57 +05:30
Johan Mellin
9621d72628
Update requirements.txt (fixed typo) ( #24 )
...
Fixed typo transformers dependancy from 0.45.2 to 4.45.2
2024-10-11 15:16:42 +05:30
Ikko Eltociear Ashimine
09464e2d78
docs: update README.md ( #21 )
...
hyperparamters -> hyperparameters
2024-10-11 04:43:27 +05:30
Yuxuan.Zhang
e50cb9c4ab
Merge pull request #19 from a-r-r-o-w/zR-dev
...
Darft of Chinese README
2024-10-10 22:15:48 +08:00
zR
13083b5bbf
.github page
2024-10-10 22:14:25 +08:00
zR
353e4fc885
Update README_zh.md
2024-10-10 21:49:31 +08:00
zR
8b7a886e72
Merge branch 'main' into zR-dev
2024-10-10 21:49:22 +08:00
zR
0bd88a4f4d
zh readme
2024-10-10 21:38:21 +08:00
Sayak Paul
4fc0ec6453
Merge pull request #17 from a-r-r-o-w/a-r-r-o-w-patch-1
...
Update README.md
2024-10-10 18:09:45 +05:30
Aryan
00a80930a7
Update README.md
2024-10-10 18:06:57 +05:30
Aryan
ee02219eac
readme updates + refactor ( #14 )
...
* refactor
* update
* update
* Update README.md
Co-authored-by: Sayak Paul <spsayakpaul@gmail.com >
* Update README.md
Co-authored-by: Sayak Paul <spsayakpaul@gmail.com >
* address review comments part i
* fix adapter name causing peft error
* prompts -> prompt
* update
* add video
* update
---------
Co-authored-by: Sayak Paul <spsayakpaul@gmail.com >
2024-10-10 18:03:45 +05:30
glide-the
2a70b16d3c
add "max_sequence_length": model_config.max_text_seq_length, ( #15 )
...
Co-authored-by: --unset <--unset>
2024-10-10 17:08:40 +05:30
glide-the
00f7519f0e
add VideoDatasetWithResizeAndRectangleCrop dataset resize crop ( #13 )
...
* fix issue 12: train dataset _preprocess_data crop video
* add_argument video_reshape_mode input videos are reshaped to this mode
* arg enable_model_cpu_offloading
* Add imageio-ffmpeg and imageio to requirements.txt
* apply changes to i2v lora script
* make style
---------
Co-authored-by: --unset <--unset>
Co-authored-by: Aryan <aryan@huggingface.co >
2024-10-10 01:55:12 +05:30
Aryan
b1b72c0e38
CogVideoX I2V; CPU offloading; Model README descriptions ( #11 )
...
* update
* update
* update readme
* update
* update
* update model desc
* update
* Update training/prepare_dataset.py
2024-10-09 17:44:59 +05:30
Yuxuan.Zhang
cc1d2e759f
Multi-GPU parallel encoding support for training videos. ( #6 )
...
* Multi-GPU parallel encoding support for training videos.
* revert
* make style
* update
---------
Co-authored-by: Aryan <aryan@huggingface.co >
2024-10-09 04:19:59 +05:30
Aryan
3a519f5443
Full finetuning memory requirements ( #9 )
...
* update
* update
* model cpu offloading
2024-10-09 03:11:04 +05:30
Aryan
be9d99a1b0
DeepSpeed and DDP Configs ( #10 )
...
* add configs
* remove compiled ddp config
* add coauthor
Co-Authored-By: Sayak Paul <spsayakpaul@gmail.com >
* update
* deepspeed numbers nd fixes
---------
Co-authored-by: Sayak Paul <spsayakpaul@gmail.com >
2024-10-09 02:54:18 +05:30
Sayak Paul
795f2c24d4
Merge pull request #8 from a-r-r-o-w/readme-ii
...
refactor readme i.
2024-10-08 17:36:18 +05:30
sayakpaul
27af518053
readme i.
2024-10-08 16:55:52 +05:30
Aryan
0937e923e3
update ( #7 )
2024-10-08 16:04:22 +05:30
Aryan
d792cbbb64
update ( #5 )
2024-10-06 21:53:43 +05:30
zR
1d5bf447b3
Multi-GPU parallel encoding support for training videos.
2024-10-05 22:28:45 +08:00
Aryan
d81209bf49
Low-bit memory optimizers, CpuOffloadOptimizer, Memory Reports ( #3 )
...
* update
* update
* update
* update
* update
* update
* update
* update
* update
2024-10-04 18:18:54 +05:30
Aryan
512a8a6854
CogVideoX LoRA and full finetuning ( #1 )
...
* update
* update
* update
* update
* update
* update
* update
* update
* update
* update
* update
* update
* update
* update
* update
2024-10-01 08:05:10 +05:30
Aryan
95774133eb
Initial commit
2024-09-25 13:35:24 +05:30