Commits · 16a32c9dab03d41204120b63cdae71c40b279bdf · renzhc / diffusers_dcu

23 Nov, 2022 1 commit

[Versatile Diffusion] Add versatile diffusion model (#1283) · 2625fb59

Patrick von Platen authored Nov 23, 2022



* up

* convert dual unet

* revert dual attn

* adapt for vd-official

* test the full pipeline

* mixed inference

* mixed inference for text2img

* add image prompting

* fix clip norm

* split text2img and img2img

* fix format

* refactor text2img

* mega pipeline

* add optimus

* refactor image var

* wip text_unet

* text unet end to end

* update tests

* reshape

* fix image to text

* add some first docs

* dual guided pipeline

* fix token ratio

* propose change

* dual transformer as a native module

* DualTransformer(nn.Module)

* DualTransformer(nn.Module)

* correct unconditional image

* save-load with mega pipeline

* remove image to text

* up

* uP

* fix

* up

* final fix

* remove_unused_weights

* test updates

* save progress

* uP

* fix dual prompts

* some fixes

* finish

* style

* finish renaming

* up

* fix

* fix

* fix

* finish
Co-authored-by: anton-l <anton@huggingface.co>

2625fb59

04 Nov, 2022 1 commit
- fix the parameter naming in `self.downsamplers` (#1108) · 5b20d3b3
  Chenguo Lin authored Nov 05, 2022
```
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>
```
  5b20d3b3
03 Nov, 2022 1 commit

VQ-diffusion (#658) · ef2ea33c

Will Berman authored Nov 03, 2022



* Changes for VQ-diffusion VQVAE

Add specify dimension of embeddings to VQModel:
`VQModel` will by default set the dimension of embeddings to the number
of latent channels. The VQ-diffusion VQVAE has a smaller
embedding dimension, 128, than number of latent channels, 256.

Add AttnDownEncoderBlock2D and AttnUpDecoderBlock2D to the up and down
unet block helpers. VQ-diffusion's VQVAE uses those two block types.

* Changes for VQ-diffusion transformer

Modify attention.py so SpatialTransformer can be used for
VQ-diffusion's transformer.

SpatialTransformer:
- Can now operate over discrete inputs (classes of vector embeddings) as well as continuous.
- `in_channels` was made optional in the constructor so two locations where it was passed as a positional arg were moved to kwargs
- modified forward pass to take optional timestep embeddings

ImagePositionalEmbeddings:
- added to provide positional embeddings to discrete inputs for latent pixels

BasicTransformerBlock:
- norm layers were made configurable so that the VQ-diffusion could use AdaLayerNorm with timestep embeddings
- modified forward pass to take optional timestep embeddings

CrossAttention:
- now may optionally take a bias parameter for its query, key, and value linear layers

FeedForward:
- Internal layers are now configurable

ApproximateGELU:
- Activation function in VQ-diffusion's feedforward layer

AdaLayerNorm:
- Norm layer modified to incorporate timestep embeddings

* Add VQ-diffusion scheduler

* Add VQ-diffusion pipeline

* Add VQ-diffusion convert script to diffusers

* Add VQ-diffusion dummy objects

* Add VQ-diffusion markdown docs

* Add VQ-diffusion tests

* some renaming

* some fixes

* more renaming

* correct

* fix typo

* correct weights

* finalize

* fix tests

* Apply suggestions from code review
Co-authored-by: Anton Lozhkov <aglozhkov@gmail.com>

* Apply suggestions from code review
Co-authored-by: Pedro Cuenca <pedro@huggingface.co>

* finish

* finish

* up
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>
Co-authored-by: Anton Lozhkov <aglozhkov@gmail.com>
Co-authored-by: Pedro Cuenca <pedro@huggingface.co>

ef2ea33c

02 Nov, 2022 1 commit

Up to 2x speedup on GPUs using memory efficient attention (#532) · 98c42134

MatthieuTPHR authored Nov 02, 2022



* 2x speedup using memory efficient attention

* remove einops dependency

* Swap K, M in op instantiation

* Simplify code, remove unnecessary maybe_init call and function, remove unused self.scale parameter

* make xformers a soft dependency

* remove one-liner functions

* change one letter variable to appropriate names

* Remove Env variable dependency, remove MemoryEfficientCrossAttention class and use enable_xformers_memory_efficient_attention method

* Add memory efficient attention toggle to img2img and inpaint pipelines

* Clearer management of xformers' availability

* update optimizations markdown to add info about memory efficient attention

* add benchmarks for TITAN RTX

* More detailed explanation of how the mem eff benchmark were ran

* Removing autocast from optimization markdown

* import_utils: import torch only if is available
Co-authored-by: Nouamane Tazi <nouamane98@gmail.com>

98c42134

31 Oct, 2022 1 commit

Remove some unused parameter in CrossAttnUpBlock2D (#1034) · 7fb4b882

Laurent Mazare authored Oct 31, 2022

Remove some unused parameter

The `downsample_padding` parameter does not seem to be used in `CrossAttnUpBlock2D` (or by any up block for that matter) so removing it.

7fb4b882

25 Oct, 2022 1 commit

[Dance Diffusion] Add dance diffusion (#803) · 88fa6b7d

Patrick von Platen authored Oct 25, 2022



* start

* add more logic

* Update src/diffusers/models/unet_2d_condition_flax.py

* match weights

* up

* make model work

* making class more general, fixing missed file rename

* small fix

* make new conversion work

* up

* finalize conversion

* up

* first batch of variable renamings

* remove c and c_prev var names

* add mid and out block structure

* add pipeline

* up

* finish conversion

* finish

* upload

* more fixes

* Apply suggestions from code review

* add attr

* up

* uP

* up

* finish tests

* finish

* uP

* finish

* fix test

* up

* naming consistency in tests

* Apply suggestions from code review
Co-authored-by: Suraj Patil <surajp815@gmail.com>
Co-authored-by: Pedro Cuenca <pedro@huggingface.co>
Co-authored-by: Nathan Lambert <nathan@huggingface.co>
Co-authored-by: Anton Lozhkov <anton@huggingface.co>

* remove hardcoded 16

* Remove bogus

* fix some stuff

* finish

* improve logging

* docs

* upload
Co-authored-by: Nathan Lambert <nol@berkeley.edu>
Co-authored-by: Suraj Patil <surajp815@gmail.com>
Co-authored-by: Pedro Cuenca <pedro@huggingface.co>
Co-authored-by: Nathan Lambert <nathan@huggingface.co>
Co-authored-by: Anton Lozhkov <anton@huggingface.co>

88fa6b7d

12 Oct, 2022 1 commit
- add or fix license formatting in models directory (#808) · 5afc2b60
  Nathan Lambert authored Oct 12, 2022
```
* add or fix license formatting

* fix quality
```
  5afc2b60
30 Sep, 2022 1 commit

Allow resolutions that are not multiples of 64 (#505) · a784be2e

Josh Achiam authored Sep 30, 2022



* Allow resolutions that are not multiples of 64

* ran black

* fix bug

* add test

* more explanation

* more comments
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>

a784be2e

22 Sep, 2022 1 commit

[UNet2DConditionModel] add gradient checkpointing (#461) · e7120bae

Suraj Patil authored Sep 22, 2022

* add grad ckpt to downsample blocks

* make it work

* don't pass gradient_checkpointing to upsample block

* add tests for UNet2DConditionModel

* add test_gradient_checkpointing

* add gradient_checkpointing for up and down blocks

* add functions to enable and disable grad ckpt

* remove the forward argument

* better naming

* make supports_gradient_checkpointing private

e7120bae

16 Sep, 2022 1 commit

Fix typos and add Typo check GitHub Action (#483) · 76d492ea

Yuta Hayashibe authored Sep 16, 2022

* Fix typos

* Add a typo check action

* Fix a bug

* Changed to manual typo check currently

Ref: https://github.com/huggingface/diffusers/pull/483#pullrequestreview-1104468010

Co-authored-by: Anton Lozhkov <aglozhkov@gmail.com>

* Removed a confusing message

* Renamed "nin_shortcut" to "in_shortcut"

* Add memo about NIN
Co-authored-by: Anton Lozhkov <aglozhkov@gmail.com>

76d492ea

15 Sep, 2022 1 commit

[UNet2DConditionModel, UNet2DModel] pass norm_num_groups to all the blocks (#442) · d144c46a

Suraj Patil authored Sep 15, 2022

* pass norm_num_groups to unet blocs and attention

* fix UNet2DConditionModel

* add norm_num_groups arg in vae

* add tests

* remove comment

* Apply suggestions from code review

d144c46a

08 Sep, 2022 1 commit
- [Black] Update black (#433) · b2b3b1a8
  Patrick von Platen authored Sep 08, 2022
```
* Update black

* update table
```
  b2b3b1a8
06 Sep, 2022 1 commit

Efficient Attention (#366) · 5c4ea00d

Patrick von Platen authored Sep 06, 2022



* up

* add tests

* correct

* up

* finish

* better naming

* Update README.md
Co-authored-by: Pedro Cuenca <pedro@huggingface.co>
Co-authored-by: Pedro Cuenca <pedro@huggingface.co>

5c4ea00d

04 Sep, 2022 1 commit
- Fix typo in unet_blocks.py (#353) · 6c0ca5ef
  Yuntian Deng authored Sep 04, 2022
```
Update unet_blocks.py

fix typo
```
  6c0ca5ef
25 Aug, 2022 1 commit
- [Clean up] Clean unused code (#245) · c1efda70
  Patrick von Platen authored Aug 25, 2022
```
* CleanResNet

* refactor more

* correct
```
  c1efda70
10 Aug, 2022 1 commit
- add attention up/down blocks for VAE (#161) · b344c953
  Suraj Patil authored Aug 10, 2022
  
  b344c953
05 Aug, 2022 1 commit
- [UNet2DConditionModel] add cross_attention_dim as an argument (#155) · c4a3b09a
  Suraj Patil authored Aug 05, 2022
```
add cross_attention_dim as an argument
```
  c4a3b09a
28 Jul, 2022 1 commit

[Vae and AutoencoderKL] Final clean of LDM checkpoints (#137) · 3100bc96

Patrick von Platen authored Jul 28, 2022

* [Vae and AutoencoderKL clean]

* save intermediate finished work

* more progress

* more progress

* finish modeling code

* save intermediate

* finish

* Correct tests

3100bc96

20 Jul, 2022 1 commit

Big Model Renaming (#109) · 9c3820d0

Patrick von Platen authored Jul 21, 2022

* up

* change model name

* renaming

* more changes

* up

* up

* up

* save checkpoint

* finish api / naming

* finish config renaming

* rename all weights

* finish really

9c3820d0

19 Jul, 2022 2 commits
- Get diffusers ready 🚀🚀🚀 (#101) · 8c31925b
  Patrick von Platen authored Jul 19, 2022
```
* big purge

* more fixes

* finish for now
```
  8c31925b
- Finalize ldm (#96) · d5acb411
  Patrick von Platen authored Jul 19, 2022
```
* upload

* make checkpoint work

* finalize
```
  d5acb411
18 Jul, 2022 1 commit

[SDE] Merge to unconditional model (#89) · ba3c9a9a

Patrick von Platen authored Jul 18, 2022

* up

* more

* uP

* make dummy test pass

* save intermediate

* p

* p

* finish

* finish

* finish

ba3c9a9a

14 Jul, 2022 2 commits
- [DDPM] Make DDPM work (#88) · 6d5ef87e
  Patrick von Platen authored Jul 14, 2022
```
* up

* finish

* uP
```
  6d5ef87e
- save intermediate (#87) · e7fe901e
  Patrick von Platen authored Jul 14, 2022
```
* save intermediate

* up

* up
```
  e7fe901e
13 Jul, 2022 1 commit
- Clean uncond unet more (#85) · 5e12d5c6
  Patrick von Platen authored Jul 13, 2022
```
* up

* finished clean up

* remove @
```
  5e12d5c6
12 Jul, 2022 1 commit
- Add unconditional image generation (#79) · 06c79730
  Patrick von Platen authored Jul 12, 2022
```
* uP

* finish downsampling layers

* finish major refactor

* remove bugus file
```
  06c79730
05 Jul, 2022 1 commit
- [MidBlock] Fix mid block (#78) · ea8d58ea
  Patrick von Platen authored Jul 05, 2022
```
* upload files

* finish
```
  ea8d58ea
04 Jul, 2022 2 commits
- Add MidBlock to Grad-TTS (#74) · c352faea
  Patrick von Platen authored Jul 04, 2022
```
Finish
```
  c352faea
- update mid block (#70) · 94566e6d
  Patrick von Platen authored Jul 04, 2022
```
* update mid block

* finish mid block
```
  94566e6d