Commits · 527430d0a4aee332dc354f2fc61e2f7df1eb4d59 · renzhc / diffusers_dcu

25 Jul, 2024 1 commit

[LoRA] introduce LoraBaseMixin to promote reusability. (#8774) · 527430d0

Sayak Paul authored Jul 25, 2024



* introduce  to promote reusability.

* up

* add more tests

* up

* remove comments.

* fix fuse_nan test

* clarify the scope of fuse_lora and unfuse_lora

* remove space

* rewrite fuse_lora a bit.

* feedback

* copy over load_lora_into_text_encoder.

* address dhruv's feedback.

* fix-copies

* fix issubclass.

* num_fused_loras

* fix

* fix

* remove mapping

* up

* fix

* style

* fix-copies

* change to SD3TransformerLoRALoadersMixin

* Apply suggestions from code review
Co-authored-by: Dhruv Nair <dhruv.nair@gmail.com>

* up

* handle wuerstchen

* up

* move lora to lora_pipeline.py

* up

* fix-copies

* fix documentation.

* comment set_adapters().

* fix-copies

* fix set_adapters() at the model level.

* fix?

* fix

---------
Co-authored-by: Dhruv Nair <dhruv.nair@gmail.com>

527430d0

18 Jul, 2024 1 commit

[docs] pipeline docs for latte (#8844) · 12625c1c

Aryan authored Jul 18, 2024

* add pipeline docs for latte

* add inference time to latte docs

* apply review suggestions

12625c1c

12 Jul, 2024 3 commits

add PAG support sd15 controlnet (#8820) · d704b3bf

Nguyễn Công Tú Anh authored Jul 12, 2024



* add pag support sd15 controlnet

* fix quality import

* remove unecessary import

* remove if state

* fix tests

* remove useless function

* add sd1.5 controlnet pag docs

---------
Co-authored-by: anhnct8 <anhnct8@fpt.com>

d704b3bf

[Docs] add AuraFlow docs (#8851) · 973a62d4

Sayak Paul authored Jul 12, 2024

* add pipeline documentation.

* add api spec for pipeline

* model documentation

* model spec

973a62d4

Add single file loading support for AnimateDiff (#8819) · 11d18f32
Dhruv Nair authored Jul 12, 2024
```
* update

* update

* update

* update
```
11d18f32

11 Jul, 2024 2 commits

[Core] Add Kolors (#8812) · 87b9db64
Álvaro Somoza authored Jul 11, 2024
```
* initial draft
```
87b9db64

Latte: Latent Diffusion Transformer for Video Generation (#8404) · b8cf84a3

Xin Ma authored Jul 11, 2024



* add Latte to diffusers

* remove print

* remove print

* remove print

* remove unuse codes

* remove layer_norm_latte and add a flag

* remove layer_norm_latte and add a flag

* update latte_pipeline

* update latte_pipeline

* remove unuse squeeze

* add norm_hidden_states.ndim == 2: # for Latte

* fixed test latte pipeline bugs

* fixed test latte pipeline bugs

* delete sh

* add doc for latte

* add licensing

* Move Transformer3DModelOutput to modeling_outputs

* give a default value to sample_size

* remove the einops dependency

* change norm2 for latte

* modify pipeline of latte

* update test for Latte

* modify some codes for latte

* modify for Latte pipeline

* modify for Latte pipeline

* modify for Latte pipeline

* modify for Latte pipeline

* modify for Latte pipeline

* modify for Latte pipeline

* modify for Latte pipeline

* modify for Latte pipeline

* modify for Latte pipeline

* modify for Latte pipeline

* modify for Latte pipeline

* modify for Latte pipeline

* modify for Latte pipeline

* modify for Latte pipeline

* modify for Latte pipeline

* modify for Latte pipeline

* modify for Latte pipeline

* modify for Latte pipeline

* modify for Latte pipeline

* modify for Latte pipeline

* modify for Latte pipeline

* modify for Latte pipeline

* modify for Latte pipeline

* modify for Latte pipeline

* modify for Latte pipeline

* modify for Latte pipeline

* modify for Latte pipeline

* video_length -> num_frames; update prepare_latents copied from

* make fix-copies

* make style

* typo: videe -> video

* update

* modify for Latte pipeline

* modify latte pipeline

* modify latte pipeline

* modify latte pipeline

* modify latte pipeline

* modify for Latte pipeline

* Delete .vscode directory

* make style

* make fix-copies

* add latte transformer 3d to docs _toctree.yml

* update example

* reduce frames for test

* fixed bug of _text_preprocessing

* set num frame to 1 for testing

* remove unuse print

* add text = self._clean_caption(text) again

---------
Co-authored-by: Sayak Paul <spsayakpaul@gmail.com>
Co-authored-by: YiYi Xu <yixu310@gmail.com>
Co-authored-by: Aryan <contact.aryanvs@gmail.com>
Co-authored-by: Aryan <aryan@huggingface.co>

b8cf84a3

08 Jul, 2024 1 commit

[Alpha-VLLM Team] Add Lumina-T2X to diffusers (#8652) · 98388670

PommesPeter authored Jul 08, 2024




---------
Co-authored-by: zhuole1025 <zhuole1025@gmail.com>
Co-authored-by: YiYi Xu <yixu310@gmail.com>

98388670

03 Jul, 2024 2 commits

Revert "[LoRA] introduce `LoraBaseMixin` to promote reusability." (#8773) · 984d3405
Sayak Paul authored Jul 03, 2024
```
Revert "[LoRA] introduce `LoraBaseMixin` to promote reusability. (#8670)"

This reverts commit a2071a18.
```
984d3405

[LoRA] introduce `LoraBaseMixin` to promote reusability. (#8670) · a2071a18

Sayak Paul authored Jul 03, 2024

* introduce  to promote reusability.

* up

* add more tests

* up

* remove comments.

* fix fuse_nan test

* clarify the scope of fuse_lora and unfuse_lora

* remove space

a2071a18

01 Jul, 2024 2 commits

Remove legacy single file model loading mixins (#8754) · 0368483b
Dhruv Nair authored Jul 01, 2024
```
update
```
0368483b

[doc] add a tip about using SDXL refiner with hunyuan-dit and pixart (#8735) · ddb9d854

YiYi Xu authored Jul 01, 2024



* up

* Apply suggestions from code review
Co-authored-by: Steven Liu <59462357+stevhliu@users.noreply.github.com>

---------
Co-authored-by: Sayak Paul <spsayakpaul@gmail.com>
Co-authored-by: Steven Liu <59462357+stevhliu@users.noreply.github.com>

ddb9d854

29 Jun, 2024 1 commit
- add PAG support for SD architecture (#8725) · 8690e8b9
  Shauray Singh authored Jun 30, 2024
```
* add pag to sd pipelines
```
  8690e8b9
26 Jun, 2024 2 commits
- [Chore] remove deprecation from transformer2d regarding the output class. (#8698) · 10b4e354
  Sayak Paul authored Jun 26, 2024
```
* remove deprecation from transformer2d regarding the output class.

* up

* deprecate more
```
  10b4e354
- [Tencent Hunyuan Team] Add Hunyuan-DiT ControlNet Inference (#8694) · fa2abfdb
  XCL authored Jun 26, 2024
```
* add controlnet support

---------
Co-authored-by: xingchaoliu <xingchaoliu@tencent.com>
Co-authored-by: yiyixuxu <yixu310@gmail,com>
```
  fa2abfdb
25 Jun, 2024 2 commits

[Docs] SD3 T5 Token limit doc (#8654) · 14d224d4

Álvaro Somoza authored Jun 25, 2024



* doc for max_sequence_length

* better position and changed note to tip

* apply suggestions

---------
Co-authored-by: Sayak Paul <spsayakpaul@gmail.com>

14d224d4

add PAG support (#7944) · 540399f5

YiYi Xu authored Jun 25, 2024



* first draft


---------
Co-authored-by: yiyixuxu <yixu310@gmail,com>
Co-authored-by: Junhwa Song <ethan9867@gmail.com>
Co-authored-by: Ahn Donghoon (안동훈 / suno) <suno.vivid@gmail.com>
Co-authored-by: Sayak Paul <spsayakpaul@gmail.com>
Co-authored-by: Steven Liu <59462357+stevhliu@users.noreply.github.com>

540399f5

24 Jun, 2024 4 commits

[docs] Add note for float8 (#8685) · 675be88f
Steven Liu authored Jun 24, 2024
```
add note
```
675be88f

Errata - Fix typos and improve style (#8571) · f040c27d

Tolga Cangöz authored Jun 24, 2024



* Fix typos

* Fix typos & up style

* chore: Update numbers

---------
Co-authored-by: Sayak Paul <spsayakpaul@gmail.com>

f040c27d

Discourage using deprecated `revision` parameter (#8573) · 138fac70

Tolga Cangöz authored Jun 24, 2024



* Discourage using `revision`

* `make style && make quality`

* Refactor code to use 'variant' instead of 'revision'

* `revision="bf16"` -> `variant="bf16"`

---------
Co-authored-by: Sayak Paul <spsayakpaul@gmail.com>

138fac70

Errata - Trim trailing white space in the whole repo (#8575) · 468ae09e

Tolga Cangöz authored Jun 24, 2024



* Trim all the trailing white space in the whole repo

* Remove unnecessary empty places

* make style && make quality

* Trim trailing white space

* trim

---------
Co-authored-by: Sayak Paul <spsayakpaul@gmail.com>

468ae09e

21 Jun, 2024 1 commit

[LoRA] get rid of the legacy lora remnants and make our codebase lighter (#8623) · 8eb17315

Sayak Paul authored Jun 21, 2024

* get rid of the legacy lora remnants and make our codebase lighter

* fix depcrecated lora argument

* fix

* empty commit to trigger ci

* remove print

* empty

8eb17315

19 Jun, 2024 1 commit
- Support SD3 ControlNet and Multi-ControlNet. (#8566) · e5564d45
  王奇勋 authored Jun 19, 2024
```
* sd3 controlnet



---------
Co-authored-by: haofanwang <haofanwang.ai@gmail.com>
```
  e5564d45
18 Jun, 2024 3 commits

[SD3 Docs] Corrected title about loading model with T5 "without" -> "with" (#8602) · 34fab8b5

Vasco Ramos authored Jun 18, 2024

[SD3 Docs] Corrected title about loading model with T5

Corrected the documentation title to "Loading the single file checkpoint with T5" Previously, it incorrectly stated "Loading the single file checkpoint without T5" which contradicted the code snippet showing how to load the SD3 checkpoint with the T5 model

34fab8b5

[Core] Add `shift_factor` to SD3 tiny autoencoder (#8618) · cd308200
Sayak Paul authored Jun 18, 2024
```
* shift factor argument to tiny

* remove shift factor rejigging from the sd3 docs
```
cd308200

[SD3] TAESD3 docs (#8607) · d2b10b1f

Álvaro Somoza authored Jun 18, 2024



* tased3 docs

* apply suggestion

---------
Co-authored-by: Sayak Paul <spsayakpaul@gmail.com>

d2b10b1f

13 Jun, 2024 1 commit
- Expand Single File support in SD3 Pipeline (#8517) · b1a2c0d5
  Dhruv Nair authored Jun 13, 2024
```
* update

* update
```
  b1a2c0d5
12 Jun, 2024 2 commits

Fix small typo (#8498) · 95e0c375
Radamés Ajna authored Jun 12, 2024

95e0c375

Add Stable Diffusion 3 (#8483) · 04717fd8

Dhruv Nair authored Jun 13, 2024



* up

* add sd3

* update

* update

* add tests

* fix copies

* fix docs

* update

* add dreambooth lora

* add LoRA

* update

* update

* update

* update

* import fix

* update

* Update src/diffusers/pipelines/stable_diffusion_3/pipeline_stable_diffusion_3.py
Co-authored-by: YiYi Xu <yixu310@gmail.com>

* import fix 2

* update

* Update src/diffusers/models/autoencoders/autoencoder_kl.py
Co-authored-by: YiYi Xu <yixu310@gmail.com>

* Update src/diffusers/models/autoencoders/autoencoder_kl.py
Co-authored-by: YiYi Xu <yixu310@gmail.com>

* Update src/diffusers/models/autoencoders/autoencoder_kl.py
Co-authored-by: YiYi Xu <yixu310@gmail.com>

* Update src/diffusers/models/autoencoders/autoencoder_kl.py
Co-authored-by: YiYi Xu <yixu310@gmail.com>

* Update src/diffusers/models/autoencoders/autoencoder_kl.py
Co-authored-by: YiYi Xu <yixu310@gmail.com>

* Update src/diffusers/models/autoencoders/autoencoder_kl.py
Co-authored-by: YiYi Xu <yixu310@gmail.com>

* Update src/diffusers/models/autoencoders/autoencoder_kl.py
Co-authored-by: YiYi Xu <yixu310@gmail.com>

* Update src/diffusers/models/autoencoders/autoencoder_kl.py
Co-authored-by: YiYi Xu <yixu310@gmail.com>

* Update src/diffusers/models/autoencoders/autoencoder_kl.py
Co-authored-by: YiYi Xu <yixu310@gmail.com>

* Update src/diffusers/models/autoencoders/autoencoder_kl.py
Co-authored-by: YiYi Xu <yixu310@gmail.com>

* Update src/diffusers/models/autoencoders/autoencoder_kl.py
Co-authored-by: YiYi Xu <yixu310@gmail.com>

* update

* update

* update

* fix ckpt id

* fix more ids

* update

* missing doc

* Update src/diffusers/schedulers/scheduling_flow_match_euler_discrete.py
Co-authored-by: YiYi Xu <yixu310@gmail.com>

* Update src/diffusers/schedulers/scheduling_flow_match_euler_discrete.py
Co-authored-by: YiYi Xu <yixu310@gmail.com>

* Update docs/source/en/api/pipelines/stable_diffusion/stable_diffusion_3.md
Co-authored-by: Sayak Paul <spsayakpaul@gmail.com>

* Update docs/source/en/api/pipelines/stable_diffusion/stable_diffusion_3.md
Co-authored-by: Sayak Paul <spsayakpaul@gmail.com>

* update'

* fix

* update

* Update src/diffusers/models/autoencoders/autoencoder_kl.py

* Update src/diffusers/models/autoencoders/autoencoder_kl.py

* note on gated access.

* requirements

* licensing

---------
Co-authored-by: sayakpaul <spsayakpaul@gmail.com>
Co-authored-by: YiYi Xu <yixu310@gmail.com>

04717fd8

06 Jun, 2024 3 commits

Optimize test files by fixing CPU-offloading usage (#8409) · ec1aded1

Tolga Cangöz authored Jun 06, 2024

* Refactor code to remove unnecessary calls to `to(torch_device)`

* Refactor code to remove unnecessary calls to `to("cuda")`

* Update pipeline_stable_diffusion_diffedit.py

ec1aded1

[docs] Single file usage (#8412) · 151a56b8
Steven Liu authored Jun 06, 2024
```
* single file usage

* edit
```
151a56b8

[Hunyuan] add optimization related sections to the hunyuan dit docs. (#8402) · 867a2b0c

Sayak Paul authored Jun 06, 2024



* optimizations to the hunyuan dit docs.

* Apply suggestions from code review
Co-authored-by: Steven Liu <59462357+stevhliu@users.noreply.github.com>

* Update docs/source/en/api/pipelines/hunyuandit.md
Co-authored-by: Steven Liu <59462357+stevhliu@users.noreply.github.com>

---------
Co-authored-by: Steven Liu <59462357+stevhliu@users.noreply.github.com>

867a2b0c

05 Jun, 2024 2 commits

Errata (#8322) · 98730c5d

Tolga Cangöz authored Jun 05, 2024

* Fix typos

* Trim trailing whitespaces

* Remove a trailing whitespace

* chore: Update MarigoldDepthPipeline checkpoint to prs-eth/marigold-lcm-v1-0

* Revert "chore: Update MarigoldDepthPipeline checkpoint to prs-eth/marigold-lcm-v1-0"

This reverts commit fd742b30b4258106008a6af4d0dd4664904f8595.

* pokemon -> naruto

* `DPMSolverMultistep` -> `DPMSolverMultistepScheduler`

* Improve Markdown stylization

* Improve style

* Improve style

* Refactor pipeline variable names for consistency

* up style

98730c5d

[Hunyuan] allow Hunyuan DiT to run under 6GB for GPU VRAM (#8399) · 2f6f426f
Sayak Paul authored Jun 05, 2024
```
* allow hunyuan dit to run under 6GB for GPU VRAM

* add section in the docs/
```
2f6f426f

04 Jun, 2024 2 commits

[HunyuanDiT] minor docs changes in hunyuandit (#8395) · 3ff39e8e
Sayak Paul authored Jun 04, 2024
```
minor docs changes in hunyuandit
```
3ff39e8e

Update transformer2d.md title (#8375) · dc89434b

Marçal Comajoan Cara authored Jun 03, 2024

* Update transformer2d.md title

For the other classes (e.g., UNet2DModel) the title of the documentation coincides with the name of the class, but that was not the case for Transformer2DModel.

* Update model docs titles for consistency with class names

dc89434b

03 Jun, 2024 1 commit

Tencent Hunyuan Team - Updated Doc for HunyuanDiT (#8383) · 174cf868

XCL authored Jun 03, 2024

* add hunyuandit doc

* update hunyuandit doc

* update hunyuandit 2d model

* update toctree.yml for hunyuandit

174cf868

31 May, 2024 1 commit

[Core] Introduce class variants for `Transformer2DModel` (#7647) · 983dec3b

Sayak Paul authored May 31, 2024

* init for patches

* finish patched model.

* continuous transformer

* vectorized transformer2d.

* style.

* inits.

* fix-copies.

* introduce DiTTransformer2DModel.

* fixes

* use REMAPPING as suggested by @DN6

* better logging.

* add pixart transformer model.

* inits.

* caption_channels.

* attention masking.

* fix use_additional_conditions.

* remove print.

* debug

* flatten

* fix: assertion for sigma

* handle remapping for modeling_utils

* add tests for dit transformer2d

* quality

* placeholder for pixart tests

* pixart tests

* add _no_split_modules

* add docs.

* check

* check

* check

* check

* fix tests

* fix tests

* move Transformer output to modeling_output

* move errors better and bring back use_additional_conditions attribute.

* add unnecessary things from DiT.

* clean up pixart

* fix remapping

* fix device_map things in pixart2d.

* replace Transformer2DModel with appropriate classes in dit, pixart tests

* empty

* legacy mixin classes./

* use a remapping dict for fetching class names.

* change to specifc model types in the pipeline implementations.

* move _fetch_remapped_cls_from_config to modeling_loading_utils.py

* fix dependency problems.

* add deprecation note.

983dec3b

29 May, 2024 1 commit
- move `vqmodel` to `models.autoencoders`. (#8292) · 5edd0b34
  Sayak Paul authored May 29, 2024
```
move vqmodel to models.autoencoders.
```
  5edd0b34
27 May, 2024 1 commit

[Pipeline] Marigold depth and normals estimation (#7847) · b3d10d6d

Anton Obukhov authored May 27, 2024



* implement marigold depth and normals pipelines in diffusers core

* remove bibtex

* remove deprecations

* remove save_memory argument

* remove validate_vae

* remove config output

* remove batch_size autodetection

* remove presets logic
move default denoising_steps and processing_resolution into the model config
make default ensemble_size 1

* remove no_grad

* add fp16 to the example usage

* implement is_matplotlib_available
use is_matplotlib_available, is_scipy_available for conditional imports in the marigold depth pipeline

* move colormap, visualize_depth, and visualize_normals into export_utils.py

* make the denoising loop more lucid
fix the outputs to always be 4d tensors or lists of pil images
support a 4d input_image case
attempt to support model_cpu_offload_seq
move check_inputs into a separate function
change default batch_size to 1, remove any logic to make it bigger implicitly

* style

* rename denoising_steps into num_inference_steps

* rename input_image into image

* rename input_latent into latents

* remove decode_image
change decode_prediction to use the AutoencoderKL.decode method

* move clean_latent outside of progress_bar

* refactor marigold-reusable image processing bits into MarigoldImageProcessor class

* clean up the usage example docstring

* make ensemble functions members of the pipelines

* add early checks in check_inputs
rename E into ensemble_size in depth ensembling

* fix vae_scale_factor computation

* better compatibility with torch.compile
better variable naming

* move export_depth_to_png to export_utils

* remove encode_prediction

* improve visualize_depth and visualize_normals to accept multi-dimensional data and lists
remove visualization functions from the pipelines
move exporting depth as 16-bit PNGs functionality from the depth pipeline
update example docstrings

* do not shortcut vae.config variables

* change all asserts to raise ValueError

* rename output_prediction_type to output_type

* better variable names
clean up variable deletion code

* better variable names

* pass desc and leave kwargs into the diffusers progress_bar
implement nested progress bar for images and steps loops

* implement scale_invariant and shift_invariant flags in the ensemble_depth function
add scale_invariant and shift_invariant flags readout from the model config
further refactor ensemble_depth
support ensembling without alignment
add ensemble_depth docstring

* fix generator device placement checks

* move encode_empty_text body into the pipeline call

* minor empty text encoding simplifications

* adjust pipelines' class docstrings to explain the added construction arguments

* improve the scipy failure condition
add comments
improve docstrings
change the default use_full_z_range to True

* make input image values range check configurable in the preprocessor
refactor load_image_canonical in preprocessor to reject unknown types and return the image in the expected 4D format of tensor and on right device
support a list of everything as inputs to the pipeline, change type to PipelineImageInput
implement a check that all input list elements have the same dimensions
improve docstrings of pipeline outputs
remove check_input pipeline argument

* remove forgotten print

* add prediction_type model config

* add uncertainty visualization into export utils
fix NaN values in normals uncertainties

* change default of output_uncertainty to False
better handle the case of an attempt to export or visualize none

* fix `output_uncertainty=False`

* remove kwargs
fix check_inputs according to the new inputs of the pipeline

* rename prepare_latent into prepare_latents as in other pipelines
annotate prepare_latents in normals pipeline with "Copied from"
annotate encode_image in normals pipeline with "Copied from"

* move nested-capable `progress_bar` method into the pipelines
revert the original `progress_bar` method in pipeline_utils

* minor message improvement

* fix cpu offloading

* move colormap, visualize_depth, export_depth_to_16bit_png, visualize_normals, visualize_uncertainty to marigold_image_processing.py
update example docstrings

* fix missing comma

* change torch.FloatTensor to torch.Tensor

* fix importing of MarigoldImageProcessor

* fix vae offloading
fix batched image encoding
remove separate encode_image function and use vae.encode instead

* implement marigold's intial tests
relax generator checks in line with other pipelines
implement return_dict __call__ argument in line with other pipelines

* fix num_images computation

* remove MarigoldImageProcessor and outputs from import structure
update tests

* update docstrings

* update init

* update

* style

* fix

* fix

* up

* up

* up

* add simple test

* up

* update expected np input/output to be channel last

* move expand_tensor_or_array into the MarigoldImageProcessor

* rewrite tests to follow conventions - hardcoded slices instead of image artifacts
write more smoke tests

* add basic docs.

* add anton's contribution statement

* remove todos.

* fix assertion values for marigold depth slow tests

* fix assertion values for depth normals.

* remove print

* support AutoencoderTiny in the pipelines

* update documentation page
add Available Pipelines section
add Available Checkpoints section
add warning about num_inference_steps

* fix missing import in docstring
fix wrong value in visualize_depth docstring

* [doc] add marigold to pipelines overview

* [doc] add section "usage examples"

* fix an issue with latents check in the pipelines

* add "Frame-by-frame Video Processing with Consistency" section

* grammarly

* replace tables with images with css-styled images (blindly)

* style

* print

* fix the assertions.

* take from the github runner.

* take the slices from action artifacts

* style.

* update with the slices from the runner.

* remove unnecessary code blocks.

* Revert "[doc] add marigold to pipelines overview"

This reverts commit a505165150afd8dab23c474d1a054ea505a56a5f.

* remove invitation for new modalities

* split out marigold usage examples

* doc cleanup

---------
Co-authored-by: yiyixuxu <yixu310@gmail.com>
Co-authored-by: yiyixuxu <yixu310@gmail,com>
Co-authored-by: sayakpaul <spsayakpaul@gmail.com>

b3d10d6d