Commits · ba8241410646ffd3b81afa8eb0650d904de32593 · renzhc / diffusers_dcu

"vscode:/vscode.git/clone" did not exist on "d052f4c8a9fb7e135ca0f0b09f6feead93db9e01"

28 May, 2024 1 commit

[docs] Add controlnet example to marigold (#8289) · ba824141

Álvaro Somoza authored May 28, 2024



* initial doc

* fix wrong LCM sentence

* implement binary colormap without requiring matplotlib
update section about Marigold for ControlNet
update formatting of marigold_usage.md

* fix indentation

---------
Co-authored-by: anton <anton.obukhov@gmail.com>

ba824141

27 May, 2024 5 commits

install wget. (#8285) · fe5f035f
Sayak Paul authored May 27, 2024

fe5f035f

[Pipeline] Marigold depth and normals estimation (#7847) · b3d10d6d

Anton Obukhov authored May 27, 2024



* implement marigold depth and normals pipelines in diffusers core

* remove bibtex

* remove deprecations

* remove save_memory argument

* remove validate_vae

* remove config output

* remove batch_size autodetection

* remove presets logic
move default denoising_steps and processing_resolution into the model config
make default ensemble_size 1

* remove no_grad

* add fp16 to the example usage

* implement is_matplotlib_available
use is_matplotlib_available, is_scipy_available for conditional imports in the marigold depth pipeline

* move colormap, visualize_depth, and visualize_normals into export_utils.py

* make the denoising loop more lucid
fix the outputs to always be 4d tensors or lists of pil images
support a 4d input_image case
attempt to support model_cpu_offload_seq
move check_inputs into a separate function
change default batch_size to 1, remove any logic to make it bigger implicitly

* style

* rename denoising_steps into num_inference_steps

* rename input_image into image

* rename input_latent into latents

* remove decode_image
change decode_prediction to use the AutoencoderKL.decode method

* move clean_latent outside of progress_bar

* refactor marigold-reusable image processing bits into MarigoldImageProcessor class

* clean up the usage example docstring

* make ensemble functions members of the pipelines

* add early checks in check_inputs
rename E into ensemble_size in depth ensembling

* fix vae_scale_factor computation

* better compatibility with torch.compile
better variable naming

* move export_depth_to_png to export_utils

* remove encode_prediction

* improve visualize_depth and visualize_normals to accept multi-dimensional data and lists
remove visualization functions from the pipelines
move exporting depth as 16-bit PNGs functionality from the depth pipeline
update example docstrings

* do not shortcut vae.config variables

* change all asserts to raise ValueError

* rename output_prediction_type to output_type

* better variable names
clean up variable deletion code

* better variable names

* pass desc and leave kwargs into the diffusers progress_bar
implement nested progress bar for images and steps loops

* implement scale_invariant and shift_invariant flags in the ensemble_depth function
add scale_invariant and shift_invariant flags readout from the model config
further refactor ensemble_depth
support ensembling without alignment
add ensemble_depth docstring

* fix generator device placement checks

* move encode_empty_text body into the pipeline call

* minor empty text encoding simplifications

* adjust pipelines' class docstrings to explain the added construction arguments

* improve the scipy failure condition
add comments
improve docstrings
change the default use_full_z_range to True

* make input image values range check configurable in the preprocessor
refactor load_image_canonical in preprocessor to reject unknown types and return the image in the expected 4D format of tensor and on right device
support a list of everything as inputs to the pipeline, change type to PipelineImageInput
implement a check that all input list elements have the same dimensions
improve docstrings of pipeline outputs
remove check_input pipeline argument

* remove forgotten print

* add prediction_type model config

* add uncertainty visualization into export utils
fix NaN values in normals uncertainties

* change default of output_uncertainty to False
better handle the case of an attempt to export or visualize none

* fix `output_uncertainty=False`

* remove kwargs
fix check_inputs according to the new inputs of the pipeline

* rename prepare_latent into prepare_latents as in other pipelines
annotate prepare_latents in normals pipeline with "Copied from"
annotate encode_image in normals pipeline with "Copied from"

* move nested-capable `progress_bar` method into the pipelines
revert the original `progress_bar` method in pipeline_utils

* minor message improvement

* fix cpu offloading

* move colormap, visualize_depth, export_depth_to_16bit_png, visualize_normals, visualize_uncertainty to marigold_image_processing.py
update example docstrings

* fix missing comma

* change torch.FloatTensor to torch.Tensor

* fix importing of MarigoldImageProcessor

* fix vae offloading
fix batched image encoding
remove separate encode_image function and use vae.encode instead

* implement marigold's intial tests
relax generator checks in line with other pipelines
implement return_dict __call__ argument in line with other pipelines

* fix num_images computation

* remove MarigoldImageProcessor and outputs from import structure
update tests

* update docstrings

* update init

* update

* style

* fix

* fix

* up

* up

* up

* add simple test

* up

* update expected np input/output to be channel last

* move expand_tensor_or_array into the MarigoldImageProcessor

* rewrite tests to follow conventions - hardcoded slices instead of image artifacts
write more smoke tests

* add basic docs.

* add anton's contribution statement

* remove todos.

* fix assertion values for marigold depth slow tests

* fix assertion values for depth normals.

* remove print

* support AutoencoderTiny in the pipelines

* update documentation page
add Available Pipelines section
add Available Checkpoints section
add warning about num_inference_steps

* fix missing import in docstring
fix wrong value in visualize_depth docstring

* [doc] add marigold to pipelines overview

* [doc] add section "usage examples"

* fix an issue with latents check in the pipelines

* add "Frame-by-frame Video Processing with Consistency" section

* grammarly

* replace tables with images with css-styled images (blindly)

* style

* print

* fix the assertions.

* take from the github runner.

* take the slices from action artifacts

* style.

* update with the slices from the runner.

* remove unnecessary code blocks.

* Revert "[doc] add marigold to pipelines overview"

This reverts commit a505165150afd8dab23c474d1a054ea505a56a5f.

* remove invitation for new modalities

* split out marigold usage examples

* doc cleanup

---------
Co-authored-by: yiyixuxu <yixu310@gmail.com>
Co-authored-by: yiyixuxu <yixu310@gmail,com>
Co-authored-by: sayakpaul <spsayakpaul@gmail.com>

b3d10d6d

Add zip package to doc builder image (#8284) · b82f9f56
Dhruv Nair authored May 27, 2024
```
update
```
b82f9f56

[Workflows] add a more secure way to run tests from a PR. (#7969) · 6a5ba1b7

Sayak Paul authored May 27, 2024



* add a more secure way to run tests from a PR.

* make pytest more secure.

* address dhruv's comments.

* improve validation check.

* Update .github/workflows/run_tests_from_a_pr.yml
Co-authored-by: Dhruv Nair <dhruv.nair@gmail.com>

---------
Co-authored-by: Dhruv Nair <dhruv.nair@gmail.com>

6a5ba1b7

Add details about 1-stage implementation in I2VGen-XL docs (#8282) · 4d40c914
Dhaivat Bhatt authored May 27, 2024
```
* Add details about 1-stage implementation

* Add details about 1-stage implementation
```
4d40c914

24 May, 2024 9 commits

Fix CPU Offloading Usage & Typos (#8230) · 0ab63ff6

Tolga Cangöz authored May 24, 2024

* Fix typos

* Fix `pipe.enable_model_cpu_offload()` usage

* Fix cpu offloading

* Update numbers

0ab63ff6

Fix a grammatical error in the `raise` messages (#8272) · db33af06
Tolga Cangöz authored May 24, 2024
```
Fix grammatical error
```
db33af06

sampling bug fix in diffusers tutorial "basic_training.md" (#8223) · 1096f88e

Yue Wu authored May 24, 2024

sampling bug fix in basic_training.md

In the diffusers basic training tutorial, setting the manual seed argument (generator=torch.manual_seed(config.seed)) in the pipeline call inside evaluate() function rewinds the dataloader shuffling, leading to overfitting due to the model seeing same sequence of training examples after every evaluation call. Using generator=torch.Generator(device='cpu').manual_seed(config.seed) avoids this.

1096f88e

Clean up `from_single_file` docs (#8268) · cef4a512
Dhruv Nair authored May 24, 2024
```
* update

* update
```
cef4a512
Respect `resume_download` deprecation V2 (#8267) · edf5ba6a
Lucain authored May 24, 2024
```
* Fix resume_downoad FutureWarning

* only resume download
```
edf5ba6a
[Chore] run the documentation workflow in a custom container. (#8266) · 9941f1f6
Sayak Paul authored May 24, 2024
```
run the documentation workflow in a custom container.
```
9941f1f6

[Community Pipeline] FRESCO: Spatial-Temporal Correspondence for Zero-Shot... · 46a9db03

Yifan Zhou authored May 24, 2024


[Community Pipeline] FRESCO: Spatial-Temporal Correspondence for Zero-Shot Video Translation (#8239)

* code and doc

* update paper link

* remove redundant codes

* add example video

---------
Co-authored-by: Sayak Paul <spsayakpaul@gmail.com>

46a9db03

Use `freedesktop_os_release()` in diffusers cli for Python >=3.10 (#8235) · 370146e4
Dhruv Nair authored May 24, 2024
```
* update

* update
```
370146e4
Create custom container for doc builder (#8263) · 5cd45c24
Dhruv Nair authored May 24, 2024
```
* update

* update
```
5cd45c24

23 May, 2024 1 commit
- Fix resize issue in SVD pipeline with VideoProcessor (#8229) · 67b3fe0a
  Dhruv Nair authored May 23, 2024
```
update
Co-authored-by: Sayak Paul <spsayakpaul@gmail.com>
```
  67b3fe0a
22 May, 2024 2 commits
- Remove unnecessary single file tests for SD Cascade UNet (#7996) · baab0656
  Dhruv Nair authored May 22, 2024
```
update
```
  baab0656
- fix: Attribute error in Logger object (logger.warning) (#8183) · 509741ae
  BootesVoid authored May 21, 2024
  
  509741ae
21 May, 2024 2 commits
- Use HF_TOKEN env var in CI (#7993) · e1df77ee
  Lucain authored May 21, 2024
  
  e1df77ee
- [docs] VideoProcessor (#7965) · fdb1baa0
  Steven Liu authored May 20, 2024
```
* fix?

* fix?

* fix
```
  fdb1baa0
20 May, 2024 6 commits

Make VAE compatible to torch.compile() (#7984) · 6529ee67
Vinh H. Pham authored May 21, 2024
```
make VAE compatible to torch.compile()
Co-authored-by: YiYi Xu <yixu310@gmail.com>
```
6529ee67
fix: Fixed few `docstrings` according to the Google Style Guide (#7717) · df2bc5ef
Sai-Suraj-27 authored May 20, 2024
```
Fixed few docstrings according to the Google Style Guide.
```
df2bc5ef

Passing `cross_attention_kwargs` to `StableDiffusionInstructPix2PixPipeline` (#7961) · a7bf77fc

Aleksei Zhuravlev authored May 20, 2024

* Update pipeline_stable_diffusion_instruct_pix2pix.py

Add `cross_attention_kwargs` to `__call__` method of `StableDiffusionInstructPix2PixPipeline`, which are passed to UNet.

* Update documentation for pipeline_stable_diffusion_instruct_pix2pix.py

* Update docstring

* Update docstring

* Fix typing import

a7bf77fc

[docs] add doc for PixArtSigmaPipeline (#7857) · 0f0defdb

Junsong Chen authored May 21, 2024



* 1. add doc for PixArtSigmaPipeline;

---------
Co-authored-by: Sayak Paul <spsayakpaul@gmail.com>
Co-authored-by: Steven Liu <59462357+stevhliu@users.noreply.github.com>
Co-authored-by: Guillaume LEGENDRE <glegendre01@gmail.com>
Co-authored-by: Álvaro Somoza <asomoza@users.noreply.github.com>
Co-authored-by: Bagheera <59658056+bghira@users.noreply.github.com>
Co-authored-by: bghira <bghira@users.github.com>
Co-authored-by: Hyoungwon Cho <jhw9811@korea.ac.kr>
Co-authored-by: yiyixuxu <yixu310@gmail.com>
Co-authored-by: Tolga Cangöz <46008593+standardAI@users.noreply.github.com>
Co-authored-by: Philip Pham <phillypham@google.com>

0f0defdb

Update pipeline_controlnet_inpaint_sd_xl.py (#7983) · 19df9f3e
Nikita authored May 20, 2024

19df9f3e
Fix typo in "attention" (#7977) · d6ca1209
Jacob Marks authored May 20, 2024

d6ca1209

19 May, 2024 1 commit
- [tests] fix Pixart Sigma tests (#7966) · fb7ae018
  Sayak Paul authored May 19, 2024
```
* checking tests

* checking ii.

* remove prints.

* test_pixart_1024

* fix 1024.
```
  fb7ae018
17 May, 2024 1 commit
- remove unsafe workflow. (#7967) · 70f8d4b4
  Sayak Paul authored May 17, 2024
  
  70f8d4b4
16 May, 2024 4 commits

Consistent SDXL Controlnet callback tensor inputs (#7958) · 6c60e430

Álvaro Somoza authored May 16, 2024

* make _callback_tensor_inputs consistent between sdxl pipelines

* forgot this one

* fix failing test

* fix test_components_function

* fix controlnet inpaint tests

6c60e430

Fix AttributeError in train_lcm_distill_lora_sdxl_wds.py (#7923) · 1221b28e

Alphin Jain authored May 16, 2024



Fix conditional teacher model check in train_lcm_distill_lora_sdxl_wds.py
Co-authored-by: Sayak Paul <spsayakpaul@gmail.com>

1221b28e

Fix the text tokenizer name in logger warning of PixArt pipelines (#7912) · 746f603b
Liang Hou authored May 16, 2024
```
Fix CLIP to T5 in logger warning
```
746f603b

refactor: Refactored code by Merging `isinstance` calls (#7710) · 2afea72d

Sai-Suraj-27 authored May 16, 2024



* Merged isinstance calls to make the code simpler.

* Corrected formatting errors using ruff.

---------
Co-authored-by: Sayak Paul <spsayakpaul@gmail.com>
Co-authored-by: YiYi Xu <yixu310@gmail.com>

2afea72d

15 May, 2024 4 commits

[Workflows] add a workflow that can be manually triggered on a PR. (#7942) · 0f111ab7
Sayak Paul authored May 15, 2024
```
* add a workflow that can be manually triggered on a PR.

* remove sudo

* add command

* small fixes.
```
0f111ab7
move to GH hosted M1 runner (#7949) · 4dd7aaa0
Guillaume LEGENDRE authored May 15, 2024

4dd7aaa0

Adding VQGAN Training script (#5483) · d27e996c

Isamu Isozaki authored May 15, 2024



* Init commit

* Removed einops

* Added default movq config for training

* Update explanation of prompts

* Fixed inheritance of discriminator and init_tracker

* Fixed incompatible api between muse and here

* Fixed output

* Setup init training

* Basic structure done

* Removed attention for quick tests

* Style fixes

* Fixed vae/vqgan styles

* Removed redefinition of wandb

* Fixed log_validation and tqdm

* Nothing commit

* Added commit loss to lookup_from_codebook

* Update src/diffusers/models/vq_model.py
Co-authored-by: Sayak Paul <spsayakpaul@gmail.com>

* Adding perliminary README

* Fixed one typo

* Local changes

* Fixed main issues

* Merging

* Update src/diffusers/models/vq_model.py
Co-authored-by: Sayak Paul <spsayakpaul@gmail.com>

* Testing+Fixed bugs in training script

* Some style fixes

* Added wandb to docs

* Fixed timm test

* get testing suite ready.

* remove return loss

* remove return_loss

* Remove diffs

* Remove diffs

* fix ruff format

---------
Co-authored-by: Sayak Paul <spsayakpaul@gmail.com>
Co-authored-by: Dhruv Nair <dhruv.nair@gmail.com>

d27e996c

[tests] decorate StableDiffusion21PipelineSingleFileSlowTests with slow. (#7941) · 72780ff5
Sayak Paul authored May 15, 2024
```
decorate StableDiffusion21PipelineSingleFileSlowTests with slow.
```
72780ff5

14 May, 2024 4 commits
- [Pipeline] Adding BoxDiff to community examples (#7947) · 69fdb872
  Jingyang Zhang authored May 14, 2024
```
add boxdiff to community examples
```
  69fdb872
- Fix `added_cond_kwargs` when using IP-Adapter in StableDiffusionXLControlNetInpaintPipeline (#7924) · b2140a89
  Nikita authored May 14, 2024
```
Fix `added_cond_kwargs` when using IP-Adapter

Fix error when using IP-Adapter in pipeline and passing `ip_adapter_image_embeds` instead of `ip_adapter_image`
Co-authored-by: YiYi Xu <yixu310@gmail.com>
```
  b2140a89
- [Core] separate the loading utilities in modeling similar to pipelines. (#7943) · e0e8c58f
  Sayak Paul authored May 14, 2024
```
separate the loading utilities in modeling similar to pipelines.
```
  e0e8c58f
- update to use hf-workflows for reporting the Docker build statuses (#7938) · cbea5d17
  Sayak Paul authored May 14, 2024
```
update to use hf-workflows for reporting
```
  cbea5d17