diffusers

mirror of https://github.com/huggingface/diffusers.git synced 2026-01-27 17:22:53 +03:00

Author	SHA1	Message	Date
Sayak Paul	f685981ed0	[docs] minor updates to dtype map docs. (#11237 ) minor updates to dtype map docs.	2025-04-09 08:38:17 +05:30
Sayak Paul	b924251dd8	minor update to sana sprint docs. (#11236 )	2025-04-09 08:17:45 +05:30
Sayak Paul	1a04812439	[bistandbytes] improve replacement warnings for bnb (#11132 ) * improve replacement warnings for bnb * updates to docs.	2025-04-08 21:18:34 +05:30
Sayak Paul	4b27c4a494	[feat] implement `record_stream` when using CUDA streams during group offloading (#11081 ) * implement record_stream for better performance. * fix * style. * merge #11097 * Update src/diffusers/hooks/group_offloading.py Co-authored-by: Aryan <aryan@huggingface.co> * fixes * docstring. * remaining todos in low_cpu_mem_usage * tests * updates to docs. --------- Co-authored-by: Aryan <aryan@huggingface.co>	2025-04-08 21:17:49 +05:30
hlky	5d49b3e83b	Flux quantized with lora (#10990 ) * Flux quantized with lora * fix * changes * Apply suggestions from code review Co-authored-by: Sayak Paul <spsayakpaul@gmail.com> * Apply style fixes * enable model cpu offload() * Update src/diffusers/loaders/lora_pipeline.py Co-authored-by: hlky <hlky@hlky.ac> * update * Apply suggestions from code review * update * add peft as an additional dependency for gguf --------- Co-authored-by: Sayak Paul <spsayakpaul@gmail.com> Co-authored-by: github-actions[bot] <github-actions[bot]@users.noreply.github.com> Co-authored-by: Dhruv Nair <dhruv.nair@gmail.com>	2025-04-08 21:17:03 +05:30
Linoy Tsaban	71f34fc5a4	[Flux LoRA] fix issues in flux lora scripts (#11111 ) * remove custom scheduler * update requirements.txt * log_validation with mixed precision * add intermediate embeddings saving when checkpointing is enabled * remove comment * fix validation * add unwrap_model for accelerator, torch.no_grad context for validation, fix accelerator.accumulate call in advanced script * revert unwrap_model change temp * add .module to address distributed training bug + replace accelerator.unwrap_model with unwrap model * changes to align advanced script with canonical script * make changes for distributed training + unify unwrap_model calls in advanced script * add module.dtype fix to dreambooth script * unify unwrap_model calls in dreambooth script * fix condition in validation run * mixed precision * Update examples/advanced_diffusion_training/train_dreambooth_lora_flux_advanced.py Co-authored-by: Sayak Paul <spsayakpaul@gmail.com> * smol style change * change autocast * Apply style fixes --------- Co-authored-by: Sayak Paul <spsayakpaul@gmail.com> Co-authored-by: github-actions[bot] <github-actions[bot]@users.noreply.github.com>	2025-04-08 17:40:30 +03:00
Yao Matrix	c51b6bd837	introduce compute arch specific expectations and fix test_sd3_img2img_inference failure (#11227 ) * add arch specfic expectations support, to support different arch's numerical characteristics Signed-off-by: YAO Matrix <matrix.yao@intel.com> * fix typo Signed-off-by: YAO Matrix <matrix.yao@intel.com> * Apply suggestions from code review * Apply style fixes * Update src/diffusers/utils/testing_utils.py --------- Signed-off-by: YAO Matrix <matrix.yao@intel.com> Co-authored-by: hlky <hlky@hlky.ac> Co-authored-by: github-actions[bot] <github-actions[bot]@users.noreply.github.com>	2025-04-08 14:57:49 +01:00
Benjamin Bossan	fb54499614	[LoRA] Implement hot-swapping of LoRA (#9453 ) * [WIP][LoRA] Implement hot-swapping of LoRA This PR adds the possibility to hot-swap LoRA adapters. It is WIP. Description As of now, users can already load multiple LoRA adapters. They can offload existing adapters or they can unload them (i.e. delete them). However, they cannot "hotswap" adapters yet, i.e. substitute the weights from one LoRA adapter with the weights of another, without the need to create a separate LoRA adapter. Generally, hot-swapping may not appear not super useful but when the model is compiled, it is necessary to prevent recompilation. See #9279 for more context. Caveats To hot-swap a LoRA adapter for another, these two adapters should target exactly the same layers and the "hyper-parameters" of the two adapters should be identical. For instance, the LoRA alpha has to be the same: Given that we keep the alpha from the first adapter, the LoRA scaling would be incorrect for the second adapter otherwise. Theoretically, we could override the scaling dict with the alpha values derived from the second adapter's config, but changing the dict will trigger a guard for recompilation, defeating the main purpose of the feature. I also found that compilation flags can have an impact on whether this works or not. E.g. when passing "reduce-overhead", there will be errors of the type: > input name: arg861_1. data pointer changed from 139647332027392 to 139647331054592 I don't know enough about compilation to determine whether this is problematic or not. Current state This is obviously WIP right now to collect feedback and discuss which direction to take this. If this PR turns out to be useful, the hot-swapping functions will be added to PEFT itself and can be imported here (or there is a separate copy in diffusers to avoid the need for a min PEFT version to use this feature). Moreover, more tests need to be added to better cover this feature, although we don't necessarily need tests for the hot-swapping functionality itself, since those tests will be added to PEFT. Furthermore, as of now, this is only implemented for the unet. Other pipeline components have yet to implement this feature. Finally, it should be properly documented. I would like to collect feedback on the current state of the PR before putting more time into finalizing it. * Reviewer feedback * Reviewer feedback, adjust test * Fix, doc * Make fix * Fix for possible g++ error * Add test for recompilation w/o hotswapping * Make hotswap work Requires https://github.com/huggingface/peft/pull/2366 More changes to make hotswapping work. Together with the mentioned PEFT PR, the tests pass for me locally. List of changes: - docstring for hotswap - remove code copied from PEFT, import from PEFT now - adjustments to PeftAdapterMixin.load_lora_adapter (unfortunately, some state dict renaming was necessary, LMK if there is a better solution) - adjustments to UNet2DConditionLoadersMixin._process_lora: LMK if this is even necessary or not, I'm unsure what the overall relationship is between this and PeftAdapterMixin.load_lora_adapter - also in UNet2DConditionLoadersMixin._process_lora, I saw that there is no LoRA unloading when loading the adapter fails, so I added it there (in line with what happens in PeftAdapterMixin.load_lora_adapter) - rewritten tests to avoid shelling out, make the test more precise by making sure that the outputs align, parametrize it - also checked the pipeline code mentioned in this comment: https://github.com/huggingface/diffusers/pull/9453#issuecomment-2418508871; when running this inside the with torch._dynamo.config.patch(error_on_recompile=True) context, there is no error, so I think hotswapping is now working with pipelines. * Address reviewer feedback: - Revert deprecated method - Fix PEFT doc link to main - Don't use private function - Clarify magic numbers - Add pipeline test Moreover: - Extend docstrings - Extend existing test for outputs != 0 - Extend existing test for wrong adapter name * Change order of test decorators parameterized.expand seems to ignore skip decorators if added in last place (i.e. innermost decorator). * Split model and pipeline tests Also increase test coverage by also targeting conv2d layers (support of which was added recently on the PEFT PR). * Reviewer feedback: Move decorator to test classes ... instead of having them on each test method. * Apply suggestions from code review Co-authored-by: hlky <hlky@hlky.ac> * Reviewer feedback: version check, TODO comment * Add enable_lora_hotswap method * Reviewer feedback: check _lora_loadable_modules * Revert changes in unet.py * Add possibility to ignore enabled at wrong time * Fix docstrings * Log possible PEFT error, test * Raise helpful error if hotswap not supported I.e. for the text encoder * Formatting * More linter * More ruff * Doc-builder complaint * Update docstring: - mention no text encoder support yet - make it clear that LoRA is meant - mention that same adapter name should be passed * Fix error in docstring * Update more methods with hotswap argument - SDXL - SD3 - Flux No changes were made to load_lora_into_transformer. * Add hotswap argument to load_lora_into_transformer For SD3 and Flux. Use shorter docstring for brevity. * Extend docstrings * Add version guards to tests * Formatting * Fix LoRA loading call to add prefix=None See: https://github.com/huggingface/diffusers/pull/10187#issuecomment-2717571064 * Run make fix-copies * Add hot swap documentation to the docs * Apply suggestions from code review Co-authored-by: Steven Liu <59462357+stevhliu@users.noreply.github.com> --------- Co-authored-by: Sayak Paul <spsayakpaul@gmail.com> Co-authored-by: hlky <hlky@hlky.ac> Co-authored-by: YiYi Xu <yixu310@gmail.com> Co-authored-by: Steven Liu <59462357+stevhliu@users.noreply.github.com>	2025-04-08 17:05:31 +05:30
Álvaro Somoza	723dbdd363	[Training] Better image interpolation in training scripts (#11206 ) * initial * Update examples/dreambooth/train_dreambooth_lora_sdxl.py Co-authored-by: hlky <hlky@hlky.ac> * update --------- Co-authored-by: Sayak Paul <spsayakpaul@gmail.com> Co-authored-by: hlky <hlky@hlky.ac>	2025-04-08 12:26:07 +05:30
Bhavay Malhotra	fbf61f465b	[train_controlnet.py] Fix the LR schedulers when num_train_epochs is passed in a distributed training env (#8461 ) * Create diffusers.yml * fix num_train_epochs * Delete diffusers.yml * Fixed Changes --------- Co-authored-by: Sayak Paul <spsayakpaul@gmail.com> Co-authored-by: YiYi Xu <yixu310@gmail.com>	2025-04-08 12:10:09 +05:30
Inigo Goiri	841504bb1a	Add support to pass image embeddings to the WAN I2V pipeline. (#11175 ) * Add support to pass image embeddings to the pipeline. --------- Co-authored-by: hlky <hlky@hlky.ac> Co-authored-by: github-actions[bot] <github-actions[bot]@users.noreply.github.com> Co-authored-by: YiYi Xu <yixu310@gmail.com>	2025-04-07 15:47:06 -10:00
Steven Liu	fc7a867ae5	[docs] MPS update (#11212 ) mps	2025-04-07 14:32:27 -10:00
alex choi	5ded26cdc7	ensure dtype match between diffused latents and vae weights (#8391 )	2025-04-07 12:59:10 -10:00
Yao Matrix	506f39af3a	enable 1 case on XPU (#11219 ) enable case on XPU: 1. tests/quantization/bnb/test_mixed_int8.py::BnB8bitTrainingTests::test_training Signed-off-by: YAO Matrix <matrix.yao@intel.com>	2025-04-07 08:24:21 +01:00
Mikko Tukiainen	8ad68c1393	Add missing MochiEncoder3D.gradient_checkpointing attribute (#11146 ) * Add missing 'gradient_checkpointing = False' attr * Add (limited) tests for Mochi autoencoder * Apply style fixes * pass 'conv_cache' as arg instead of kwarg --------- Co-authored-by: github-actions[bot] <github-actions[bot]@users.noreply.github.com>	2025-04-06 02:46:45 +05:30
Edna	41afb6690c	Add Wan with STG as a community pipeline (#11184 ) * Add stg wan to community pipelines * remove debug prints * remove unused comment * Update doc * Add credit + fix typo * Apply style fixes --------- Co-authored-by: github-actions[bot] <github-actions[bot]@users.noreply.github.com>	2025-04-05 04:00:40 +02:00
Tolga Cangöz	13e48492f0	[LTX0.9.5] Refactor `LTXConditionPipeline` for text-only conditioning (#11174 ) * Refactor `LTXConditionPipeline` to add text-only conditioning * style * up * Refactor `LTXConditionPipeline` to streamline condition handling and improve clarity * Improve condition checks * Simplify latents handling based on conditioning type * Refactor rope_interpolation_scale preparation for clarity and efficiency * Update LTXConditionPipeline docstring to clarify supported input types * Add LTX Video 0.9.5 model to documentation * Clarify documentation to indicate support for text-only conditioning without passing `conditions` * refactor: comment out unused parameters in LTXConditionPipeline * fix: restore previously commented parameters in LTXConditionPipeline * fix: remove unused parameters from LTXConditionPipeline * refactor: remove unnecessary lines in LTXConditionPipeline	2025-04-04 16:43:15 +02:00
Suprhimp	94f2c48d58	[feat]Add strength in flux_fill pipeline (denoising strength for fluxfill) (#10603 ) * [feat]add strength in flux_fill pipeline * Update src/diffusers/pipelines/flux/pipeline_flux_fill.py * Update src/diffusers/pipelines/flux/pipeline_flux_fill.py * Update src/diffusers/pipelines/flux/pipeline_flux_fill.py * [refactor] refactor after review * [fix] change comment * Apply style fixes * empty * fix * update prepare_latents from flux.img2img pipeline * style * Update src/diffusers/pipelines/flux/pipeline_flux_fill.py ---------	2025-04-04 11:23:30 -03:00
Dhruv Nair	aabf8ce20b	Fix Single File loading for LTX VAE (#11200 ) update	2025-04-04 18:02:39 +05:30
Kenneth Gerald Hamilton	f10775b1b5	Fixed requests.get function call by adding timeout parameter. (#11156 ) * Fixed requests.get function call by adding timeout parameter. * declare DIFFUSERS_REQUEST_TIMEOUT in constants and import when needed * remove unneeded os import * Apply style fixes --------- Co-authored-by: Sai-Suraj-27 <sai.suraj.27.729@gmail.com> Co-authored-by: github-actions[bot] <github-actions[bot]@users.noreply.github.com>	2025-04-04 07:23:14 +01:00
célina	6edb774b5e	Update Style Bot workflow (#11202 ) update style bot workflow	2025-04-03 19:31:49 +02:00
Basile Lewandowski	480510ada9	Change KolorsPipeline LoRA Loader to StableDiffusion (#11198 ) Change LoRA Loader to StableDiffusion Replace the SDXL LoRA Loader Mixin inheritance with the StableDiffusion one	2025-04-03 11:21:11 -03:00
Abhipsha Das	d9023a671a	[Model Card] standardize advanced diffusion training sdxl lora (#7615 ) * model card gen code * push modelcard creation * remove optional from params * add import * add use_dora check * correct lora var use in tags * make style && make quality --------- Co-authored-by: Aryan <aryan@huggingface.co> Co-authored-by: Sayak Paul <spsayakpaul@gmail.com>	2025-04-03 07:43:01 +05:30
Eliseu Silva	c4646a3931	feat: [Community Pipeline] - FaithDiff Stable Diffusion XL Pipeline (#11188 ) * feat: [Community Pipeline] - FaithDiff Stable Diffusion XL Pipeline for Image SR. * added pipeline	2025-04-02 11:33:19 -10:00
Dhruv Nair	c97b709afa	Add CacheMixin to Wan and LTX Transformers (#11187 ) * update * update * update	2025-04-02 10:16:31 -10:00
lakshay sharma	b0ff822ed3	Update import_utils.py (#10329 ) added onnxruntime-vitisai for custom build onnxruntime pkg	2025-04-02 20:47:10 +01:00
hlky	78c2fdc52e	SchedulerMixin from_pretrained and ConfigMixin Self type annotation (#11192 )	2025-04-02 08:24:02 -10:00
hlky	54dac3a87c	Fix enable_sequential_cpu_offload in CogView4Pipeline (#11195 ) * Fix enable_sequential_cpu_offload in CogView4Pipeline * make fix-copies	2025-04-02 16:51:23 +01:00
hlky	e5c6027ef8	[docs] `torch_dtype` map (#11194 )	2025-04-02 12:46:28 +01:00
hlky	da857bebb6	Revert `save_model` in ModelMixin save_pretrained and use safe_serialization=False in test (#11196 )	2025-04-02 12:45:36 +01:00
Fanli Lin	52b460feb9	[tests] HunyuanDiTControlNetPipeline inference precision issue on XPU (#11197 ) * add xpu part * fix more cases * remove some cases * no canny * format fix	2025-04-02 12:45:02 +01:00
hlky	d8c617ccb0	allow models to run with a user-provided dtype map instead of a single dtype (#10301 ) * allow models to run with a user-provided dtype map instead of a single dtype * make style * Add warning, change `_` to `default` * make style * add test * handle shared tensors * remove warning --------- Co-authored-by: Sayak Paul <spsayakpaul@gmail.com>	2025-04-02 09:05:46 +01:00
Bruno Magalhaes	fe2b397426	remove unnecessary call to `F.pad` (#10620 ) * rewrite memory count without implicitly using dimensions by @ic-synth * replace F.pad by built-in padding in Conv3D * in-place sums to reduce memory allocations * fixed trailing whitespace * file reformatted * in-place sums * simpler in-place expressions * removed in-place sum, may affect backward propagation logic * removed in-place sum, may affect backward propagation logic * removed in-place sum, may affect backward propagation logic * reverted change	2025-04-02 08:19:51 +01:00
Eliseu Silva	be0b7f55cc	fix: for checking mandatory and optional pipeline components (#11189 ) fix: optional componentes verification on load	2025-04-02 08:07:24 +01:00
jiqing-feng	4d5a96e40a	fix autocast (#11190 ) Signed-off-by: jiqing-feng <jiqing.feng@intel.com>	2025-04-02 07:26:27 +01:00
Yao Matrix	a7f07c1ef5	map BACKEND_RESET_MAX_MEMORY_ALLOCATED to reset_peak_memory_stats on XPU (#11191 ) Signed-off-by: YAO Matrix <matrix.yao@intel.com>	2025-04-02 07:25:48 +01:00
Dhruv Nair	df1d7b01f1	[WIP] Add Wan Video2Video (#11053 ) * update * update * update * update * update * update * update * update * update * update * update * update * update * update * update * update	2025-04-01 17:22:11 +05:30
Fanli Lin	5a6edac087	[tests] no hard-coded cuda (#11186 ) no cuda only	2025-04-01 12:14:31 +01:00
kakukakujirori	e8fc8b1f81	Bug fix in LTXImageToVideoPipeline.prepare_latents() when latents is already set (#10918 ) * Bug fix in ltx * Assume packed latents. --------- Co-authored-by: Dhruv Nair <dhruv.nair@gmail.com> Co-authored-by: YiYi Xu <yixu310@gmail.com>	2025-03-31 12:15:43 -10:00
hlky	d6f4774c1c	Add `latents_mean` and `latents_std` to `SDXLLongPromptWeightingPipeline` (#11034 )	2025-03-31 11:32:29 -10:00
Mark	eb50defff2	[Docs] Fix environment variables in `installation.md` (#11179 )	2025-03-31 09:15:25 -07:00
Aryan	2c59af7222	Raise warning and round down if Wan num_frames is not 4k + 1 (#11167 ) * update * raise warning and round to nearest multiple of scale factor	2025-03-31 13:33:28 +05:30
hlky	75d7e5cc45	Fix LatteTransformer3DModel dtype mismatch with enable_temporal_attentions (#11139 )	2025-03-29 15:52:56 +01:00
Dhruv Nair	617c208bb4	[Docs] Update Wan Docs with memory optimizations (#11089 ) * update * update	2025-03-28 19:05:56 +05:30
hlky	5d970a4aa9	WanI2V encode_image (#11164 ) * WanI2V encode_image	2025-03-28 18:05:34 +05:30
kentdan3msu	de6a88c2d7	Set self._hf_peft_config_loaded to True when LoRA is loaded using `load_lora_adapter` in PeftAdapterMixin class (#11155 ) set self._hf_peft_config_loaded to True on successful lora load Sets the `_hf_peft_config_loaded` flag if a LoRA is successfully loaded in `load_lora_adapter`. Fixes bug huggingface/diffusers/issues/11148 Co-authored-by: Sayak Paul <spsayakpaul@gmail.com>	2025-03-26 18:31:18 +01:00
Dhruv Nair	7dc52ea769	[Quantization] dtype fix for GGUF + fix BnB tests (#11159 ) * update * update * update * update	2025-03-26 22:22:16 +05:30
Junsong Chen	739d6ec731	add a timestep scale for sana-sprint teacher model (#11150 )	2025-03-25 08:47:39 -10:00
Aryan	1ddf3f3a19	Improve information about group offloading and layerwise casting (#11101 ) * update * Update docs/source/en/optimization/memory.md * Apply suggestions from code review Co-authored-by: Dhruv Nair <dhruv.nair@gmail.com> * apply review suggestions * update --------- Co-authored-by: Dhruv Nair <dhruv.nair@gmail.com>	2025-03-24 23:25:59 +05:30
Jun Yeop Na	7aac77affa	[doc] Fix Korean Controlnet Train doc (#11141 ) * remove typo from korean controlnet train doc * removed more paragraphs to remain in sync with the english document	2025-03-24 09:38:21 -07:00

1 2 3 4 5 ...

5317 Commits