diffusers

mirror of https://github.com/huggingface/diffusers.git synced 2026-01-27 17:22:53 +03:00

Author	SHA1	Message	Date
Sayak Paul	cefa28f449	[docs] Promote `AutoModel` usage (#11300 ) * docs: promote the usage of automodel. * bitsandbytes * Apply suggestions from code review Co-authored-by: Steven Liu <59462357+stevhliu@users.noreply.github.com> --------- Co-authored-by: Steven Liu <59462357+stevhliu@users.noreply.github.com>	2025-04-15 09:25:40 +05:30
Beinsezii	8819cda6c0	Add `skrample` section to `community_projects.md` (#11319 ) Update community_projects.md https://github.com/huggingface/diffusers/discussions/11158#discussioncomment-12681691	2025-04-14 12:12:59 -10:00
hlky	dcf836cf47	Use float32 on mps or npu in transformer_hidream_image's rope (#11316 )	2025-04-14 20:19:21 +01:00
Álvaro Somoza	1cb73cb19f	import for FlowMatchLCMScheduler (#11318 ) * add * fix-copies	2025-04-14 06:28:57 -10:00
Linoy Tsaban	ba6008abfe	[HiDream] code example (#11317 )	2025-04-14 16:19:30 +01:00
Sayak Paul	a8f5134c11	[LoRA] support more SDXL loras. (#11292 ) * support more SDXL loras. * update --------- Co-authored-by: hlky <hlky@hlky.ac>	2025-04-14 17:09:59 +05:30
Fanli Lin	c7f2d239fe	make `KolorsPipelineFastTests::test_inference_batch_single_identical` pass on XPU (#11313 ) adjust diff	2025-04-14 11:02:02 +01:00
Yao Matrix	fa1ac50a66	make test_stable_diffusion_karras_sigmas pass on XPU (#11310 ) Signed-off-by: Matrix Yao <matrix.yao@intel.com>	2025-04-14 08:15:38 +01:00
Yao Matrix	aa541b9fab	make KandinskyV22PipelineInpaintCombinedFastTests::test_float16_inference pass on XPU (#11308 ) loose expected_max_diff from 5e-1 to 8e-1 to make KandinskyV22PipelineInpaintCombinedFastTests::test_float16_inference pass on XPU Signed-off-by: Matrix Yao <matrix.yao@intel.com>	2025-04-14 07:49:20 +01:00
Ishan Modi	f1f38ffbee	[ControlNet] Adds controlnet for SanaTransformer (#11040 ) * added controlnet for sana transformer * improve code quality * addressed PR comments * bug fixes * added test cases * update * added dummy objects * addressed PR comments * update * Forcing update * add to docs * code quality * addressed PR comments * addressed PR comments * update * addressed PR comments * added proper styling * update * Revert "added proper styling" This reverts commit `344ee8a701`. * manually ordered * Apply suggestions from code review --------- Co-authored-by: Aryan <contact.aryanvs@gmail.com>	2025-04-13 19:19:39 +05:30
Tuna Tuncer	36538e1135	Fix incorrect tile_latent_min_width calculations (#11305 )	2025-04-13 18:21:50 +05:30
Aryan	97e0ef4db4	Hidream refactoring follow ups (#11299 ) * HiDream Image * update * -einops * py3.8 * fix -einops * mixins, offload_seq, option_components * docs * Apply style fixes * trigger tests * Apply suggestions from code review Co-authored-by: Aryan <contact.aryanvs@gmail.com> * joint_attention_kwargs -> attention_kwargs, fixes * fast tests * -_init_weights * style tests * move reshape logic * update slice 😴 * supports_dduf * 🤷🏻‍♂️ * Update src/diffusers/models/transformers/transformer_hidream_image.py Co-authored-by: Aryan <contact.aryanvs@gmail.com> * address review comments * update tests * doc updates * update * Update src/diffusers/models/transformers/transformer_hidream_image.py * Apply style fixes --------- Co-authored-by: hlky <hlky@hlky.ac> Co-authored-by: github-actions[bot] <github-actions[bot]@users.noreply.github.com>	2025-04-13 09:45:19 +05:30
Adrien B	ed41db8525	Update autoencoderkl_allegro.md (#11303 ) Correction typo	2025-04-13 09:41:30 +05:30
Nikita Starodubcev	ec0b2b3947	flow matching lcm scheduler (#11170 ) * add flow matching lcm scheduler * stochastic sampling * upscaling for scale-wise generation * Apply style fixes * Apply suggestions from code review Co-authored-by: hlky <hlky@hlky.ac> --------- Co-authored-by: github-actions[bot] <github-actions[bot]@users.noreply.github.com> Co-authored-by: YiYi Xu <yixu310@gmail.com> Co-authored-by: hlky <hlky@hlky.ac>	2025-04-12 11:14:57 -10:00
hlky	0ef29355c9	HiDream Image (#11231 ) * HiDream Image --------- Co-authored-by: github-actions[bot] <github-actions[bot]@users.noreply.github.com> Co-authored-by: Aryan <contact.aryanvs@gmail.com> Co-authored-by: Aryan <aryan@huggingface.co>	2025-04-11 06:31:34 -10:00
Tuna Tuncer	bc261058ee	Fix incorrect tile_latent_min_width calculation in AutoencoderKLMochi (#11294 )	2025-04-11 19:11:08 +05:30
Sayak Paul	7054a34978	do not use `DIFFUSERS_REQUEST_TIMEOUT` for notification bot (#11273 ) fix to a constant	2025-04-11 14:19:46 +05:30
Sayak Paul	511d738121	[CI] relax tolerance for unclip further (#11268 ) relax tolerance for unclip further.	2025-04-11 14:06:52 +05:30
Sayak Paul	ea5a6a8b7c	[Tests] Cleanup lora tests utils (#11276 ) * start cleaning up lora test utils for reusability * update * updates * updates	2025-04-10 15:50:34 +05:30
hlky	b8093e6665	Fix LTX 0.9.5 single file (#11271 )	2025-04-10 07:06:13 +01:00
Yuqian Hong	e121d0ef67	[BUG] Fix convert_vae_pt_to_diffusers bug (#11078 ) * fix attention * Apply style fixes --------- Co-authored-by: github-actions[bot] <github-actions[bot]@users.noreply.github.com>	2025-04-10 06:59:45 +01:00
Yao Matrix	31c4f24fc1	make test_instant_style_multiple_masks pass on XPU (#11266 ) Signed-off-by: Matrix Yao <matrix.yao@intel.com>	2025-04-10 06:23:00 +01:00
xieofxie	0efdf411fb	add onnxruntime-qnn & onnxruntime-cann (#11269 ) Co-authored-by: hualxie <hualxie@microsoft.com>	2025-04-10 06:00:23 +01:00
Yao Matrix	450dc48a2c	make test_dict_tuple_outputs_equivalent pass on XPU (#11265 ) Signed-off-by: Matrix Yao <matrix.yao@intel.com>	2025-04-10 05:56:28 +01:00
Yao Matrix	77b4f66b9e	make test_stable_diffusion_inpaint_fp16 pass on XPU (#11264 ) Signed-off-by: Matrix Yao <matrix.yao@intel.com>	2025-04-10 05:55:54 +01:00
Yao Matrix	68663f8a17	fix test_vanilla_funetuning failure on XPU and A100 (#11263 ) * fix test_vanilla_funetuning failure on XPU and A100 Signed-off-by: Matrix Yao <matrix.yao@intel.com> * change back to 5e-2 Signed-off-by: Matrix Yao <matrix.yao@intel.com> --------- Signed-off-by: Matrix Yao <matrix.yao@intel.com>	2025-04-10 05:55:07 +01:00
Sayak Paul	ffda8735be	[LoRA] support musubi wan loras. (#11243 ) * support musubi wan loras. * Update src/diffusers/loaders/lora_conversion_utils.py Co-authored-by: hlky <hlky@hlky.ac> * support i2v loras from musubi too. --------- Co-authored-by: hlky <hlky@hlky.ac>	2025-04-10 09:50:22 +05:30
YiYi Xu	0706786e53	fix wan ftfy import (#11262 )	2025-04-09 09:08:34 -10:00
Sayak Paul	5b27f8aba8	fix consisid imports (#11254 ) * fix consisid imports * fix opencv import * fix	2025-04-09 18:49:32 +05:30
Sayak Paul	d1387ecee5	fix timeout constant (#11252 ) * fix timeout constant * style * fix	2025-04-09 17:48:52 +05:30
Ilya Drobyshevskiy	6a7c2d0afa	fix flux controlnet bug (#11152 ) Before this if txt_ids was 3d tensor, line with txt_ids[:1] concat txt_ids by batch dim. Now we first check that txt_ids is 2d tensor (or take first batch element) and then concat by token dim	2025-04-09 13:01:07 +01:00
Dhruv Nair	edc154da09	Update Ruff to latest Version (#10919 ) * update * update * update * update	2025-04-09 16:51:34 +05:30
hlky	552cd32058	[docs] AutoModel (#11250 ) Co-authored-by: Sayak Paul <spsayakpaul@gmail.com>	2025-04-09 16:42:23 +05:30
Yao Matrix	c36c745ceb	fix FluxReduxSlowTests::test_flux_redux_inference case failure on XPU (#11245 ) * loose test_float16_inference's tolerance from 5e-2 to 6e-2, so XPU can pass UT Signed-off-by: Matrix Yao <matrix.yao@intel.com> * fix test_pipeline_flux_redux fail on XPU Signed-off-by: Matrix Yao <matrix.yao@intel.com> --------- Signed-off-by: Matrix Yao <matrix.yao@intel.com>	2025-04-09 11:41:15 +01:00
hlky	437cb36c65	AutoModel (#11115 ) * AutoModel * ... * lol * ... * add test * update * make fix-copies --------- Co-authored-by: Dhruv Nair <dhruv.nair@gmail.com>	2025-04-09 15:20:07 +05:30
hlky	9ee3dd3862	AudioLDM2 Fixes (#11244 )	2025-04-09 14:12:00 +05:30
Sayak Paul	fd02aad402	fix: SD3 ControlNet validation so that it runs on a A100. (#11238 ) * fix: SD3 ControlNet validation so that it runs on a A100. * use backend-agnostic cache and pass devide.	2025-04-09 12:12:53 +05:30
Sayak Paul	6bfacf0418	[LoRA] support more comyui loras for Flux 🚨 (#10985 ) * support more comyui loras. * fix * fixes * revert changes in LoRA base. * no position_embedding * 🚨 introduce a breaking change to let peft handle module ambiguity * styling * remove position embeddings. * improvements. * style * make info instead of NotImplementedError * Update src/diffusers/loaders/peft.py Co-authored-by: hlky <hlky@hlky.ac> * add example. * robust checks * updates --------- Co-authored-by: hlky <hlky@hlky.ac>	2025-04-09 09:17:05 +05:30
Sayak Paul	f685981ed0	[docs] minor updates to dtype map docs. (#11237 ) minor updates to dtype map docs.	2025-04-09 08:38:17 +05:30
Sayak Paul	b924251dd8	minor update to sana sprint docs. (#11236 )	2025-04-09 08:17:45 +05:30
Sayak Paul	1a04812439	[bistandbytes] improve replacement warnings for bnb (#11132 ) * improve replacement warnings for bnb * updates to docs.	2025-04-08 21:18:34 +05:30
Sayak Paul	4b27c4a494	[feat] implement `record_stream` when using CUDA streams during group offloading (#11081 ) * implement record_stream for better performance. * fix * style. * merge #11097 * Update src/diffusers/hooks/group_offloading.py Co-authored-by: Aryan <aryan@huggingface.co> * fixes * docstring. * remaining todos in low_cpu_mem_usage * tests * updates to docs. --------- Co-authored-by: Aryan <aryan@huggingface.co>	2025-04-08 21:17:49 +05:30
hlky	5d49b3e83b	Flux quantized with lora (#10990 ) * Flux quantized with lora * fix * changes * Apply suggestions from code review Co-authored-by: Sayak Paul <spsayakpaul@gmail.com> * Apply style fixes * enable model cpu offload() * Update src/diffusers/loaders/lora_pipeline.py Co-authored-by: hlky <hlky@hlky.ac> * update * Apply suggestions from code review * update * add peft as an additional dependency for gguf --------- Co-authored-by: Sayak Paul <spsayakpaul@gmail.com> Co-authored-by: github-actions[bot] <github-actions[bot]@users.noreply.github.com> Co-authored-by: Dhruv Nair <dhruv.nair@gmail.com>	2025-04-08 21:17:03 +05:30
Linoy Tsaban	71f34fc5a4	[Flux LoRA] fix issues in flux lora scripts (#11111 ) * remove custom scheduler * update requirements.txt * log_validation with mixed precision * add intermediate embeddings saving when checkpointing is enabled * remove comment * fix validation * add unwrap_model for accelerator, torch.no_grad context for validation, fix accelerator.accumulate call in advanced script * revert unwrap_model change temp * add .module to address distributed training bug + replace accelerator.unwrap_model with unwrap model * changes to align advanced script with canonical script * make changes for distributed training + unify unwrap_model calls in advanced script * add module.dtype fix to dreambooth script * unify unwrap_model calls in dreambooth script * fix condition in validation run * mixed precision * Update examples/advanced_diffusion_training/train_dreambooth_lora_flux_advanced.py Co-authored-by: Sayak Paul <spsayakpaul@gmail.com> * smol style change * change autocast * Apply style fixes --------- Co-authored-by: Sayak Paul <spsayakpaul@gmail.com> Co-authored-by: github-actions[bot] <github-actions[bot]@users.noreply.github.com>	2025-04-08 17:40:30 +03:00
Yao Matrix	c51b6bd837	introduce compute arch specific expectations and fix test_sd3_img2img_inference failure (#11227 ) * add arch specfic expectations support, to support different arch's numerical characteristics Signed-off-by: YAO Matrix <matrix.yao@intel.com> * fix typo Signed-off-by: YAO Matrix <matrix.yao@intel.com> * Apply suggestions from code review * Apply style fixes * Update src/diffusers/utils/testing_utils.py --------- Signed-off-by: YAO Matrix <matrix.yao@intel.com> Co-authored-by: hlky <hlky@hlky.ac> Co-authored-by: github-actions[bot] <github-actions[bot]@users.noreply.github.com>	2025-04-08 14:57:49 +01:00
Benjamin Bossan	fb54499614	[LoRA] Implement hot-swapping of LoRA (#9453 ) * [WIP][LoRA] Implement hot-swapping of LoRA This PR adds the possibility to hot-swap LoRA adapters. It is WIP. Description As of now, users can already load multiple LoRA adapters. They can offload existing adapters or they can unload them (i.e. delete them). However, they cannot "hotswap" adapters yet, i.e. substitute the weights from one LoRA adapter with the weights of another, without the need to create a separate LoRA adapter. Generally, hot-swapping may not appear not super useful but when the model is compiled, it is necessary to prevent recompilation. See #9279 for more context. Caveats To hot-swap a LoRA adapter for another, these two adapters should target exactly the same layers and the "hyper-parameters" of the two adapters should be identical. For instance, the LoRA alpha has to be the same: Given that we keep the alpha from the first adapter, the LoRA scaling would be incorrect for the second adapter otherwise. Theoretically, we could override the scaling dict with the alpha values derived from the second adapter's config, but changing the dict will trigger a guard for recompilation, defeating the main purpose of the feature. I also found that compilation flags can have an impact on whether this works or not. E.g. when passing "reduce-overhead", there will be errors of the type: > input name: arg861_1. data pointer changed from 139647332027392 to 139647331054592 I don't know enough about compilation to determine whether this is problematic or not. Current state This is obviously WIP right now to collect feedback and discuss which direction to take this. If this PR turns out to be useful, the hot-swapping functions will be added to PEFT itself and can be imported here (or there is a separate copy in diffusers to avoid the need for a min PEFT version to use this feature). Moreover, more tests need to be added to better cover this feature, although we don't necessarily need tests for the hot-swapping functionality itself, since those tests will be added to PEFT. Furthermore, as of now, this is only implemented for the unet. Other pipeline components have yet to implement this feature. Finally, it should be properly documented. I would like to collect feedback on the current state of the PR before putting more time into finalizing it. * Reviewer feedback * Reviewer feedback, adjust test * Fix, doc * Make fix * Fix for possible g++ error * Add test for recompilation w/o hotswapping * Make hotswap work Requires https://github.com/huggingface/peft/pull/2366 More changes to make hotswapping work. Together with the mentioned PEFT PR, the tests pass for me locally. List of changes: - docstring for hotswap - remove code copied from PEFT, import from PEFT now - adjustments to PeftAdapterMixin.load_lora_adapter (unfortunately, some state dict renaming was necessary, LMK if there is a better solution) - adjustments to UNet2DConditionLoadersMixin._process_lora: LMK if this is even necessary or not, I'm unsure what the overall relationship is between this and PeftAdapterMixin.load_lora_adapter - also in UNet2DConditionLoadersMixin._process_lora, I saw that there is no LoRA unloading when loading the adapter fails, so I added it there (in line with what happens in PeftAdapterMixin.load_lora_adapter) - rewritten tests to avoid shelling out, make the test more precise by making sure that the outputs align, parametrize it - also checked the pipeline code mentioned in this comment: https://github.com/huggingface/diffusers/pull/9453#issuecomment-2418508871; when running this inside the with torch._dynamo.config.patch(error_on_recompile=True) context, there is no error, so I think hotswapping is now working with pipelines. * Address reviewer feedback: - Revert deprecated method - Fix PEFT doc link to main - Don't use private function - Clarify magic numbers - Add pipeline test Moreover: - Extend docstrings - Extend existing test for outputs != 0 - Extend existing test for wrong adapter name * Change order of test decorators parameterized.expand seems to ignore skip decorators if added in last place (i.e. innermost decorator). * Split model and pipeline tests Also increase test coverage by also targeting conv2d layers (support of which was added recently on the PEFT PR). * Reviewer feedback: Move decorator to test classes ... instead of having them on each test method. * Apply suggestions from code review Co-authored-by: hlky <hlky@hlky.ac> * Reviewer feedback: version check, TODO comment * Add enable_lora_hotswap method * Reviewer feedback: check _lora_loadable_modules * Revert changes in unet.py * Add possibility to ignore enabled at wrong time * Fix docstrings * Log possible PEFT error, test * Raise helpful error if hotswap not supported I.e. for the text encoder * Formatting * More linter * More ruff * Doc-builder complaint * Update docstring: - mention no text encoder support yet - make it clear that LoRA is meant - mention that same adapter name should be passed * Fix error in docstring * Update more methods with hotswap argument - SDXL - SD3 - Flux No changes were made to load_lora_into_transformer. * Add hotswap argument to load_lora_into_transformer For SD3 and Flux. Use shorter docstring for brevity. * Extend docstrings * Add version guards to tests * Formatting * Fix LoRA loading call to add prefix=None See: https://github.com/huggingface/diffusers/pull/10187#issuecomment-2717571064 * Run make fix-copies * Add hot swap documentation to the docs * Apply suggestions from code review Co-authored-by: Steven Liu <59462357+stevhliu@users.noreply.github.com> --------- Co-authored-by: Sayak Paul <spsayakpaul@gmail.com> Co-authored-by: hlky <hlky@hlky.ac> Co-authored-by: YiYi Xu <yixu310@gmail.com> Co-authored-by: Steven Liu <59462357+stevhliu@users.noreply.github.com>	2025-04-08 17:05:31 +05:30
Álvaro Somoza	723dbdd363	[Training] Better image interpolation in training scripts (#11206 ) * initial * Update examples/dreambooth/train_dreambooth_lora_sdxl.py Co-authored-by: hlky <hlky@hlky.ac> * update --------- Co-authored-by: Sayak Paul <spsayakpaul@gmail.com> Co-authored-by: hlky <hlky@hlky.ac>	2025-04-08 12:26:07 +05:30
Bhavay Malhotra	fbf61f465b	[train_controlnet.py] Fix the LR schedulers when num_train_epochs is passed in a distributed training env (#8461 ) * Create diffusers.yml * fix num_train_epochs * Delete diffusers.yml * Fixed Changes --------- Co-authored-by: Sayak Paul <spsayakpaul@gmail.com> Co-authored-by: YiYi Xu <yixu310@gmail.com>	2025-04-08 12:10:09 +05:30
Inigo Goiri	841504bb1a	Add support to pass image embeddings to the WAN I2V pipeline. (#11175 ) * Add support to pass image embeddings to the pipeline. --------- Co-authored-by: hlky <hlky@hlky.ac> Co-authored-by: github-actions[bot] <github-actions[bot]@users.noreply.github.com> Co-authored-by: YiYi Xu <yixu310@gmail.com>	2025-04-07 15:47:06 -10:00
Steven Liu	fc7a867ae5	[docs] MPS update (#11212 ) mps	2025-04-07 14:32:27 -10:00

1 2 3 4 5 ...

5355 Commits