Fix bnb dequantization #33546

SunMarc · 2024-09-17T17:23:04Z

What does this PR do ?

This PR fixes a small error on how dequantization is handled in bnb. We need to set has_been_replaced to True if the dequantization happens.

Fixes an issue raised by @sayakpaul

HuggingFaceDocBuilderDev · 2024-09-17T17:47:53Z

The docs for this PR live here. All of your documentation changes will be reflected on that endpoint. The docs are available until 30 days after the last update.

sayakpaul

Should we maybe add any dequantization tests to https://github.com/huggingface/transformers/blob/main/tests/quantization/bnb/test_4bit.py? I know that we have one that checks for inference but can you think of something else? No big deal, though.

SunMarc · 2024-09-18T02:12:27Z

We do have a test to check the quality of the dequantized model. In this case, it was just the flag that was not set, hence triggering a false alarm.

* quantization config. * fix-copies * fix * modules_to_not_convert * add bitsandbytes utilities. * make progress. * fixes * quality * up * up rotary embedding refactor 2: update comments, fix dtype for use_real=False (#9312) fix notes and dtype up up * minor * up * up * fix * provide credits where due. * make configurations work. * fixes * fix * update_missing_keys * fix * fix * make it work. * fix * provide credits to transformers. * empty commit * handle to() better. * tests * change to bnb from bitsandbytes * fix tests fix slow quality tests SD3 remark fix complete int4 tests add a readme to the test files. add model cpu offload tests warning test * better safeguard. * change merging status * courtesy to transformers. * move upper. * better * make the unused kwargs warning friendlier. * harmonize changes with huggingface/transformers#33122 * style * trainin tests * feedback part i. * Add Flux inpainting and Flux Img2Img (#9135) --------- Co-authored-by: yiyixuxu <[email protected]> Update `UNet2DConditionModel`'s error messages (#9230) * refactor [CI] Update Single file Nightly Tests (#9357) * update * update feedback. improve README for flux dreambooth lora (#9290) * improve readme * improve readme * improve readme * improve readme fix one uncaught deprecation warning for accessing vae_latent_channels in VaeImagePreprocessor (#9372) deprecation warning vae_latent_channels add mixed int8 tests and more tests to nf4. [core] Freenoise memory improvements (#9262) * update * implement prompt interpolation * make style * resnet memory optimizations * more memory optimizations; todo: refactor * update * update animatediff controlnet with latest changes * refactor chunked inference changes * remove print statements * update * chunk -> split * remove changes from incorrect conflict resolution * remove changes from incorrect conflict resolution * add explanation of SplitInferenceModule * update docs * Revert "update docs" This reverts commit c55a50a. * update docstring for freenoise split inference * apply suggestions from review * add tests * apply suggestions from review quantization docs. docs. * Revert "Add Flux inpainting and Flux Img2Img (#9135)" This reverts commit 5799954. * tests * don * Apply suggestions from code review Co-authored-by: Steven Liu <[email protected]> * contribution guide. * changes * empty * fix tests * harmonize with huggingface/transformers#33546. * numpy_cosine_distance * config_dict modification. * remove if config comment. * note for load_state_dict changes. * float8 check. * quantizer. * raise an error for non-True low_cpu_mem_usage values when using quant. * low_cpu_mem_usage shenanigans when using fp32 modules. * don't re-assign _pre_quantization_type. * make comments clear. * remove comments. * handle mixed types better when moving to cpu. * add tests to check if we're throwing warning rightly. * better check. * fix 8bit test_quality. * handle dtype more robustly. * better message when keep_in_fp32_modules. * handle dtype casting. * fix dtype checks in pipeline. * fix warning message. * Update src/diffusers/models/modeling_utils.py Co-authored-by: YiYi Xu <[email protected]> * mitigate the confusing cpu warning --------- Co-authored-by: Vishnu V Jaddipal <[email protected]> Co-authored-by: Steven Liu <[email protected]> Co-authored-by: YiYi Xu <[email protected]>

fix bnb dq

5bb7a51

SunMarc requested review from sayakpaul and LysandreJik September 17, 2024 17:23

sayakpaul approved these changes Sep 18, 2024

View reviewed changes

sayakpaul added a commit to huggingface/diffusers that referenced this pull request Sep 18, 2024

harmonize with huggingface/transformers#33546.

971305b

LysandreJik approved these changes Sep 18, 2024

View reviewed changes

SunMarc merged commit 6019f3f into huggingface:main Sep 18, 2024
21 checks passed

itazap pushed a commit to NielsRogge/transformers that referenced this pull request Sep 20, 2024

Fix bnb dequantization (huggingface#33546)

8405ec5

amyeroberts pushed a commit to amyeroberts/transformers that referenced this pull request Oct 2, 2024

Fix bnb dequantization (huggingface#33546)

10e019c

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Fix bnb dequantization #33546

Fix bnb dequantization #33546

SunMarc commented Sep 17, 2024 •

edited

Loading

HuggingFaceDocBuilderDev commented Sep 17, 2024

sayakpaul left a comment

SunMarc commented Sep 18, 2024

Fix bnb dequantization #33546

Fix bnb dequantization #33546

Conversation

SunMarc commented Sep 17, 2024 • edited Loading

What does this PR do ?

HuggingFaceDocBuilderDev commented Sep 17, 2024

sayakpaul left a comment

Choose a reason for hiding this comment

SunMarc commented Sep 18, 2024

SunMarc commented Sep 17, 2024 •

edited

Loading