1351 Commits

Author SHA1 Message Date
zhoufan2956
0f7eb12d62 mix_ascend_bf16_infer_err (#6794) 2025-02-12 06:48:11 -05:00
comfyanonymous
16bd3b8441 Add add_weight_wrapper function to model patcher.
Functions can now easily be added to wrap/modify model weights.
2025-02-12 05:55:35 -05:00
comfyanonymous
2a521a2e33 Cleanup. 2025-02-11 17:17:03 -05:00
HishamC
f49d8d3d1e Fix for running via DirectML (#6542)
* Fix for running via DirectML

Fix DirectML empty image generation issue with Flux1. add CPU fallback for unsupported path. Verified the model works on AMD GPUs

* fix formating

* update casual mask calculation
2025-02-11 17:11:32 -05:00
comfyanonymous
a7e5e662ab Make --force-fp16 actually force the diffusion model to be fp16. 2025-02-11 08:33:09 -05:00
comfyanonymous
6e2027a69b Make lumina model work with any latent resolution. 2025-02-10 00:24:20 -05:00
comfyanonymous
c80b2a425c Remove useless function. 2025-02-09 07:02:57 -05:00
Pam
a1cd544ace res_multistep: Fix cfgpp and add ancestral samplers (#6731) 2025-02-08 19:39:58 -05:00
comfyanonymous
84a5e51448 Make error more clear to user. 2025-02-08 18:57:24 -05:00
catboxanon
ca2a621e0b Allow FP16 accumulation with --fast (#6453)
Currently only applies to PyTorch nightly releases. (>=20250208)
2025-02-08 17:00:56 -05:00
comfyanonymous
6865c73a03 Don't compress http response by default.
Remove argument to disable it.

Add new --enable-compress-response-body argument to enable it.
2025-02-07 03:29:21 -05:00
comfyanonymous
1908648a5b Remove some useless code. 2025-02-06 05:00:37 -05:00
comfyanonymous
50a778e1ce Set the shift for Lumina back to 6. 2025-02-05 14:49:52 -05:00
comfyanonymous
ff7449847e Upcasting rope to fp32 seems to make no difference in this model. 2025-02-05 04:32:47 -05:00
comfyanonymous
f210ee69df Use regular numbers for rope in lumina model. 2025-02-05 04:17:25 -05:00
comfyanonymous
eaf581f1ed Fix lumina 2 batches. 2025-02-04 21:48:11 -05:00
comfyanonymous
0467641f12 Lower the default shift of lumina to reduce artifacts. 2025-02-04 06:50:37 -05:00
comfyanonymous
24b0e00898 Fix on python 3.9 2025-02-04 04:20:56 -05:00
comfyanonymous
e03a640306 Support Lumina 2 model. 2025-02-04 04:16:30 -05:00
comfyanonymous
fc1a24c402 Use maximum negative value instead of -inf for masks in text encoders.
This is probably more correct.
2025-02-02 09:46:00 -05:00
Dr.Lt.Data
4a3500c3ef better guide message for sageattention (#6634) 2025-02-02 09:26:47 -05:00
KarryCharon
e12555bdbf add disable-compres-response-body cli args; add compress middleware; (#6672) 2025-02-02 09:24:55 -05:00
comfyanonymous
af91112845 Only use stable cascade lora format with cascade model. 2025-02-01 06:35:22 -05:00
comfyanonymous
dcacbab9b0 Allow batch of different sigmas when noise scaling. 2025-01-30 06:49:52 -05:00
filtered
acf1f98438 Allow changing folder_paths.base_path via command line argument. (#6600)
* Reimpl. CLI arg directly inside folder_paths.

* Update tests to use CLI arg mocking.

* Revert last-minute refactor.

* Fix test state polution.
2025-01-29 08:06:28 -05:00
comfyanonymous
9dbb2b6d89 More friendly error messages for corrupted safetensors files. 2025-01-28 09:41:09 -05:00
comfyanonymous
5da1230892 Lower minimum ratio of loaded weights on Nvidia. 2025-01-27 05:26:51 -05:00
comfyanonymous
00e27c6202 Remove redundant code. 2025-01-25 19:04:53 -05:00
comfyanonymous
6da2db0e10 Remove useless code. 2025-01-24 06:15:54 -05:00
comfyanonymous
9cc63be0a0 Remove useless code. 2025-01-23 05:56:23 -05:00
Chenlei Hu
25f4e415da Remove unused function lcm in conds.py (#6572) 2025-01-23 05:54:09 -05:00
comfyanonymous
e3f3830dde Remove support for python 3.8. 2025-01-22 17:04:30 -05:00
chaObserv
4d5a8e8525 Add gradient estimation sampler (#6554) 2025-01-22 05:29:40 -05:00
comfyanonymous
eb64454c08 Add FluxDisableGuidance node to disable using the guidance embed. 2025-01-20 14:50:24 -05:00
comfyanonymous
71281dcc56 Cleanup old TODO. 2025-01-20 03:44:13 -05:00
Sergii Dymchenko
441dfbf8d9 Use torch.special.expm1 (#6388)
* Use `torch.special.expm1`

This function provides greater precision than `exp(x) - 1` for small values of `x`.

Found with TorchFix https://github.com/pytorch-labs/torchfix/

* Use non-alias
2025-01-19 04:54:32 -05:00
catboxanon
ccec4da418 Remove comfy.samplers self-import (#6506) 2025-01-18 17:49:51 -05:00
comfyanonymous
39b2c8eeee Uni pc sampler now works with audio and video models. 2025-01-18 05:27:58 -05:00
comfyanonymous
b8fbb7c807 Add warning when using old pytorch versions. 2025-01-17 18:47:27 -05:00
comfyanonymous
d439446067 Fix some cosmos fp8 issues. 2025-01-16 17:45:37 -05:00
comfyanonymous
63af9e9d90 Fix cosmos VAE failing with videos longer than 121 frames. 2025-01-16 16:30:06 -05:00
comfyanonymous
5650bfe1a9 Code refactor. 2025-01-16 07:23:54 -05:00
comfyanonymous
9ea569d2b0 Tweak hunyuan memory usage factor. 2025-01-16 06:31:03 -05:00
comfyanonymous
868a4d1551 Clean up some debug lines. 2025-01-16 04:24:39 -05:00
comfyanonymous
ad3ed4154c More accurate memory estimation for cosmos and hunyuan video. 2025-01-16 03:48:40 -05:00
comfyanonymous
220f470bfa Slightly lower hunyuan video memory usage. 2025-01-16 00:23:01 -05:00
comfyanonymous
3423ffcf70 Lower cosmos diffusion model memory usage. 2025-01-15 23:46:42 -05:00
comfyanonymous
d20c6506b5 Lower cosmos VAE memory usage by a bit. 2025-01-15 22:57:52 -05:00
comfyanonymous
8bf2059401 Optimize first attention block in cosmos VAE. 2025-01-15 21:48:46 -05:00
comfyanonymous
fd4d654129 Remove unsafe embedding load for very old pytorch. 2025-01-15 04:32:23 -05:00