comfyanonymous
451a692a3d
Add a way to set a different compute dtype for the model at runtime.
...
Currently only works for diffusion models.
2025-02-13 20:34:03 -05:00
comfyanonymous
1f657be9a7
Better memory estimation for ROCm that support mem efficient attention.
...
There is no way to check if the card actually supports it so it assumes
that it does if you use --use-pytorch-cross-attention with yours.
2025-02-13 08:32:36 -05:00
comfyanonymous
66d579cd19
Fix ruff.
2025-02-12 06:49:16 -05:00
zhoufan2956
0f7eb12d62
mix_ascend_bf16_infer_err ( #6794 )
2025-02-12 06:48:11 -05:00
comfyanonymous
16bd3b8441
Add add_weight_wrapper function to model patcher.
...
Functions can now easily be added to wrap/modify model weights.
2025-02-12 05:55:35 -05:00
comfyanonymous
2a521a2e33
Cleanup.
2025-02-11 17:17:03 -05:00
HishamC
f49d8d3d1e
Fix for running via DirectML ( #6542 )
...
* Fix for running via DirectML
Fix DirectML empty image generation issue with Flux1. add CPU fallback for unsupported path. Verified the model works on AMD GPUs
* fix formating
* update casual mask calculation
2025-02-11 17:11:32 -05:00
comfyanonymous
a7e5e662ab
Make --force-fp16 actually force the diffusion model to be fp16.
2025-02-11 08:33:09 -05:00
comfyanonymous
6e2027a69b
Make lumina model work with any latent resolution.
2025-02-10 00:24:20 -05:00
comfyanonymous
c80b2a425c
Remove useless function.
2025-02-09 07:02:57 -05:00
Pam
a1cd544ace
res_multistep: Fix cfgpp and add ancestral samplers ( #6731 )
2025-02-08 19:39:58 -05:00
comfyanonymous
84a5e51448
Make error more clear to user.
2025-02-08 18:57:24 -05:00
catboxanon
ca2a621e0b
Allow FP16 accumulation with --fast ( #6453 )
...
Currently only applies to PyTorch nightly releases. (>=20250208)
2025-02-08 17:00:56 -05:00
comfyanonymous
6865c73a03
Don't compress http response by default.
...
Remove argument to disable it.
Add new --enable-compress-response-body argument to enable it.
2025-02-07 03:29:21 -05:00
comfyanonymous
1908648a5b
Remove some useless code.
2025-02-06 05:00:37 -05:00
comfyanonymous
50a778e1ce
Set the shift for Lumina back to 6.
2025-02-05 14:49:52 -05:00
comfyanonymous
ff7449847e
Upcasting rope to fp32 seems to make no difference in this model.
2025-02-05 04:32:47 -05:00
comfyanonymous
f210ee69df
Use regular numbers for rope in lumina model.
2025-02-05 04:17:25 -05:00
comfyanonymous
eaf581f1ed
Fix lumina 2 batches.
2025-02-04 21:48:11 -05:00
comfyanonymous
0467641f12
Lower the default shift of lumina to reduce artifacts.
2025-02-04 06:50:37 -05:00
comfyanonymous
24b0e00898
Fix on python 3.9
2025-02-04 04:20:56 -05:00
comfyanonymous
e03a640306
Support Lumina 2 model.
2025-02-04 04:16:30 -05:00
comfyanonymous
fc1a24c402
Use maximum negative value instead of -inf for masks in text encoders.
...
This is probably more correct.
2025-02-02 09:46:00 -05:00
Dr.Lt.Data
4a3500c3ef
better guide message for sageattention ( #6634 )
2025-02-02 09:26:47 -05:00
KarryCharon
e12555bdbf
add disable-compres-response-body cli args; add compress middleware; ( #6672 )
2025-02-02 09:24:55 -05:00
comfyanonymous
af91112845
Only use stable cascade lora format with cascade model.
2025-02-01 06:35:22 -05:00
comfyanonymous
dcacbab9b0
Allow batch of different sigmas when noise scaling.
2025-01-30 06:49:52 -05:00
filtered
acf1f98438
Allow changing folder_paths.base_path via command line argument. ( #6600 )
...
* Reimpl. CLI arg directly inside folder_paths.
* Update tests to use CLI arg mocking.
* Revert last-minute refactor.
* Fix test state polution.
2025-01-29 08:06:28 -05:00
comfyanonymous
9dbb2b6d89
More friendly error messages for corrupted safetensors files.
2025-01-28 09:41:09 -05:00
comfyanonymous
5da1230892
Lower minimum ratio of loaded weights on Nvidia.
2025-01-27 05:26:51 -05:00
comfyanonymous
00e27c6202
Remove redundant code.
2025-01-25 19:04:53 -05:00
comfyanonymous
6da2db0e10
Remove useless code.
2025-01-24 06:15:54 -05:00
comfyanonymous
9cc63be0a0
Remove useless code.
2025-01-23 05:56:23 -05:00
Chenlei Hu
25f4e415da
Remove unused function lcm in conds.py ( #6572 )
2025-01-23 05:54:09 -05:00
comfyanonymous
e3f3830dde
Remove support for python 3.8.
2025-01-22 17:04:30 -05:00
chaObserv
4d5a8e8525
Add gradient estimation sampler ( #6554 )
2025-01-22 05:29:40 -05:00
comfyanonymous
eb64454c08
Add FluxDisableGuidance node to disable using the guidance embed.
2025-01-20 14:50:24 -05:00
comfyanonymous
71281dcc56
Cleanup old TODO.
2025-01-20 03:44:13 -05:00
Sergii Dymchenko
441dfbf8d9
Use torch.special.expm1 ( #6388 )
...
* Use `torch.special.expm1`
This function provides greater precision than `exp(x) - 1` for small values of `x`.
Found with TorchFix https://github.com/pytorch-labs/torchfix/
* Use non-alias
2025-01-19 04:54:32 -05:00
catboxanon
ccec4da418
Remove comfy.samplers self-import ( #6506 )
2025-01-18 17:49:51 -05:00
comfyanonymous
39b2c8eeee
Uni pc sampler now works with audio and video models.
2025-01-18 05:27:58 -05:00
comfyanonymous
b8fbb7c807
Add warning when using old pytorch versions.
2025-01-17 18:47:27 -05:00
comfyanonymous
d439446067
Fix some cosmos fp8 issues.
2025-01-16 17:45:37 -05:00
comfyanonymous
63af9e9d90
Fix cosmos VAE failing with videos longer than 121 frames.
2025-01-16 16:30:06 -05:00
comfyanonymous
5650bfe1a9
Code refactor.
2025-01-16 07:23:54 -05:00
comfyanonymous
9ea569d2b0
Tweak hunyuan memory usage factor.
2025-01-16 06:31:03 -05:00
comfyanonymous
868a4d1551
Clean up some debug lines.
2025-01-16 04:24:39 -05:00
comfyanonymous
ad3ed4154c
More accurate memory estimation for cosmos and hunyuan video.
2025-01-16 03:48:40 -05:00
comfyanonymous
220f470bfa
Slightly lower hunyuan video memory usage.
2025-01-16 00:23:01 -05:00
comfyanonymous
3423ffcf70
Lower cosmos diffusion model memory usage.
2025-01-15 23:46:42 -05:00