2620 Commits

Author SHA1 Message Date
comfyanonymous
23922a2fbd Do RMSNorm in native type. 2024-08-27 02:41:56 -04:00
Chenlei Hu
daf975da91 Cleanup empty dir if frontend zip download failed (#4574) 2024-08-27 02:07:25 -04:00
Robin Huang
2f382e6c0f Add relative file path to the progress report. (#4621) 2024-08-27 02:06:12 -04:00
comfyanonymous
f2a13e96b4 Make the stochastic fp8 rounding reproducible. 2024-08-26 15:12:06 -04:00
comfyanonymous
6c7971a4d7 Use less memory in float8 lora patching by doing calculations in fp16. 2024-08-26 14:45:58 -04:00
comfyanonymous
0144d62aa9 Fix potential memory leak. 2024-08-26 02:07:32 -04:00
comfyanonymous
0911b077b3 Fix some controlnets OOMing when loading. 2024-08-25 05:54:29 -04:00
guill
42a3c54e60 [Bug #4529] Fix graph partial validation failure (#4588)
Currently, if a graph partially fails validation (i.e. some outputs are
valid while others have links from missing nodes), the execution loop
could get an exception resulting in server lockup.

This isn't actually possible to reproduce via the default UI, but is a
potential issue for people using the API to construct invalid graphs.
2024-08-24 15:34:58 -04:00
comfyanonymous
662b2e0d76 Clarify how to use high quality previews. 2024-08-24 02:31:03 -04:00
comfyanonymous
b9590ea077 Fix onnx export. 2024-08-23 17:52:47 -04:00
comfyanonymous
49cd4f5123 Fix dora. 2024-08-23 04:58:59 -04:00
comfyanonymous
65a577e0db Cleanup. 2024-08-23 04:06:27 -04:00
comfyanonymous
103d705247 Fix. 2024-08-23 04:04:55 -04:00
Simon Lui
24af9ec35a Rework IPEX check for future inclusion of XPU into Pytorch upstream and do a bit more optimization of ipex.optimize(). (#4562) 2024-08-23 03:59:57 -04:00
Scorpinaus
77841395ab Added SD15_Inpaint_Diffusers model support for unet_config_from_diffusers_unet function (#4565) 2024-08-23 03:57:08 -04:00
Chenlei Hu
f5b7f9bcfc Fix task.status.status_str caused by #2666 (#4551)
* Fix task.status.status_str caused by 2666 regression

* fix

* fix
2024-08-22 17:38:30 -04:00
comfyanonymous
c6f6298c52 Missing imports. 2024-08-22 17:20:51 -04:00
comfyanonymous
59d0907512 Move calculate function to comfy.lora 2024-08-22 17:12:00 -04:00
comfyanonymous
23836e6596 Code cleanups. 2024-08-22 17:05:12 -04:00
comfyanonymous
7bad18b66d Fix issue with models staying loaded in memory. 2024-08-22 15:58:20 -04:00
comfyanonymous
262abf2310 Try to fix some lora issues. 2024-08-22 15:32:18 -04:00
David
ff022fccc2 Generalize MacOS version check for force-upcast-attention (#4548)
This code automatically forces upcasting attention for MacOS versions 14.5 and 14.6. My computer returns the string "14.6.1" for `platform.mac_ver()[0]`, so this generalizes the comparison to catch more versions.

I am running MacOS Sonoma 14.6.1 (latest version) and was seeing black image generation on previously functional workflows after recent software updates. This PR solved the issue for me.

See comfyanonymous/ComfyUI#3521
2024-08-22 13:24:21 -04:00
comfyanonymous
2e13ae7d8c Fix issue. 2024-08-22 10:38:24 -04:00
guill
d47390e3af Fix a bug where cached outputs affected IS_CHANGED (#4535)
This change fixes a bug where non-constant values could be passed to the
IS_CHANGED function. This would result in workflows taking an extra
execution before they acted as if they were cached.

The actual change is like 4 characters -- the rest is adding unit tests.
2024-08-21 23:38:46 -04:00
comfyanonymous
f420b9c67d Add a shortcut to the nightly package to run with --fast. 2024-08-21 23:36:58 -04:00
comfyanonymous
190a03c892 fp16 is actually faster than fp32 on a GTX 1080. 2024-08-21 23:23:50 -04:00
comfyanonymous
883e5a9aa5 Fix slow performance on 10 series Nvidia GPUs. 2024-08-21 16:39:02 -04:00
comfyanonymous
9c847187cb Try a different type of flux fp16 fix. 2024-08-21 16:17:15 -04:00
comfyanonymous
a1a6c213a5 Make --fast work on pytorch nightly. 2024-08-21 14:01:41 -04:00
Svein Ove Aas
b4bedcdc11 Replace use of .view with .reshape (#4522)
When generating images with fp8_e4_m3 Flux and batch size >1, using --fast, ComfyUI throws a "view size is not compatible with input tensor's size and stride" error pointing at the first of these two calls to view.

As reshape is semantically equivalent to view except for working on a broader set of inputs, there should be no downside to changing this. The only difference is that it clones the underlying data in cases where .view would error out. I have confirmed that the output still looks as expected, but cannot confirm that no mutable use is made of the tensors anywhere.

Note that --fast is only marginally faster than the default.
2024-08-21 11:21:48 -04:00
Alex "mcmonkey" Goodwin
b5173d39d6 add a get models list api route (#4519)
* get models list api route

* remove copypasta
2024-08-21 02:04:42 -04:00
Robin Huang
db97ea767f Add GET /internal/files. (#4295)
* Create internal route table.

* List files.

* Add GET /internal/files.

Retrieves list of files in models, output, and user directories.

* Refactor file names.

* Use typing_extensions for Python 3.8

* Fix tests.

* Remove print statements.

* Update README.

* Add output and user to valid directory test.

* Add missing type hints.
2024-08-21 01:25:06 -04:00
comfyanonymous
d8af5d17b5 Remove useless line, adjust windows default reserved vram. 2024-08-21 00:47:19 -04:00
Chenlei Hu
a6944c1139 Add optional deprecated/experimental flag to node class (#4506)
* Add optional deprecated flag to node class

* nit

* Add experimental flag
2024-08-21 00:01:34 -04:00
Chenlei Hu
18e296d8da Update frontend to 1.2.30 (#4513) 2024-08-21 00:00:49 -04:00
comfyanonymous
6c68634351 Speed up fp8 matrix mult by using better code. 2024-08-20 22:53:26 -04:00
comfyanonymous
8addad7f2c Simpletrainer lycoris format. 2024-08-20 12:05:13 -04:00
comfyanonymous
7c0c89e70c Add --fast argument to enable experimental optimizations.
Optimizations that might break things/lower quality will be put behind
this flag first and might be enabled by default in the future.

Currently the only optimization is float8_e4m3fn matrix multiplication on
4000/ADA series Nvidia cards or later. If you have one of these cards you
will see a speed boost when using fp8_e4m3fn flux for example.
2024-08-20 11:55:51 -04:00
comfyanonymous
d30678a05e Support loading long clipl model with the CLIP loader node. 2024-08-20 10:46:36 -04:00
comfyanonymous
2062f06e50 Properly set if clip text pooled projection instead of using hack. 2024-08-20 10:46:36 -04:00
comfyanonymous
05938809a4 Make cast_to a nop if weight is already good. 2024-08-20 10:46:36 -04:00
comfyanonymous
c3bec99a9b Fix potential issue with multi devices. 2024-08-20 10:46:36 -04:00
Chenlei Hu
44a4791b6e Update README.md (Add shield badges) (#4490) 2024-08-19 18:25:20 -04:00
comfyanonymous
f9c6ff16f7 New load_text_encoder_state_dicts function.
Now you can load text encoders straight from a list of state dicts.
2024-08-19 17:36:35 -04:00
comfyanonymous
5115d926eb Add a --reserve-vram argument if you don't want comfy to use all of it.
--reserve-vram 1.0 for example will make ComfyUI try to keep 1GB vram free.

This can also be useful if workflows are failing because of OOM errors but
in that case please report it if --reserve-vram improves your situation.
2024-08-19 17:16:18 -04:00
comfyanonymous
f2d98aa557 Bug fixes. 2024-08-19 16:28:55 -04:00
comfyanonymous
0f356a9212 Use better dtype for the lowvram lora system. 2024-08-19 15:35:25 -04:00
comfyanonymous
48335ad8ca Remove duplication. 2024-08-19 15:26:50 -04:00
comfyanonymous
7aec00e4a2 CheckpointSave node name. 2024-08-19 15:06:12 -04:00
Yoland Yan
165e9f4b35 Update issue template config.yml to direct frontend issues to frontend repos (#4486)
* Update config.yml

* Typos
2024-08-19 13:41:30 -04:00