ComfyUI

mirror of https://github.com/comfyanonymous/ComfyUI.git synced 2025-07-13 19:07:01 +08:00

Author	SHA1	Message	Date
comfyanonymous	8af9a91e0c	A few improvements to #5937 .	2024-12-06 05:49:15 -05:00
Jedrzej Kosinski	0ee322ec5f	ModelPatcher Overhaul and Hook Support (#5583 ) * Added hook_patches to ModelPatcher for weights (model) * Initial changes to calc_cond_batch to eventually support hook_patches * Added current_patcher property to BaseModel * Consolidated add_hook_patches_as_diffs into add_hook_patches func, fixed fp8 support for model-as-lora feature * Added call to initialize_timesteps on hooks in process_conds func, and added call prepare current keyframe on hooks in calc_cond_batch * Added default_conds support in calc_cond_batch func * Added initial set of hook-related nodes, added code to register hooks for loras/model-as-loras, small renaming/refactoring * Made CLIP work with hook patches * Added initial hook scheduling nodes, small renaming/refactoring * Fixed MaxSpeed and default conds implementations * Added support for adding weight hooks that aren't registered on the ModelPatcher at sampling time * Made Set Clip Hooks node work with hooks from Create Hook nodes, began work on better Create Hook Model As LoRA node * Initial work on adding 'model_as_lora' lora type to calculate_weight * Continued work on simpler Create Hook Model As LoRA node, started to implement ModelPatcher callbacks, attachments, and additional_models * Fix incorrect ref to create_hook_patches_clone after moving function * Added injections support to ModelPatcher + necessary bookkeeping, added additional_models support in ModelPatcher, conds, and hooks * Added wrappers to ModelPatcher to facilitate standardized function wrapping * Started scaffolding for other hook types, refactored get_hooks_from_cond to organize hooks by type * Fix skip_until_exit logic bug breaking injection after first run of model * Updated clone_has_same_weights function to account for new ModelPatcher properties, improved AutoPatcherEjector usage in partially_load * Added WrapperExecutor for non-classbound functions, added calc_cond_batch wrappers * Refactored callbacks+wrappers to allow storing lists by id * Added forward_timestep_embed_patch type, added helper functions on ModelPatcher for emb_patch and forward_timestep_embed_patch, added helper functions for removing callbacks/wrappers/additional_models by key, added custom_should_register prop to hooks * Added get_attachment func on ModelPatcher * Implement basic MemoryCounter system for determing with cached weights due to hooks should be offloaded in hooks_backup * Modified ControlNet/T2IAdapter get_control function to receive transformer_options as additional parameter, made the model_options stored in extra_args in inner_sample be a clone of the original model_options instead of same ref * Added create_model_options_clone func, modified type annotations to use __future__ so that I can use the better type annotations * Refactored WrapperExecutor code to remove need for WrapperClassExecutor (now gone), added sampler.sample wrapper (pending review, will likely keep but will see what hacks this could currently let me get rid of in ACN/ADE) * Added Combine versions of Cond/Cond Pair Set Props nodes, renamed Pair Cond to Cond Pair, fixed default conds never applying hooks (due to hooks key typo) * Renamed Create Hook Model As LoRA nodes to make the test node the main one (more changes pending) * Added uuid to conds in CFGGuider and uuids to transformer_options to allow uniquely identifying conds in batches during sampling * Fixed models not being unloaded properly due to current_patcher reference; the current ComfyUI model cleanup code requires that nothing else has a reference to the ModelPatcher instances * Fixed default conds not respecting hook keyframes, made keyframes not reset cache when strength is unchanged, fixed Cond Set Default Combine throwing error, fixed model-as-lora throwing error during calculate_weight after a recent ComfyUI update, small refactoring/scaffolding changes for hooks * Changed CreateHookModelAsLoraTest to be the new CreateHookModelAsLora, rename old ones as 'direct' and will be removed prior to merge * Added initial support within CLIP Text Encode (Prompt) node for scheduling weight hook CLIP strength via clip_start_percent/clip_end_percent on conds, added schedule_clip toggle to Set CLIP Hooks node, small cleanup/fixes * Fix range check in get_hooks_for_clip_schedule so that proper keyframes get assigned to corresponding ranges * Optimized CLIP hook scheduling to treat same strength as same keyframe * Less fragile memory management. * Make encode_from_tokens_scheduled call cleaner, rollback change in model_patcher.py for hook_patches_backup dict * Fix issue. * Remove useless function. * Prevent and detect some types of memory leaks. * Run garbage collector when switching workflow if needed. * Moved WrappersMP/CallbacksMP/WrapperExecutor to patcher_extension.py * Refactored code to store wrappers and callbacks in transformer_options, added apply_model and diffusion_model.forward wrappers * Fix issue. * Refactored hooks in calc_cond_batch to be part of get_area_and_mult tuple, added extra_hooks to ControlBase to allow custom controlnets w/ hooks, small cleanup and renaming * Fixed inconsistency of results when schedule_clip is set to False, small renaming/typo fixing, added initial support for ControlNet extra_hooks to work in tandem with normal cond hooks, initial work on calc_cond_batch merging all subdicts in returned transformer_options * Modified callbacks and wrappers so that unregistered types can be used, allowing custom_nodes to have their own unique callbacks/wrappers if desired * Updated different hook types to reflect actual progress of implementation, initial scaffolding for working WrapperHook functionality * Fixed existing weight hook_patches (pre-registered) not working properly for CLIP * Removed Register/Direct hook nodes since they were present only for testing, removed diff-related weight hook calculation as improved_memory removes unload_model_clones and using sample time registered hooks is less hacky * Added clip scheduling support to all other native ComfyUI text encoding nodes (sdxl, flux, hunyuan, sd3) * Made WrapperHook functional, added another wrapper/callback getter, added ON_DETACH callback to ModelPatcher * Made opt_hooks append by default instead of replace, renamed comfy.hooks set functions to be more accurate * Added apply_to_conds to Set CLIP Hooks, modified relevant code to allow text encoding to automatically apply hooks to output conds when apply_to_conds is set to True * Fix cached_hook_patches not respecting target_device/memory_counter results * Fixed issue with setting weights from hooks instead of copying them, added additional memory_counter check when caching hook patches * Remove unnecessary torch.no_grad calls for hook patches * Increased MemoryCounter minimum memory to leave free by 2 until a better way to get inference memory estimate of currently loaded models exists For encode_from_tokens_scheduled, allow start_percent and end_percent in add_dict to limit which scheduled conds get encoded for optimization purposes * Removed a .to call on results of calculate_weight in patch_hook_weight_to_device that was screwing up the intermediate results for fp8 prior to being passed into stochastic_rounding call * Made encode_from_tokens_scheduled work when no hooks are set on patcher * Small cleanup of comments * Turn off hook patch caching when only 1 hook present in sampling, replace some current_hook = None with calls to self.patch_hooks(None) instead to avoid a potential edge case * On Cond/Cond Pair nodes, removed opt_ prefix from optional inputs * Allow both FLOATS and FLOAT for floats_strength input * Revert change, does not work * Made patch_hook_weight_to_device respect set_func and convert_func * Make discard_model_sampling True by default * Add changes manually from 'master' so merge conflict resolution goes more smoothly * Cleaned up text encode nodes with just a single clip.encode_from_tokens_scheduled call * Make sure encode_from_tokens_scheduled will respect use_clip_schedule on clip * Made nodes in nodes_hooks be marked as experimental (beta) * Add get_nested_additional_models for cases where additional_models could have their own additional_models, and add robustness for circular additional_models references * Made finalize_default_conds area math consistent with other sampling code * Changed 'opt_hooks' input of Cond/Cond Pair Set Default Combine nodes to 'hooks' * Remove a couple old TODO's and a no longer necessary workaround	2024-12-02 14:51:02 -05:00
comfyanonymous	497db6212f	Alternative fix for #5767	2024-11-26 17:53:04 -05:00
comfyanonymous	5818f6cf51	Remove print.	2024-11-22 10:49:15 -05:00
comfyanonymous	5e16f1d24b	Support Lightricks LTX-Video model.	2024-11-22 08:46:39 -05:00
comfyanonymous	8f0009aad0	Support new flux model variants.	2024-11-21 08:38:23 -05:00
comfyanonymous	b699a15062	Refactor inpaint/ip2p code.	2024-11-19 03:25:25 -05:00
comfyanonymous	5cbb01bc2f	Basic Genmo Mochi video model support. To use: "Load CLIP" node with t5xxl + type mochi "Load Diffusion Model" node with the mochi dit file. "Load VAE" with the mochi vae file. EmptyMochiLatentVideo node for the latent. euler + linear_quadratic in the KSampler node.	2024-10-26 06:54:00 -04:00
comfyanonymous	0075c6d096	Mixed precision diffusion models with scaled fp8. This change allows supports for diffusion models where all the linears are scaled fp8 while the other weights are the original precision.	2024-10-21 18:12:51 -04:00
comfyanonymous	a68bbafddb	Support diffusion models with scaled fp8 weights.	2024-10-19 23:47:42 -04:00
comfyanonymous	e38c94228b	Add a weight_dtype fp8_e4m3fn_fast to the Diffusion Model Loader node. This is used to load weights in fp8 and use fp8 matrix multiplication.	2024-10-09 19:43:17 -04:00
comfyanonymous	9953f22fce	Add --fast argument to enable experimental optimizations. Optimizations that might break things/lower quality will be put behind this flag first and might be enabled by default in the future. Currently the only optimization is float8_e4m3fn matrix multiplication on 4000/ADA series Nvidia cards or later. If you have one of these cards you will see a speed boost when using fp8_e4m3fn flux for example.	2024-08-20 11:55:51 -04:00
comfyanonymous	e9589d6d92	Add a way to set model dtype and ops from load_checkpoint_guess_config.	2024-08-11 08:50:34 -04:00
comfyanonymous	b334605a66	Fix OOMs happening in some cases. A cloned model patcher sometimes reported a model was loaded on a device when it wasn't.	2024-08-06 13:36:04 -04:00
comfyanonymous	0a6b008117	Fix issue with some custom nodes.	2024-08-04 10:03:33 -04:00
comfyanonymous	ba9095e5bd	Automatically use fp8 for diffusion model weights if: Checkpoint contains weights in fp8. There isn't enough memory to load the diffusion model in GPU vram.	2024-08-03 13:45:19 -04:00
comfyanonymous	ea03c9dcd2	Better per model memory usage estimations.	2024-08-02 18:09:24 -04:00
comfyanonymous	3a9ee995cf	Tweak regular SD memory formula.	2024-08-02 17:34:30 -04:00
comfyanonymous	47da42d928	Better Flux vram estimation.	2024-08-02 17:02:35 -04:00
comfyanonymous	d420bc792a	Tweak the memory usage formulas for Flux and SD.	2024-08-01 17:53:45 -04:00
comfyanonymous	1589b58d3e	Basic Flux Schnell and Flux Dev model implementation.	2024-08-01 09:49:29 -04:00
comfyanonymous	a5f4292f9f	Basic hunyuan dit implementation. (#4102 ) * Let tokenizers return weights to be stored in the saved checkpoint. * Basic hunyuan dit implementation. * Fix some resolutions not working. * Support hydit checkpoint save. * Init with right dtype. * Switch to optimized attention in pooler. * Fix black images on hunyuan dit.	2024-07-25 18:21:08 -04:00
comfyanonymous	9f291d75b3	AuraFlow model implementation.	2024-07-11 16:52:26 -04:00
comfyanonymous	8ceb5a02a3	Support saving stable audio checkpoint that can be loaded back.	2024-06-27 11:06:52 -04:00
comfyanonymous	bb1969cab7	Initial support for the stable audio open model.	2024-06-15 12:14:56 -04:00
comfyanonymous	0ec513d877	Add a --force-channels-last to inference models in channel last mode.	2024-06-15 01:08:12 -04:00
comfyanonymous	1ddf512fdc	Don't auto convert clip and vae weights to fp16 when saving checkpoint.	2024-06-12 01:07:58 -04:00
comfyanonymous	694e0b48e0	SD3 better memory usage estimation.	2024-06-12 00:49:00 -04:00
comfyanonymous	9424522ead	Reuse code.	2024-06-11 07:20:26 -04:00
comfyanonymous	8c4a9befa7	SD3 Support.	2024-06-10 14:06:23 -04:00
comfyanonymous	cd07340d96	Typo fix.	2024-05-08 18:36:56 -04:00
comfyanonymous	1088d1850f	Support for CosXL models.	2024-04-05 10:53:41 -04:00
comfyanonymous	575acb69e4	IP2P model loading support. This is the code to load the model and inference it with only a text prompt. This commit does not contain the nodes to properly use it with an image input. This supports both the original SD1 instructpix2pix model and the diffusers SDXL one.	2024-03-31 03:10:28 -04:00
comfyanonymous	94a5a67c32	Cleanup to support different types of inpaint models.	2024-03-29 14:44:13 -04:00
comfyanonymous	40e124c6be	SV3D support.	2024-03-18 16:54:13 -04:00
comfyanonymous	0ed72befe1	Change log levels. Logging level now defaults to info. --verbose sets it to debug.	2024-03-11 13:54:56 -04:00
comfyanonymous	65397ce601	Replace prints with logging and add --verbose argument.	2024-03-10 12:14:23 -04:00
comfyanonymous	51df846598	Let conditioning specify custom concat conds.	2024-03-02 11:44:06 -05:00
comfyanonymous	cb7c3a2921	Allow image_only_indicator to be None.	2024-02-29 13:11:30 -05:00
comfyanonymous	8daedc5bf2	Auto detect playground v2.5 model.	2024-02-27 18:03:03 -05:00
comfyanonymous	0d0fbabd1d	Pass pooled CLIP to stage b.	2024-02-20 04:24:45 -05:00
comfyanonymous	667c92814e	Stable Cascade Stage B.	2024-02-16 13:02:03 -05:00
comfyanonymous	f83109f09b	Stable Cascade Stage C.	2024-02-16 10:55:08 -05:00
comfyanonymous	25a4805e51	Add a way to set different conditioning for the controlnet.	2024-02-09 14:13:31 -05:00
comfyanonymous	4871a36458	Cleanup some unused imports.	2024-01-21 21:51:22 -05:00
comfyanonymous	d76a04b6ea	Add unfinished ImageOnlyCheckpointSave node to save a SVD checkpoint. This node is unfinished, SVD checkpoints saved with this node will work with ComfyUI but not with anything else.	2024-01-17 19:46:21 -05:00
comfyanonymous	2395ae740a	Make unclip more deterministic. Pass a seed argument note that this might make old unclip images different.	2024-01-14 17:28:31 -05:00
comfyanonymous	10f2609fdd	Add InpaintModelConditioning node. This is an alternative to VAE Encode for inpaint that should work with lower denoise. This is a different take on #2501	2024-01-11 03:15:27 -05:00
comfyanonymous	8c6493578b	Implement noise augmentation for SD 4X upscale model.	2024-01-03 14:27:11 -05:00
comfyanonymous	a7874d1a8b	Add support for the stable diffusion x4 upscaling model. This is an old model. Load the checkpoint like a regular one and use the new SD_4XUpscale_Conditioning node.	2024-01-03 03:37:56 -05:00

1 2

93 Commits