InvokeAI

mirror of https://github.com/invoke-ai/InvokeAI synced 2025-07-25 21:05:37 +00:00

Author	SHA1	Message	Date
Ryan Dick	87261bdbc9	FLUX memory management improvements (#6791 ) ## Summary This PR contains several improvements to memory management for FLUX workflows. It is now possible to achieve better FLUX model caching performance, but this still requires users to manually configure their `ram`/`vram` settings. E.g. a `vram` setting of 16.0 should allow for all quantized FLUX models to be kept in memory on the GPU. Changes: - Check the size of a model on disk and free the requisite space in the model cache before loading it. (This behaviour existed previously, but was removed in https://github.com/invoke-ai/InvokeAI/pull/6072/files. The removal did not seem to be intentional). - Removed the hack to free 24GB of space in the cache before loading the FLUX model. - Split the T5 embedding and CLIP embedding steps into separate functions so that the two models don't both have to be held in RAM at the same time. - Fix a bug in `InvokeLinear8bitLt` that was causing some tensors to be left on the GPU when the model was offloaded to the CPU. (This class is getting very messy due to the non-standard state_dict handling in `bnb.nn.Linear8bitLt`. ) - Tidy up some dtype handling in FluxTextToImageInvocation to avoid situations where we hold references to two copies of the same tensor unnecessarily. - (minor) Misc cleanup of ModelCache: improve docs and remove unused vars. Future: We should revisit our default ram/vram configs. The current defaults are very conservative, and users could see major performance improvements from tuning these values. ## QA Instructions I tested the FLUX workflow with the following configurations and verified that the cache hit rates and memory usage matched the expected behaviour: - `ram = 16` and `vram = 16` - `ram = 16` and `vram = 1` - `ram = 1` and `vram = 1` Note that the changes in this PR are not isolated to FLUX. Since we now check the size of models on disk, we may see slight changes in model cache offload patterns for other models as well. ## Checklist - [x] _The PR has a short but descriptive title, suitable for a changelog_ - [x] _Tests added / updated (if applicable)_ - [x] _Documentation added / updated (if applicable)_	2024-08-29 15:17:45 -04:00
Ryan Dick	4e4b6c6dbc	Tidy variable management and dtype handling in FluxTextToImageInvocation.	2024-08-29 19:08:18 +00:00
Ryan Dick	5e8cf9fb6a	Remove hack to clear cache from the FluxTextToImageInvocation. We now clear the cache based on the on-disk model size.	2024-08-29 19:08:18 +00:00
Ryan Dick	c738fe051f	Split T5 encoding and CLIP encoding into separate functions to ensure that all model references are locally-scoped so that the two models don't have to be help in memory at the same time.	2024-08-29 19:08:18 +00:00
Ryan Dick	29fe1533f2	Fix bug in InvokeLinear8bitLt that was causing old state information to persist after loading from a state dict. This manifested as state tensors being left on the GPU even when a model had been offloaded to the CPU cache.	2024-08-29 19:08:18 +00:00
Ryan Dick	77090070bd	Check the size of a model on disk and make room for it in the cache before loading it.	2024-08-29 19:08:18 +00:00
Ryan Dick	6ba9b1b6b0	Tidy up GIG -> GB and remove unused GIG constant.	2024-08-29 19:08:18 +00:00
Ryan Dick	c578b8df1e	Improve ModelCache docs.	2024-08-29 19:08:18 +00:00
Ryan Dick	cad9a41433	Remove unused MOdelCache.exists(...) function.	2024-08-29 19:08:18 +00:00
Ryan Dick	5fefb3b0f4	Remove unused param from ModelCache.	2024-08-29 19:08:18 +00:00
Ryan Dick	5284a870b0	Remove unused constructor params from ModelCache.	2024-08-29 19:08:18 +00:00
Ryan Dick	e064377c05	Remove default model cache sizes from model_cache_default.py. These defaults were misleading, because the config defaults take precedence over them.	2024-08-29 19:08:18 +00:00
Mary Hipp	3e569c8312	feat(ui): add fields for CLIP embed models and Flux VAE models in workflows	2024-08-29 11:52:51 -04:00
maryhipp	16825ee6e9	feat(nodes): bump version of flux model node, update default workflow	2024-08-29 11:52:51 -04:00
Mary Hipp	3f5340fa53	feat(nodes): add submodels as inputs to FLUX main model node instead of hardcoded names	2024-08-29 11:52:51 -04:00
chainchompa	f2a1a39b33	Add selectedStylePreset to app parameters (#6787 ) ## Summary - Add selectedStylePreset to app parameters <!--A description of the changes in this PR. Include the kind of change (fix, feature, docs, etc), the "why" and the "how". Screenshots or videos are useful for frontend changes.--> ## Related Issues / Discussions <!--WHEN APPLICABLE: List any related issues or discussions on github or discord. If this PR closes an issue, please use the "Closes #1234" format, so that the issue will be automatically closed when the PR merges.--> ## QA Instructions <!--WHEN APPLICABLE: Describe how you have tested the changes in this PR. Provide enough detail that a reviewer can reproduce your tests.--> ## Merge Plan <!--WHEN APPLICABLE: Large PRs, or PRs that touch sensitive things like DB schemas, may need some care when merging. For example, a careful rebase by the change author, timing to not interfere with a pending release, or a message to contributors on discord after merging.--> ## Checklist - [ ] _The PR has a short but descriptive title, suitable for a changelog_ - [ ] _Tests added / updated (if applicable)_ - [ ] _Documentation added / updated (if applicable)_	2024-08-28 10:53:07 -04:00
chainchompa	326de55d3e	remove api changes and only preselect style preset	2024-08-28 09:53:29 -04:00
chainchompa	b2df909570	added selectedStylePreset to preload presets when app loads	2024-08-28 09:50:44 -04:00
chainchompa	026ac36b06	Revert "added selectedStylePreset to preload presets when app loads" This reverts commit `e97fd85904`.	2024-08-28 09:44:08 -04:00
chainchompa	92125e5fd2	bug fixes	2024-08-27 16:13:38 -04:00
chainchompa	c0c139da88	formatting ruff	2024-08-27 15:46:51 -04:00
chainchompa	404ad6a7fd	cleanup	2024-08-27 15:42:42 -04:00
chainchompa	fc39086fb4	call stylePresetSelected	2024-08-27 15:34:31 -04:00
chainchompa	cd215700fe	added route for selecting style preset	2024-08-27 15:34:07 -04:00
chainchompa	e97fd85904	added selectedStylePreset to preload presets when app loads	2024-08-27 15:33:24 -04:00
Brandon Rising	0a263fa5b1	chore: bump version to v4.2.9rc1 v4.2.9rc1	2024-08-27 12:09:27 -04:00
Mary Hipp	fae3836a8d	fix CLIP	2024-08-27 10:29:10 -04:00
Mary Hipp	b3d2eb4178	add translations for new model types in MM, remove clip vision from filter since its not displayed in list	2024-08-27 10:29:10 -04:00
psychedelicious	576f1cbb75	build: remove broken scripts These two scripts are broken and can cause data loss. Remove them. They are not in the launcher script, but _are_ available to users in the terminal/file browser. Hopefully, when we removing them here, `pip` will delete them on next installation of the package...	2024-08-27 22:01:45 +10:00
Ryan Dick	50085b40bb	Update starter model size estimates.	2024-08-26 20:17:50 -04:00
Mary Hipp	cff382715a	default workflow: add steps to exposed fields, add more notes	2024-08-26 20:17:50 -04:00
Brandon Rising	54d54d1bf2	Run ruff	2024-08-26 20:17:50 -04:00
Mary Hipp	e84ea68282	remove prompt	2024-08-26 20:17:50 -04:00
Mary Hipp	160dd36782	update default workflow for flux	2024-08-26 20:17:50 -04:00
Brandon Rising	65bb46bcca	Rename params for flux and flux vae, add comments explaining use of the config_path in model config	2024-08-26 20:17:50 -04:00
Brandon Rising	2d185fb766	Run ruff	2024-08-26 20:17:50 -04:00
Brandon Rising	2ba9b02932	Fix type error in tsc	2024-08-26 20:17:50 -04:00
Brandon Rising	849da67cc7	Remove no longer used code in the flux denoise function	2024-08-26 20:17:50 -04:00
Brandon Rising	3ea6c9666e	Remove in progress images until we're able to make the valuable	2024-08-26 20:17:50 -04:00
Brandon Rising	cf633e4ef2	Only install starter models if not already installed	2024-08-26 20:17:50 -04:00
Ryan Dick	bbf934d980	Remove outdated TODO.	2024-08-26 20:17:50 -04:00
Ryan Dick	620f733110	ruff format	2024-08-26 20:17:50 -04:00
Ryan Dick	67928609a3	Downgrade accelerate and huggingface-hub deps to original versions.	2024-08-26 20:17:50 -04:00
Ryan Dick	5f15afb7db	Remove flux repo dependency	2024-08-26 20:17:50 -04:00
Ryan Dick	635d2f480d	ruff	2024-08-26 20:17:50 -04:00
Brandon Rising	70c278c810	Remove dependency on flux config files	2024-08-26 20:17:50 -04:00
Brandon Rising	56b9906e2e	Setup scaffolding for in progress images and add ability to cancel the flux node	2024-08-26 20:17:50 -04:00
Ryan Dick	a808ce81fd	Replace swish() with torch.nn.functional.silu(h). They are functionally equivalent, but in my test VAE deconding was ~8% faster after the change.	2024-08-26 20:17:50 -04:00
Ryan Dick	83f82c5ddf	Switch the CLIP-L start model to use our hosted version - which is much smaller.	2024-08-26 20:17:50 -04:00
Brandon Rising	101de8c25d	Update t5 encoder formats to accurately reflect the quantization strategy and data type	2024-08-26 20:17:50 -04:00

1 2 3 4 5 ...

12660 Commits