InvokeAI

mirror of https://github.com/invoke-ai/InvokeAI.git synced 2026-02-02 08:15:08 -05:00

Author	SHA1	Message	Date
psychedelicious	e09cf64779	feat: more updates to first run view	2025-01-09 11:20:05 +11:00
psychedelicious	e6deaa2d2f	feat(ui): minor layout tweaks for first run screen	2025-01-09 11:20:05 +11:00
psychedelicious	5246b31347	feat(ui): add low vram link to first run page	2025-01-09 11:20:05 +11:00
psychedelicious	89b576f10d	fix(ui): prevent canvas & main panel content from scrolling Hopefully fixes issues where, when run via the launcher, the main panel kinda just scrolls out of bounds.	2025-01-09 09:14:22 +11:00
psychedelicious	d7893a52c3	tweak(ui): whats new copy	2025-01-08 15:26:26 +11:00
Mary Hipp	b9c45c3232	Whats new update	2025-01-08 15:26:26 +11:00
David Burnett	afc9d3b98f	more ruff formating	2025-01-07 20:18:19 -05:00
David Burnett	7ddc757bdb	ruff format changes	2025-01-07 20:18:19 -05:00
David Burnett	d8da9b45cc	Fix for DEIS / DPM clash	2025-01-07 20:18:19 -05:00
Ryan Dick	607d19f4dd	We should not trust the value of since the model could be partially-loaded.	2025-01-07 19:22:31 -05:00
Ryan Dick	974b4671b1	Deprecate the `ram` and `vram` configs to make the migration to dynamic memory limits smoother for users who had previously overriden these values.	2025-01-07 16:45:29 +00:00
Ryan Dick	85eb4f0312	Fix an edge case with model offloading from VRAM to RAM. If a GGML-quantized model is offloaded from VRAM inside of a torch.inference_mode() context manager, this will cause the following error: 'RuntimeError: Cannot set version_counter for inference tensor'.	2025-01-07 15:59:50 +00:00
psychedelicious	67e948b50d	chore: bump version to v5.6.0rc1	2025-01-07 19:41:56 +11:00
Riccardo Giovanetti	d9a20f319f	translationBot(ui): update translation (Italian) Currently translated at 99.3% (1639 of 1649 strings) Co-authored-by: Riccardo Giovanetti <riccardo.giovanetti@gmail.com> Translate-URL: https://hosted.weblate.org/projects/invokeai/web-ui/it/ Translation: InvokeAI/Web UI	2025-01-07 19:32:50 +11:00
Riku	38d4863e09	translationBot(ui): update translation (German) Currently translated at 71.7% (1181 of 1645 strings) Co-authored-by: Riku <riku.block@gmail.com> Translate-URL: https://hosted.weblate.org/projects/invokeai/web-ui/de/ Translation: InvokeAI/Web UI	2025-01-07 19:32:50 +11:00
Nik Nikovsky	cd7ba14adc	translationBot(ui): update translation (Polish) Currently translated at 16.5% (273 of 1645 strings) translationBot(ui): update translation (Polish) Currently translated at 15.4% (254 of 1645 strings) translationBot(ui): update translation (Polish) Currently translated at 10.8% (178 of 1645 strings) Co-authored-by: Nik Nikovsky <zejdzztegomaila@gmail.com> Translate-URL: https://hosted.weblate.org/projects/invokeai/web-ui/pl/ Translation: InvokeAI/Web UI	2025-01-07 19:32:50 +11:00
Linos	e5b6beb24d	translationBot(ui): update translation (Vietnamese) Currently translated at 100.0% (1649 of 1649 strings) translationBot(ui): update translation (Vietnamese) Currently translated at 100.0% (1645 of 1645 strings) translationBot(ui): update translation (Vietnamese) Currently translated at 100.0% (1645 of 1645 strings) translationBot(ui): update translation (Vietnamese) Currently translated at 100.0% (1645 of 1645 strings) Co-authored-by: Linos <linos.coding@gmail.com> Translate-URL: https://hosted.weblate.org/projects/invokeai/web-ui/vi/ Translation: InvokeAI/Web UI	2025-01-07 19:32:50 +11:00
Ryan Dick	d7ab464176	Offload the current model when locking if it is already partially loaded and we have insufficient VRAM.	2025-01-07 02:53:44 +00:00
Ryan Dick	548b3eddb8	pnpm typegen	2025-01-07 01:20:15 +00:00
Ryan Dick	5b42b7bd45	Add a utility to help with determining the working memory required for expensive operations.	2025-01-07 01:20:15 +00:00
Ryan Dick	71b97ce7be	Reduce the likelihood of encountering https://github.com/invoke-ai/InvokeAI/issues/7513 by elminating places where the door was left open for this to happen.	2025-01-07 01:20:15 +00:00
Ryan Dick	b343f81644	Use torch.cuda.memory_allocated() rather than torch.cuda.memory_reserved() to be more conservative in setting dynamic VRAM cache limits.	2025-01-07 01:20:15 +00:00
Ryan Dick	4abfb35321	Tune SD3 VAE decode working memory estimate.	2025-01-07 01:20:15 +00:00
Ryan Dick	cba6528ea7	Add a 20% buffer to all VAE decode working memory estimates.	2025-01-07 01:20:15 +00:00
Ryan Dick	6a5cee61be	Tune the working memory estimate for FLUX VAE decoding.	2025-01-07 01:20:15 +00:00
Ryan Dick	bd8017ecd5	Update working memory estimate for VAE decoding when tiling is being applied.	2025-01-07 01:20:15 +00:00
Ryan Dick	299eb94a05	Estimate the working memory required for VAE decoding, since this operations tends to be memory intensive.	2025-01-07 01:20:15 +00:00
Ryan Dick	fc4a22fe78	Allow expensive operations to request more working memory.	2025-01-07 01:20:13 +00:00
Ryan Dick	a167632f09	Calculate model cache size limits dynamically based on the available RAM / VRAM.	2025-01-07 01:14:20 +00:00
Ryan Dick	1321fac8f2	Remove get_cache_size() and set_cache_size() endpoints. These were unused by the frontend and refer to cache fields that are no longer accessible.	2025-01-07 01:06:20 +00:00
Ryan Dick	6a9de1fcf3	Change definition of VRAM in use for the ModelCache from sum of model weights to the total torch.cuda.memory_allocated().	2025-01-07 00:31:53 +00:00
Ryan Dick	e5180c4e6b	Add get_effective_device(...) utility to aid in determining the effective device of models that are partially loaded.	2025-01-07 00:31:00 +00:00
Ryan Dick	2619ef53ca	Handle device casting in ia2_layer.py.	2025-01-07 00:31:00 +00:00
Ryan Dick	bcd29c5d74	Remove all cases where we check the 'model.device'. This is no longer trustworthy now that partial loading is permitted.	2025-01-07 00:31:00 +00:00
Ryan Dick	1b7bb70bde	Improve handling of cases when application code modifies the size of a model after registering it with the model cache.	2025-01-07 00:31:00 +00:00
Ryan Dick	7127040c3a	Remove unused function set_nested_attr(...).	2025-01-07 00:31:00 +00:00
Ryan Dick	ceb2498a67	Add log prefix to model cache logs.	2025-01-07 00:31:00 +00:00
Ryan Dick	d0bfa019be	Add 'enable_partial_loading' config flag.	2025-01-07 00:31:00 +00:00
Ryan Dick	535e45cedf	First pass at adding partial loading support to the ModelCache.	2025-01-07 00:30:58 +00:00
Ryan Dick	c579a218ef	Allow models to be locked in VRAM, even if they have been dropped from the RAM cache (related: https://github.com/invoke-ai/InvokeAI/issues/7513 ).	2025-01-06 23:02:52 +00:00
Riku	f4f7415a3b	fix(app): remove obsolete DEFAULT_PRECISION variable	2025-01-06 11:14:58 +11:00
Mary Hipp	7d6c443d6f	fix(api): limit board_name length to 300 characters	2025-01-06 10:49:49 +11:00
psychedelicious	4815b4ea80	feat(ui): tweak verbiage for model install errors	2025-01-03 11:21:23 -05:00
psychedelicious	d77a6ccd76	fix(ui): model install error toasts not updating correctly	2025-01-03 11:21:23 -05:00
psychedelicious	3e860c8338	feat(ui): starter models filter works with model base For example, "flux" now matches any starter model with a model base of "FLUX".	2025-01-03 11:21:23 -05:00
psychedelicious	4f2ef7ce76	refactor(ui): handle hf vs civitai/other url model install errors separately Previously, we didn't differentiate between model install errors for different types of model install sources, resulting in a buggy UX: - If a HF model install failed, but it was a HF URL install and not a repo id install, the link to the HF model page was incorrect. - If a non-HF URL install (e.g. civitai) failed, we treated it as a HF URL install. In this case, if the user's HF token was invalid or unset, we directed the user to set it. If the HF token was valid, we displayed an empty red toast. If it's not a HF URL install, then of course neither of these are correct. Also, the logic for handling the toasts was a bit complicated. This change does a few things: - Consolidate the model install error toasts into one place - the socket.io event handler for the model install error event. There is no more global state for the toasts and there are no hooks managing them. - Handling the different cases for errors, including all combinations of HF/non-HF and unauthorized/forbidden/unknown.	2025-01-03 11:21:23 -05:00
psychedelicious	d7e9ad52f9	chore(ui): typegen	2025-01-03 11:21:23 -05:00
psychedelicious	b6d7a44004	refactor(events): include full model source in model install events This is required to fix an issue with the MM UI's error handling. Previously, we only included the model source as a string. That could be an arbitrary URL, file path or HF repo id, but the frontend has no parsing logic to differentiate between these different model sources. Without access to the type of model source, it is difficult to determine how the user should proceed. For example, if it's HF URL with an HTTP unauthorized error, we should direct the user to log in to HF. But if it's a civitai URL with the same error, we should not direct the user to HF. There are a variety of related edge cases. With this change, the full `ModelSource` object is included in each model install event, including error events. I had to fix some circular import issues, hence the import changes to files other than `events_common.py`.	2025-01-03 11:21:23 -05:00
psychedelicious	e18100ae7e	refactor(ui): move model install error event handling to own file No logic change.	2025-01-03 11:21:23 -05:00
psychedelicious	ad0aa0e6b2	feat(ui): reset canvas layers only resets the layers	2025-01-03 11:02:04 -05:00

1 2 3 4 5 ...

10199 Commits