Kohya-ss-sd-scripts

mirror of https://github.com/kohya-ss/sd-scripts.git synced 2026-04-10 15:00:23 +00:00

Author	SHA1	Message	Date
Matthew Turnshek	a10ca40d37	Merge `6f7498f959` into `feb38356ea`	2026-03-27 18:04:42 +00:00
Kohya S.	feb38356ea	Merge pull request #2291 from kohya-ss/fix-anima-validation-with-text-encoder-output-cache fix: Anima validation dataset not working with Text Encoder output cache	2026-03-22 22:23:27 +09:00
Kohya S	cdb49f9fe7	fix: Anima validation dataset not working with Text Encoder output caching due to caption dropout	2026-03-22 22:19:47 +09:00
Kohya S	bd19e4c15d	Merge branch 'main' into sd3	2026-03-22 21:10:51 +09:00
Kohya S.	b2abe873a5	Merge pull request #2283 from kozistr/deps/pytorch-optimizer Bump `pytorch-optimizer` into 3.10.0	2026-03-19 09:18:06 +09:00
Kohya S.	7c159291e9	docs: add skip_image_resolution to config README (#2288 ) * docs: add skip_image_resolution option to config README Document the skip_image_resolution dataset option added in PR #2273. Add option description, multi-resolution dataset TOML example, and command-line argument entry to both Japanese and English config READMEs. Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com> * docs: clarify `skip_image_resolution` functionality in dataset config --------- Co-authored-by: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-19 09:17:29 +09:00
woctordho	1cd95b2d8b	Add `skip_image_resolution` to deduplicate multi-resolution dataset (#2273 ) * Add min_orig_resolution and max_orig_resolution * Rename min_orig_resolution to skip_image_resolution; remove max_orig_resolution * Change skip_image_resolution to tuple * Move filtering to __init__ * Minor fix	2026-03-19 08:43:39 +09:00
kozistr	1bd0b0faf1	build(deps): bump pytorch-optimizer into 3.10.0	2026-03-02 14:39:48 +09:00
Kohya S	d633b51126	Merge branch 'dev' into sd3	2026-02-26 08:22:30 +09:00
Kohya S.	1a3ec9ea74	Merge pull request #2280 from kohya-ss/fix-main-wd14-tagger-unbound-local-error fix: rename character_tags to img_character_tags to fix UnboundLocalError	2026-02-26 08:21:42 +09:00
Kohya S	e1aedceffa	fix: rename character_tags to img_character_tags to fix unboundlocalerror	2026-02-26 08:18:45 +09:00
Kohya S.	2217704ce1	feat: Support LoKr/LoHa for SDXL and Anima (#2275 ) * feat: Add LoHa/LoKr network support for SDXL and Anima - networks/network_base.py: shared AdditionalNetwork base class with architecture auto-detection (SDXL/Anima) and generic module injection - networks/loha.py: LoHa (Low-rank Hadamard Product) module with HadaWeight custom autograd, training/inference classes, and factory functions - networks/lokr.py: LoKr (Low-rank Kronecker Product) module with factorization, training/inference classes, and factory functions - library/lora_utils.py: extend weight merge hook to detect and merge LoHa/LoKr weights alongside standard LoRA Linear and Conv2d 1x1 layers only; Conv2d 3x3 (Tucker decomposition) support will be added separately. Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com> * feat: Enhance LoHa and LoKr modules with Tucker decomposition support - Added Tucker decomposition functionality to LoHa and LoKr modules. - Implemented new methods for weight rebuilding using Tucker decomposition. - Updated initialization and weight handling for Conv2d 3x3+ layers. - Modified get_diff_weight methods to accommodate Tucker and non-Tucker modes. - Enhanced network base to include unet_conv_target_modules for architecture detection. * fix: rank dropout handling in LoRAModule for Conv2d and Linear layers, see #2272 for details * doc: add dtype comment for load_safetensors_with_lora_and_fp8 function * fix: enhance architecture detection to support InferSdxlUNet2DConditionModel for gen_img.py * doc: update model support structure to include Lumina Image 2.0, HunyuanImage-2.1, and Anima-Preview * doc: add documentation for LoHa and LoKr fine-tuning methods * Update networks/network_base.py Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com> * Update docs/loha_lokr.md Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com> * fix: refactor LoHa and LoKr imports for weight merging in load_safetensors_with_lora_and_fp8 function --------- Co-authored-by: Claude Opus 4.6 <noreply@anthropic.com> Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com>	2026-02-23 22:09:00 +09:00
Kohya S.	f90fa1a89a	feat: backward compatibility for SD/SDXL latent cache (#2276 ) * fix: improve handling of legacy npz files and add logging for fallback scenarios * fix: simplify fallback handling in SdSdxlLatentsCachingStrategy	2026-02-23 21:44:51 +09:00
Kohya S.	98a42e4cd6	Merge pull request #2277 from kohya-ss/feat-stability-with-fp16-for-anima feat: Stability with fp16 for anima	2026-02-23 21:15:49 +09:00
Kohya S	892f8be78f	fix: cast input tensor to float32 for improved numerical stability in residual connections	2026-02-23 21:12:57 +09:00
woctordho	50694df3cf	Multi-resolution dataset for SD1/SDXL (#2269 ) * Multi-resolution dataset for SD1/SDXL * Add fallback to legacy key without resolution suffix * Support numpy 2.2	2026-02-23 15:30:36 +09:00
duongve13112002	609d1292f6	Fix the LoRA dropout issue in the Anima model during LoRA training. (#2272 ) * Support network_reg_alphas and fix bug when setting rank_dropout in training lora for anima model * Update anima_train_network.md * Update anima_train_network.md * Remove network_reg_alphas * Update document	2026-02-23 15:13:40 +09:00
Kohya S.	48d368fa55	Merge pull request #2268 from kohya-ss/sd3 merge sd3 to main	2026-02-16 08:07:29 +09:00
Kohya S.	3265f2edfb	Merge pull request #2267 from kohya-ss/fix-github-actions-error fix: `str is not "no"` to `str != "no"`	2026-02-16 08:01:20 +09:00
Kohya S	ef051427df	fix: `str is not "no"` to `str != "no"`	2026-02-16 07:58:15 +09:00
Kohya S.	573a7fa06c	Merge pull request #2262 from duongve13112002/fix_lumina Fix bug and optimization for Lumina model	2026-02-16 07:54:49 +09:00
Kohya S.	ae72efb92b	Merge pull request #2264 from kohya-ss/release/v0.10.1 Release 0.10.1 v0.10.1	2026-02-13 08:34:42 +09:00
Kohya S	449e70b4cf	README: Update change history for version 0.10.1 with Anima model support	2026-02-13 08:31:22 +09:00
Kohya S.	b237b8deb3	Merge pull request #2263 from kohya-ss/sd3 feat: Anima support	2026-02-13 08:20:44 +09:00
Kohya S.	34e7138b6a	Add/modify some implementation for anima (#2261 ) * fix: update extend-exclude list in _typos.toml to include configs * fix: exclude anima tests from pytest * feat: add entry for 'temperal' in extend-words section of _typos.toml for Qwen-Image VAE * fix: update default value for --discrete_flow_shift in anima training guide * feat: add Qwen-Image VAE * feat: simplify encode_tokens * feat: use unified attention module, add wrapper for state dict compatibility * feat: loading with dynamic fp8 optimization and LoRA support * feat: add anima minimal inference script (WIP) * format: format * feat: simplify target module selection by regular expression patterns * feat: kept caption dropout rate in cache and handle in training script * feat: update train_llm_adapter and verbose default values to string type * fix: use strategy instead of using tokenizers directly * feat: add dtype property and all-zero mask handling in cross-attention in LLMAdapterTransformerBlock * feat: support 5d tensor in get_noisy_model_input_and_timesteps * feat: update loss calculation to support 5d tensor * fix: update argument names in anima_train_utils to align with other archtectures * feat: simplify Anima training script and update empty caption handling * feat: support LoRA format without `net.` prefix * fix: update to work fp8_scaled option * feat: add regex-based learning rates and dimensions handling in create_network * fix: improve regex matching for module selection and learning rates in LoRANetwork * fix: update logging message for regex match in LoRANetwork * fix: keep latents 4D except DiT call * feat: enhance block swap functionality for inference and training in Anima model * feat: refactor Anima training script * feat: optimize VAE processing by adjusting tensor dimensions and data types * fix: wait all block trasfer before siwtching offloader mode * feat: update Anima training guide with new argument specifications and regex-based module selection. Thank you Claude! * feat: support LORA for Qwen3 * feat: update Anima SAI model spec metadata handling * fix: remove unused code * feat: split CFG processing in do_sample function to reduce memory usage * feat: add VAE chunking and caching options to reduce memory usage * feat: optimize RMSNorm forward method and remove unused torch_attention_op * Update library/strategy_anima.py Use torch.all instead of all. Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com> * Update library/safetensors_utils.py Fix duplicated new_key for concat_hook. Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com> * Update anima_minimal_inference.py Remove unused code. Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com> * Update anima_train.py Remove unused import. Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com> * Update library/anima_train_utils.py Remove unused import. Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com> * fix: review with Copilot * feat: add script to convert LoRA format to ComfyUI compatible format (WIP, not tested yet) * feat: add process_escape function to handle escape sequences in prompts * feat: enhance LoRA weight handling in model loading and add text encoder loading function * feat: improve ComfyUI conversion script with prefix constants and module name adjustments * feat: update caption dropout documentation to clarify cache regeneration requirement * feat: add clarification on learning rate adjustments * feat: add note on PyTorch version requirement to prevent NaN loss --------- Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com>	2026-02-13 08:15:06 +09:00
Kohya S	9144463f7b	Merge branch 'dev' into sd3	2026-02-13 08:14:21 +09:00
Duoong	1640e53392	Fix bug and optimization Lumina training	2026-02-12 22:52:28 +07:00
duongve13112002	e21a7736f8	Support Anima model (#2260 ) * Support Anima model * Update document and fix bug * Fix latent normlization * Fix typo * Fix cache embedding * fix typo in tests/test_anima_cache.py * Remove redundant argument apply_t5_attn_mask * Improving caching with argument caption_dropout_rate * Fix W&B logging bugs * Fix discrete_flow_shift default value	2026-02-08 10:18:55 +09:00
Kohya S.	8b5ce3e641	Merge pull request #2255 from cgcalatrava/fix-diffusers-unet-import Fix AttributeError for UNet2DConditionModel with newer diffusers versions	2026-01-20 07:50:04 +09:00
cgcalatrava	da07e4c617	Make UNet2DConditionModel import compatible with old and new diffusers versions	2026-01-19 20:53:00 +01:00
Kohya S.	966e9d7f6b	Merge pull request #2254 from kohya-ss/dev Merge the changes from the sd3 branch into main v0.10.0	2026-01-19 22:00:25 +09:00
Kohya S.	2a2760e702	Merge pull request #1374 from kohya-ss/sd3 support SD3	2026-01-19 21:50:22 +09:00
Kohya S.	b996440c5f	Doc update sd3 branch documentation (#2253 ) * doc: move sample prompt file documentation, and remove history for branch * doc: remove outdated FLUX.1 and SD3 training information from README * doc: update README and training documentation for clarity and structure	2026-01-19 21:38:46 +09:00
Kohya S.	a9af52692a	feat: add pyramid noise and noise offset options to generation script (#2252 ) * feat: add pyramid noise and noise offset options to generation script * fix: fix to work with SD1.5 models * doc: update to match with latest gen_img.py * doc: update README to clarify script capabilities and remove deprecated sections	2026-01-18 16:56:48 +09:00
Kohya S.	c6bc632ec6	fix: metadata dataset degradation and make it work (#2186 ) * fix: support dataset with metadata * feat: support another tagger model * fix: improve handling of image size and caption/tag processing in FineTuningDataset * fix: enhance metadata loading to support JSONL format in FineTuningDataset * feat: enhance image loading and processing in ImageLoadingPrepDataset with batch support and output options * fix: improve image path handling and memory management in dataset classes * Update finetune/tag_images_by_wd14_tagger.py Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com> * fix: add return type annotation for process_tag_replacement function and ensure tags are returned * feat: add artist category threshold for tagging * doc: add comment for clarification --------- Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com>	2026-01-18 15:17:07 +09:00
Kohya S.	f7f971f50d	Merge pull request #2251 from kohya-ss/fix-pytest-for-lumina fix(tests): add ip_noise_gamma args for MockArgs in pytest	2026-01-18 15:09:47 +09:00
Kohya S	c4be615f69	fix(tests): add ip_noise_gamma args for MockArgs in pytest	2026-01-18 15:05:57 +09:00
Kohya S.	e06e063970	Merge pull request #2225 from urlesistiana/sd3_lumina2_ts_fix fix: lumina 2 timesteps handling	2026-01-18 14:39:04 +09:00
Kohya S.	94e3dbebea	Merge pull request #2246 from kozistr/deps/pytorch-optimizer Bump `pytorch-optimizer` version to v3.9.0	2025-12-21 22:51:32 +09:00
kozistr	95a65b89a5	build(deps): bump pytorch-optimizer to v3.9.0	2025-12-21 15:53:47 +09:00
matthew	6f7498f959	make stable validation loss stable and deterministic for uncached flux latents	2025-11-12 15:11:03 -05:00
Kohya S.	a5a162044c	Merge pull request #2226 from kohya-ss/fix-hunyuan-image-batch-gen-error fix: error on batch generation closes #2209	2025-10-15 21:57:45 +09:00
Kohya S	a33cad714e	fix: error on batch generation closes #2209	2025-10-15 21:57:11 +09:00
urlesistiana	f7fc7ddda2	fix #2201 : lumina 2 timesteps handling	2025-10-13 16:08:28 +08:00
Kohya S.	5e366acda4	Merge pull request #2003 from laolongboy/sd3-dev Fix missing parameters in model conversion script	2025-10-01 21:03:12 +09:00
Kohya S	5462a6bb24	Merge branch 'dev' into sd3	2025-09-29 21:02:02 +09:00
Kohya S	63711390a0	Merge branch 'main' into dev	2025-09-29 20:56:07 +09:00
Kohya S.	206adb6438	Merge pull request #2216 from kohya-ss/fix-sdxl-textual-inversion-training-disable-mmap fix: disable_mmap_safetensors not defined in SDXL TI training	2025-09-29 20:55:02 +09:00
Kohya S	60bfa97b19	fix: disable_mmap_safetensors not defined in SDXL TI training	2025-09-29 20:52:48 +09:00
Kohya S.	f0c767e0f2	Merge pull request #2213 from kohya-ss/doc-hunyuan-image-training-text-encoder-cpu-note docs: enhance text encoder CPU usage instructions for HunyuanImage-2.…	2025-09-28 18:32:11 +09:00

1 2 3 4 5 ...

2506 Commits