jab editionPreviously on /sdg/: >>109720592 >Beginner UIEasyDiffusion: https://easydiffusion.github.ioSwarmUI: https://github.com/mcmonkeyprojects/SwarmUI>Advanced UIComfyUI: https://github.com/comfyanonymous/ComfyUIForge Classic: https://github.com/Haoming02/sd-webui-forge-classicStability Matrix: https://github.com/LykosAI/StabilityMatrix>Minimax H3https://huggingface.co/Comfy-Org/MiniMax-H3>Krea 2https://huggingface.co/krea/Krea-2-Rawhttps://huggingface.co/krea/Krea-2-Turbohttps://lumenastrum.github.io/clio-style-preview/gallery/>Z-Imagehttps://comfyanonymous.github.io/ComfyUI_examples/z_image>Flux.2 Dev/Kleinhttps://comfyanonymous.github.io/ComfyUI_examples/flux2>Chromahttps://comfyanonymous.github.io/ComfyUI_examples/chroma>Animahttps://huggingface.co/circlestone-labs/Anima>Models, LoRAs & upscalinghttps://civitai.comhttps://huggingface.cohttps://tungsten.runhttps://yodayo.com/modelshttps://www.diffusionarc.comhttps://miyukiai.comhttps://civitaiarchive.comhttps://civitasbay.orghttps://www.stablebay.orghttps://openmodeldb.info>Index of guides and other toolshttps://rentry.org/sdg-link>Related boards>>>/aco/csdg/>>>/b/degen>>>/d/ddg>>>/e/edg>>>/gif/vdg>>>/h/hdg>>>/tg/slop>>>/trash/sdg>>>/u/udg>>>/vp/napt>>>/vt/vtaiOP https://rentry.co/twkuk8tz
>mfw Resource news09/04/2026>lightx2v/Minimax-h3-Turbo · FL2V Turbo 4-step v1.2 (768p)https://huggingface.co/lightx2v/Minimax-h3-Turbo/discussions/52#6a9a890895a616c64799324f>ComfyUI NVIDIA DLSS 5 Visual Enhancerhttps://github.com/Konohamaru04/ComfyUI-NVIDIA-DLSS-Frame-Interpolation>Viggle-Animate: Character Replacement in Video from a Single Repainted Frame https://huggingface.co/Viggle/Viggle-Animate>DSAQuant: Denoising-Stage-Aligned Quantization-Aware Training for Video Generationhttps://robbyant-research.github.io/DSAQuant>Do Video Generators Track the World Across Segments? A Benchmark and Method for World-State Reasoning in Video Continuationhttps://github.com/AMAP-ML/StateAgent>FlashRender: Few-Step Generative Rendering via Camera-Controlled Video MeanFlowhttps://byeongjun-park.github.io/FlashRender>LLaDA-Image: Building Strong Image Generators with Fully Open Training Recipeshttps://huggingface.co/inclusionAI/LLaDA-Image>ComfyUI-VDN-H3: v1.4.0 — Faster streaming, VRAM-aware buffer retention Latesthttps://github.com/Saganaki22/ComfyUI-VDN-H3/releases/tag/v1.4.0>AetherScale for ComfyUI: GPU-native NVIDIA video enhancementhttps://github.com/vizart-vj/ComfyUI-AetherScale09/03/2026>lightx2v Minimax-h3-Turbo ref2v Lora v1.0https://huggingface.co/lightx2v/Minimax-h3-Turbo/blob/main/minimax_h3_ref2v_turbo_8step_v1.0_768p_comfyui_bf16.safetensors>DreamX-Creator 1.0: Model Weightshttps://huggingface.co/GD-ML/DreamX-Creator>ComfyUI MiniMaxH3 CLIPCached: disk cache for MiniMax H3 conditioninghttps://github.com/Mu5hr00moO/ComfyUI-MiniMaxH3-CLIPCached>VDN-Minimax-H3 (VDN-H3): Hybrid Attention to Speed Up Video Models with Near-Lossless Qualityhttps://huggingface.co/OpenVDN/vdn-minimax-h3>SolarWM: Open Data and Scalable Training for Long-Horizon Video World Modelshttps://junchao-cs.github.io/SolarWM-Web>TAME: Temporal-Aware Mixture-of-Experts for Text-Video Retrievalhttps://github.com/sejong-rcv/TAME
>mfw Research news09/04/2026>OctWorld: Long-Range World-Consistent Video Generation with Octree-Based 3D Mappinghttps://maxtirerror.github.io/octworldpage>One Editor, Many Edits: A Unified Training-Free Framework for Diverse Video Editinghttps://plan-lab.github.io/editvid>Building Pretraining Data for World Models: An Unreal Engine-Based Pipeline for Action-Conditioned Video Generationhttps://arxiv.org/abs/2609.03557>ToPO: Token-Conditioned Preference Routing for Attention-Based Latent Diffusion Modelshttps://arxiv.org/abs/2609.03688>SVG-Score: Human-Aligned Evaluation of Text-to-SVG Generationhttps://arxiv.org/abs/2609.03806>Select, Compress, Reinvest: A Controlled Study of Visual-Token Allocation in Long-Video MLLMshttps://arxiv.org/abs/2609.03820>Persistent Identity Preservation in Generative Image Models: A Benchmark and Evaluation Systemhttps://arxiv.org/abs/2609.04151>LLaDA-Image: Building Strong Image Generators with Fully Open Training Recipeshttps://arxiv.org/abs/2609.03796>SPARK: Input-Conditioned Sparse Activation Modulation for Frozen DiT-based Super-Resolutionhttps://arxiv.org/abs/2609.03813>Seeing Before Synthesizing: VLM-Guided Transition Event Discovery for Weakly-Supervised Dense Video Captioninghttps://arxiv.org/abs/2609.04183>When Do Frozen VLMs Respond to Image-Free Object-Token Edits? An Answer-Key-Free Protocol and What It Revealshttps://arxiv.org/abs/2609.03429>Who Speaks for the Pruned? Visual Token Pruning as Coverage Optimizationhttps://arxiv.org/abs/2609.03158>The impact of phase information for few-shot fine-grained image classificationhttps://arxiv.org/abs/2609.03829>The Shape of Time: Video-Token Contrast for Temporal Understanding in VideoLMshttps://arxiv.org/abs/2609.04110>Editable Visual Designhttps://arxiv.org/abs/2609.04034
>>109731933>ComfyUI NVIDIA DLSS 5 Visual Enhancer>Requirements>Windowscunts
>>109731967its NVIDIA's fault. they hate freedom
>>109732067it's just a matter of timeare you using it? how are you getting those details? you used to struggle with ships right
>>109731967they ported the slopengine to comfy? lmao
>>109732106it was inevitable
>>109732120some guy on twitter has been running 10x feedback loops of dlss5 on persona
https://x.com/hatty__/status/2094197695527170483no idea what game this is but lol
why are the gens in this thread always to soulless and disgusting looking.
>>109732092its nothing fancy. just krea2 + a goofy RTX upscaler pass. krea2 is just that good
>>109732156oh hellohappy friday>>109732198show us your superior gen. we'll shut down sdg for good if you trump us
>>109732140>>109732156that's nutty>>109732209yah k2+rtx upscale is my go to combopush it to 4x or 8x, it shouldnt take any longer
>>109732214yo.krea2 stuck on woke harley what a shame
>>109732324chub harley a qt
>>109732357i'll excise the birds of prey slop even if it kills me. just putting that name in makes it do margo robbie era garbage.
it just can't fucking help itself
>>109732370>>109732404i wonder if adding "btas" or "alan burnett style" would help
>>109732410this is btas style rewrite, but "harley quinn" has been completely subsumed by the margo shit. might as well give up at this point and just go woke (already broke)
>>109732429gross
lucky get
>>109732463just say hair covered or hidden then lel
>>109732485somethin shifted in the latents. praise be
>>109732490based bible verse saved the day
broke the spell
>>109732685they cant all be bangers
OMG I FORGOT TO SAVE SO IT WAS USING __std/xl/girls/base_harley_quinn_tas_zit_curvy_nb2__LITERALLY WITHOUT EXPANSION! FML
>>109732700heh
>>109732700>>109732733she still has the nu-harley face shape thoqt harley had a bit of a rounder/wider face i think
>>109732749yeah, i tried to clean room it... i'd need a lora to make it work perfectly.
too much cheese
name my new kawaii metal album
gn all
>>109733101night
>>109733265hey! ur late
>>109731933>>ComfyUI NVIDIA DLSS 5 Visual Enhancer>https://github.com/Konohamaru04/ComfyUI-NVIDIA-DLSS-Frame-InterpolationInteresting.. Which "dlss_model_preset" is best [J,K,L,M]?
blah blah blah
i miss schizo anon
batmanhttps://suno.com/s/9DrJWWt0Ufajrjf9https://youtu.be/DxMJ0I_JrTQ
Is there a reliable and lightweight way (running this on a potato) of getting t2i to correctly output 2 specific characters interacting without mixing their features? Using comfyui and a fine tuned illustrious model.The first thing I naively did was just have a main prompt that had "2girls" and the generic quality keywords for that model, and then using 2 prompts that just had "[character keywords], [pose]" connected to a "Set Area with Percentage", combined in a "combine conditioning" node chain, giving each prompt half of the latent image (to start, I made the 2 areas not overlap and the pose was independent from the other character), but that somehow still managed to mix the character features, even though the "set area" strength was indeed set to 1.What does work reasonably well is making a 3d render of said characters (I have a few 3d models) in a pose I want and just doing i2i with a single prompt mentioning all the characters involved and the specific pose and shit, with a 0.5 denoising value. Apparently that's enough clues to prevent feature mixing. Since it's a single pass i2i it doesn't take long on my potato sans the time I actually spend manipulating 3d models.But obviously i2i takes manual setup while t2i would be preferable for less manual input.Should I just generate a base image and do inpainting for this sorta thing? If so, do I like generate the characters bald or something so I don't have to fiddle with masking out hair too much and such?
>>109734328>output 2 specific characters interacting without mixing their featuresThat's pretty much impossible.
>>109734514So my only options are inpainting and i2i at moderate denoise from a picture with all the elements already in place? I guess it's understandable, it can't read minds. There's some kinda auto-masking node that generates an image mask based on what you tell it to look for, right? I should probably use that one for speeding up inpainting.Thanks.
>>109734608I'm being facetious. All you need is to identify each character and assign specific looks, actions, and properties to each.Like pic related.