[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
Settings Mobile Home
/g/ - Technology

Name
Options
Comment
Verification
4chan Pass users can bypass this verification. [Learn More] [Login]
File
  • Please read the Rules and FAQ before posting.
  • You may highlight syntax and preserve whitespace by using [code] tags.

08/21/20New boards added: /vrpg/, /vmg/, /vst/ and /vm/
05/04/17New trial board added: /bant/ - International/Random
10/04/16New board for 4chan Pass users: /vip/ - Very Important Posts
[Hide] [Show All]


[Advertise on 4chan]


File: 1788467048429888.png (1.48 MB, 896x1152)
1.48 MB PNG
jab edition

Previously on /sdg/: >>109720592

>Beginner UI
EasyDiffusion: https://easydiffusion.github.io
SwarmUI: https://github.com/mcmonkeyprojects/SwarmUI

>Advanced UI
ComfyUI: https://github.com/comfyanonymous/ComfyUI
Forge Classic: https://github.com/Haoming02/sd-webui-forge-classic
Stability Matrix: https://github.com/LykosAI/StabilityMatrix

>Minimax H3
https://huggingface.co/Comfy-Org/MiniMax-H3

>Krea 2
https://huggingface.co/krea/Krea-2-Raw
https://huggingface.co/krea/Krea-2-Turbo
https://lumenastrum.github.io/clio-style-preview/gallery/

>Z-Image
https://comfyanonymous.github.io/ComfyUI_examples/z_image

>Flux.2 Dev/Klein
https://comfyanonymous.github.io/ComfyUI_examples/flux2

>Chroma
https://comfyanonymous.github.io/ComfyUI_examples/chroma

>Anima
https://huggingface.co/circlestone-labs/Anima

>Models, LoRAs & upscaling
https://civitai.com
https://huggingface.co
https://tungsten.run
https://yodayo.com/models
https://www.diffusionarc.com
https://miyukiai.com
https://civitaiarchive.com
https://civitasbay.org
https://www.stablebay.org
https://openmodeldb.info

>Index of guides and other tools
https://rentry.org/sdg-link

>Related boards
>>>/aco/csdg/
>>>/b/degen
>>>/d/ddg
>>>/e/edg
>>>/gif/vdg
>>>/h/hdg
>>>/tg/slop
>>>/trash/sdg
>>>/u/udg
>>>/vp/napt
>>>/vt/vtai

OP https://rentry.co/twkuk8tz
>>
>mfw Resource news

09/04/2026

>lightx2v/Minimax-h3-Turbo · FL2V Turbo 4-step v1.2 (768p)
https://huggingface.co/lightx2v/Minimax-h3-Turbo/discussions/52#6a9a890895a616c64799324f

>ComfyUI NVIDIA DLSS 5 Visual Enhancer
https://github.com/Konohamaru04/ComfyUI-NVIDIA-DLSS-Frame-Interpolation

>Viggle-Animate: Character Replacement in Video from a Single Repainted Frame
https://huggingface.co/Viggle/Viggle-Animate

>DSAQuant: Denoising-Stage-Aligned Quantization-Aware Training for Video Generation
https://robbyant-research.github.io/DSAQuant

>Do Video Generators Track the World Across Segments? A Benchmark and Method for World-State Reasoning in Video Continuation
https://github.com/AMAP-ML/StateAgent

>FlashRender: Few-Step Generative Rendering via Camera-Controlled Video MeanFlow
https://byeongjun-park.github.io/FlashRender

>LLaDA-Image: Building Strong Image Generators with Fully Open Training Recipes
https://huggingface.co/inclusionAI/LLaDA-Image

>ComfyUI-VDN-H3: v1.4.0 — Faster streaming, VRAM-aware buffer retention Latest
https://github.com/Saganaki22/ComfyUI-VDN-H3/releases/tag/v1.4.0

>AetherScale for ComfyUI: GPU-native NVIDIA video enhancement
https://github.com/vizart-vj/ComfyUI-AetherScale

09/03/2026

>lightx2v Minimax-h3-Turbo ref2v Lora v1.0
https://huggingface.co/lightx2v/Minimax-h3-Turbo/blob/main/minimax_h3_ref2v_turbo_8step_v1.0_768p_comfyui_bf16.safetensors

>DreamX-Creator 1.0: Model Weights
https://huggingface.co/GD-ML/DreamX-Creator

>ComfyUI MiniMaxH3 CLIPCached: disk cache for MiniMax H3 conditioning
https://github.com/Mu5hr00moO/ComfyUI-MiniMaxH3-CLIPCached

>VDN-Minimax-H3 (VDN-H3): Hybrid Attention to Speed Up Video Models with Near-Lossless Quality
https://huggingface.co/OpenVDN/vdn-minimax-h3

>SolarWM: Open Data and Scalable Training for Long-Horizon Video World Models
https://junchao-cs.github.io/SolarWM-Web

>TAME: Temporal-Aware Mixture-of-Experts for Text-Video Retrieval
https://github.com/sejong-rcv/TAME
>>
>mfw Research news

09/04/2026

>OctWorld: Long-Range World-Consistent Video Generation with Octree-Based 3D Mapping
https://maxtirerror.github.io/octworldpage

>One Editor, Many Edits: A Unified Training-Free Framework for Diverse Video Editing
https://plan-lab.github.io/editvid

>Building Pretraining Data for World Models: An Unreal Engine-Based Pipeline for Action-Conditioned Video Generation
https://arxiv.org/abs/2609.03557

>ToPO: Token-Conditioned Preference Routing for Attention-Based Latent Diffusion Models
https://arxiv.org/abs/2609.03688

>SVG-Score: Human-Aligned Evaluation of Text-to-SVG Generation
https://arxiv.org/abs/2609.03806

>Select, Compress, Reinvest: A Controlled Study of Visual-Token Allocation in Long-Video MLLMs
https://arxiv.org/abs/2609.03820

>Persistent Identity Preservation in Generative Image Models: A Benchmark and Evaluation System
https://arxiv.org/abs/2609.04151

>LLaDA-Image: Building Strong Image Generators with Fully Open Training Recipes
https://arxiv.org/abs/2609.03796

>SPARK: Input-Conditioned Sparse Activation Modulation for Frozen DiT-based Super-Resolution
https://arxiv.org/abs/2609.03813

>Seeing Before Synthesizing: VLM-Guided Transition Event Discovery for Weakly-Supervised Dense Video Captioning
https://arxiv.org/abs/2609.04183

>When Do Frozen VLMs Respond to Image-Free Object-Token Edits? An Answer-Key-Free Protocol and What It Reveals
https://arxiv.org/abs/2609.03429

>Who Speaks for the Pruned? Visual Token Pruning as Coverage Optimization
https://arxiv.org/abs/2609.03158

>The impact of phase information for few-shot fine-grained image classification
https://arxiv.org/abs/2609.03829

>The Shape of Time: Video-Token Contrast for Temporal Understanding in VideoLMs
https://arxiv.org/abs/2609.04110

>Editable Visual Design
https://arxiv.org/abs/2609.04034
>>
>>109731933
>ComfyUI NVIDIA DLSS 5 Visual Enhancer
>Requirements
>Windows
cunts
>>
>>
>>
File: debo_sw_k2_00041_.png (3.33 MB, 1664x1280)
3.33 MB PNG
>>109731967
its NVIDIA's fault. they hate freedom
>>
>>109732067
it's just a matter of time
are you using it? how are you getting those details? you used to struggle with ships right
>>
>>109731967
they ported the slopengine to comfy? lmao
>>
>>109732106
it was inevitable
>>
File: file.png (3.32 MB, 2560x1440)
3.32 MB PNG
>>109732120
some guy on twitter has been running 10x feedback loops of dlss5 on persona
>>
>>
https://x.com/hatty__/status/2094197695527170483
no idea what game this is but lol
>>
why are the gens in this thread always to soulless and disgusting looking.
>>
File: debo_sw_k2_00045_.png (3.52 MB, 1664x1280)
3.52 MB PNG
>>109732092
its nothing fancy. just krea2 + a goofy RTX upscaler pass. krea2 is just that good
>>
File: debo_sw_k2_00047_.png (3.23 MB, 1664x1280)
3.23 MB PNG
>>109732156
oh hello
happy friday

>>109732198
show us your superior gen. we'll shut down sdg for good if you trump us
>>
>>109732140
>>109732156
that's nutty

>>109732209
yah k2+rtx upscale is my go to combo
push it to 4x or 8x, it shouldnt take any longer
>>
>>
>>
>>
>>109732214
yo.
krea2 stuck on woke harley what a shame
>>
>>109732324
chub harley a qt
>>
>>
>>109732357
i'll excise the birds of prey slop even if it kills me. just putting that name in makes it do margo robbie era garbage.
>>
it just can't fucking help itself
>>
>>109732370
>>109732404
i wonder if adding "btas" or "alan burnett style" would help
>>
>>109732410
this is btas style rewrite, but "harley quinn" has been completely subsumed by the margo shit. might as well give up at this point and just go woke (already broke)
>>
>>109732429
gross
>>
lucky get
>>
>>109732463
just say hair covered or hidden then lel
>>
>>109732485
somethin shifted in the latents. praise be
>>
>>109732490
based bible verse saved the day
>>
>>
>>
>>
File: debo_sw_k2_00066_.png (3.33 MB, 1664x1280)
3.33 MB PNG
>>
>>
>>
broke the spell
>>
>>109732685
they cant all be bangers
>>
>>
OMG I FORGOT TO SAVE SO IT WAS USING
__std/xl/girls/base_harley_quinn_tas_zit_curvy_nb2__
LITERALLY WITHOUT EXPANSION!

FML
>>
>>109732700
heh
>>
>>
>>109732700
>>109732733
she still has the nu-harley face shape tho
qt harley had a bit of a rounder/wider face i think
>>
>>
>>
>>109732749
yeah, i tried to clean room it... i'd need a lora to make it work perfectly.
>>
>>
>>
>>
>>
>>
>>
>>
File: debo_sw_k2_00071_.png (3.41 MB, 1664x1280)
3.41 MB PNG
>>
>>
too much cheese
>>
>>
name my new kawaii metal album
>>
>>
>>
>>
gn all
>>
>>109733101
night
>>
File: 039195632030145253.png (3.66 MB, 1212x1804)
3.66 MB PNG
>>
File: kenner_noise0_scale.png (3.84 MB, 1178x1766)
3.84 MB PNG
>>
>>109733265
hey! ur late
>>
File: 🐸00000_78759_.png (3.23 MB, 1145x1421)
3.23 MB PNG
>>109731933
>>ComfyUI NVIDIA DLSS 5 Visual Enhancer
>https://github.com/Konohamaru04/ComfyUI-NVIDIA-DLSS-Frame-Interpolation
Interesting.. Which "dlss_model_preset" is best [J,K,L,M]?
>>
blah blah blah
>>
i miss schizo anon
>>
File: batman.jpg (187 KB, 1920x1080)
187 KB JPG
batman
https://suno.com/s/9DrJWWt0Ufajrjf9
https://youtu.be/DxMJ0I_JrTQ
>>
>>
Is there a reliable and lightweight way (running this on a potato) of getting t2i to correctly output 2 specific characters interacting without mixing their features? Using comfyui and a fine tuned illustrious model.

The first thing I naively did was just have a main prompt that had "2girls" and the generic quality keywords for that model, and then using 2 prompts that just had "[character keywords], [pose]" connected to a "Set Area with Percentage", combined in a "combine conditioning" node chain, giving each prompt half of the latent image (to start, I made the 2 areas not overlap and the pose was independent from the other character), but that somehow still managed to mix the character features, even though the "set area" strength was indeed set to 1.

What does work reasonably well is making a 3d render of said characters (I have a few 3d models) in a pose I want and just doing i2i with a single prompt mentioning all the characters involved and the specific pose and shit, with a 0.5 denoising value. Apparently that's enough clues to prevent feature mixing. Since it's a single pass i2i it doesn't take long on my potato sans the time I actually spend manipulating 3d models.

But obviously i2i takes manual setup while t2i would be preferable for less manual input.
Should I just generate a base image and do inpainting for this sorta thing? If so, do I like generate the characters bald or something so I don't have to fiddle with masking out hair too much and such?
>>
>>109734328
>output 2 specific characters interacting without mixing their features

That's pretty much impossible.
>>
>>109734514
So my only options are inpainting and i2i at moderate denoise from a picture with all the elements already in place? I guess it's understandable, it can't read minds. There's some kinda auto-masking node that generates an image mask based on what you tell it to look for, right? I should probably use that one for speeding up inpainting.
Thanks.
>>
>>109734608
I'm being facetious. All you need is to identify each character and assign specific looks, actions, and properties to each.
Like pic related.
>>
File: comfyui_00580_.png (1.49 MB, 1152x896)
1.49 MB PNG



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.