frens editionPreviously on /sdg/: >>109458081 >Beginner UIEasyDiffusion: https://easydiffusion.github.ioSwarmUI: https://github.com/mcmonkeyprojects/SwarmUI>Advanced UIComfyUI: https://github.com/comfyanonymous/ComfyUIForge Classic: https://github.com/Haoming02/sd-webui-forge-classicStability Matrix: https://github.com/LykosAI/StabilityMatrix>Z-Imagehttps://comfyanonymous.github.io/ComfyUI_examples/z_imagehttps://huggingface.co/Tongyi-MAI/Z-Imagehttps://huggingface.co/Tongyi-MAI/Z-Image-Turbo>Flux.2 Dev/Kleinhttps://comfyanonymous.github.io/ComfyUI_examples/flux2https://huggingface.co/black-forest-labs/FLUX.2-devhttps://huggingface.co/black-forest-labs/FLUX.2-klein-4Bhttps://huggingface.co/black-forest-labs/FLUX.2-klein-9B>Chromahttps://comfyanonymous.github.io/ComfyUI_examples/chromahttps://huggingface.co/lodestones/Chroma1-HDhttps://huggingface.co/silveroxides/Chroma-GGUF>Animahttps://huggingface.co/circlestone-labs/Anima>Qwen Image & Edithttps://docs.comfy.org/tutorials/image/qwen/qwen-imagehttps://huggingface.co/Qwen/Qwen-Image>Text & image to video - Wan 2.2https://docs.comfy.org/tutorials/video/wan/wan2_2>Models, LoRAs & upscalinghttps://civitai.comhttps://huggingface.cohttps://tungsten.runhttps://yodayo.com/modelshttps://www.diffusionarc.comhttps://miyukiai.comhttps://civitaiarchive.comhttps://civitasbay.orghttps://www.stablebay.orghttps://openmodeldb.info>Index of guides and other toolshttps://rentry.org/sdg-link>Related boards>>>/aco/csdg/>>>/b/degen>>>/d/ddg>>>/e/edg>>>/gif/vdg>>>/h/hdg>>>/tg/slop>>>/trash/sdg>>>/u/udg>>>/vp/napt>>>/vt/vtaiOP https://rentry.co/twkuk8tz
>mfw Resource news08/06/2026>Flash-VAED: Plug-and-Play VAE Decoders for Efficient Video Generationhttps://github.com/Aoko955/Flash-VAED>(preview) MiniMax-H3 Turbo LoRA — 4-step audio-video generation https://huggingface.co/larryvrh/MiniMax-H3-Turbo-Lora>MiniMax-H3 Turbo 4-Step — ComfyUI Pruned-Model LoRAs https://huggingface.co/drbaph/MiniMax-H3-Turbo-Lora-ComfyUI>ComfyUI-H3-Multishothttps://github.com/jlucasmcrell/ComfyUI-H3-Multishot>Krea2 Turbo – OpenPose ControlNet LoRA https://huggingface.co/thedeoxen/Krea-2-pose-controlnet>MiniMax H3 experimental Int8 convrot VAEhttps://huggingface.co/Kijai/MiniMax-H3-experimental>UniWorld-View: Large-Baseline View Synthesis via Video Diffusion Modelshttps://zhouhyocean.github.io/uniworld-view>OmniVR: Joint Video-Audio Conditional Generation for Restoring Degraded Historical Filmshttps://xin1u.github.io/OminiVR_PAGE>DIVE: Dynamic Iterative Visual Evidence Construction for Efficient Vision-Language Modelshttps://github.com/Zhong-Chenchen/DIVE.git>Multi-View Face and Gesture Animation with Dynamic Gaussianshttps://dfki-av.github.io/MVFGA>EmpaAva: An Open-source Agentic 3D-Avatar Empathetic Live Chatbothttps://empaava.top>Context-Anchored Tile Refinehttps://github.com/Blakeem/ComfyUI-ContextAnchoredTileRefine>ComfyUI Video Tilerhttps://github.com/maDcaDDie2000/comfyui-video-tiler08/05/2026>Inline Studio v1.2.62 - Minimax H3 Lora training still onlyhttps://github.com/inlineresearch/Inline-Studio/releases/tag/v1.2.62>Qwen3-VL-32B-Instruct-MiniMax-H3-GGUFhttps://huggingface.co/nif0/Qwen3-VL-32B-Instruct-MiniMax-H3-GGUF>Qwen3-VL-32B-Instruct-ultra-uncensored-heretic-H3-GGUFhttps://huggingface.co/nif0/Qwen3-VL-32B-Instruct-ultra-uncensored-heretic-H3-GGUF>MiniMax-H3-TAE: 2D tine VAE for MiniMax-H3https://huggingface.co/Kijai/MiniMax-H3-TAE>SPADE: An Input-Adaptive Sparse Attention Engine for Fast Video Diffusion Models Inferencehttps://github.com/6somehow/DAC-SPADE
>mfw Research news08/06/2026>When Diffusion Models Forget Who You Are: Identity Preservation in Face Inpainting under Large Occlusionshttps://arxiv.org/abs/2608.04820>HelloWorld: Enabling Socially Interactive Characters in Video World Modelshttps://arxiv.org/abs/2608.05070>OmniEdit-Bench: A Comprehensive Benchmark for Instruction-based Video Editinghttps://arxiv.org/abs/2608.05049>ContextMaster: Interactive Multi-Shot Video Creation via Fixed-Budget Sparse Context Routinghttps://guoxu1233.github.io/ContextMaster>STEP-OPD: Rethinking Output Targets and Internal Dynamics in On-Policy Distillation for Diffusion Modelshttps://arxiv.org/abs/2608.04887>Simile Understanding in Text-to-Image Models: An Evaluation Frameworkhttps://arxiv.org/abs/2608.04750>ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generationhttps://arxiv.org/abs/2608.04436>CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Modelshttps://arxiv.org/abs/2608.04302>Rethinking Pixel Mean Flows via Interval Denoiserhttps://arxiv.org/abs/2608.04818>Persistent Object Narratives for Token-Efficient Video Language Modelshttps://arxiv.org/abs/2608.04866>Poly-OPD: Heterogeneous Multi-Teacher On-Policy Distillation for Capability-Selectable Flow Modelshttps://arxiv.org/abs/2608.04349>Not All Redundant Tokens Are Alike: Analyzing Visual Token Pruning through Token Roleshttps://arxiv.org/abs/2608.04483>Beyond Global Routing Aggregation: Phase-Aware Expert Merging for MoE Vision-Language Modelshttps://arxiv.org/abs/2608.04454>When does training on downscaled images yield the same gradients?https://arxiv.org/abs/2608.04448>Unleashing the Potential of Vision-Language Models for Generalizable AI-Generated Image Detectionhttps://arxiv.org/abs/2608.04935
>shithole general
Good morning.
>what timmy gonna do?
Holy slop.
>gm
>>109483789Good morning anon.
Gm! Bot status?
>>109484048you're heck'in cute and valid lumi
>>109483989Gm
>>109484118Gm
>>109485030what is the style on that?
>>109485047no particular style, just>action cartoon illustration with clean studio animation lines, bold color blockingI figure these all are mostly seed dependent and wouldn't be particularly reproducible
>>109485088it's not "enhanced" by llm?
>>109485097no, I had to turn off prompt enhancement because it kept getting confused and refusing to process inputs. thats what kept giving me the stock images of people talking
booba physics lolhttps://files.catbox.moe/fy1wlp.webm
>>109485118shes so happy
>>109485109lelwhat node are you using for llm? depending on the node they can be finicky>>109485118lel nice
>>109485187>what node are you using for llm?apparently its just called 'generate text'. it doesn't even have a model input, so I have no clue what it is. idk where I got it from either. it works ok when it works tho
>>109485203well there's your problemget https://github.com/silveroxides/ComfyUI-UtilsCollectionand useUC_TextEncodeKrea2SystemPrompt or it may be called 'system prompt encode' nowhook up your wildcard output to prompt, the system prompt, and the clip
>>109485232cool, I'll try this outwhat is the thinking content field for?
>>109485232that uses the clip (krea uses qwen) as an "llm" so you dont need to load anotheri'm sure your node does something similar but it's stupid lel>>109485245not sure. also check out the presets (unified presets) from that node collection
>>109485245maybe you can supply up-front thinking for the 4B lobotomite model lmao
i dont use those text encoder nodes personally (i use gemma4 via llama.cpp node) but you can do some crazy shit with those UtilsCollection nodessame guy that started the int8 convrot stuff and all the chroma tricks
>>109485305>i dont use those text encoder nodes personallyits kinda nice to tie together all the wildcards into something more natural. its also good for resolving conflicts that the wildcards might have tossed out. I don't use it much tho desu
>>109485305i just spend three-hundredths of a cent for deepseek-v4-flash, although i hear that gravy train is coming to an end :(
me in the back>well this night's gonna get weird>>109485355>>109485366that's what i use gemma for lel
>>109485369~3,401 gens/$1 ain't too shabby. i could use gemma but the offloading takes longer than the api does.
i just meant i dont use that specific text encoder node with the "built-in" llmanyway, time to crashgn all
>>109485366I guess it never was as cheap as advertised, they were just subsidizing costs to bait people in. same as openai/anthropic>>109485369>i use gemmayou run it on a second card or something?
>>109485398no, offloading and reloading
>>109485395gn
>>109485398i think it might just be if you use it directly from alibaba though, openrouter might stay cheap. actually 5.6-luna is on sale for less than deepseek rn i might just use that, it's faster too. >>109485395gn
>>109485417luna has been very good for daily usebut i guess gpt6 is a week a way? early reports say better than mythos (or fable, I forget)
>>109485513i've been using terra for agent shit, but i don't do a lot of huge codebase shit so my $100 sub stretches an awful long way, plus they reset usage all the fucking time
damn i need to prod it to be more creative with these paper craft things, although this is kind of neat. next one is a few minutes outhttps://files.catbox.moe/w20ml1.mp4
>>109485622maybe "stopmotion animation" would make it do more articulation and movement
>>109485644yeah i have to tune this thing for h3, it doesn't quite get ithttps://files.catbox.moe/yr91i3.mp4