Discussion and Development of Local Image, Video, and Audio ModelsPrevious: >>109902994https://rentry.org/ldg-lazy-getting-started-guide>UIComfyUI: https://github.com/comfyanonymous/ComfyUISwarmUI: https://github.com/mcmonkeyprojects/SwarmUISDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineageWan2GP: https://github.com/deepbeepmeep/Wan2GPNeural-Pixel (sd.cpp): https://github.com/Luiz-Alcantara/Neural-Pixel>Checkpoints, LoRAs, & Upscalershttps://huggingface.co/modelshttps://huggingbay.xyzhttps://civitai.comhttps://civitaiarchive.comhttps://openmodeldb.info>Tuninghttps://github.com/spacepxl/demystifying-sd-finetuninghttps://github.com/ostris/ai-toolkithttps://github.com/Nerogar/OneTrainerhttps://github.com/tdrussell/diffusion-pipehttps://github.com/kohya-ss/sd-scriptshttps://github.com/kohya-ss/musubi-tuner>Qwen Image 2.1https://huggingface.co/Qwen/Qwen-Image-2.1>Minimax H3https://huggingface.co/Comfy-Org/MiniMax-H3https://neta.art/use-cases/en/h3-1000-prompt-list>Animahttps://huggingface.co/circlestone-labs/Animahttps://animastyles.thetacursed.comhttps://tagexplorer.github.io/https://animadex.net>Krea 2https://huggingface.co/krea/Krea-2-Rawhttps://huggingface.co/krea/Krea-2-Turbohttps://lumenastrum.github.io/clio-style-preview/gallery/>Kleinhttps://huggingface.co/collections/black-forest-labs/flux2>MiscShare Metadata: https://catbox.moe | https://litterbox.catbox.moe/Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusionArchive: https://rentry.org/sdg-linkCollage: https://rentry.org/neo_collage>Neighbors>>>/aco/csdg>>>/b/degen>>>/gif/vdg>>>/d/ddg>>>/e/edg>>>/h/hdg>>>/trash/slop>>>/vt/vtai>>>/u/udg>Local Text>>>/g/lmg>Maintain Thread Qualityhttps://rentry.org/debohttps://rentry.org/animanon
it's so over
>>109908496Damn, still baking schizo rentries. Anons need to try harder to get the schizo to kill herself
>>109908496Yuck
>mfw Resource news09/25/2026>Fizgig H3 Tweaks: Training-free tweaks for MiniMax H3 in ComfyUIhttps://github.com/shootthesound/ComfyUI-Fizgig-H3-Tweaks>Krea 2 inpaint edithttps://huggingface.co/Cierpliwy/krea2-inpaint-edit>Pruna-Qwen-Image-2.1: Few-step LoRA adapters for Qwen-Image-2.1https://huggingface.co/PrunaAI/Pruna-Qwen-Image-2.1>AV-GRPO: Modality-Anchored Decoupling Diffusion Reinforcement Learning for Joint Audio-Video Generationhttps://github.com/zhiyuxu03/AV-GRPO>Kijai Qwen Image 2.1 Fun_controlnet_union (BF16/int8_convrot)https://huggingface.co/Kijai/QwenImage_experimental/tree/main/model_patches09/24/2026>Making the MiniMax H3 Video VAE 2x Fasterhttps://blog.comfy.org/p/making-the-minimax-h3-video-vae-2x>Qwen-Image-2.1 Text Encoder (Heretic) — GGUF · FP8 · bf16https://huggingface.co/pottokao/Qwen-Image-2.1-Text-Encoder-Heretic-GGUF>unsloth/Qwen-Image-2.1-GGUFhttps://huggingface.co/unsloth/Qwen-Image-2.1-GGUF>Ming-Image-0.1-Design GGUFhttps://huggingface.co/realrebelai/Ming-Image_GGUFs>MiMo-V2.6 series: Frontier intelligence, all the modalities, built in publichttps://mimo.xiaomi.com/mimo-v2-6>ComfyUI-Qwen-Image-2.1-PromptEnhancer-MTPhttps://github.com/mozophe/ComfyUI-Qwen-Image-2.1-PromptEnhancer-MTP>Qwen-Image-2.1-Fun-Controlnet-Unionhttps://huggingface.co/alibaba-pai/Qwen-Image-2.1-Fun-Controlnet-Union>Latent evolving World Action Modelhttps://github.com/XuejiFang/LeWAM>Prompt Studio: Multimodal prompt studio for ComfyUIhttps://github.com/tngklp/ComfyUI-Prompt-Studio09/23/2026>Qwen-Image-2.1-viggle-turbo — v0.2 (preview) https://huggingface.co/Viggle/Qwen-Image-2.1-viggle-turbo>Ming-Image-0.1-Design: 6B text-to-image model for UI, infographics, posters, and other text-rich visual designshttps://huggingface.co/inclusionAI/Ming-Image-0.1-Design>MiniMax-H3-Fun-Controlnet-Union-2.0 https://huggingface.co/alibaba-pai/MiniMax-H3-Fun-Controlnet-Union-2.0
Blessed thread of frenship
>mfw Research news09/25/2026>ViRDM: Taming Representation Distribution Matching for Few-Step Causal Video Generationhttps://arxiv.org/abs/2609.28923>WanPE: Towards Cinematic Prompt Enhancement for Modern Text-to-Video Generationhttps://arxiv.org/abs/2609.30221>ComplexSync: High-Fidelity and Real-Time Lip Sync in Complex Scenarioshttps://arxiv.org/abs/2609.29225>SALI: Shot-Aware Late Interaction for Cross-Shot Relation Matching in Text-to-Video Retrieval using Film-Grammar Knowledgehttps://arxiv.org/abs/2609.29721>OmniFabric: Coherent UV Space Texture Synthesis for 3D Garment Reconstructionhttps://humansensinglab.github.io/OmniFabric>CARE: Condition-Aware Representation Regularization for Diffusion Modelshttps://arxiv.org/abs/2609.28561>EIB-Net: Entropy-Guided Information Bottleneck for Generalizable AI-Generated Image Detectionhttps://arxiv.org/abs/2609.29064>Accelerating Video Diffusion via Training-Free Trajectory Routinghttps://arxiv.org/abs/2609.30096>TOLA: Text-aware One-Step Latent Adaptation for Diffusion-based Text Image Super-Resolutionhttps://arxiv.org/abs/2609.29240>Spectral Amplitude Purification in Distribution Matching for Diffusion Distillationhttps://arxiv.org/abs/2609.29116>Domain Recentering and Confidence-Weighted Prior Calibration for Vision-Language Modelshttps://arxiv.org/abs/2609.29358>Training Object Permanence in World Modelshttps://object-permanence.world>CinematicVQA: Benchmarking Film-Grammar Reasoning in Large Vision-Language Modelshttps://arxiv.org/abs/2609.28813>Where Hallucinations Live: A Cross-Architecture Circuit in VQ-Tokenized Vision-Language Modelshttps://shamanthak-hegde.github.io/where-hallucinations-live>Pistis Technical Reporthttps://arxiv.org/abs/2609.28554>JoyAI-VL-Interaction: Real-Time Vision-Language Interaction Intelligencehttps://arxiv.org/abs/2606.14777
>>109908533Friends don't let you schizo out in the OP
>>109908533nigger
>>109908496Thank you for baking this thread, anon >>109908533Thank you for blessing this thread, anon
>>109908550We already know he is one
>>109908512its just getting started
need kino
Thanks so much to catjack specifically for making this a trans safe space! You are valid sister!
lilbro is crashing out over the bake *skull*
>>109908585The OP has had schizo coped baked in since the beginning
>>109908604Catjack doesn't know how filters work or how ignoring people works so she has a shitty tantrum and ruins everything. It's the only logical endgame for /ldg/ and catjack should just stay in the containment thread /sdg/ for namefags like him
i think diffusion models are slop and can never not be slopi think ultimately the role of diffusion will be a rendering layer for llm's to orchestratebut on its own the diffusion model is just not capable of producing kino
>>109908661good thing we pivoted to flow based models, then
>>109908661very deep anon, I came twice
>>109908672qrd
>>109908672flow matching is also slopmore specifically the ways that we interface with these models is sloptext, image, control etcall these methods are bad and awfulto truly ascend you need and entity can just play the latent space like a musician plays and instrumentwe will never be able to do this
>>109908585It's funny that he replied to your post continuing to crash out keeeeeek
Every time someone says "ai will never do X" some months later it starts doing X. Rookie thinking.
i stayed up all night making kinos again
>>109908648I thought troonjack started /ldg/.
>>109908714Well anon, I'm all ears
gettin real sick of your shit around here
>>109908687play me like one of your french trumpets
>>109908773kek
>>109908698AI you will never be a woman
>>109908801AI will never be my girlfriend
krea's censorship is starting to annoy me, even more sfw but ecchi concepts like holding a cup between breasts it will just refuseIs there a best recommend lora that can fix that while not impacting the intelligence too much and not just a straight up porn lora that makes everything nude
>>109908808the textfusion refusal lora but just use qwen instead
>>109908773i don't get it
>>109908760No, anon likes to make stuff up to trick newfrens.
>>109908828lol what a prankster :P
>>109908808Skeelshoe?
every schizo argument requires multiple sides
TIL Ernie Base is a gorillion times better than the Turbo version
>>109908922>Ernieget a load of this oldfag
>>109908937Kek nice
cozy breas