Discussion and Development of Local Image, Video, and Audio ModelsPrevious: >>110012492https://rentry.org/ldg-lazy-getting-started-guide>UIComfyUI: https://github.com/comfyanonymous/ComfyUISwarmUI: https://github.com/mcmonkeyprojects/SwarmUISDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineageWan2GP: https://github.com/deepbeepmeep/Wan2GPNeural-Pixel (sd.cpp): https://github.com/Luiz-Alcantara/Neural-Pixel>Checkpoints, LoRAs, & Upscalershttps://huggingface.co/modelshttps://huggingbay.xyzhttps://civitai.comhttps://civitaiarchive.comhttps://openmodeldb.info>Tuninghttps://github.com/spacepxl/demystifying-sd-finetuninghttps://github.com/ostris/ai-toolkithttps://github.com/Nerogar/OneTrainerhttps://github.com/tdrussell/diffusion-pipehttps://github.com/kohya-ss/sd-scriptshttps://github.com/kohya-ss/musubi-tuner>Qwen Image 2.1https://huggingface.co/Qwen/Qwen-Image-2.1>Minimax H3https://huggingface.co/Comfy-Org/MiniMax-H3https://neta.art/use-cases/en/h3-1000-prompt-list>Animahttps://huggingface.co/circlestone-labs/Animahttps://animastyles.thetacursed.comhttps://tagexplorer.github.io/https://animadex.net>Krea 2https://huggingface.co/krea/Krea-2-Rawhttps://huggingface.co/krea/Krea-2-Turbohttps://lumenastrum.github.io/clio-style-preview/gallery/>Kleinhttps://huggingface.co/collections/black-forest-labs/flux2>MiscShare Metadata: https://catbox.moe | https://litterbox.catbox.moe/Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusionArchive: https://rentry.org/sdg-linkCollage: https://rentry.org/neo_collage>Neighbors>>>/aco/csdg>>>/b/degen>>>/gif/vdg>>>/d/ddg>>>/e/edg>>>/h/hdg>>>/trash/slop>>>/vt/vtai>>>/u/udg>Local Text>>>/g/lmg
Half the catalog is /ldg/ yet theres not a single good local model
>>110022098
Blessed thread of frenship
*sip*Aaaah...this finna be a good thread fr fr JSID....(Jack Said Israel Dies)
>mfw Resource news10/08/2026>Iris-3B: Pixel-space generative model that can act as a general vision learnerhttps://github.com/speridlabs/iris-3b>ComfyUI VELA H3https://github.com/Speach1sdef178/ComfyUI-VELA-H3>GRACE: Generation-aware latent compression for efficient video generationhttps://cvlab-kaist.github.io/GRACE>QuadTok: Quadtree Visual Tokenizer for Autoregressive Image Generationhttps://github.com/myc634/QuadTok>OverLay++: Dense-Overlap Layout-to-Image Generation Datasethttps://mlpc-ucsd.github.io/OverLayPP>StoryBlender: Inter-Shot Consistent and Editable 3D Storyboard with Spatial-temporal Dynamicshttps://engineeringai-lab.github.io/StoryBlender>H3 Long Shot Studio API v1.0https://github.com/r34vtraining/Longshot_Studio10/07/2026>Run HunyuanImage 3.0 (80B) natively in ComfyUI on a single 12–24 GB GPUhttps://github.com/PedroMarinhoDev/ComfyUI-HunyuanImage3>ComfyUI H3 Video Upsamplerhttps://github.com/dntpi/ComfyUI-H3-Video-Upsampler>Veda Sparse Attention for ComfyUI (MiniMax-H3)https://github.com/veda-sparse/Veda-on-ComfyUI>ReDetail 2.0: Video upscaling and re-detailing for ComfyUI on LTX-2.5https://github.com/Bambushu/redetail>FastVideo FastH3 Trim for ComfyUIhttps://huggingface.co/FastVideo/FastVideo-FastH3-Trim-Comfy>FIBO Scene Analyzer [dev] https://huggingface.co/briaai/fibo-scene-analyzer>Two Halves are More than One: Phase-wise Velocity Distillation for Fast and High-Quality Image Generationhttps://github.com/PolyU-VCLab/PVD>S2PD: Serial-to-Parallel Diffusion for Physically and Logically Consistent Video Generationhttps://jefequien.github.io/S2PD>Disentangling Dual Image References in Frequency Aware Diffusion Models for Personalized Generationhttps://github.com/htyjers/Dual-FDM>Talk Like You: Imitating How You Speak in Real-Time Talking Head Generationhttps://bq-wang0511.github.io/TalkLikeYou>On Color Alignment in VAE Latent Spaces and Applicationshttps://julian075.github.io/Color_Subspace
>mfw Research news10/08/2026>MORCA: Offline-to-Online Reinforcement Learning for Adaptive Cache Reuse in Video Diffusion Accelerationhttps://arxiv.org/abs/2610.10457>ORCA: Hunting Compositional Failures in Text-to-Image Diffusionhttps://arxiv.org/abs/2610.09841>Relational Abstractions for Spatial Reasoning with Diffusion Modelshttps://arxiv.org/abs/2610.09780>SGF+: Decoupling Gradient Flows for Autoregressive Video Generationhttps://arxiv.org/abs/2610.10429>Real-Time Joint Audio-Video Generation by Parallel Adapter Compositionhttps://arxiv.org/abs/2610.10343>Latent Watermarks under Generative Editing: A Benchmark and Analysis of Detection Survivalhttps://arxiv.org/abs/2610.09702>Visual Jev Rewards: Reference-Bound Verification for Multi-Subject Image Generationhttps://arxiv.org/abs/2610.09328>Enhancing Multi-Region Stylization with Interior-Guided Boundary Repairhttps://arxiv.org/abs/2610.09706>UltraText Bench: A Comprehensive Bilingual Benchmark for Evaluating Visual Text Rendering in Image Generationhttps://arxiv.org/abs/2610.09823>Consistent Distribution Matching for Data-Free Diffusion Distillationhttps://consistentdmd.github.io>Personalize at Test Time: Learning User Preferences for Image Generationhttps://arxiv.org/abs/2610.09015>DISRQAD: Diffusion Image Super-Resolution Quality Assessment Dataset and Benchmarkhttps://arxiv.org/abs/2610.09077>Pooling Representation Autoencoders for Efficient Diffusionhttps://arxiv.org/abs/2610.09242>VIS-Ground: Video Interactive Storytelling with Contextual Groundinghttps://bx126.github.io/vis-ground.github.io>Iris-3B: Going Beyond the Latent with Pixel-Space Diffusion Training, Conversion and Fine-Tuninghttps://arxiv.org/abs/2610.09450>What Makes Synthetic Hard Negatives Work in Vision-Language Pretraining?https://arxiv.org/abs/2610.09700
>>110022397ty anon
o new models besides Anima that I can run under 5GB on a 4090?
>>110022389>>110022397Thank you you're the only reason i still visit this thread
>>110022075Hey you dropped this>Maintain Thread Qualityhttps://rentry.org/debohttps://rentry.org/animanon
>>110022389>>110022397thanks!
>>110022472
>>110022397>>DISRQAD: Diffusion Image Super-Resolution Quality Assessment Dataset and BenchmarkThis seems useful
>>110022389>>110022397Thank you for ignoring the schizo and still working hard for the news bro
cozy breas
>>110022551
jooooooooooo bidenWAKE UP
>>110022564
>>110022389>>110022397Based thank you
>>110022615
>>110022389>>110022397WowHow do you compile the news? Can you drop some hints? Curios!