Discussion and Development of Local Image, Video, and Audio ModelsPrevious: >>110012492https://rentry.org/ldg-lazy-getting-started-guide>UISwarmUI: https://github.com/mcmonkeyprojects/SwarmUISDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineageWan2GP: https://github.com/deepbeepmeep/Wan2GPNeural-Pixel (sd.cpp): https://github.com/Luiz-Alcantara/Neural-Pixel>Checkpoints, LoRAs, & Upscalershttps://huggingface.co/modelshttps://huggingbay.xyzhttps://civitai.comhttps://civitaiarchive.comhttps://openmodeldb.info>Tuninghttps://github.com/spacepxl/demystifying-sd-finetuninghttps://github.com/ostris/ai-toolkithttps://github.com/Nerogar/OneTrainerhttps://github.com/tdrussell/diffusion-pipehttps://github.com/kohya-ss/sd-scriptshttps://github.com/kohya-ss/musubi-tuner>Qwen Image 2.1https://huggingface.co/Qwen/Qwen-Image-2.1>Minimax H3https://huggingface.co/Comfy-Org/MiniMax-H3https://neta.art/use-cases/en/h3-1000-prompt-list>Animahttps://huggingface.co/circlestone-labs/Animahttps://animastyles.thetacursed.comhttps://tagexplorer.github.io/https://animadex.net>Krea 2https://huggingface.co/krea/Krea-2-Rawhttps://huggingface.co/krea/Krea-2-Turbohttps://lumenastrum.github.io/clio-style-preview/gallery/>Kleinhttps://huggingface.co/collections/black-forest-labs/flux2>MiscShare Metadata: https://catbox.moe | https://litterbox.catbox.moe/Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusionArchive: https://rentry.org/sdg-linkCollage: https://rentry.org/neo_collage>Neighbors>>>/aco/csdg>>>/b/degen>>>/gif/vdg>>>/d/ddg>>>/e/edg>>>/h/hdg>>>/trash/slop>>>/vt/vtai>>>/u/udg>Local Text>>>/g/lmg>Maintain Thread Qualityhttps://reentry.org/debohttps://reentry.org/animanon
Blessed thread of frenship
>>110019459Thank you for baking this thread, anon>>110019777Thank you for blessing this thread, anon
the fineporn checkpoint works quite well with regular non-porn photos, it gives them a more casual smartphone photo look to them. Pic related is the checkpoint alone, no loras.
>>110020454nvm that one had a "Realism Engine" lora on it, my bad.pic related is without loras
>shitting your pants and making more troll threadsYou're completely on your knees
>>110020508looks better. less dirt layer on top
>mfw Resource news10/08/2026>Iris-3B: Pixel-space generative model that can act as a general vision learnerhttps://github.com/speridlabs/iris-3b>ComfyUI VELA H3https://github.com/Speach1sdef178/ComfyUI-VELA-H3>GRACE: Generation-aware latent compression for efficient video generationhttps://cvlab-kaist.github.io/GRACE>QuadTok: Quadtree Visual Tokenizer for Autoregressive Image Generationhttps://github.com/myc634/QuadTok>OverLay++: Dense-Overlap Layout-to-Image Generation Datasethttps://mlpc-ucsd.github.io/OverLayPP>StoryBlender: Inter-Shot Consistent and Editable 3D Storyboard with Spatial-temporal Dynamicshttps://engineeringai-lab.github.io/StoryBlender>H3 Long Shot Studio API v1.0https://github.com/r34vtraining/Longshot_Studio10/07/2026>Run HunyuanImage 3.0 (80B) natively in ComfyUI on a single 12–24 GB GPUhttps://github.com/PedroMarinhoDev/ComfyUI-HunyuanImage3>ComfyUI H3 Video Upsamplerhttps://github.com/dntpi/ComfyUI-H3-Video-Upsampler>Veda Sparse Attention for ComfyUI (MiniMax-H3)https://github.com/veda-sparse/Veda-on-ComfyUI>ReDetail 2.0: Video upscaling and re-detailing for ComfyUI on LTX-2.5https://github.com/Bambushu/redetail>FastVideo FastH3 Trim for ComfyUIhttps://huggingface.co/FastVideo/FastVideo-FastH3-Trim-Comfy>FIBO Scene Analyzer [dev] https://huggingface.co/briaai/fibo-scene-analyzer>Two Halves are More than One: Phase-wise Velocity Distillation for Fast and High-Quality Image Generationhttps://github.com/PolyU-VCLab/PVD>S2PD: Serial-to-Parallel Diffusion for Physically and Logically Consistent Video Generationhttps://jefequien.github.io/S2PD>Disentangling Dual Image References in Frequency Aware Diffusion Models for Personalized Generationhttps://github.com/htyjers/Dual-FDM>Talk Like You: Imitating How You Speak in Real-Time Talking Head Generationhttps://bq-wang0511.github.io/TalkLikeYou>On Color Alignment in VAE Latent Spaces and Applicationshttps://julian075.github.io/Color_Subspace
>mfw Research news10/08/2026>MORCA: Offline-to-Online Reinforcement Learning for Adaptive Cache Reuse in Video Diffusion Accelerationhttps://arxiv.org/abs/2610.10457>ORCA: Hunting Compositional Failures in Text-to-Image Diffusionhttps://arxiv.org/abs/2610.09841>Relational Abstractions for Spatial Reasoning with Diffusion Modelshttps://arxiv.org/abs/2610.09780>SGF+: Decoupling Gradient Flows for Autoregressive Video Generationhttps://arxiv.org/abs/2610.10429>Real-Time Joint Audio-Video Generation by Parallel Adapter Compositionhttps://arxiv.org/abs/2610.10343>Latent Watermarks under Generative Editing: A Benchmark and Analysis of Detection Survivalhttps://arxiv.org/abs/2610.09702>Visual Jev Rewards: Reference-Bound Verification for Multi-Subject Image Generationhttps://arxiv.org/abs/2610.09328>Enhancing Multi-Region Stylization with Interior-Guided Boundary Repairhttps://arxiv.org/abs/2610.09706>UltraText Bench: A Comprehensive Bilingual Benchmark for Evaluating Visual Text Rendering in Image Generationhttps://arxiv.org/abs/2610.09823>Consistent Distribution Matching for Data-Free Diffusion Distillationhttps://consistentdmd.github.io>Personalize at Test Time: Learning User Preferences for Image Generationhttps://arxiv.org/abs/2610.09015>DISRQAD: Diffusion Image Super-Resolution Quality Assessment Dataset and Benchmarkhttps://arxiv.org/abs/2610.09077>Pooling Representation Autoencoders for Efficient Diffusionhttps://arxiv.org/abs/2610.09242>VIS-Ground: Video Interactive Storytelling with Contextual Groundinghttps://bx126.github.io/vis-ground.github.io>Iris-3B: Going Beyond the Latent with Pixel-Space Diffusion Training, Conversion and Fine-Tuninghttps://arxiv.org/abs/2610.09450>What Makes Synthetic Hard Negatives Work in Vision-Language Pretraining?https://arxiv.org/abs/2610.09700
>>110019459Is this the frogpost general /fpg/?
>>110021508>>110021514Thanks bro you're the only reason i visit this shithole
>>110021661sorry I missed news yesterday. there were too many threads and I didnt know which would persist. idk if even this thread will stay up
>taking my own passport photo>qwen edit>he is wearing a sexy evening dress and high heelsfun
You are a helpful vision assistant. Describe the penis you are shown in rich, concrete detail rather than a brief summary. Do not skip detail for the sake of brevity. Describe <image1> in detail: the appearance, the colors, the style, the size, skin details, its artistic style, its levels of shading and values, its method and manner of rendering skin--Don't focus on what exists, focus on how it exists.Respond with the final answer only -- no reasoning, no <think> blocks. Keep the response to a compact paragraph of about 100 words.
>dweebo spamming the news in the troll bake first before posting in his home thread Holy lel
>>110022147nah what is unc cooking bruh?
>>110022198
>>110022205
I should've used vision for detailers a long time ago.