Previous: >>109540195https://rentry.org/ldg-lazy-getting-started-guide>UIComfyUI: https://github.com/comfyanonymous/ComfyUISwarmUI: https://github.com/mcmonkeyprojects/SwarmUISDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineageWan2GP: https://github.com/deepbeepmeep/Wan2GP>Checkpoints, LoRAs, & Upscalershttps://huggingface.co/modelshttps://huggingbay.xyzhttps://civitai.comhttps://civitaiarchive.comhttps://openmodeldb.info>Tuninghttps://github.com/spacepxl/demystifying-sd-finetuninghttps://github.com/ostris/ai-toolkithttps://github.com/Nerogar/OneTrainerhttps://github.com/tdrussell/diffusion-pipehttps://github.com/kohya-ss/sd-scriptshttps://github.com/kohya-ss/musubi-tuner>Minimax H3https://huggingface.co/Comfy-Org/MiniMax-H3>Krea 2https://huggingface.co/krea/Krea-2-Rawhttps://huggingface.co/krea/Krea-2-Turbohttps://lumenastrum.github.io/clio-style-preview/gallery/>Animahttps://huggingface.co/circlestone-labs/Animahttps://tagexplorer.github.io/https://animadex.net>Kleinhttps://huggingface.co/collections/black-forest-labs/flux2>MiscLocal Model Meta: https://rentry.org/localmodelsmetaShare Metadata: https://catbox.moe | https://litterbox.catbox.moe/Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusionArchive: https://rentry.org/sdg-linkCollage: https://rentry.org/ldgcollage_v2>Neighbors>>>/aco/csdg>>>/b/degen>>>/gif/vdg>>>/d/ddg>>>/e/edg>>>/h/hdg>>>/trash/slop>>>/vt/vtai>>>/u/udg>Local Text>>>/g/lmg>Maintain Thread Qualityhttps://rentry.org/debohttps://rentry.org/animanon
balloon boobs
blessed thread of frenship
animators status?
>>109542154RAPED
Can someone contact minimaxAI and tell them to train their image model on danbooru? Better yet, make an anime finetune for it in like a day or two.
>>109542150Nah, 10s is plenty, 12s is great, 15s is perfect. 20 is too much right now
>>109542158It can do anime just fine?
Now that the dust settled, has anyone else started feeling H3’s 10-15s length limit being just too short?LTX spoiled me with 25s
>>109542171>>109542159
https://litter.catbox.moe/fty2n9bs1c6d06hl.mp4
>>109542159What are you generally genning?
>>109542154Im become jobless soon. What should i do ?
idk why I'm getting filtered so hard by this gen.
>>109542144>mfw Resource news08/12/2026>LTX-2.5 22B IC-LoRA Pixel Spatial Upscaler https://huggingface.co/Lightricks/LTX-2.5-22b-IC-LoRA-Pixel-Spatial-Upscaler>LTX-2.5 22B Distilled — NVFP4, ComfyUI-ready https://huggingface.co/BennyDaBall/LTX-2.5-22b-distilled-nvfp4-comfy>LTX-2.5 22B — GGUFhttps://huggingface.co/realrebelai/LTX-2.5_GGUFs>ComfyUI NVIDIA RTX VSR Prohttps://github.com/whmc76/ComfyUI-NVIDIA-RTX-VSR-Pro>Stable Layers: Decomposing Images into Editable RGBA Layers https://huggingface.co/StabilityLabs/Stable-Layers>PEAK: Precise and Persistent Concept Erasure via k-Sparse Autoencodershttps://github.com/manmanTAT/PEAK>Flow Straight to Reality: Perceptually Consistent Flow Matching for Efficient Image Restorationhttps://github.com/aiimaginglab/PCFlow>MMArt A Multi-Perspective Multimodal Dataset for Visual Art Understandinghttps://shuaiwang97.github.io/MMArt>ComfyUI H3 Studio: Video editor for MiniMax H3 inside a single ComfyUI nodehttps://github.com/shootthesound/ComfyUI-H3Studio>MINIMAX H3 Prompt Studio: Build structured MiniMax H3 video-generation promptshttps://github.com/lololerigolo60/Minimax-H3-prompt-studio/tree/main>ComfyUI Image Conveyor v1.4 — now with MiniMax H3 multi-reference supporthttps://github.com/xmarre/ComfyUI-Image-Conveyor/releases/tag/v1.4.0>VPIPE: Real-time multimodal AI pipelines on Apple Siliconhttps://github.com/tgo-app-dev/vpipe08/11/2026>LTX-2.5https://ltx.io/model/ltx-2-5>LIGHTX2V v1.0 4-step/8-step Turbo Minimax H3 lorashttps://huggingface.co/lightx2v/Minimax-h3-Turbo/tree/main>ComfyUI-DoRA-Dynamic-LoRA-Loaderhttps://github.com/xmarre/ComfyUI-DoRA-Dynamic-LoRA-Loader>DeepFreqMark: End-To-End Learnable Frequency-Domain Watermarking with Spherical Attack Simulation for Latent Diffusion Modelshttps://github.com/chenhsiu48/DeepFreqMark>Staying True to the Origin: Continuous Image Stylization with Smooth Transitionshttps://reychiaro.github.io/StyleController
>>109542171you can go further with h3. i haven't tested what length it starts to fall apart since i don't have enough vram
>mfw Research news08/12/2026>UniProbe: A Learnable Token-Level Hallucination Detector for Large VLMs using Multi-Structural Internal Representationshttps://research.nvidia.com/labs/par/uniprobe>SparSTAR: Sparse Attention for SpaceTime AutoRegressive Video Synthesishttps://arxiv.org/abs/2608.10519>Watching Synthetic Videos: Aligning Cross-modal Representations with Visual Synthesis for Zero-shot Video Captioninghttps://arxiv.org/abs/2608.11013>Beyond Pixels: From Video Priors to 4D Worldshttps://hayd-zju.github.io/Beyond-Pixels>Bridging Event Streams and DiT: Event-Guided Video Frame Interpolationhttps://joseph-lin-tech.github.io/BridgeEventDiT-VFI>Stream Forcing: Constructing Unified Training Trajectory for Robust Streaming Video Generationhttps://arxiv.org/abs/2608.10439>NullEdit: Stealthy Image Protection via VLM Condition Redirectionhttps://arxiv.org/abs/2608.10870>Where To Look? : Causal Tracing of Vision Encoders in VLMhttps://arxiv.org/abs/2608.10758>Rethinking Text-Based Image Retrieval in Specific Domainhttps://arxiv.org/abs/2608.10524>Putting Registers to Work: Task Registers for Token Pruning in Vision Transformershttps://arxiv.org/abs/2608.10989>Human versus Computer Visionhttps://arxiv.org/abs/2608.10181>Dynamic Context Adapters: Efficiently Infusing History into Vision-and-Language Modelshttps://arxiv.org/abs/2608.10525>When Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language Modelshttps://arxiv.org/abs/2608.11024>Once Poisoned, Arbitrarily Controlled: A Programmable Backdoor in VLMshttps://arxiv.org/abs/2608.10959>Meshy T2: Fast Native Mesh Generation with Flow Matchinghttps://arxiv.org/abs/2607.28675>FoR-SALE: Frame of Reference-guided Spatial Adjustment in LLM-based Diffusion Editinghttps://arxiv.org/abs/2509.23452
>>109542171I can 20 secs with 8step turbo lora. It took 15 minutes though
Gotta try how it looks if I spin a character around with just frontal view ref.My imagen model doesn't have that many face shots from multiple angles, at least bot ones that look alike
>>109542200>>109542194Not with any decent res.
>>109542188porn
>>109542200resolution and gpu?
I'm just generating sexy anime videos.
>>109542211>>10954221620 secs 0.7mp + RTX upscale + Lightx2v 8step turbo lora. took 15 minutes on my 5070ti
>>109542220just as god intended
most of the game is about genning spank bank materialthe true endgame, however, is turning my LLM ERP's into movies/OVA's. with lots of sex in them obviously.
>>109542223my 5080 gpu farts and freezes when I go for 20s 0.7mp 8 steps (128 ram)maybe one of the cope nodes is bad, could you share the wf?
>>109542162it can't do stuff that finetuned anime model would.
>>109542235I only use sage attention
>>109542158>>109542239>Can someone contact minimaxAI and tell them to train their image model on danbooru? Is this just referring to H3? Is there some sort of image workflow for it?
>>109542241startup flag or the node? all up to date?
why do a 20s gen when you can do 2 10s gens
My uncles boyfriend told me that my uncle, a researcher at Lightricks, participated in a company-wide suicide pact yesterday.
>>109542251one long video with a story line is more kino than multiple short ones
>he can't tell a story across two gensngmi
is it able to not sound like sand paper whenever there's fucking?
Why does the shit from civitai generator look way better than what I do locally with the same models/loras? What's their secret sauce?
>>109542263describe the sound as bones breaking no im not kidding
>>109542261They think they're going to destroy Hollywood but can't even make two gens. Rough.
which turbo lora do I use for H3?there are like 20
im glad to be here with all of youwho remembers 2023 when todd was letting us use his model for free chats on the tavern
>>109542278Todd was the best model I’ve ever tried, still no clue what it was though
>>109542260do you think when they make a movie, they film the whole thing in one take?
>>109542276Theres only two of them Larry and Lightx2v
crazy how clankers can make you basic ass video editors in fucking htmlhttps://files.catbox.moe/oymtau.webm
>>109542269I went hunting on civ, the best sounding videos had this>overall_soundscape:>Slimy squelching, wet fleshy slaps, soft female moans.otherwise I'll try>loud sound of bones breaking, with wet sounds of blood gushing and gore splattering
>>109542288depends on the moviehttps://www.youtube.com/watch?v=ucspfmRM7vI&list=PLb4iXJOYkKszmT5dwMzuf38hzCxeqjCgJ&index=6
>>109542297good style, good gen.
>>109542288the best ones do
>>109542283What an evil wizard
>>109542297It’s actually kind of creepy how many mundane tasks I’ve offloaded to llms. I don’t have to think that far back to when I’d have to convince myself not to spend a hundred or so bucks on software I’m sure I absolutely need. Now it’s totally possible to ask an llm to make a basic functional piece of software and to do the exact job you need and then forget about it entirely. I think once it fully sinks in that a majority of software can be completely replaced by a text prompt and a few minutes, there will be some upheaval.
>>109542288Russian Ark did it
>>109542297?????
>>109542297>>109542315Can you guys recommend local prompt builder for H3? I’ve been writing manuallyAlso do local models have vision? need for r2v
>>109542239which are?
>>109542339nice lol
Does Torch compile work with H3 ?
>overall_soundscape:>Slimy squelching, wet fleshy slaps, soft female moans.>loud sound of bones breaking, with wet sounds of blood gushing and gore splatteringhttps://files.catbox.moe/hdvgue.mp4I think my model is cursed.
>>109542235>>109542223it'll be slower but I'll try on the mac ultras and see if I can do full res over 20 seconds. I have the vram
>>109542246https://www.reddit.com/r/StableDiffusion/comments/1vh9rtw/comment/p29u3x8/ >Regarding single-frame image generation, we are deriving a dedicated image model from a common ancestor in the H3 model lineage, and we expect to make it available to the community.It will use the same VAE encoder as H3. Since the temporal encoder is causal, we can obtain a 2D VAE encoder through weight slicing. We also plan to provide a dedicated VAE decoder specifically designed for image generation.>>109542335most anime concepts? 99% of anime characters you can find on danbooru? 99.9% of artist artstyles? Come on anon.
>>109542401the VAE is ass though
>>109542365why this looks so sloppa
>>109542412the concept and its execution are so good that i can forgive the slopped realism
>>109542339oddly enough I have trouble getting krea and h3 to do dark rooms that dont have an imaginary studio light
>>109542412he probably used a fried image for it
>>109542412has turbo lora stink on it
This is the endgame of Local Video gen right because i cant see this get improved much more. And theres no way we can afford a new GPU with the current prices
>>109542334>Also do local models have vision?yes. most do. and it's typically in the info exposed on either the UI/CLI or on hf or wherever you get your models.> Can you guys recommend local prompt builder for H3? See if you're already running a decently powerful LLM with vision then you can actually hand it the prompting guide https://huggingface.co/MiniMaxAI/MiniMax-H3/blob/main/docs/VIDEO_PROMPT_WRITING_GUIDE_ref_en.md and tell it to write a prompt where x y happens with z.
yeah ok. I'm pretty convinced we'll need a sex sounds lora.
>>109542449>can’t see this get improvedI agree but I want more than 15 seconds of slop
>the original DiT paper was published four years ago >the first model to use this new arc was pixart alpha >only recently has DiT became the norm for the average user Incredible