Previous: >>109540195https://rentry.org/ldg-lazy-getting-started-guide>UIComfyUI: https://github.com/comfyanonymous/ComfyUISwarmUI: https://github.com/mcmonkeyprojects/SwarmUISDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineageWan2GP: https://github.com/deepbeepmeep/Wan2GP>Checkpoints, LoRAs, & Upscalershttps://huggingface.co/modelshttps://huggingbay.xyzhttps://civitai.comhttps://civitaiarchive.comhttps://openmodeldb.info>Tuninghttps://github.com/spacepxl/demystifying-sd-finetuninghttps://github.com/ostris/ai-toolkithttps://github.com/Nerogar/OneTrainerhttps://github.com/tdrussell/diffusion-pipehttps://github.com/kohya-ss/sd-scriptshttps://github.com/kohya-ss/musubi-tuner>Minimax H3https://huggingface.co/Comfy-Org/MiniMax-H3>Krea 2https://huggingface.co/krea/Krea-2-Rawhttps://huggingface.co/krea/Krea-2-Turbohttps://lumenastrum.github.io/clio-style-preview/gallery/>Animahttps://huggingface.co/circlestone-labs/Animahttps://tagexplorer.github.io/https://animadex.net>Kleinhttps://huggingface.co/collections/black-forest-labs/flux2>MiscLocal Model Meta: https://rentry.org/localmodelsmetaShare Metadata: https://catbox.moe | https://litterbox.catbox.moe/Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusionArchive: https://rentry.org/sdg-linkCollage: https://rentry.org/ldgcollage_v2>Neighbors>>>/aco/csdg>>>/b/degen>>>/gif/vdg>>>/d/ddg>>>/e/edg>>>/h/hdg>>>/trash/slop>>>/vt/vtai>>>/u/udg>Local Text>>>/g/lmg>Maintain Thread Qualityhttps://rentry.org/debohttps://rentry.org/animanon
balloon boobs
blessed thread of frenship
animators status?
>>109542154RAPED
Can someone contact minimaxAI and tell them to train their image model on danbooru? Better yet, make an anime finetune for it in like a day or two.
>>109542150Nah, 10s is plenty, 12s is great, 15s is perfect. 20 is too much right now
>>109542158It can do anime just fine?
Now that the dust settled, has anyone else started feeling H3’s 10-15s length limit being just too short?LTX spoiled me with 25s
>>109542171>>109542159
https://litter.catbox.moe/fty2n9bs1c6d06hl.mp4
>>109542159What are you generally genning?
>>109542154Im become jobless soon. What should i do ?
idk why I'm getting filtered so hard by this gen.
>>109542144>mfw Resource news08/12/2026>LTX-2.5 22B IC-LoRA Pixel Spatial Upscaler https://huggingface.co/Lightricks/LTX-2.5-22b-IC-LoRA-Pixel-Spatial-Upscaler>LTX-2.5 22B Distilled — NVFP4, ComfyUI-ready https://huggingface.co/BennyDaBall/LTX-2.5-22b-distilled-nvfp4-comfy>LTX-2.5 22B — GGUFhttps://huggingface.co/realrebelai/LTX-2.5_GGUFs>ComfyUI NVIDIA RTX VSR Prohttps://github.com/whmc76/ComfyUI-NVIDIA-RTX-VSR-Pro>Stable Layers: Decomposing Images into Editable RGBA Layers https://huggingface.co/StabilityLabs/Stable-Layers>PEAK: Precise and Persistent Concept Erasure via k-Sparse Autoencodershttps://github.com/manmanTAT/PEAK>Flow Straight to Reality: Perceptually Consistent Flow Matching for Efficient Image Restorationhttps://github.com/aiimaginglab/PCFlow>MMArt A Multi-Perspective Multimodal Dataset for Visual Art Understandinghttps://shuaiwang97.github.io/MMArt>ComfyUI H3 Studio: Video editor for MiniMax H3 inside a single ComfyUI nodehttps://github.com/shootthesound/ComfyUI-H3Studio>MINIMAX H3 Prompt Studio: Build structured MiniMax H3 video-generation promptshttps://github.com/lololerigolo60/Minimax-H3-prompt-studio/tree/main>ComfyUI Image Conveyor v1.4 — now with MiniMax H3 multi-reference supporthttps://github.com/xmarre/ComfyUI-Image-Conveyor/releases/tag/v1.4.0>VPIPE: Real-time multimodal AI pipelines on Apple Siliconhttps://github.com/tgo-app-dev/vpipe08/11/2026>LTX-2.5https://ltx.io/model/ltx-2-5>LIGHTX2V v1.0 4-step/8-step Turbo Minimax H3 lorashttps://huggingface.co/lightx2v/Minimax-h3-Turbo/tree/main>ComfyUI-DoRA-Dynamic-LoRA-Loaderhttps://github.com/xmarre/ComfyUI-DoRA-Dynamic-LoRA-Loader>DeepFreqMark: End-To-End Learnable Frequency-Domain Watermarking with Spherical Attack Simulation for Latent Diffusion Modelshttps://github.com/chenhsiu48/DeepFreqMark>Staying True to the Origin: Continuous Image Stylization with Smooth Transitionshttps://reychiaro.github.io/StyleController
>>109542171you can go further with h3. i haven't tested what length it starts to fall apart since i don't have enough vram
>mfw Research news08/12/2026>UniProbe: A Learnable Token-Level Hallucination Detector for Large VLMs using Multi-Structural Internal Representationshttps://research.nvidia.com/labs/par/uniprobe>SparSTAR: Sparse Attention for SpaceTime AutoRegressive Video Synthesishttps://arxiv.org/abs/2608.10519>Watching Synthetic Videos: Aligning Cross-modal Representations with Visual Synthesis for Zero-shot Video Captioninghttps://arxiv.org/abs/2608.11013>Beyond Pixels: From Video Priors to 4D Worldshttps://hayd-zju.github.io/Beyond-Pixels>Bridging Event Streams and DiT: Event-Guided Video Frame Interpolationhttps://joseph-lin-tech.github.io/BridgeEventDiT-VFI>Stream Forcing: Constructing Unified Training Trajectory for Robust Streaming Video Generationhttps://arxiv.org/abs/2608.10439>NullEdit: Stealthy Image Protection via VLM Condition Redirectionhttps://arxiv.org/abs/2608.10870>Where To Look? : Causal Tracing of Vision Encoders in VLMhttps://arxiv.org/abs/2608.10758>Rethinking Text-Based Image Retrieval in Specific Domainhttps://arxiv.org/abs/2608.10524>Putting Registers to Work: Task Registers for Token Pruning in Vision Transformershttps://arxiv.org/abs/2608.10989>Human versus Computer Visionhttps://arxiv.org/abs/2608.10181>Dynamic Context Adapters: Efficiently Infusing History into Vision-and-Language Modelshttps://arxiv.org/abs/2608.10525>When Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language Modelshttps://arxiv.org/abs/2608.11024>Once Poisoned, Arbitrarily Controlled: A Programmable Backdoor in VLMshttps://arxiv.org/abs/2608.10959>Meshy T2: Fast Native Mesh Generation with Flow Matchinghttps://arxiv.org/abs/2607.28675>FoR-SALE: Frame of Reference-guided Spatial Adjustment in LLM-based Diffusion Editinghttps://arxiv.org/abs/2509.23452
>>109542171I can 20 secs with 8step turbo lora. It took 15 minutes though
Gotta try how it looks if I spin a character around with just frontal view ref.My imagen model doesn't have that many face shots from multiple angles, at least bot ones that look alike
>>109542200>>109542194Not with any decent res.
>>109542188porn
>>109542200resolution and gpu?
I'm just generating sexy anime videos.
>>109542211>>10954221620 secs 0.7mp + RTX upscale + Lightx2v 8step turbo lora. took 15 minutes on my 5070ti
>>109542220just as god intended
most of the game is about genning spank bank materialthe true endgame, however, is turning my LLM ERP's into movies/OVA's. with lots of sex in them obviously.
>>109542223my 5080 gpu farts and freezes when I go for 20s 0.7mp 8 steps (128 ram)maybe one of the cope nodes is bad, could you share the wf?
>>109542162it can't do stuff that finetuned anime model would.
>>109542235I only use sage attention
>>109542158>>109542239>Can someone contact minimaxAI and tell them to train their image model on danbooru? Is this just referring to H3? Is there some sort of image workflow for it?
>>109542241startup flag or the node? all up to date?
why do a 20s gen when you can do 2 10s gens
My uncles boyfriend told me that my uncle, a researcher at Lightricks, participated in a company-wide suicide pact yesterday.
>>109542251one long video with a story line is more kino than multiple short ones
>he can't tell a story across two gensngmi
is it able to not sound like sand paper whenever there's fucking?
Why does the shit from civitai generator look way better than what I do locally with the same models/loras? What's their secret sauce?
>>109542263describe the sound as bones breaking no im not kidding
>>109542261They think they're going to destroy Hollywood but can't even make two gens. Rough.
which turbo lora do I use for H3?there are like 20
im glad to be here with all of youwho remembers 2023 when todd was letting us use his model for free chats on the tavern
>>109542278Todd was the best model I’ve ever tried, still no clue what it was though
>>109542260do you think when they make a movie, they film the whole thing in one take?
>>109542276Theres only two of them Larry and Lightx2v
crazy how clankers can make you basic ass video editors in fucking htmlhttps://files.catbox.moe/oymtau.webm
>>109542269I went hunting on civ, the best sounding videos had this>overall_soundscape:>Slimy squelching, wet fleshy slaps, soft female moans.otherwise I'll try>loud sound of bones breaking, with wet sounds of blood gushing and gore splattering
>>109542288depends on the moviehttps://www.youtube.com/watch?v=ucspfmRM7vI&list=PLb4iXJOYkKszmT5dwMzuf38hzCxeqjCgJ&index=6
>>109542297good style, good gen.
>>109542288the best ones do
>>109542283What an evil wizard
>>109542297It’s actually kind of creepy how many mundane tasks I’ve offloaded to llms. I don’t have to think that far back to when I’d have to convince myself not to spend a hundred or so bucks on software I’m sure I absolutely need. Now it’s totally possible to ask an llm to make a basic functional piece of software and to do the exact job you need and then forget about it entirely. I think once it fully sinks in that a majority of software can be completely replaced by a text prompt and a few minutes, there will be some upheaval.
>>109542288Russian Ark did it
>>109542297?????
>>109542297>>109542315Can you guys recommend local prompt builder for H3? I’ve been writing manuallyAlso do local models have vision? need for r2v
>>109542239which are?
>>109542339nice lol
Does Torch compile work with H3 ?
>overall_soundscape:>Slimy squelching, wet fleshy slaps, soft female moans.>loud sound of bones breaking, with wet sounds of blood gushing and gore splatteringhttps://files.catbox.moe/hdvgue.mp4I think my model is cursed.
>>109542235>>109542223it'll be slower but I'll try on the mac ultras and see if I can do full res over 20 seconds. I have the vram
>>109542246https://www.reddit.com/r/StableDiffusion/comments/1vh9rtw/comment/p29u3x8/ >Regarding single-frame image generation, we are deriving a dedicated image model from a common ancestor in the H3 model lineage, and we expect to make it available to the community.It will use the same VAE encoder as H3. Since the temporal encoder is causal, we can obtain a 2D VAE encoder through weight slicing. We also plan to provide a dedicated VAE decoder specifically designed for image generation.>>109542335most anime concepts? 99% of anime characters you can find on danbooru? 99.9% of artist artstyles? Come on anon.
>>109542401the VAE is ass though
>>109542365why this looks so sloppa
>>109542412the concept and its execution are so good that i can forgive the slopped realism
>>109542339oddly enough I have trouble getting krea and h3 to do dark rooms that dont have an imaginary studio light
>>109542412he probably used a fried image for it
>>109542412has turbo lora stink on it
This is the endgame of Local Video gen right because i cant see this get improved much more. And theres no way we can afford a new GPU with the current prices
>>109542334>Also do local models have vision?yes. most do. and it's typically in the info exposed on either the UI/CLI or on hf or wherever you get your models.> Can you guys recommend local prompt builder for H3? See if you're already running a decently powerful LLM with vision then you can actually hand it the prompting guide https://huggingface.co/MiniMaxAI/MiniMax-H3/blob/main/docs/VIDEO_PROMPT_WRITING_GUIDE_ref_en.md and tell it to write a prompt where x y happens with z.
yeah ok. I'm pretty convinced we'll need a sex sounds lora.
>>109542449>can’t see this get improvedI agree but I want more than 15 seconds of slop
>the original DiT paper was published four years ago >the first model to use this new arc was pixart alpha >only recently has DiT became the norm for the average user Incredible
reference to video is a slow but magic AI video upscaler...what the fuck. You plug the video in and use the same prompt or go promptless, not sure which is better yet. It even fixes the low res fucked up faces within reason.I keep getting slopped gens when genning beyond standard def. Its better to gen low res quickly then queue up your good gens for upscaling.
Sex sex sex blah blah blah. How about you gen some fantasy kino? Cmon we all know how much isekai slop you have watched and enjoyed. Gen your waifu picking up a new quest at the guild tavern, picking a fight and winning against some thug.
>>109542449It will taper off eventually but the default should be expecting four next years to have the same level of improvement as the last four years
>>109542477>just generate the same video two times
>>109542478I Don't even like this gen, all I want to do is get back to my actual good none porn gens. but I refuse let it fucking defeat me like that.
>>109542477Bro at this point just use the turbo lora 2pass workflow.
>>109542449>i cant see this get improved much moreyou can't? one-shot models we have now are non-starters. we need a continuous gen model so instead of being capped at 10s/15s/20s, we can instead prompt continuously in chunks. these current gen models will never tell a story. they're just toy novelties.
>>109542503yep, real time models are the next step. ltx is almost there, but the resolution is too low on consumer hardware
>>109542487Some prompts are AI slop seed lotteries or will only produce realism at low res. Its either that or you delete most of your gens instead of a low res one that took 5 minutes. Also you can upscale to 2 MP and know the time is spent on something worth it
>>109542449Nah, these things still have a long way to go and they're still relatively easy to run on consumer hardware.
>>109542171your post gave me something to think aboutso i went to the free movie websitefired up the new spooderman and told the ai to make something to count shotsthis movie made a billion dollars in a few days, but just watching the first five minutesthe average time for a shot starts to dip below 3 seconds!i say we emulate this for some quick money and then we use that money to buy 6000s and make the actual kino we wanted to all along
>>109542267anyone know?
>>109542499I asked about a 2 pass earlier and nobody said anything
>>109542518it isn't just about how long each shot is, but you need to maintain scene consistency between all of them. generating a longer video lets you keep the scene in memory longer, especially if you are prompting multiple shots within it
>>109542518shot length isn't a useful metric. scene length is. dozens of shots all have to maintain the same scene, location, tone, emotion, framing, pacing, trajectory, etc. viable scene continuation is very hard
>>109542518whats funny is when feature length films are entirely made by AI I bet the average shot length will be longer
>>109542452Isn't there an official mcp for h3 prompts?
>>109542525https://files.catbox.moe/8asrfu.mp4It's a bit messy but the important bit is clean.
>>109542554kek
>>109542477>But magic AI video upscaler...what the fuck.Link? Weren't anons using regular upscalers?
I feel like my brain has been calcified by the limitations of previous models and I have to retrain my expectations and reconsider my limitations because of just how powerful h3 is.
>>109542608>I have to retrain my expectations and reconsider my limitationsyou spend too many hours around marketing retards
is there a way to overlay audio with ref model? Like I give it a song to play, but want it to be faint in the background while the subject is speaking
>>109542608I'm in the same boat, currently ripping apart half of my pipeline as its not necessary to tard wrangle the model anymore
>>109542297what is this?
>>109542576neat thanks
>>109542648You're one of my favorite genner
>>109542571yes you can also gen via API, its just not the focus of this bread
so far I used>int8 convrot>comfy kitchen attention>distill 8step loraanything else can speed up H3?
>>109542666checkd and thanks anon
>>109542276They all fuck up prompt adherence and output quality.