Discussion and Development of Local Image, Video, and Music ModelsPrevious: >>109483964https://rentry.org/ldg-lazy-getting-started-guide>UIComfyUI: https://github.com/comfyanonymous/ComfyUISwarmUI: https://github.com/mcmonkeyprojects/SwarmUISDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineageWan2GP: https://github.com/deepbeepmeep/Wan2GP>Checkpoints, LoRAs, & Upscalershttps://huggingface.co/modelshttps://huggingbay.xyzhttps://civitai.comhttps://civitaiarchive.comhttps://openmodeldb.info>Tuninghttps://github.com/spacepxl/demystifying-sd-finetuninghttps://github.com/ostris/ai-toolkithttps://github.com/Nerogar/OneTrainerhttps://github.com/tdrussell/diffusion-pipehttps://github.com/kohya-ss/sd-scriptshttps://github.com/kohya-ss/musubi-tuner>Minimax H3https://huggingface.co/Comfy-Org/MiniMax-H3>Krea 2https://huggingface.co/krea/Krea-2-Rawhttps://huggingface.co/krea/Krea-2-Turbohttps://lumenastrum.github.io/clio-style-preview/gallery/>Animahttps://huggingface.co/circlestone-labs/Animahttps://tagexplorer.github.io/https://animadex.net>Kleinhttps://huggingface.co/collections/black-forest-labs/flux2>MiscLocal Model Meta: https://rentry.org/localmodelsmetaShare Metadata: https://catbox.moe | https://litterbox.catbox.moe/Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusionArchive: https://rentry.org/sdg-linkCollage: https://rentry.org/ldgcollage_v2>Neighbors>>>/aco/csdg>>>/b/degen>>>/gif/vdg>>>/d/ddg>>>/e/edg>>>/h/hdg>>>/trash/slop>>>/vt/vtai>>>/u/udg>Local Text>>>/g/lmg>Maintain Thread Qualityhttps://rentry.org/debohttps://rentry.org/animanon
>>109485064fax
>>109485123Damn i got error every time i tried to update my comfyui to nightly
>>109485128>collagefinally, thanks for the bake anonhttps://files.catbox.moe/xbzkob.mp4>>>/wsg/6209573
most kino collage in years
>mfw Resource news08/06/2026>Flash-VAED: Plug-and-Play VAE Decoders for Efficient Video Generationhttps://github.com/Aoko955/Flash-VAED>(preview) MiniMax-H3 Turbo LoRA — 4-step audio-video generation https://huggingface.co/larryvrh/MiniMax-H3-Turbo-Lora>MiniMax-H3 Turbo 4-Step — ComfyUI Pruned-Model LoRAs https://huggingface.co/drbaph/MiniMax-H3-Turbo-Lora-ComfyUI>ComfyUI-H3-Multishothttps://github.com/jlucasmcrell/ComfyUI-H3-Multishot>Krea2 Turbo – OpenPose ControlNet LoRA https://huggingface.co/thedeoxen/Krea-2-pose-controlnet>MiniMax H3 experimental Int8 convrot VAEhttps://huggingface.co/Kijai/MiniMax-H3-experimental>UniWorld-View: Large-Baseline View Synthesis via Video Diffusion Modelshttps://zhouhyocean.github.io/uniworld-view>OmniVR: Joint Video-Audio Conditional Generation for Restoring Degraded Historical Filmshttps://xin1u.github.io/OminiVR_PAGE>DIVE: Dynamic Iterative Visual Evidence Construction for Efficient Vision-Language Modelshttps://github.com/Zhong-Chenchen/DIVE.git>Multi-View Face and Gesture Animation with Dynamic Gaussianshttps://dfki-av.github.io/MVFGA>EmpaAva: An Open-source Agentic 3D-Avatar Empathetic Live Chatbothttps://empaava.top>Context-Anchored Tile Refinehttps://github.com/Blakeem/ComfyUI-ContextAnchoredTileRefine>ComfyUI Video Tilerhttps://github.com/maDcaDDie2000/comfyui-video-tiler08/05/2026>Inline Studio v1.2.62 - Minimax H3 Lora training still onlyhttps://github.com/inlineresearch/Inline-Studio/releases/tag/v1.2.62>Qwen3-VL-32B-Instruct-MiniMax-H3-GGUFhttps://huggingface.co/nif0/Qwen3-VL-32B-Instruct-MiniMax-H3-GGUF>Qwen3-VL-32B-Instruct-ultra-uncensored-heretic-H3-GGUFhttps://huggingface.co/nif0/Qwen3-VL-32B-Instruct-ultra-uncensored-heretic-H3-GGUF>MiniMax-H3-TAE: 2D tine VAE for MiniMax-H3https://huggingface.co/Kijai/MiniMax-H3-TAE>SPADE: An Input-Adaptive Sparse Attention Engine for Fast Video Diffusion Models Inferencehttps://github.com/6somehow/DAC-SPADE
BLESSED COLLAGE THREAD
>mfw Research news08/06/2026>When Diffusion Models Forget Who You Are: Identity Preservation in Face Inpainting under Large Occlusionshttps://arxiv.org/abs/2608.04820>HelloWorld: Enabling Socially Interactive Characters in Video World Modelshttps://arxiv.org/abs/2608.05070>OmniEdit-Bench: A Comprehensive Benchmark for Instruction-based Video Editinghttps://arxiv.org/abs/2608.05049>ContextMaster: Interactive Multi-Shot Video Creation via Fixed-Budget Sparse Context Routinghttps://guoxu1233.github.io/ContextMaster>STEP-OPD: Rethinking Output Targets and Internal Dynamics in On-Policy Distillation for Diffusion Modelshttps://arxiv.org/abs/2608.04887>Simile Understanding in Text-to-Image Models: An Evaluation Frameworkhttps://arxiv.org/abs/2608.04750>ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generationhttps://arxiv.org/abs/2608.04436>CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Modelshttps://arxiv.org/abs/2608.04302>Rethinking Pixel Mean Flows via Interval Denoiserhttps://arxiv.org/abs/2608.04818>Persistent Object Narratives for Token-Efficient Video Language Modelshttps://arxiv.org/abs/2608.04866>Poly-OPD: Heterogeneous Multi-Teacher On-Policy Distillation for Capability-Selectable Flow Modelshttps://arxiv.org/abs/2608.04349>Not All Redundant Tokens Are Alike: Analyzing Visual Token Pruning through Token Roleshttps://arxiv.org/abs/2608.04483>Beyond Global Routing Aggregation: Phase-Aware Expert Merging for MoE Vision-Language Modelshttps://arxiv.org/abs/2608.04454>When does training on downscaled images yield the same gradients?https://arxiv.org/abs/2608.04448>Unleashing the Potential of Vision-Language Models for Generalizable AI-Generated Image Detectionhttps://arxiv.org/abs/2608.04935
>>109485143funny
for me its 4 and 10
Nope, deleting my comfy install. It's too addictive, this is a dead end.
>>109485138did you ask claude or chatgpt to fix it?
ok I think I'll try ref model
for me it's 2 and 12
excluding mine, 1 and 3
How do I trick H3 to generate a single image?
>>109485158I get it. I uninstalled all of my shit earlier this year and deleted all of my gens. Came back though, vid gen just too tempting rn :'(
reminder that with local AI, even if the entire world ended tomorrow, you would have infinite entertainment at your disposal.
>>109485173Try ref image size on max mode.I hear it's better trying to test it to see if it's true
>>109485201is animate inanimate a fetish of yours? because it's patrician tier shit you're making here pal.
>>109485188>describe frame you want> set video length to like, 2 frames>>109485201damn, it does stop motion quite well
why didnt they do the same frame injection system as ltx? that would solve the problem of seamless video continuations
>>109485208I thought it was supposed to be good at anime
>>109485156>for me its 4 and 10same, assuming you're talking about ages
>>109485225Seems fine.
>>109485205>is animate inanimate a fetish of yours?no i just like how it can do stop motion. but what would be an example of something less patrician in that fetish? >>109485215indeed https://pastebin.com/WBbYXvPE
max might be the way to go
I've got a question for anyone who knows the ins and outs of generating loras.When compiling a big collection of images to train the lora on, do I want to crop it to remove all empty space?
>>109485225I never claimed it was a good prompt.
Time to sleep. I will gen less tomorrow. Time to do my work
>>109485253No.
he's not wrong though, euler sucks!https://files.catbox.moe/61dif8.mp4
>>109485247If you want her to suck it like one of those gravure videos dont mention that ice cream as food
you weren't kidding, ref is like 30% slower
>>109485268>the incoherent rambly gibberish when they start fighting like toddlers lol
>>109485268funny
>>109485270I intended for her to eat it because it's in a public space.I should change the icy shape if I want to do that.>>109485277change the pixel setting to max too and you get more speed losses>>109485268I've had great success with euler, what's wrong with it?
>>109485295good gen debo.
>>109485208same prompt but with ref model and the reference panels, (gemini already shat me a ref prompt)>>109485288>change the pixel setting to max too and you get more speed lossesThe 30% was already with max. I'm ok with it being slower.
>>109485318holy kino, can you share it with sound?
https://files.catbox.moe/0n93fm.mp4
>>109485318She got paid $10k by the Olympics for just showing up btw.
https://huggingface.co/drbaph/MiniMax-H3-Turbo-Lora-ComfyUI/blob/main/minimax_h3_turbo_4step_ckpt850_pruned_comfyui.safetensorsit's overcooked, so go for a strength of ~0.75 it's still better than 500 steps at strength 1
>>109485268kek, nice one.>>109485225It is
>>109485327https://streamable.com/cdjsv4
>>109485340It's better if you throw it in the trash and stick to 20 steps with optimisation nodes.
>>109485347kek, nice
>>109485299tyty>>109485318my goat
>>109485340i find that my gens become a bit overcooked as they go past ~7 seconds
the latest comfy update fucked lots of things and is causing lots of glitches>random noise widget gets stuck and using the same seed.>preview gets stuck>image loader hangs until you refresh the tab>job queue gets stuck at the end of a gen, causing the run button not to work the fuck did comfy do to mess it this badly in one update?
>>109485340are we really supposed to use 4 steps for these 4 step loras?the lora from the other day where it was recommended to use 10 steps I think was producing better videos.i mean i guess it's obvious that 10 steps would be better than 4 steps?
for once. I'm not pooling.
>>109485396for the moment not really, I go for 8
LLMs help with prompts a lot if you give them the prompt guide docshttps://files.catbox.moe/ul1krv.mp4
yeah sorry tranime pedos are not allowed to comment on the turbo lora. it's complete dogshit for anything live action and immediately gives it that "AI slop" look
>>109485208>>109485326What's the manga/series? I've seen (presumably) your gens in the past and I like the style.
Is unloading and loading models often a bad thing for your video card?
>>109485427i love clown girls so this speaks to me on a deep level
>>109485427me irl
>sarrs live action i need the marvel superhero 1 bob 1 vagene family meetup kindly sir
>>109485427>pc exploded because my malicious comfyui node blew it up when it detected you forgot asian in the prompt
>>109485427I look like this and I do this!
I just love H<3>https://files.catbox.moe/z2fygc.mp4
>>109485443>it detected you forgot masterpiece, best quality, absurdres, highres, realistic, ultra realistic, score_10,
>>109485454Negative: Bad hands, deformed hands, no more than 5 fingers, no less than 5 fingers
>>109485452that is fucking completely insane what the fuckthis concept is so far out there can any other model even pull this off? will have a cheeky wank to this later anyway.
>>109485452>H<3aktually H < 3 would be H2 and H2 sucks
>>109485452she shouldve sucked the little one up inside her cookie
T2VA is pretty nice with an LLM's help.20steps res_multi/simple on 306012gb in under 5mins with just sageattention and spectrum
>>109485432It's the first time I do gens like this, but I think I know the anon too.manga is BLAME! highly recommend it.ok ref model is insane.
>>109485470susgotta be a turbo lora in there
>>109485475>BLAME!Very cool gens, anon. Thanks. I'll check it out tomorrow.
Aiiie someone (kijai probably) renamed or moved time shift slope in the minimax model implementation and it broke like every cache/attention/optimization node. A couple nodes have been updated but not all.
>>109485475referencewith sound >>>/wsg/6209600
is there a consensus on the best sage attention node and the best cache node?
Minimax h3h3https://files.catbox.moe/xjrnx7.mp4
>>109485459none whatsoever, even image models struggle with such niche concepts and here we have a blown video model doing crazy concepts with simple prompting.>>109485464this is only the start
cozy breas
https://huggingface.co/SexGod1979/PinkCherry_MiniMax-H3
>>109485519>Trained for excellent rabbit motion, along with pink flowers with nectars glisteningBASED
>high quality furry rabbits, rainbows and cherry trees (pink flowers open). Trained for excellent rabbit motion, along with pink flowers with nectars glistening
>>109485478generating at 0.4mp and using rtx upscaler as well
>>109485519>a whole ass checkpointactually, not based at all wtf? just do a lora hot damn.
>>109485519Why did vidrel suddenly pop up in my browser when i opened that link?https://www.youtube.com/watch?v=mJmjljQP3oY&pp=ygUYaSBsb3ZlIGJlaWppbmcgdGlhbmFubWVu
>>109485524>BASEDhe's based in China after all
>>109485475mappa on suicide watch
>>109485492[beta] utilscollection :: MiniMax H3 Cache @ 0.05 / 0.15 / 0.90 / 2 / auto / true>Skipped 7/15 block-stack executions (1.88x theoretical block-stack speedup).easycache 0.10 / 0.30 /0.70>EasyCache - skipped 0/15 steps (1.00x speedup).easycache 0.30 / 0.20 / 0.90 >EasyCache - skipped 4/15 steps (1.36x speedup).>speed wise just a tiny bit slower than the MiniMax H3 Cache outcome, about 5%, quality mostly the same between them both, though i prefer the first one
>>109485533https://huggingface.co/SexGod1979/PinkFluffyBunny-MiniMax-H3/tree/main
>>109485574so it's just the same lora as the one he has on civit?
>>109485519>linking this in the threads lee checks daily for anything to send license infractions toYou're going to hurt the Chinese person's credit score at this rate.
>>109485092Isn't ref the one that can do everything fl can do, but with more image inputs?
>>109485580It is just for generating better bunnies
>>109485581FL can do multiple references too, just feed it multiple characters in one image and prompt for the scene to change after the first few frames to whatever you want.
>minimax_h3_turbo_4step_ema_ckpt850.safetensors>recommended — current final checkpoint (time-averaged EMA, sharp at 4 steps)Is this in a usable state?
>>109485592yea but dont use comfy converted, use originals with https://github.com/Larryvrh/ComfyUI-MiniMax-H3-Turbo
Wan animate 2 is comming (lol)https://github.com/Comfy-Org/ComfyUI/commit/a464ac33588ae182f81a090d910cfbf21e255b73
>>109485607hey you just reminded me one of the first things i genned with ltx 2.3 was boxxy going "MY NAME IS BOXXY AAND.. THESE ARE MY TITS" and then lifting her shirtgonna remake that tomorrow with h3.>>109485611lmao
>>109485611>based on Wan2.1-I2V-14B
Can the reference given to H3 be a voice? (Can't test shit until the weekend).
>>109485620yes, you can use audio clips as references
https://files.catbox.moe/00a1o2.mp4
>>109485128Link to Morrigan/the succubus please? I didn't find it in the old thread and can't tell if it is one of the expired link or bake issues.
>>109485627kino
>>109485628It's in one of the last three threads if you follow the previous thread link.
What sampler should I use for h3 then?
>>109485470>20steps res_multi/simple on 306012gb in under 5mins with just sageattention and spectrumyou should try kijais new low vram and feed forward nodes after your sage attention patch and before spectrum and see if you get any speed improvement (there is no quality loss)
>>109485626Very nice.
https://files.catbox.moe/l0qcq6.mp4
>>109485492>is there a consensus on the best sage attention node and the best cache node?best sage attention is kijai's mem eff sage attention afaikbest cache node is either spectrum or the utilscollection node. no clear thread favorite yet
>8 steps >turbo lora 850 strength 0.65https://files.catbox.moe/5hn9vh.mp4>turbo lora 500 strength 1https://files.catbox.moe/t8ts0v.mp4I don't like it, they make the video slopped af
>>109485499i chuckled
>>109485654I'm gonna become a master lip reader by the end of the H3 era. I was able to get "read the fucking" from walt and "memes" at the end from jesse kek
>>109485654It's true.
>>109485599>MiniMax-H3 denoises the video and audio streams on two different flow schedules (video shift 12, audio shift 3). ComfyUI's stock samplers step both streams on one schedule, which is fine at ~20 steps but badly over-steps the audio at 4 steps — the audio comes out distorted or blown out. This sampler steps each stream on its own schedule, so audio stays clean at 4 steps. If you load the LoRA and use a stock sampler at 4 steps and the audio is broken, this is why.This is not true though???? Changing audio shift while keeping everything else the same changes audio.
>>109485699true or not it works. Use either the original or 500 one though, the 850 is too overcooked
>>109485639>threeAh thanks, gave up too early.
>>109485711its in here >>109480220
I would not bother with the turbo lora yet. let it cook.
>>109485423The original repo even gives skill files, but those seem to be for specific styles. I wonder how many users know about either of these things.
>>109485423>>109485727lol it's immediately obvious from the markdown that LLMs were used to write the prompt guideif you're writing prompts by hand for H3 that's pretty crazy
>>109485741I assume that LLMs have no concept of pacing, and I like having well paced dialogue so I do it myself
8 steps for 850 step lora seems goodhttps://files.catbox.moe/wfcxps.mp4
>>109485741>if you're writing prompts by hand for H3 that's pretty crazyhaha... yeah...
>>109485758what lora strength did you go for?
>>109485705Ema or non-ema for 500 steps one?
>>109485741>>109485759writing a prompt to hand off to an llm right
>>109485756>I assume that LLMs have no concept of pacing, and I like having well paced dialogue so I do it myselffrom my testing it seems to be alright. gotta remember to let it know the duration of videos you're interested in so it doesnt make prompts that are too long or too short
I think the only way Black Forest Labs can win now is if they forget current Flux 3 Dev and distill a new one from their best model, into at most 22b params not including TE/VAE, add NSFW/Copyrighted data to the dataset, and do their best to actually make a good model to publish in order to compete with H3, everything less than that will just be DOA.
>>1094857631.0, ema
patiently waiting for things to chill the fuck out. in the mean time any ideas on what i should gen in the mean time?
>>109485777thanks anon, did you test both non-ema and ema?
>>1094857781girl, asian, huge breasts
850 EMA4 vs 6 stepshttps://files.catbox.moe/f7089e.mp4>>109485782ema is far better
Has anyone has done any tests on how many languages the model know and how accurate it is at each language?I wonder if it can do like Albanian and stuff like that.
>>109485785>ema is far bettercool, because I've read in the model card that ema was worse at the begining of the training, but I guess 850 steps is the moment when ema takes the cake
>>109485741>>109485770You vastly overestimate the IQ, patience, and skill level of the average prompter.
>>109485756You know you can prompt for the pacing too right? anyways, it's not like you're not allowed to go and edit the timestamps it came up with yourself.
you get what i was going for at least
>>109485785How'd you make that cool overlay? I'll touch tips with you if you tell me
First time trying this on my 10 year old laptop. It's actually not that bad. It's fun to keep genning until I got a style I like.
>>109485794it can do portuguese, so it can probably do most
just deleted ltx and wanAMA
>wondering why the gen didn't use the reference.>I gave it the wrong image.>>>/wsg/6209623
>>109485828looks like GUTS or GANTS whatevert the fuck that animo was named
>>109485828This style is a crazy hack because the gen artifacts just end up looking like they're meant to be there.
>>109485811genuinely cutelove it anon
https://www.reddit.com/r/StableDiffusion/comments/1vhloyz/walter_white_and_the_minimax_h3_official/
>>109485828with the actual reference.>>>/wsg/6209626>>109485833GANTZ? I could see that. this is BLAME! tho. tbtbh one or the other could have easily been influenced by the other.
>>109485826>just deleted ltx and wanbased >>>/wsg/6208736
>>109485870All right, Damn Mr Anon... I just wanted to generate memes :(
my head hurts from how good this model ishttps://voca.ro/1lmz6OUgFbal
>>109485919link to inappropriate girls in question?
>>109485929aww that's cute
>>109485929>shadman
>>109485930>link to inappropriate girls in question?i'm rerunning the prompt i changed it a bit because the LLM embellished with some visual stuff that didnt look good e.g. leaving a lipstick marki also made their outfits sluttier
>>109485919i don't know your source audio but i am actually probably expecting even a bit better from omnivoice/fish/moss/...it's still nice tho, did you get this from h3?
>>109485981i had no source audio, that was from a text to video h3 gen
Flux 3 max VS Minimax H3 comparisonhttps://streamable.com/l47iwi
>>109485616i wanted the camera to stay affixed to her helmet but gave up after two attempts.
>start reading the manual>its starting to make senseoh no, its fucking over for me isnt it
https://files.catbox.moe/l5laxv.mp4
https://files.catbox.moe/81r3tu.mp4
>tfw I'm still running cuda 12 Fuck why didn't you tell me anon
>>109486006just wait until your teacher covers writing, that's gonna go crazy in the special ed classroom
>>109486014>Fuck why didn't you tell me anonGPT 5.6 luna which costs 20 cents per million tokens told me and did it all for me
I have so much to meme I feel overwhelmed..
Should I be using this node? I’ve been on an ealy H3 workflow and it was fine without. And why do people have different values on this node?
>>109486014Cu130 has been out for a year anon. Of course newly trained models and designed opts would need or at minimum benefit from it anon.Do you want a reminder that the sky is blue as well?
>>109486029kek
>>109486016i have never read a book in my life, not even a children's book. i still dont know how the hungry caterpillar endson the other hand i read documentation and manuals back to back if its autistic enough of a device.
>>109486029Lel>>109486031If you have to ask remove it and forget about it.
>>109486000really good gen
>>109486032>Do you want a reminder that the sky is blue as well?I guess I need one :(
>>109486031This node is a curse because once you know what it does you'll be tweaking the values for every single gen.
>>109486036>i still dont know how the hungry caterpillar endshe eats the entire universe
Any good video upscaling workflows?I saw comfyui at openmodeldatabase and want to find a simple video workflow but they aren't around.
https://files.catbox.moe/khcen8.webmVideo references is the best human invention in history!
>>>/wsg/6209638>>>/wsg/6209639
>>109486048i approve any and all doro posting
>>109486048kek quicker than me to post my own shit well done
>>109486045i refuse to believe you and i will not read it to find out if youre bullshitting me or not
>>109486031It's really hard to tune. Only need to mess with it if you're doing low step stuff.
>>109486013>saliva trailkino
my uncle works at BFL and said they just installed suicide nets on the sides of the building.
Interesting, good way to test how it interprets multiple characters in split image. Motion is pretty much identical. I guess I need a much more in depth prompt, very basic one.
>>109486014>>109486032what are the benefits of cuda 13 concretely?
>>109486067whats the benefit of running newer software?
>>109486073More tracking and data collection?
>>109485990oh. that's very convenient then!>>109485934ty. needed to give the cube a better ending.
Is there a turbo lora + lora weight + step count + scheduler + shift combo that eliminates the blur without producing artifacts or make it extremely strongly opinionated?I can't get the turbo lora to work properly.
>>109486062kekd
https://civitaiarchive.com/models/2839513?modelVersionId=3205117The wait is over boys. Finally we have the LoRA we needed to gen actual good shit
>>109486031bypass, forget about it. just enabling it is gonna fuck up the pace of your videos because it does not actually default to 12. using a lower value will cause the actions to speed up in the vid and the longer your vid the more fucked up it gets in the end
>>109486097For a second I thought she was May Li
Why doesn't the OP have a rentry listing all H3 optimizations? Why are you niggers so lazy? They are not like his in /lmg/
>>109486150>They are not like hishttps://youtu.be/H58vbez_m4E?t=111
>>109486150No one is stopping you from writing a rentry. We'll add it to the OP if it's good.
Zoom in and out test for recognizing subjects and memory, very impressive.
Anyone here tried using either of these new director nodes? Or should I stick with basic proompting
>>109486167Those are useless and retarded.Takes more time for something I can literally just write.More so when 90% of my prompts are written by my local LLM.
keekhttps://files.catbox.moe/ft7mhy.mp4>>>/wsg/6209659
>>109486186Quality on par with the latest seasons desu
>>109486167the point of AI is to not do work anon
For AceStep anon>>>/wsg/6209660>>>/wsg/6209661
Is this overkill?
>>109486217Post output and references when you're done and we'll see
>>109486223I'm talking about all the optimizer nodes at the top. I got the from some dude on reddit but it seems like overkill and they could be working against each other on some level.
>>109486217No is not.A good varied character video ref at high res is way better than pics and basically perfect but it also tanks performance like crazy, x2 or more.
>>109486225Was mostly curious how good it took several references desu. Haven't tried with more than 2
big speed gains inc
>>109486217can't wait to see the results, thank you for exploring the frontier
With Krea2, is it possible to gen anime-style images with both a man and a woman where the man has the same or a slightly lighter skin tone? I mean pasty white guy X white woman, not white guy on black girl.Basically what I'm asking is whether there's a way to prompt around the inherent bias in the dataset where the guy usually has darker skin when he's shown touching the woman.
VRAMlet (8GB) and RAMlet (16GB) status? Will I be able to gen funny memes?
Adding a new character while retaining the original style and lighting of the input frame, very impressed.
>>109486256>--yuse-sage -attinitinkek
>>109486271>very impressed.yeah it's really good at inserting a character in the same style as the one from the input image >>>/wsg/6207411
>>109486271Real video of Me in my Apartment
Is the turbo H3 lora only for fl2v model or also supports ref2v?Or it’s shit either way and I shouldn’t use it
So fucking cool. I know sd2 etc can do this shit, but now it's local.>>109486285That's with the reference model? Regular irl photo of costanza, prompted to use the style of another?
>>109486217you decide if you need it, but it can work. there were 9 ref image + prompt demos even with the early access users
sloooooow moooootttiiiooonnnnnn
>>109486031i think a shift_video value of 8.0 makes the movement more rapid, while 12.0 makes the movement more chill.being too rapid isn't a good thing because it can look unnatural and shitty, but being slow can be boring. I've set it to 10.0 now
so whats the duration limit exactly
>>109486323>That's with the reference model?no, I2V, the model already knows costanza
>>10948633515 seconds officially, I've had it up to 30 but it loses all coherency after 40.
Erm... Can't really tell you what's up other than I have a 2 minute reference video for my r2v for one of them, and two parallel r2v happenings.
>>109486347>I have a 2 minute reference videoI hope you have at least a pro 6000 and the video it's scaled to 240p
>>109486370Indeed, took a while to chunk, it was about to be a 100 minute wait, so I stopped that, yes shorter vids.
Which of all the social medias is good for farming ai memes for views?
>>109486312Last time I tried it made the output on ref look worse than without it, so probably. Though it's weird that nobody specifies it.
>>109486346so what youre saying is i need to chain several videos together
>>109486390>MiniMax_H3_00243_.webmah fuck thats a really good one kek
>>109486390Truth social
>>109486393that depends, but a single shot can't be more than 30 or so seconds. I'm sure it's simple to stitch shots together, character consistency is a solved problem
>>109486390x is good for it.For bigger accounts to farm you for millions of impressions while you get none that is.
I don't get it, I'm on 16gb vram/32gb ram, and everyone seems to be running the int8 h3 model. Meanwhile when I run it something hits swap and gen speed falls off a cliff
>>109486439Does your hardware accelerate int8?If this is intel/amd, did comfy merge dynamic memory shit for them?Do you have any fancy cli args that might interfere?Lastly you are a LinuxGOD and using ZRAM, right anon?
Huh, it doesn't know hitler, shame. Easily fixed with ref model.>>109486407Good point, watermark my shit.
>>109485930>link to inappropriate girls in question?i hope you now see why i only posted the audio the first timethis is a pretty common fantasy for highschool girls too
Thread based beyond belief
>>109486454>>109486458these need to go on /r/stablediffusion asap
This model has brought gooning to 2D to a whole new level.
>>109485128How did this stuff get so good
>>109486453>Does your hardware accelerate int8?I think this might be it, apparently the venv runs on an older cuda build that runs int8 in software. Thanks for that.Other than that amd/nvidia, no and I really should be
>>109486467omg yaaaas
assuming youre not using the turbo shit, and arent using sage attention, whats the expected gen time for 4:3 0.4mp thats 15 seconds long?
>>109486485thank the chinese
>>109486315ok so it seems to work, and pretty fast too.3 ref images, 0.4mp (768x640) 15 seconds clip takes 230s on my 5080
Patch Sage attention never works for me. Am I retarded? Flash attention v2 works for me. I'm on AMD
wow, h3 is able to reconstruct impossible 2D hentai body proportion into 3D, which WAN fails to do.https://files.catbox.moe/ovigqf.mp4
>>109486512oops, meant for >>109486392no sage, no torch patching
https://files.catbox.moe/5kyksf.mp4>>>/wsg/6209683
Just a heads up I got kino in the oven.
Feels like my gens are especially low res with ref2v, what do you think?0.4 mp, sage attn 2, res_multistep, 20 steps. nothing else special, just seems like everyone elses 0.4's are better looking.
>>109486585Try beta scheduler
im dying
>not tappingrespect to that alpha male
>>109486603one of best gens i've seen so fartopkek
>>109486589thanks, posting in case anyone else wants to judge the difference.
>>109486523>the thigh pubes
Anyone using this wf?https://civitai.com/models/2834514/minimax-h3-t2v-i2v-ref2v-advanced-filmmaking-workflow-or-all-speedups-qol-featuresgetting a 5sec gen done at 437 secs30906 steps1.4mpusing turbo lora at 1.40 and another loraUPDATEenabled sol-attn and the gen was done at 384.93noticed that the sound is very bad.. I added a nsfw lora after the turbo lora, unsure if related
>>109486665Why is demon Putin giving me weed for free?
Any tricks to getting dialogue timing right? Usually the dialogue starts after the actions that should happen after it in the prompt
>>109486679And apparently he has no midsection.
while i like the speed, turbo lora makes skin look diseased and also adds moles and stuffshift 12/4, euler/beta, 8steps
>>109486703are you using time codes for both?
>>109486709To be fair that's selling the demon look more than a starving African with glued-on deer horns
>>109486735So far I'm only using time codes for shots. I guess I'll try to manually time literally everything
>>109485128Where's that ass video? Can somebody reupload it?
Can you extend videos with h3?
best way to get anima to generate JUST pictures of clothing items? It always wants to put a person in it.
This is fucking insane. Literally grok tier.
It feels like my gens got slower today. Like I'm not loading as many layers of the model into VRAM as before.
>>109486825Mine definitely are after I realized yesterday that my GPU was dying. The temperatures were relatively okay but I hadn’t noticed that the hotspot was going over 100°C during generations because HWiNFO wasn’t showing it for my card. I had to undervolt it and adjust the power and fan settings quite a bit.
testing turbo lora. it seems like the prompt adherence is weaker with the turbo.
>>109486876lol 100c hotspot is finea 5090 has a 130c hotspot at 50% tdp
>>109486886I had some other symptoms before that as well like crashing which didn’t seem to be an OOM issue because it went away immediately after I power-limited the GPU.
>>109486217https://i.4cdn.org/wsg/1786097011798158.mp4
>>109486901that's a blitzball sphere
Also, gimi
>>109486913The sphere was one of the reference images and sometimes it genned it correctly, sometimes not, in all gens it was referenced directly in the prompt.The magic of seed numbers i guess.
is there a proper dialogue format? when I add some "dialogue line" the character then won’t shut up and continues blabbering nonsense
I'd like to use another browser to launch comfyui than my default. Anyone? The edited "Main.py" from Github isn't doing it.
>>109487029there is an easier way to do it but i don't remember what it is.
>>109487009I'm experiencing the same problem. Lots of gibberish. Also hard to control which caracter says what.
>https://civitai.red/models/2840051/speculum?modelVersionId=3205793https://civitai.red/models/2840051/speculum?modelVersionId=3205793>https://civitai.red/models/2840051/speculum?modelVersionId=3205793https://civitai.red/models/2840051/speculum?modelVersionId=3205793IMAGINE THE 20 SECOND LONG VIDEOS MADE AFTER
>>109487029just type http://localhost:8188 in the other browser
>>109487046yeah, gemini solved it.I am sad, at one point we won't have anything to say to each other if we keep asking to fucking robots.
>>109487029Just copy the address and paste it in another browser.
>>109487009Is your video long enough to fit the dialog?Ref video is better for this with a very specific syntax.
>>109486497Stop using "megapixels", no one ever used that term in computer graphics except r-eddit tier normies when they got their first digital cameras.Comfyanonymous (he has literal zero real world experience in digital image manipulation) and his merry folk of slop retards are so incredibly stupid for trying to normalise anything like this.
>start saving my dogshit qwen enhanced prompts>3 million characters in lengthshan't scroll past or read any of that
>>109487087*group
>>109487083Yes but how do i generate a reference video that contains the correct dialogue?
>>109487094Read the guide!!You have to list all the speakers in the video and assign them ids in the speaker section then when prompting them to speak in the video description you have to put<Subject #> (S#) <d>[English] TEXT </d>
>>109487104Okay, I will read the guide, but shit's hard to remember. Lots of rules in that guide.
fuuuck I think GGUF can reach a permanent state where it doesn't load into vram in ComfyUI (instead wants to load into ram).I tried a thousand different things: multiple vram cleaners, vram debuggers, closing lots of programs, restarting the workflow, cutting away at the workflow, restarting the browser, and the ONLY thing that worked was restarting the PC.It's only a 16GB gguf file while I have 24GB 3090 RTX.The reason I'm willing to blame gguf is I've heard the guy in charge of ComfyUI is a little bitch about gguf and ignoring it, so it likely has these kinds of bugs.
>>109486390Instagram reels 100%
>>109487104will try that, thanks. but I think most of this is placebo as QwenVL 32b should be smart enough to handle your garbage prompt with any structured format
Would upgrading 64gb to 128 increase speed or just get rid of OOMs?
Sigma shift? Are we still doing that?
>>109487169>getting rid of OOMslol. Nah. I think there's memory management issues that throwing more RAM at won't solve.
>>109487183We have been doing that non-stop last two years.
>>109487131GGUF is much slower than int4convrot from my experience.
https://huggingface.co/lightx2v/MiniMax-H3-Prompt-Rewriter-LoRA>writing boomer prompts gives the best resultskek
>>109487305
anons update spectrum if you use it cause it's better nowNew default settingsdegree = 1warmup_steps = 1bootstrap_first_forecast = truetail_actual_steps = 1
>>109487104>>109487149Reporting back. This format works, no more additional gobbledygook
>>109487337>soulful video vs corposlop (ours)
>>109487339heres a test:https://files.catbox.moe/u8bua0.mp4
Updated Comfyui and sol-attn and now I just get OOMs trying to generate 17 seconds video using references.>[WARNING] [MiniMaxH3-SolAttn] kernel 执行失败 (OutOfMemoryError: Allocation on device 0 would exceed allowed memory. (out of memory)
>>109487340What is (S#)?
>>109485998So flux is just sloppa generation, all here is slop like fuck and have the same flux girl face
For some reason, gen speed heavily depends on reference resolution even dough resolution of actual gens remains the same
>>109487393# = number so S1... S2 etc which is the Subject
>>109485998Okay but can flux do bobs and begana when it's not even local?
How do I use reference videos? Just the load video node doesn't seem to work, unless it can't take .mp4s or something
>>109487422Use the get video componements node
>>109487385Both Sol implementations are still highly experimental.Can't say much besides roll both comfy and the extension back or wait until it gets patched.
what a time to be alive. 15s in 218 seconds with the new spectrum update and patch sage kj (auto), 0.3mp, no turboThe setting is the TV show South Park.0 to 5s: the 4 characters Cartman, Kyle, Stan, and Kenny are standing together at a bus stop, in front of a widescreen TV that is off. Cartman says "guys, why do THEY control everything? You know what I mean Kyle.". Kyle, in his classic green hat says "no, I don't.". 5 to 10s: Cartman turns on the widescreen TV that shows a chart that says "Jew ownership in Blackrock", with a pie chart that shows "95% Jewish". Cartman says "see THIS is why I can't have good video games, Kyle."10s to 15s: Cartman presses a button and the TV shows a new chart saying "Shekel Shekelberg", with a South Park style rabbi beside the text. Cartman says "they are ruining it all Kyle, I just want to play my videogames."https://files.catbox.moe/fo2psg.mp4
>>109487339>>109487481how much faster?
>>109487337>>109487349for real lol, they just made a sloppifier LoRA
>fresh copy of comfy>python -m pip install triton-windows>3.13.14 python libs and include inside comfy>sageattention 2.2.0 cu130torch2,10.0andhigher.post6did I install it correctly? I'm on 3090
Can someone redpill me on sol diffusion? Is it better than sage? Can or should they be used together? Any side effects?
>>109487490"In that first post I released the MiniMax H3 Spectrum integration and was getting around 34% lower Euler sampling time and 30% lower RES sampling time with the more conservative settings I was using at the time."ive only tested a couple gens but it does seem faster than the old one, update/git pull it if using spectrum, it has new default settings, pick vram if you have 12gb+
>>109483968>>109483968>>109483968
No collage no migrate. Kys, troll.
>>109487536don't migrate to the spaghetti