Best Currently Maintained Edition Discussion and Development of Local Image, Video, and Music ModelsPrevious: >>109627980https://rentry.org/ldg-lazy-getting-started-guide>UIComfyUI: https://github.com/comfyanonymous/ComfyUISwarmUI: https://github.com/mcmonkeyprojects/SwarmUISDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineageWan2GP: https://github.com/deepbeepmeep/Wan2GP>Checkpoints, LoRAs, & Upscalershttps://huggingface.co/modelshttps://huggingbay.xyzhttps://civitai.comhttps://civitaiarchive.comhttps://openmodeldb.info>Tuninghttps://github.com/spacepxl/demystifying-sd-finetuninghttps://github.com/ostris/ai-toolkithttps://github.com/Nerogar/OneTrainerhttps://github.com/tdrussell/diffusion-pipehttps://github.com/kohya-ss/sd-scriptshttps://github.com/kohya-ss/musubi-tuner>Minimax H3https://huggingface.co/Comfy-Org/MiniMax-H3>Krea 2https://huggingface.co/krea/Krea-2-Rawhttps://huggingface.co/krea/Krea-2-Turbohttps://lumenastrum.github.io/clio-style-preview/gallery/>Animahttps://huggingface.co/circlestone-labs/Animahttps://tagexplorer.github.io/https://animadex.net>Kleinhttps://huggingface.co/collections/black-forest-labs/flux2>MiscGentleman’s Guide: https://rentry.org/ldg-gettingstartedShare Metadata: https://catbox.moe | https://litterbox.catbox.moe/Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusionArchive: https://rentry.org/sdg-linkCollage: https://rentry.org/ldgcollage_v2>Neighbors>>>/aco/csdg>>>/b/degen>>>/gif/vdg>>>/d/ddg>>>/e/edg>>>/h/hdg>>>/trash/slop>>>/vt/vtai>>>/u/udg>Local Text>>>/g/lmg>Maintain Thread Qualityhttps://rentry.org/debohttps://rentry.org/animanon
breast thread of breasts
>>109630182>>https://rentry.org/tkdupekk>>109629403>https://rentry.org/ldg-gettingstartedAdded to https://rentry.org/ldg-lazy-getting-started-guide#anon-guides-and-resources
what does everyone think of WAN 3.0?
>only 4Fuck off
Recommend h3 workflow pls, so many different snake oil nodes and tech got released since I first tried, idk what's actually good. Also, is ref2v still much slower than t2v and i2v?
>posted my gens on some other board>people are going out of their way to be butthurt about itwhy is this always the case?why cant people appreciate AI art?
>>109630601just stack 10 h3 cache nodes. saves a lot of time looks very good
>>109630601here you go lad, I know, they're hard to find
>>109630634Some seethe at superiority. Simpleas.
>>109630545very kewl and based
>>109630652and the gens I posted werent total slop either but people still went out of their way saying "this is AI slop", "stop posting AI slop" etc.like what is their issue? Art is Art and people post human art all the time and some of that shit looks ten thousand times worse than my gens.
forgive me for not posting the best vids everI can only do so many iterations with a 5070 ti
Might as well leave "it" here toohttps://litter.catbox.moe/r2zma0.mp4
>>109630601>>109630650I know anon was being a dickhead but he isn't wrong. Any other H3 workflow that you don't build upon or just use from the original are going to be shit and absolutely not for your setup. I tried a couple and they all crash or do weird shit, the comfyui recommended workflow just works. If you see nodes that might help you, you should try them and see if they improve.
>>109630729That's why I asked. I guess I'll just keep using patch sage attention kj and the 8turbo lora and that's that...
>>109630703pretty cool. that's Blizzard tier animation in the palm of your hand.
>>109630740>that's Blizzard tier animationyou don't really believe that do you? you're just being nice right?
>>109630703i will now download your cookie clicker on the app store
>>109630722top kek
>>109630729>>109630601This is sound and solid advice, the comfyui workflows are usually really solid and don't require any nodes, they even have all the download links in those black notes.What I recommend if you want to check out other workflows, especially those that needlessly use 300 nodes to "make the workflow look pretty", make a fresh comfy portable instance and try them out there, see what you like and what you don't like.But for videogen there isn't much you can do rn,>RTX Super resolution upscalingDoes almost nothing, but it does minutely improve image quality, costs also almost no resources or time>RIFE frame interpolationGreat at realistic images, smoothes out the video a bit and causes less sharpness artifacts sometimes, but can cause other artifacts to show up if you crank it up too high. Shit for anime>Comfy kitchen attentionNear lossless video prompting acceleration>Sage attentionVideo prompting acceleration>Cach patcher, lazy cache, etc etcAbysmal, destroys both videoquality and prompt coherence>Spectrum10-20% speedup, questionable if it reduces prompt adherence, try for yourself>FastVideo??? anyone tried that yet?>TurboLorause larry'S 600 turbo lora at 6-12 steps, anything else rn is a memeThat's what I got so far, but a good prompt is honestly the most important thing. And use the correct model for the correct application.Hidden pro tip: change "ref_image_size" to max instead of match and you'll get much sharper and higher quality videos when doing i2v.Hope any of that helps
I really need to start using the reference model
>>109630751well sure. the resolution perhaps isn't there. but the style is there.
>>109630722funny, but what does the Google symbol mean
>>109630542>mfw Resource news08/23/2026>Krea 2 Turbo — 4-Step Distillation LoRAhttps://huggingface.co/lvladikov/Krea2-Turbo-Distill-4step-LoRA>Alibaba to issue US$10 billion in new shares for huge AI push amid strong investor demandhttps://www.scmp.com/tech/big-tech/article/3364957/alibaba-issue-hk80-billion-new-shares-global-ai-push>Nvidia Customers Notified About AI-Related Price Hikes Above 15%https://www.bloomberg.com/news/articles/2026-08-22/nvidia-customers-notified-about-ai-related-price-hikes-above-15>H3 Motion Context Clip Stitcherhttps://github.com/noembryo/ComfyUI-noEmbryo#h3-motion-context-clip-stitcher>Comfyui-MMH3-UltimateUpscalehttps://github.com/bbaudio-2025/Comfyui-MMH3-UltimateUpscale>FastVideo-Minimax-FastH3-Preview-v0.2https://huggingface.co/FastVideo/FastVideo-Minimax-FastH3-Preview-v0.208/22/2026>ComfyUI-H3-AudioRefinehttps://github.com/Adudeguyman/ComfyUI-H3-AudioRefine>Anima-3.8B with Qwen-3.5 4Bhttps://huggingface.co/lylogummy/Anima-3.8B08/21/2026>MiniMax H3 Super Acceleration fast draft generation and high-resolution refinement, powered by Sol Enginehttps://nvlabs.github.io/Sana/Sol-Engine/H3-Super-Acceleration>4DAnyone: Create Anyone in 4D from a Casual Monocular Videohttps://4danyone.github.io>H3 Prompt Composer Version 5.37.1https://github.com/BMB12d3/minimax-h3-prompt-composer>LTX-2.5 Gemma-4 12B NVFP4 for ComfyUI https://huggingface.co/Deadshot699/ltx-2.5-gemma4-12b-comfy-nvfp4>MiniMax-H3 Pruned Ref-Delta Fused r1024 — ComfyUI Single Filehttps://huggingface.co/xmarre/MiniMax-H3-Pruned-Ref-Delta-Fused-r1024-ComfyUI>Phosphene 4.6.0 Adds Video Editor, H3, LTX 2.5https://github.com/mrbizarro/Phosphene/releases/tag/v4.6.0>H3 Optimizations: Standalone production optimization nodes for MiniMax H3 in ComfyUIhttps://github.com/Zironic/H3-Optimizations>MiniMax H3 known character listshttps://huggingface.co/datasets/malcolmrey/various
>>109630887gemma-chan is a popular 4chan mascot for the google made gemma models
>>109630722Also this was the voice I had in mind when doing the video, but it wouldn't actually change it via prompt so I now tried with a 3 sec voice sample of heavy. Unfortunately the video itself came out like shithttps://litter.catbox.moe/6ybyfs.mp4
>>109630887it's one of the symbols of our overlords here in america. it's a zionist psyop to make you love google.
so why can't I use the ref model for t2va?
>>109630912cute and slightly erotic, nice
>>109630703what are you're computer specs and average gen time?
>>109630912why does she keep shruggingshit is getting on my nerveswho even makes this bullshit
>>109630982unironically indians
>>109630951desu such a waste of electricity Minimax is amusing but keeps fucking up the detailsI need to go back to making images
>>109630993>everything was troons>now everything is indianslmao monomaniacal anons life is simple
What model (and/or workflow) are you all using for style transfers?I've been trying to take a reference image that contains the artstyle, scene, composition, expression, etc. that I want and replace the character in it with a reference image of a new character with Qwen edit 2511, but have not had much luck
Where is everyone :(((((
>>109631045reporting in
>>109631031maybe try the krea2 edit lora (on Civitai) I've seen some anons here use it, haven't tried it myself.I found some workflow on youtube with some boomer and cope nodes and it kinda worked to an extent but comfy update broke half the nodes.Anyway Krea2 was built for style transfer so it's the right model to use. I'm thinking that edit lora is a good option.
>>109631076I just recently tried it and it seems broken rn to me
i'm half jewish half indian and trans.why are people here so mean to me?I just want to locally diffuse!
>>109631082have you tried it specifically for style transfer?
>>109631045it takes time for me to come up with ideas for gens, anon
>>109631112this, people think it's all sunshine and rainbows until they have to proompt for themselves
>>109630995desu big models are more fun when you have better hardware
>>109631140I passed on getting a 5090 because I would have to buy a new chassis and they seem energy-inefficient, and 32GB probably isn't enough either. You would want the Pro
>>109630925You can, if you want slower gen times and worse quality/acting
>>109631146who hangs their shirts in the kitchen wtf
>>109630740fun fact blizzard outsourced the cinematics and didn't make them. I mean they were fucking insane what would you expect
>>109631206they didn't always do that. all the OG great world of warcraft ones are by them
>>109631112For me the biggest limitation is how much time it takes to generate a good quality video without random issues.
>>109631227Spending hours trying to figure out why the thing clearly described in your prompt isn't showing up.
>>109631093noI was getting broken outputs (some weird artifacts etc.)
cozy breas
Qwen3.8 27B reasons way too much for simple generation prompting with the default reasoning effort (xhigh). Try setting reasoning effort to medium. and note that low is not good because it often uses even more tokens than medium. The default xhigh is good for outputs that require a lot of reasoning like coding but for prompt gen it will just slow everything down.
>>109631263I see. sad. I was hoping that might be the one thing it might actually do. oh well. I guess we still have to stick to loras for now.They're time consuming and they suck the life out of you and ruin all the fun of diffusion but at least we know we have them if we really need them.
>>109631296do your own research though. my setup might be fucked, not the edit workflow. I'm very new to krea, been only doing videogen
>>109630601plaguekind on civitai>>109630995undervolt
>>109631296>>109631303The edit lora is great for changing one image, but somewhere between mediocre and lukewarm for style transfers using TWO images
>>109631221Oh neat, things used to be, in fact, not lame and gay
>>109631293wait a minute. there's a new qwen? Is it an edit model?
>>10963115432GB is plenty for doing this stuff as a hobby. The pruned Minimax models fit into VRAM comfortably, with lots of room for KV cache and the like. The RTX 6000 Pro has a lot more VRAM, but is not ahead by that much in terms of compute while being multiple times more expensive.
>>109631303>>109631296As far as I can tell they only make the style transfer available on cloud Krea 2 officially. There are custom user made nodes for style transfer but I haven't tested any.
>>109631326thanks for the feedback. i guess it's a no go then because I already use qwen and klein for image editing. Come to think about it, I haven't really tried transferring styles in those models. Maybe that's something to look into.
>>109631227with better hardware that gets easier
>>109631346yeah I've tried some custom nodes like I mentioned earlier. They get it right partially. They change the rendering style but not the character style (propotions etc)
>try to reuse Krea 2 wf from old image>it didn't save the text from the Generate Text node>generating text again gives a completely different result on the same seed>but it saved comfyanon's furry demon biker 1girl oc prompt for some reasonWhy doesn't generate text just save the text by default
https://files.catbox.moe/cjw2tq.mp4
>>109631425we're not letting go of this one huh, can't blame you either>>109631429is that H3?>>109631357Also I tried out Krea2 style transfer to test it, but semi cheated by using a base krea2 image to edit, since it already has knowledge of what it created if that makes sense.Here are my results, not sure if that would satisfy your needs
This is way too much funhttps://files.catbox.moe/vnd0me.mp4
I'm trying to get H3 to point toes inward which is an extremely common posture for women having their photo taken and it just won't fucking do it. I can type all this fantasy shit and it's like YES MASTER. But if I want dime a dozen celebrity red carpet trope its UARRRR I OINT UNDESTRAN MATE
>>109631377in the future this node is gonna save your life https://github.com/pythongosssss/ComfyUI-Custom-Scripts#show-textits like preview text but it saves it properly so when you drag the image the exact prompt is visible
My machine isn't strong enough to run anything good. I saw you can rent compute from places like runpod. Anyone have any experience with that?
>>109631460>is that H3?Yes ref2va with a forced start segment to copy voice tone that's trimmed from this since the video quality diverged from source right at the inference split.
With the effort you took to solve a captcha to confirm you are poor as FUCK, you could have placed in a job application in the same effort.
>>109631490kekd
>>109631463Classic porn plot.
>>109631460>we're not letting go of this one huh, can't blame you eitherIt's actually from earlier, just forgot to post it.
I miss my Baldachin's Blessing.
>>109631463>I'm sorry I couldn't find the milk potion anywhere>I asked you to find a book
>>109631531Great tummy
>>109631474I was considering that just so I don't have my card cooking for hours on end. I don't know what the best gpu-for-rent options are tho sry
why is krea so bad at nsfw anime content? i dont want to stack a ton of loras just to make it not suck ass
>>109631641>why is krea so bad at nsfw anime content?i lacks one of the best repos of nsfw anime content in its dataset simpleas
>>109631641>i dont want to stack a ton of loras just to make it not suck assKrea2 in general is dogshit at nsfw if you're looking for intricate poses or something like that, that's not an anime specific thing. If you tell me what you were trying to gen, I can give a go at it to see if it's just skill issue
>>109631641Real Labs are scared of making their models good at NSFW without a lot of legwork from the user.
I'm getting sick of gemma QAT dropping the ball is qwen better?
>>109631772don't use qat use normal gemmaqwen isn't better it's different
>>109631649weird how the quality is shit in [shot 1] but then surprisingly good afterwards, did you use multiple references or is that H3's doing?
So after using qwen and grok 4.6 all day, it turned out the anons in previous thread were correct, and penetration was not working well due to several things:1: Prompting language is beyond important for h3. You really need to dig down and describe things down to the second (and it will stick to them).2: Shift video actually matters a lot. If you want slower and more detailed penetration movement, dropping down to 6 or so can have a huge impact, especially if you want it to animate on two's/three's anime or western cartoon 2d style.3:Attention and XXX loras seem to work wayyy better at around 0.50 strenght vs even 0.70 or worse yet 1.00 (At least at 25 steps). Also in the lora selection, the 2nd option (VIS) does not work like how it does on image models where in a lora the two settings have to match, and its better to have it stick to 1.0 even if you drop the model strenght to 0.50 or whatnot.
>>109631793wasn't qat supposed to be better for anything under q8?
>>109631822I haven't used the shift video node at all during my gens and got great results (didn't try nsfw yet), can you give me a QRD of why I should waste my time with it?
>>109631824from my test it's almost systematically shittier despite what google claimed, and it seems to be consensus in lmg
>>109631834For example i tried an animation where I specified that there are 4x thrusts within 5 seconds. When it was at default 11 or 12 shift video, the animation instead tried to fit 5-6 thrusts to match the timing of the total lenght even when i specified otherwise. When i dropped it down to 9, i noticed it seems to often (but not 100%) have one less thrust than at 11. Dropping all the way to 6 made the thrusts match the timings i specified. Dropping bellow that made the animations sometimes almost (or at times did) stop. If you animate 3d style it might not be as important, but for anime/2d style being able to slow down the animation like this matters a lot.
>>109631822So are you prompting like this:at 00:01.000 this happensat 00:02.000 this happensat 00:03.500 another thing happenswithin a single shot?
>>109631860Yea like this: Continuous slow sex with no pauses between strokes, only four thrusts total, each a smooth circular grind rather than a straight piston. Thrust 1 from 0.00 to 1.25: his hips sink and roll in a continuous circle from her right toward her left while her hips roll up to meet him and seat fully. At 2.50 seconds her mouth shape changes from a grin with teeth to an open mouth with tongue visible. Thrust 2 from 1.25 to 2.50: same continuous right-to-left circular down-roll without stopping. Thrust 3 from 2.50 to 3.75: same. Thrust 4 from 3.75 to 5.00: same circular roll, finishing fully seated. The four strokes blend into one unbroken grinding rhythm. Her breasts move with body weight and a soft lag behind each roll, natural weighted bounce only, no exaggerated jiggle, no wild flopping, no flying droplets. Small collar-bell sway follows the roll. Wet shine and smear around the condom and contact point.
>>109631869Wow ok. And do you find that telling the model what not to do, e.g. "no wild flopping" actually works? I've tried that and it seems to have no effect or even the opposite.
>>109631848Extremely interesting, going to try it out for myself.But I do have a question, can I just pack the ModelSamplingMiniMaxH3 node between the basic guider and the load diffusion model node?Or how did you do it, cause I think I already have one built into the default MiniMax H3 Reference to Video node that sets it to 12, hidden and unchangeable
>>109631464found out this pose is just incompatible with T-posing, not even fake T-posing like "raise arms out" or "lateral raise".
>>109631875It has for me at 25 steps. I've had over 30 generations where the dick didnt flop out of her at random using the settings i mentioned above.
>>109631882I'm using this setup and it uses comfykitchen (just like 1-2% slower than sageattention but you dont need to install all the nvidia cude stuff).https://civitai.red/models/2831978/dasiwa-minimax-h3-workflows-or-t2va-or-fl2va-or-ref2va
>>109631934>Civitai workflowsJust checked it out and puked a bit in my mouth, so much stuff in there that's useless, still thanks for the effort. For anyone else wanting to try out shift_video I figured it out, you simply have to put the "ModelSamplingMiniMaxH3" node after your last Lora/ComfyKitchen or other patcher nodes and before the basic guider and scheduler, then you can play with it.And holy shit the effect is immediate, this is a good node to play with for 2d/anime animations
>>109631966Anything specific that you find useless in it? It seems to generate on my 5080 at around 7.2s/it for a 5 second clip which is pretty much the speed of a barebones h3 image2vid.
>>109631979All of the KJnodes that also install that hideous laggy interface obviously, but that's just personal preference, I dislike node bloat when they're clearly not needed.>Cache nodes>fp16 acc.>upscale 2x using meme method>upscale with image upscale modesso fucking useless, shoot him for that one>upscale rtx with a meme node instead of the original one>watermarkThese are imo useless features that are bloating the workflow and make many of the things hard to understand.I'm not directly saying the workflow is bad or that it produces bad results, simply that it's disgusting to look at with a bunch of useless stuff, while hiding important things like "reference settings" seemingly away, or I cant see it without the nodes
>>109632024make a pr that deletes all the useless shit
>>109632024Ah i see. I have no idea in regards to cache nodes and whether they do anything for h3. I might try making a clean workflow while still using the loras and kitchen attention.
i pulled.....
>>109632034I already have my own workflow so I don't really need to take another one and change it.>>109632041Like I said, the workflow is perfectly fine in function as long as you don't use the useless stuff, you don't need to change anything if you're happy with the results and you're not agitated looking at that or by the fact you're potentially loading a bunch of nodes you don't really need each time you start up comfy.I'd say if you like the workflow, keep it as isAlso it's hard to compare gen times but this 1MP 10sec video took me 442sec on my 5090, I've set shift_video to 6 as a testhttps://litter.catbox.moe/yxlwwd.mp4
Now that the dust has settled, what is the best fl2v Turbo lora for H3?
8 steps turbo lora, 11 minutes 1 mp, not badI hate agorist fags and zogbot libertarians btwhttps://files.catbox.moe/b22ko5.mp4
>>109632079What did he mess up this time?
>>109632099>https://huggingface.co/larryvrh/MiniMax-H3-Turbo-Lora/tree/mainThis one of course.
>>109632119video extensions
>>109632123>Big dick Larry 600yessir
>>109632088Man i wish i got a 5090. I got a 5080 standing in line at launch and couldve taken a 5090 but i figured id never need it. Now the damn thing costs 5k, and its not like i coudnt afford the 2k at the time...
anyone have a good video ref 2 va wf?What should the input video resolution be?I used 1M video force 24fps, and the gen time was terrible. Forcing the resolution by half seemed to be better and I really didn't see quality loss
>>109632123Which one of them?what about ref2v?
>ref2v takes a fucking long time>i2v doesnt give sufficient enough results>fl2v i dont really care for, no use case for my slop>lf2v is interesting and no use case for my slop but ill try it for fun>t2v is not coherent enough but kino for insane chaotic slopim tired bros
>>109632137comfy org has a very simple workflow that works
Does the avatarfag that hates Ani and Debo still here?
>TFW someone smarter than you goes out of their way to explain shit to you in a really concise step-by-step way and you're still too stupid to get it.H...haha...ha..
>>109632146>https://comfy.org/workflows/b34841f6789c-b34841f6789c/this one? but I'm looking for video ref, not image. I assume they are the same but video ref is different beast for me.
>>109632132Besides the extra Vram I don't think the difference is that big, the actual regret is not getting a rtx 6000 pro with 96gb back when it was still 6-7k>>109632143Try spectrum if you're not using it already, might cut your time down by another 10-30%, haven't noticed a drawback yet>>109632137>>109632146This, try that one first then expandAlso, same prompt same seed same everything but shift video 9 this time (took only 414s this time, probably because models were loaded)https://litter.catbox.moe/173ec8.mp4shift 6 comparison:https://litter.catbox.moe/yxlwwd.mp4
>>109632162>igputhe rest is solid advice but it depends so much on which igpu and what resources that fucker is using. you dont even necessarily need those specific methods to set it up. the ck attention is useful to me though but whether its better than flash attention is hard to say
>>109632172Anon, comfy has its on workflow browser. Nothing external
>>109632180I'm a retard. I think I'll finally be able to use local AI when it gets made into a single executable.
>>109632175>Try spectrumi have it already but i do not like the way it sacrifices quality for speed
Scail 2... whatever happened there?
>>109632198im retarded too, im using amd. https://github.com/CS1o/Stable-Diffusion-Info/wiki/Webui-Installation-Guides#amd-comfyui-with-rocm this was written by a german ESL, and has some bad choices in my opinion but it works and theres 2 methods, one is the "download a fuckin zip and run it" and "if you want the speedups"
>>109632210>H3ppened
>>109632210It was great, been replaced by h3 though.
>been checking for months to see if someone backed up a lora on civitarchive>nothingim going to have to track down the guy who made the lora and ask him for it.https://civitaiarchive.com/models/2454555?modelVersionId=2782555
>>109632232>>109632230can H3 replace people like that?
>>109632192nigga, read, I don't need image ref 2VA. I need video ref 2VA workflow. They are kinda the same but not really; if I put 1M input video, my computer'll explode.
>>109632247you are too dumb to gen a video
>>109632238Have you been living under a rock?>https://files.catbox.moe/d6ebyp.mp4
Ok now I got a full setvideo_shift 12:https://litter.catbox.moe/279cz0.mp4video_shift 9:https://litter.catbox.moe/173ec8.mp4video_shift 6:https://litter.catbox.moe/yxlwwd.mp4very clear difference>>109632247>nigga, readIronic that the retard who cannot read is telling me to read. I have to assume this is either bait or you're mentally challenged
>>109632238Yup.
>>109632172>>109632247just add a Load Video (Upload) node to ref_video_0I uploaded the image ref workflow to chatgpt and it did the rest
>>109632271>just add a Load Video (Upload) node to ref_video_0anon...I did that. I'm trying to say what the difference between 0.5 and 1M video inputbecause I don't see much quality change but gen time is massive
>>109632225I took a look at it. There's no command line for my GPU. It only goes up to 6900, I have a 7900.
>>109632310whats your gpu
>>109632310Those instructions are for older gpu models, the support you need is baked in already. Read it carefully
>>1096323187900xt.
>>109632323Now, In my defense, I did say I was retarded.
>>109632310Have you tried literally just downloading the comfyui folder made specifically for amd and running the "run_amd_gpu.bat" file?Or is your problem that you want all of the speedups like sageattention triton and all this other bullshit?
>>109632333Just follow the instructions there and you can skip the part that isnt relevant to you. Just make sure youre prepared for issues somewhere
>>109632341>>109632352I basically am just paranoid, and just want to make sure I'm not skipping something. This will install the thing for me to play around with it and shit, right? Or am I completely missing the part where I had to install the thing and this is just the GUI for the thing and I needed to install something else prior?
>>109632362my man just fuckin read the instructions its that easy
>increase reference images from 2 to 3>the time per step increases from 2 minutes to 10 minutesnani?
>>109632368Oh my goodness, the new frontier. I have no idea what I'm doing but I'm in.I'm probably slightly less retarded than I'm putting on, but to be fair my confidence is at 0 generally as a character trait so I am ALWAYS afraid I'll blow something up with everything no matter how thoroughly I read the instructions.
>>109632362To test it you don't have to install anything.Simply download "ComfyUI_windows_portable_amd.7z "unpack it... then double click "run_amd_gpu.bat"You'll instantly see if you can run comfyui or not, there is nothing to install and you're most likely just being retarded rn>>109632372maybe your third reference was much higher in resolution than the others, or you're running out of vram, check in task manager
Is it bad for my GPU if I generate batches and leave it running non stop for many hours? I only got into vid genning since Minimax was released. Will vids destroy my GPU
>>109632379yeah the images are 4kI'm genning at 0.6
>>109632379Yeah I ran it, works and all, see >>109632378Thanks for being patient with me. Now that I'm in I'll try to figure out how to use it from youtube tutorials.
>>109632384your gpu will die in a year
>>109632384Throttle or cap its power and make sure the temps stay cool.
>>109632378you know maybe i was wrong to not go to school and be truant, maybe i was wrong to be terminally online and learn most of the shit i know from a computer for 20 years, maybe i was wrong to not go to college (i was right on not going into debt though), but honest to god, taking out computer literacy courses in schools is the biggest mistake we ever made
>>109632385Have you set ref_image_size to "match" to avoid throwing 4k images at your puny 0.6MP video?Alternatively try downsampling them with a "scale image by" node down to half and see if that fixes your times for you. Super alternatively try restarting comfyui, that clears the vram cache and also could solve your issue.Omega alternatively, try putting two references into one picture with a cut out background and describe them well.>>109632384yes GPU's are made for roughly 2 weeks of genning before you notice a 30% decrease in performance>>109632392Now the sky is the limit, don't forget to check out civitai.red (nsfw) to see what models are currently the latest and greatest and try something out if you don't already have something specific in mind.Also I recommend the >>109632192 template browser, it usually has a workflow for everything and massive giant notes that explain everything
>>109632401I think it's a matter of curiosity, I am thankful I have been curious enough to learn everything I needed to know and more (except being social)
i see a new lora just released for something i like
>>109632401On average, though you might not believe, I'm actually more tech-literate than an average person. It's just whenever it comes to the coding side or anything related to it, it turns out as I said in that post 0 confidence is a hell of a thing.
>>109632414have you tried reading to get started like the rest of us hopefully did? im not asking you to setup a networked linux setup that you remote into to make slop, but i am expecting you to read basic instructions. you will have to let go of being spoonfed or handheld eventually. the only reason i decided to help a little bit is because you made the unfortunate choice to go amd like i did, and that makes you automatically in hard zone for this shit.
i dont trust any lora for h3. i believe any lora attempt should feed data in the same prompt structure and detail as regular h3
>>109632438Not the same guy. It's the person that just made an anecdotal remark. I already got it working and said thank you for the help, I think postmortem frustration with me might be a bit excessive but it's not like you didn't earn the right. Still if you missed me saying it, thanks for the help.
>>109632460i deal with retards that cant open zips all day and theyre like 25. my frustration comes from a very important place in my heart
>>109632468Honestly it wasn't even an issue of a zip file. The few times I dealt with github files in the past I had to download the software and the GUI separately, so when the AI itself is called "ComfyUI" in my mind it flagged as "This is the GUI, but it doesn't say anywhere how to get the core files." So I felt like I was coming into the final phase of a process I missed the initial steps for.
>>109632447certain motions don't work for t2v unless you have the lora for it
>>109632411>>109632400temp is 73C after several hours. I think my 5070 Ti has good longevity because of lower power draw
please stop botting this thread
whats the point of civitaiarchive if i still need to log in to civit to download something?
>>109632489>not getting the 5090isnt like 8k right now? i shudder to think what the 6090 will cost if they even bother to make one.
>>109632489use ref2v turbo lora, and dont generate at anything crazy like 5mp4080, 16gb, at 8 steps the gens are good and take a few minutes tops, but use the lightx2v ref turbo lora at 8 steps not 4 (4 is shit audio/blur)
5090 is a gaming card, and they get hot af, its sad that this is our best option
>>109632259Which of the shifts did you like best?
is this considered nsfw for 4chan blue boards? I know this shit is allowed on instagram which is for kidsif it's nsfw I'll delete it okay?
Be thankful that your RAM sticks don't have bad cells.
>>109632521>nipplesfuck off
>>109632513shell out 16k for the blackwell then
>>109632494No need to be shy anon, here is some more.>https://files.catbox.moe/kocgxx.mp4
>>109632526what's wrong with nipples?
>>109632520Shift 12 is the only one that stayed coherent.Shift 6 had a really nice spin.It's clear that in this case the regular setting of shift 12 is best, but different videos may utilize the increased motion
>>109632510140s for this test at 0.6mp, same setup, 8 steps totalhttps://files.catbox.moe/nhygn1.mp4
>>109632536what part of "blue board" and "sfw" are you not understanding
>>109632541but it's allowed on IG
>>109632542wow thats great this is 4chan not instagram. fuck off right back to that god forsaken normie site
>>109632528>https://files.catbox.moe/kocgxx.mp4Pure comedy
I'd say that's a really solid result and time, I recommend checking out the larry lora though and comparing it to the og 4step lora, might improve results further
>>109632542IG is not a SFW site. Just because kids browse it doesn't mean much.
>>109632548I don't use it, just got images from there
>>109632548>fuck off right back to that god forsaken normie siteinstagram is less normie than 4chan at this point
>>109632554>i dont use it>admits to using it to get images>frogpostinghonestly just nuke the internet, put is back to 1954
>>109632538>>109632551missed the reply button apparently
>>109632561what's wrong with frogs
>>109630523>ldg lazy getting started>stolill links to the heavily outdated "Local SOTA Models Meta" as its first resource, no update since June 2025>still linking comfy 1girl guide, which doesn't even mention anima
>>109632539>>109632551>>109632565Am I completely retarded or being trolled, sorry for the spam but this reply for that post...Also the thread schizo finally mentioned the guides again lmao, don't engage him
i heard minimax is super fucking slow even on a 5090, how true is that? like a few minutes or 10-20 minutes?
>>109630982>goes to ai board>gets angry at gacha rolls
>>109632573my gens are taking about 3 1/2 minutes for 15 second clips at .5 MP
>>109632494Not him but dryhumping is my ultimate fetish and this is doing it for me
>>109632550>Pure comedyJust like any other hentai.And this already has better quality than Queenbee could ever hope to have.
>>109632578i see, mind catboxing a gen so i can yoink the minimax WF you're using for that result? i'd like to try out minimax
>>1096325785090?
>>109632578that's piss btw
stop saying mp. say something like 720p instead
>>109632162I am that Anon; glad you saw it at the end of the thread.Here's a quicker alternative install method:1) Install latest AMD Adrenalin2) Download and try the latest portable AMD release of ComfyUI:https://github.com/Comfy-Org/ComfyUI/releasesorhttps://docs.comfy.org/installation/comfyui_portable_windows#amd-gpu>Double click run_amd_gpu.bat to launch ComfyUI.>Normally, ComfyUI will automatically open your default browser and navigate to http://127.0.0.1:8188. If it doesn’t open automatically, please manually open your browser and visit this address.That lets you quickly try out how well it runs on your system.The only thing is it's on a slightly older ROCm version (7.2), but it still works. The one advantage of 7.14 for me is that it makes cudnn/MIOpen worth enabling for VAE decode* on my machine, instead of crippling. If you don't use COMFYUI_ENABLE_MIOPEN = 1, then I think there's not as much difference.(Oh, and it looks like ComfyUI enables dynamic vram by default for ROCm 7.14 now, so that saves me from having to type --enable-dynamic-vram when launching.)*: VAE decode is the last step of generating, and is basically like an upscaling step. It makes VRAM usage spike, which can sometimes freeze your machine if your pic is too big. Enabling MIOpen greatly reduces the spike (after recent bugfixes; YMMV). Another workaround is to use tiled VAE decode to split up the upscaling job.>>109632180True, I can't say how well certain settings will work on other AMD hardware. Generally I'd expect newer hardware to have better support. I'm on Strix Point, aka gfx1150, a Ryzen AI HX 370 with 890M.
>>109632578>3 1/2 minutes for 15 second clips at .5 MPYeah but it also look ultra deep fried from all your cope nodes.
my GPU shows 0% work and 70 degrees when genningwhy does it show 0%?Also how do I prompt in minimax to swap clothes?Every time I try it swaps the girls as well
I think I'm starting to get away with just 4 steps now. Results are less sloppa 6-8 steps are better, but it's 100s/it already.
>>109632610That's what got me interested in trying, and a couple very patient anons guided me through absolute retardation on my part, and I got it working (Me) >>109632378 >>109632480 Now i'm here in the blank UI. Trying to figure out how to start getting it done. Seems all the templates require me to download massive files and I don't mind but I honestly don't know what I'm doing. I'll just work on it slowly with time now that I have it installed.
>>109632610>890M.Im so fucking sorry for you anon. Like genuinely.t. 7900 XTX and 7700S owner
>>109632591i would but every time i upload to catbox it tells me invalid uploader
>>109632601yeah 5090
>>109632591https://litter.catbox.moe/vs8acw.mp4
>>109631429This one is excellent
>>109632633thanks
>>109632641yeah, does it load in comfy like that? i didn't realize you could load an mp4 as a workflow?
I wonder how Anima even knows this artist @monkechrome
>>109632659yeah, it loaded fine. mp4 files can contain metadata so there should be no issues.
>>109632664For being trained off a robot world model it does pretty good.
my magnum opus got ruinedgood night
>>1096326220.5M with 6 steps,50s/it is also alrightsave about 2':30s vs 1M 4 stepsBut there is video-audio desync somewhere; I'm to tired to figure it out
>>109632608>720pAnd how many pixels is that? 100x720? 320x720? 900x720? 2156x720? Saying 0.5mp means it's 500000 pixels in any aspect ratio.
>>109632624You're on the right track with the templates, and yeah, model files are big.A few suggestions:>In the upper-right of the templates browser, you can click the "Runs on" dropdown and check "ComfyUI" to filter out templates that run on the cloud.>The current state-of-the-art are something like Krea 2 for general images, Anima for anime gens with booru tags, and Minimax H3 for video. For editing, Flux 2 Klein 9B KV. Z-Image Turbo is also good for general images, and I think that's the default template that shows up when you close all your open workflows.>SDXL-based models are older, but run faster (about 3x faster than Anima). For anime, there's Illustrious-based variants like NoobAI or WaiIllustriousSDXL (for easymode slopping; it's how I started).
which disposable email lets me make discord accounts? i found the commit that brought a regression but i have to go into their discord server to report it
>>109632695Just use your email and phone numberIt's not like they aren't keeping tabs on you anyway
>>109632693>and Minimax H3 for videoOh yeah, that's the first one I saw and figured I don't know what the fuck I'm doing so I might as well just grab the first thing I see. It's downloading a bunch of files now.
>>109632699no thanks glownigger
>>109632684What's your prompt for this?I haven't had success with replacing/editing with h3 ref2v. Often I just get the video ref regenerated on its own.
>>109632681CRUNCHCRUNCH
>>109631869don't know if you're still around anon, but i just wanted to say thanks for clueing me in to the importance of using timestamps.
>>109632711 It's a video edit workflow. Original video from /kpop/ on /gif/prompt is from LLM and is not very good. It's too long to post on 4chin. You tell clankers to replace <subject 2> from <video 1> with < subject 1> from <picture 1>
>>109631822>>109631869Why can't we just type in "humping, pumping, semen dumping." and have ti work? The world is not fair.
>>109632758It's Christian Helmsworth
>>109632762*T'Chaka O'Dinson, the white ape
>ensure the first frame of the video is <Picture 1>.editing fun:https://files.catbox.moe/m4ea5t.mp4
sick to death of the word "then"
>>109632823anon what the fuck why did you post a photo of me
>>109632736No problem>>109632759That's ideally what loras are for. Or someone would need to fuse the core model with nsfw that has those trigger words for poses.
>>109632840replace it with a comma
>>109632840Didn't you learn anything from Trey Parker? You don't write "and then" you write "therefore" or "but".
>mfw made the perfect pov cowgirl r2v workflow with 1:1 ref accuracy, complete with fast plapping sounds and a thumb in mouth [shot 2] without crunching granola or slurping ramen sfxits literally over for my benis
>>109632909Try finger and then therefor try but whole.
>>109632968i'd like to make a fully t2v pipeline for a long video, but the massive generation times are too risky since the first video in the chain could very well end up being a dud
No doubt I'm stupid but I downloaded all the models that were missing but the thing still says missing models, how do I, you know. Do I put them in a specific directory or is there a menu option to direct it to the files?
>>109633006you need to mount the files to your tensor environment
>>109632968>and he won't share it either
What is better for H3 prompt writing, Gemma 4 Heretic or Qwen3.8 Heretic?I heard Qwen is more suited for agent workflow?
im bored, give me ideas
>>1096330481girl, standing,
>>109633006The template has notes on the left that show which folders the model files go in.
>>109633051damn... thats a good one
>>109633052she's got nice perchlorates
>>109633052imagine the smell
>>109633006>>109633056Oh, and press the "R" key to refresh ComfyUI's file detection. It should be able to see the new files once you do that.
the kinoplexatorium will be open shortly
>>109633112i see you fucker tell robert to eat a dick
>>109633128who is robert?
>>109633056>>109633104You're a champ. As someone who's spend decades in tight-knit hobbies, I know how frustrating it is when faggots like me come in and can't figure out shit you've explained to motherfuckers a million times already. I fully appreciate your help.
>my gens are too stiff until i use the shift node>when i turn off the shift node, no matter how i prompt it never comes out the way i wanti hate this
>>109633148Bet it's not the only things that's stiff, eh champ, what kind of filth you generating?
>>109633171one man standing in an open field yelling slurs at various looney tunes characters
>>109633177Oh my how risque.
>>109633181its very important how well animated they are to me, please understand. i need like better than roger rabbit style mixing
>>109633136I just hate how much frequency this place is just comfyui tech support. The software is honestly shitty garbage. The only thing that makes it good is the backend speed compared to the rest of the diffusion uis and it pisses me off how unstable the entirety of the space is when it comes to this stuff. It's not your fault it's just that the entire space is run by script kiddies
>>109633187Oh dude, I'm not disparaging you at all. I'm just sitting here not even knowing what to start typing in so what you're doing is already magic to me.
>>109633190I get it. Growing pains of emergent tech and shit. For what it's worth, I'm sure it's just a matter of time before a lot of this shit gets consolidated and standardized and make the community mature a little more and spend less time being groundhog day QnA.
>>109633192type something kino in
>>109633192absurdities work usually
>>109633112i am nearly seated
>>109633200
>>109633218punch it
>>109633218i will be waiting for the results
>>109633218we're about to witness true greatness
>>109633219>>109633222>>109633230Still at 0 percent so maybe not? How long do these things usually last?
>>109633242we're gonna be waiting a while for you since you have no idea what youre doing
>>109633248I do not deny this. I am as clueless as clueless gets.
Just realized that if I want to get into creating a story with H3, I'll actually need to write a script.
>>109633039Anyone?
>>109633136Cheers!Some other random tidbits:>Reusing the same input words (prompt) can generate either an identical output if you use the same seed (a random number, see attached pic), or it can generate a different result if you use a different seed.>ComfyUI saves the whole workflow (including prompt and seed) in the metadata of the output file. If you drag-and-drop the output file into ComfyUI, it will load the embedded workflow that was used to generate the file.>4chan scrubs this metadata from uploaded images. People sometimes post catbox uploads of their gens if they want to share the workflow they used. (Or if they just want to share a video with sound, which most boards don't allow.)>Getting back to seeds: the default Minimax template has a noise_seed field with a fixed value typed into it (looks like 168866841893410). Other templates may have a separate seed node (see attached pic) with a randomizer setting, so you can rerun the same prompt and get slightly different results each time.>The text saying "control before generate" means that when you click Run, the number changes BEFORE your gen starts. I think the default is "control after generate".>Control BEFORE is arguably more intuitive, since if you get an almost-good result that you want to tweak more, you can change from "Randomize" to "Fixed" and rerun the same seed easily with a modified prompt. If you use control AFTER, you have either ctrl+z to get the previous seed back, or dig it up from the output file. This will make more sense once you play with genning more.>Anyway, the setting for Before/After is found under Settings -> Comfy -> Node Widget -> Widget Control Mode. For future reference.
>>109630890thanks!
>>109633330>he fell for the malwareits over
>>109633337proofs?
>>109633198>I'm sure it's just a matter of time before a lot of this shit gets consolidatedwebsloppers can't help but bloat. It's been getting worse every year while llm bros get binaries where you don't have to curate all this shitty bloat
>>109633242Your hardware should be much faster, but on my humble iGPU, a 4-second 0.2 megapixel video originally took about 30 minutes. Switching to --use-ck-attention knocked that down to 20 mins, and then enabling the turbo lora knocked that down to 10 mins, since it uses fewer steps.There is a newer version of the template that has a toggle switch for turbo mode (see attached pic). Drag this template into ComfyUI:https://github.com/Comfy-Org/workflow_templates/blob/main/templates/video_minimax_h3_i2v.json
>>109632493some mirror to hfdescriptions and examplesto see what loras were deleted
>>109633362Oh neat, >Switching to --use-ck-attention knocked that down to 20 mins, I don't know what that means but all the other stuff got the progress bar moving.
>>109631624NPI ended up checking it out and following some youtube guide. I was able to get comfy running on it and genned some videos. For reference, a 3 sec video took like 2 hours on my machine. I rented an RTX 4090 for two hours and it cost me a little less than $2 or like 75c/hour.A lot of that time was just setting it up and troubleshooting, but once I got it running i was able to gen pics in seconds and videos in about 7 minutes. I didn't get great results but i think that's more due to my prompting and workflows than the models. I used wan2.1, but would be open to checking other models in the future.https://litter.catbox.moe/zjih29mtqrmuivdf.mp4I'll probably try it out again next weekend and creep these threads for good models or tips. In the future I might check out the 100 VRAM GPUs and run one of those massive models. Its only like $3 an hour.
>>109633219>>109633222>>109633230https://litter.catbox.moe/y07nei0mr013egxl.mp4Don't say I have not delivered in gratitude for helping me reach this point.
>>109633426kino
>>109633426remarkable kino
>>109633426My PC just straight up shut down trying to run it again so I think I'm trying too much.
>>109633426Cute
>>109633406Use a text editor to open run_amd_gpu.bat and add --use-ck-attention to the launch command. I'm not sure what the default is, but the format should be something like:python main.py --enable-dynamic-vram --enable-manager --disable-smart-memory --use-ck-attentionEt cetera. You won't have so many flags, but it should give you an idea of what it looks like. The order of the flags doesn't matter, I think.There's also a GUI node that can be used to only enable Comfy Kitchen Attention for the model you hook up to it, instead of turning it on for everything; see attached pic. But figuring out where to hook that up in the Minimax template would be a bit annoying.
Bake a collage or don't bake at all
https://litter.catbox.moe/hjxx95sofxc78pg7.mp4neat, can mix styles fine.Use <Picture 1> for the physical identity of Denton, with the voice of <Audio 1>.the setting is <Picture 2>.medium shot of JC Denton in an office building with a "UNATCO" sign, looking exactly like the character in <Picture 1>, is looking forward and points to the camera. Denton speaks naturally in a dimly lit, gritty cyberpunk interior, saying the exact line: "Hey you, have you seen any Hatsune Miku AI generations in this location?.".camera cuts to a shot of Jerry Seinfeld from the show Seinfeld, who says "what is a Hatsune Miku?"camera cuts to Denton, who says "a virtual idol, a vocaloid, an idol. she sings music."Low-poly aesthetics, moody green and blue ambient lighting, nostalgic year 2000 PC gaming look.
>>109633136We are all sirs here actually
>>109633493thank you Mark R.
>>109633500congrats you can read embedded windows/user folders from a workflow, mr hackerman
>>109633504i appreciate it, Mark R.
>>109633511>Inspection: Anyone who extracts or inspects the JSON payload from the video file using a text editor or a tool like ComfyUI-Workflow-Inspector can read those text fields and see the Windows username.comfy is a retard for that btw, what if someone had sensitive data in that, it has no relevance to a node workflow.
Do more detailed prompts increase or decrease time it takes to generate?
>>109633563not really, if you want a super fast test do 0.3mp, when you have a good prompt then bump the resolution.
>>109633461Finishing successfully once is an encouraging sign. There may be ways to stabilize it. Video gen's one of the most demanding things you can run.>>109633563I think a negligible increase unless you're doing text encoding on CPU.
>>109633582I think it was my GPU overdrawing power. Is there a way to cap it off to stop the gen from running it at 100%
>2026>no audio model trained on R18 ASMR
>>109633594what operating system are you using?
>>109633605Windows 11.
>>109633619>>109633619>>109633619MOVE
>>109633612actually, you're using amd right? i am not familiar with those cards. just look up for popular software that limits the power for amd gpus. usually it is overclocking software, but those would give you the option to reduce power as well
>>109633627Yeah underclocking/downvolting is an option, I was just hoping there might have been an app-specific setting.
>>109633594Looks like AMD Adrenalin may have a power limit slider, or there are more general presets:https://www.amd.com/en/resources/support-articles/faqs/DH3-020.html
>>109632138holy, catbox?
>>109633629>I was just hoping there might have been an app-specific settingi don't think any GPU software is sophisticated enough to work that way. all hardware settings are global
>>109633619>>109633619>>109633619NEW
>>109633644Fair enough.
>>109633633>>109633644Adrenalin says it has per-game/app profiles, but I don't know if it will work with ComfyUI/Python. If pointing it at run_amd_gpu.bat doesn't work, maybe try pointing it at python.exe?https://www.amd.com/en/resources/support-articles/faqs/DH3-012.html#dh3-012-application
>>109633676oh. pointing it to the .bat script would definitely not work since it's just a launcher for a different program. he has to look at task manager or whatever other software that tracks GPU usage per process and use that one