[a / b / c / d / e / f / g / gif / h / hr / k / m / o / p / s / t / u / v / vg / vm / vmg / vr / vrpg / vst / w / wg] [i / ic] [r9k / s4s / vip] [cm / hm / lgbt / y] [3 / aco / adv / an / bant / biz / cgl / ck / co / diy / fa / fit / gd / hc / his / int / jp / lit / mlp / mu / n / news / out / po / pol / pw / qst / sci / soc / sp / tg / toy / trv / tv / vp / vt / wsg / wsr / x / xs] [Settings] [Search] [Mobile] [Home]
Board
Settings Mobile Home
/g/ - Technology


Thread archived.
You cannot reply anymore.


[Advertise on 4chan]


Best Currently Maintained Edition

Discussion and Development of Local Image, Video, and Music Models

Previous: >>109627980

https://rentry.org/ldg-lazy-getting-started-guide

>UI
ComfyUI: https://github.com/comfyanonymous/ComfyUI
SwarmUI: https://github.com/mcmonkeyprojects/SwarmUI
SDWebUI: https://rentry.org/ldg-lazy-getting-started-guide#the-stable-diffusion-web-ui-lineage
Wan2GP: https://github.com/deepbeepmeep/Wan2GP

>Checkpoints, LoRAs, & Upscalers
https://huggingface.co/models
https://huggingbay.xyz
https://civitai.com
https://civitaiarchive.com
https://openmodeldb.info

>Tuning
https://github.com/spacepxl/demystifying-sd-finetuning
https://github.com/ostris/ai-toolkit
https://github.com/Nerogar/OneTrainer
https://github.com/tdrussell/diffusion-pipe
https://github.com/kohya-ss/sd-scripts
https://github.com/kohya-ss/musubi-tuner

>Minimax H3
https://huggingface.co/Comfy-Org/MiniMax-H3

>Krea 2
https://huggingface.co/krea/Krea-2-Raw
https://huggingface.co/krea/Krea-2-Turbo
https://lumenastrum.github.io/clio-style-preview/gallery/

>Anima
https://huggingface.co/circlestone-labs/Anima
https://tagexplorer.github.io/
https://animadex.net

>Klein
https://huggingface.co/collections/black-forest-labs/flux2

>Misc
Gentleman’s Guide: https://rentry.org/ldg-gettingstarted
Share Metadata: https://catbox.moe | https://litterbox.catbox.moe/
Txt2Img Plugin: https://github.com/Acly/krita-ai-diffusion
Archive: https://rentry.org/sdg-link
Collage: https://rentry.org/ldgcollage_v2

>Neighbors
>>>/aco/csdg
>>>/b/degen
>>>/gif/vdg
>>>/d/ddg
>>>/e/edg
>>>/h/hdg
>>>/trash/slop
>>>/vt/vtai
>>>/u/udg

>Local Text
>>>/g/lmg

>Maintain Thread Quality
https://rentry.org/debo
https://rentry.org/animanon
>>
breast thread of breasts
>>
>>109630182
>>https://rentry.org/tkdupekk
>>109629403
>https://rentry.org/ldg-gettingstarted

Added to https://rentry.org/ldg-lazy-getting-started-guide#anon-guides-and-resources
>>
what does everyone think of WAN 3.0?
>>
>only 4
Fuck off
>>
Recommend h3 workflow pls, so many different snake oil nodes and tech got released since I first tried, idk what's actually good. Also, is ref2v still much slower than t2v and i2v?
>>
>posted my gens on some other board
>people are going out of their way to be butthurt about it

why is this always the case?
why cant people appreciate AI art?
>>
>>109630601
just stack 10 h3 cache nodes. saves a lot of time looks very good
>>
>>109630601
here you go lad, I know, they're hard to find
>>
>>109630634
Some seethe at superiority. Simpleas.
>>
>>109630545
very kewl and based
>>
>>109630652
and the gens I posted werent total slop either but people still went out of their way saying "this is AI slop", "stop posting AI slop" etc.
like what is their issue? Art is Art and people post human art all the time and some of that shit looks ten thousand times worse than my gens.
>>
File: MiniMax_H3_00170_.mp4 (3.22 MB, 1056x608)
3.22 MB
3.22 MB MP4
forgive me for not posting the best vids ever
I can only do so many iterations with a 5070 ti
>>
Might as well leave "it" here too
https://litter.catbox.moe/r2zma0.mp4
>>
>>109630601
>>109630650
I know anon was being a dickhead but he isn't wrong. Any other H3 workflow that you don't build upon or just use from the original are going to be shit and absolutely not for your setup. I tried a couple and they all crash or do weird shit, the comfyui recommended workflow just works. If you see nodes that might help you, you should try them and see if they improve.
>>
>>109630729
That's why I asked. I guess I'll just keep using patch sage attention kj and the 8turbo lora and that's that...
>>
>>109630703
pretty cool. that's Blizzard tier animation in the palm of your hand.
>>
>>109630740
>that's Blizzard tier animation
you don't really believe that do you? you're just being nice right?
>>
>>109630703
i will now download your cookie clicker on the app store
>>
>>109630722
top kek
>>
>>109630729
>>109630601
This is sound and solid advice, the comfyui workflows are usually really solid and don't require any nodes, they even have all the download links in those black notes.
What I recommend if you want to check out other workflows, especially those that needlessly use 300 nodes to "make the workflow look pretty", make a fresh comfy portable instance and try them out there, see what you like and what you don't like.
But for videogen there isn't much you can do rn,
>RTX Super resolution upscaling
Does almost nothing, but it does minutely improve image quality, costs also almost no resources or time
>RIFE frame interpolation
Great at realistic images, smoothes out the video a bit and causes less sharpness artifacts sometimes, but can cause other artifacts to show up if you crank it up too high. Shit for anime
>Comfy kitchen attention
Near lossless video prompting acceleration
>Sage attention
Video prompting acceleration
>Cach patcher, lazy cache, etc etc
Abysmal, destroys both videoquality and prompt coherence
>Spectrum
10-20% speedup, questionable if it reduces prompt adherence, try for yourself
>FastVideo
??? anyone tried that yet?
>TurboLora
use larry'S 600 turbo lora at 6-12 steps, anything else rn is a meme

That's what I got so far, but a good prompt is honestly the most important thing. And use the correct model for the correct application.
Hidden pro tip: change "ref_image_size" to max instead of match and you'll get much sharper and higher quality videos when doing i2v.
Hope any of that helps
>>
I really need to start using the reference model
>>
>>109630751
well sure. the resolution perhaps isn't there. but the style is there.
>>
>>109630722
funny, but what does the Google symbol mean
>>
>>109630542
>mfw Resource news

08/23/2026

>Krea 2 Turbo — 4-Step Distillation LoRA
https://huggingface.co/lvladikov/Krea2-Turbo-Distill-4step-LoRA

>Alibaba to issue US$10 billion in new shares for huge AI push amid strong investor demand
https://www.scmp.com/tech/big-tech/article/3364957/alibaba-issue-hk80-billion-new-shares-global-ai-push

>Nvidia Customers Notified About AI-Related Price Hikes Above 15%
https://www.bloomberg.com/news/articles/2026-08-22/nvidia-customers-notified-about-ai-related-price-hikes-above-15

>H3 Motion Context Clip Stitcher
https://github.com/noembryo/ComfyUI-noEmbryo#h3-motion-context-clip-stitcher

>Comfyui-MMH3-UltimateUpscale
https://github.com/bbaudio-2025/Comfyui-MMH3-UltimateUpscale

>FastVideo-Minimax-FastH3-Preview-v0.2
https://huggingface.co/FastVideo/FastVideo-Minimax-FastH3-Preview-v0.2

08/22/2026

>ComfyUI-H3-AudioRefine
https://github.com/Adudeguyman/ComfyUI-H3-AudioRefine

>Anima-3.8B with Qwen-3.5 4B
https://huggingface.co/lylogummy/Anima-3.8B

08/21/2026

>MiniMax H3 Super Acceleration fast draft generation and high-resolution refinement, powered by Sol Engine
https://nvlabs.github.io/Sana/Sol-Engine/H3-Super-Acceleration

>4DAnyone: Create Anyone in 4D from a Casual Monocular Video
https://4danyone.github.io

>H3 Prompt Composer Version 5.37.1
https://github.com/BMB12d3/minimax-h3-prompt-composer

>LTX-2.5 Gemma-4 12B NVFP4 for ComfyUI
https://huggingface.co/Deadshot699/ltx-2.5-gemma4-12b-comfy-nvfp4

>MiniMax-H3 Pruned Ref-Delta Fused r1024 — ComfyUI Single File
https://huggingface.co/xmarre/MiniMax-H3-Pruned-Ref-Delta-Fused-r1024-ComfyUI

>Phosphene 4.6.0 Adds Video Editor, H3, LTX 2.5
https://github.com/mrbizarro/Phosphene/releases/tag/v4.6.0

>H3 Optimizations: Standalone production optimization nodes for MiniMax H3 in ComfyUI
https://github.com/Zironic/H3-Optimizations

>MiniMax H3 known character lists
https://huggingface.co/datasets/malcolmrey/various
>>
>>109630887
gemma-chan is a popular 4chan mascot for the google made gemma models
>>
>>109630722
Also this was the voice I had in mind when doing the video, but it wouldn't actually change it via prompt so I now tried with a 3 sec voice sample of heavy. Unfortunately the video itself came out like shit
https://litter.catbox.moe/6ybyfs.mp4
>>
>>
>>109630887
it's one of the symbols of our overlords here in america. it's a zionist psyop to make you love google.
>>
so why can't I use the ref model for t2va?
>>
>>109630912
cute and slightly erotic, nice
>>
File: 07650-2642686090.png (2.81 MB, 1920x1280)
2.81 MB PNG
>>
>>109630703
what are you're computer specs and average gen time?
>>
>>109630912
why does she keep shrugging
shit is getting on my nerves
who even makes this bullshit
>>
>>109630982
unironically indians
>>
File: localkek.png (9 KB, 731x52)
9 KB PNG
>>109630951
desu such a waste of electricity
Minimax is amusing but keeps fucking up the details
I need to go back to making images
>>
>>109630993
>everything was troons
>now everything is indians
lmao monomaniacal anons life is simple
>>
What model (and/or workflow) are you all using for style transfers?
I've been trying to take a reference image that contains the artstyle, scene, composition, expression, etc. that I want and replace the character in it with a reference image of a new character with Qwen edit 2511, but have not had much luck
>>
Where is everyone :(((((
>>
>>109631045
reporting in
>>
File: i_00058_.png (2.43 MB, 1080x1920)
2.43 MB PNG
>>109631031
maybe try the krea2 edit lora (on Civitai) I've seen some anons here use it, haven't tried it myself.
I found some workflow on youtube with some boomer and cope nodes and it kinda worked to an extent but comfy update broke half the nodes.
Anyway Krea2 was built for style transfer so it's the right model to use. I'm thinking that edit lora is a good option.
>>
>>109631076
I just recently tried it and it seems broken rn to me
>>
i'm half jewish half indian and trans.
why are people here so mean to me?
I just want to locally diffuse!
>>
>>109631082
have you tried it specifically for style transfer?
>>
>>109631045
it takes time for me to come up with ideas for gens, anon
>>
>>109631112
this, people think it's all sunshine and rainbows until they have to proompt for themselves
>>
>>109630995
desu big models are more fun when you have better hardware
>>
File: maid.mp4 (1.62 MB, 1056x608)
1.62 MB
1.62 MB MP4
>>
>>109631140
I passed on getting a 5090 because I would have to buy a new chassis and they seem energy-inefficient, and 32GB probably isn't enough either. You would want the Pro
>>
>>109630925
You can, if you want slower gen times and worse quality/acting
>>
>>109631146
who hangs their shirts in the kitchen wtf
>>
>>109630740
fun fact blizzard outsourced the cinematics and didn't make them. I mean they were fucking insane what would you expect
>>
>>109631206
they didn't always do that. all the OG great world of warcraft ones are by them
>>
>>109631112
For me the biggest limitation is how much time it takes to generate a good quality video without random issues.
>>
>>109631227
Spending hours trying to figure out why the thing clearly described in your prompt isn't showing up.
>>
>>109631093
no
I was getting broken outputs (some weird artifacts etc.)
>>
cozy breas
>>
Qwen3.8 27B reasons way too much for simple generation prompting with the default reasoning effort (xhigh). Try setting reasoning effort to medium. and note that low is not good because it often uses even more tokens than medium. The default xhigh is good for outputs that require a lot of reasoning like coding but for prompt gen it will just slow everything down.
>>
>>109631263
I see. sad. I was hoping that might be the one thing it might actually do. oh well. I guess we still have to stick to loras for now.
They're time consuming and they suck the life out of you and ruin all the fun of diffusion but at least we know we have them if we really need them.
>>
>>109631296
do your own research though. my setup might be fucked, not the edit workflow. I'm very new to krea, been only doing videogen
>>
>>109630601
plaguekind on civitai
>>109630995
undervolt
>>
>>109631296
>>109631303
The edit lora is great for changing one image, but somewhere between mediocre and lukewarm for style transfers using TWO images
>>
>>109631221
Oh neat, things used to be, in fact, not lame and gay
>>
>>109631293
wait a minute. there's a new qwen? Is it an edit model?
>>
>>109631154
32GB is plenty for doing this stuff as a hobby. The pruned Minimax models fit into VRAM comfortably, with lots of room for KV cache and the like. The RTX 6000 Pro has a lot more VRAM, but is not ahead by that much in terms of compute while being multiple times more expensive.
>>
File: 391002233194197.mp4 (3.67 MB, 640x832)
3.67 MB
3.67 MB MP4
>>
>>109631303
>>109631296
As far as I can tell they only make the style transfer available on cloud Krea 2 officially. There are custom user made nodes for style transfer but I haven't tested any.
>>
>>109631326
thanks for the feedback. i guess it's a no go then because I already use qwen and klein for image editing.
Come to think about it, I haven't really tried transferring styles in those models. Maybe that's something to look into.
>>
File: 660644999154235.mp4 (3.47 MB, 608x864)
3.47 MB
3.47 MB MP4
>>
>>109631227
with better hardware that gets easier
>>
>>109631346
yeah I've tried some custom nodes like I mentioned earlier. They get it right partially. They change the rendering style but not the character style (propotions etc)
>>
>try to reuse Krea 2 wf from old image
>it didn't save the text from the Generate Text node
>generating text again gives a completely different result on the same seed
>but it saved comfyanon's furry demon biker 1girl oc prompt for some reason
Why doesn't generate text just save the text by default
>>
File: 523991199657429.mp4 (3.53 MB, 960x544)
3.53 MB
3.53 MB MP4
>>
File: theaudacity_preview.png (280 KB, 640x480)
280 KB PNG
https://files.catbox.moe/cjw2tq.mp4
>>
File: testo.jpg (639 KB, 2342x2113)
639 KB JPG
>>109631425
we're not letting go of this one huh, can't blame you either
>>109631429
is that H3?

>>109631357
Also I tried out Krea2 style transfer to test it, but semi cheated by using a base krea2 image to edit, since it already has knowledge of what it created if that makes sense.
Here are my results, not sure if that would satisfy your needs
>>
File: videoframe_29325.png (738 KB, 544x960)
738 KB PNG
This is way too much fun

https://files.catbox.moe/vnd0me.mp4
>>
I'm trying to get H3 to point toes inward which is an extremely common posture for women having their photo taken and it just won't fucking do it. I can type all this fantasy shit and it's like YES MASTER. But if I want dime a dozen celebrity red carpet trope its UARRRR I OINT UNDESTRAN MATE
>>
>>109631377
in the future this node is gonna save your life https://github.com/pythongosssss/ComfyUI-Custom-Scripts#show-text
its like preview text but it saves it properly so when you drag the image the exact prompt is visible
>>
My machine isn't strong enough to run anything good. I saw you can rent compute from places like runpod. Anyone have any experience with that?
>>
>>109631460
>is that H3?
Yes ref2va with a forced start segment to copy voice tone that's trimmed from this since the video quality diverged from source right at the inference split.
>>
With the effort you took to solve a captcha to confirm you are poor as FUCK, you could have placed in a job application in the same effort.
>>
>>109631490
kekd
>>
>>109631463
Classic porn plot.
>>
File: 101045875185258.mp4 (3.69 MB, 704x1024)
3.69 MB
3.69 MB MP4
>>109631460
>we're not letting go of this one huh, can't blame you either
It's actually from earlier, just forgot to post it.
>>
File: 1001921610261099.mp4 (3.48 MB, 768x960)
3.48 MB
3.48 MB MP4
I miss my Baldachin's Blessing.
>>
>>109631463
>I'm sorry I couldn't find the milk potion anywhere
>I asked you to find a book
>>
>>109631531
Great tummy
>>
>>109631474
I was considering that just so I don't have my card cooking for hours on end. I don't know what the best gpu-for-rent options are tho sry
>>
why is krea so bad at nsfw anime content? i dont want to stack a ton of loras just to make it not suck ass
>>
File: 546463678724258.mp4 (3.46 MB, 960x768)
3.46 MB
3.46 MB MP4
>>
>>109631641
>why is krea so bad at nsfw anime content?
i lacks one of the best repos of nsfw anime content in its dataset simpleas
>>
>>109631641
>i dont want to stack a ton of loras just to make it not suck ass
Krea2 in general is dogshit at nsfw if you're looking for intricate poses or something like that, that's not an anime specific thing. If you tell me what you were trying to gen, I can give a go at it to see if it's just skill issue
>>
>>109631641
Real Labs are scared of making their models good at NSFW without a lot of legwork from the user.
>>
I'm getting sick of gemma QAT dropping the ball is qwen better?
>>
>>109631772
don't use qat use normal gemma
qwen isn't better it's different
>>
>>109631649
weird how the quality is shit in [shot 1] but then surprisingly good afterwards, did you use multiple references or is that H3's doing?
>>
So after using qwen and grok 4.6 all day, it turned out the anons in previous thread were correct, and penetration was not working well due to several things:
1: Prompting language is beyond important for h3. You really need to dig down and describe things down to the second (and it will stick to them).
2: Shift video actually matters a lot. If you want slower and more detailed penetration movement, dropping down to 6 or so can have a huge impact, especially if you want it to animate on two's/three's anime or western cartoon 2d style.
3:Attention and XXX loras seem to work wayyy better at around 0.50 strenght vs even 0.70 or worse yet 1.00 (At least at 25 steps). Also in the lora selection, the 2nd option (VIS) does not work like how it does on image models where in a lora the two settings have to match, and its better to have it stick to 1.0 even if you drop the model strenght to 0.50 or whatnot.
>>
>>109631793
wasn't qat supposed to be better for anything under q8?
>>
>>109631822
I haven't used the shift video node at all during my gens and got great results (didn't try nsfw yet), can you give me a QRD of why I should waste my time with it?
>>
>>109631824
from my test it's almost systematically shittier despite what google claimed, and it seems to be consensus in lmg
>>
>>109631834
For example i tried an animation where I specified that there are 4x thrusts within 5 seconds. When it was at default 11 or 12 shift video, the animation instead tried to fit 5-6 thrusts to match the timing of the total lenght even when i specified otherwise. When i dropped it down to 9, i noticed it seems to often (but not 100%) have one less thrust than at 11. Dropping all the way to 6 made the thrusts match the timings i specified. Dropping bellow that made the animations sometimes almost (or at times did) stop. If you animate 3d style it might not be as important, but for anime/2d style being able to slow down the animation like this matters a lot.
>>
>>109631822
So are you prompting like this:
at 00:01.000 this happens
at 00:02.000 this happens
at 00:03.500 another thing happens

within a single shot?
>>
>>109631860
Yea like this:
Continuous slow sex with no pauses between strokes, only four thrusts total, each a smooth circular grind rather than a straight piston. Thrust 1 from 0.00 to 1.25: his hips sink and roll in a continuous circle from her right toward her left while her hips roll up to meet him and seat fully. At 2.50 seconds her mouth shape changes from a grin with teeth to an open mouth with tongue visible. Thrust 2 from 1.25 to 2.50: same continuous right-to-left circular down-roll without stopping. Thrust 3 from 2.50 to 3.75: same. Thrust 4 from 3.75 to 5.00: same circular roll, finishing fully seated. The four strokes blend into one unbroken grinding rhythm. Her breasts move with body weight and a soft lag behind each roll, natural weighted bounce only, no exaggerated jiggle, no wild flopping, no flying droplets. Small collar-bell sway follows the roll. Wet shine and smear around the condom and contact point.
>>
>>109631869
Wow ok. And do you find that telling the model what not to do, e.g. "no wild flopping" actually works? I've tried that and it seems to have no effect or even the opposite.
>>
>>109631848
Extremely interesting, going to try it out for myself.
But I do have a question, can I just pack the ModelSamplingMiniMaxH3 node between the basic guider and the load diffusion model node?
Or how did you do it, cause I think I already have one built into the default MiniMax H3 Reference to Video node that sets it to 12, hidden and unchangeable
>>
>>109631464
found out this pose is just incompatible with T-posing, not even fake T-posing like "raise arms out" or "lateral raise".
>>
File: sampler.jpg (208 KB, 688x1109)
208 KB JPG
>>109631875
It has for me at 25 steps. I've had over 30 generations where the dick didnt flop out of her at random using the settings i mentioned above.
>>
>>109631882
I'm using this setup and it uses comfykitchen (just like 1-2% slower than sageattention but you dont need to install all the nvidia cude stuff).
https://civitai.red/models/2831978/dasiwa-minimax-h3-workflows-or-t2va-or-fl2va-or-ref2va
>>
>>109631934
>Civitai workflows
Just checked it out and puked a bit in my mouth, so much stuff in there that's useless, still thanks for the effort.

For anyone else wanting to try out shift_video I figured it out, you simply have to put the "ModelSamplingMiniMaxH3" node after your last Lora/ComfyKitchen or other patcher nodes and before the basic guider and scheduler, then you can play with it.
And holy shit the effect is immediate, this is a good node to play with for 2d/anime animations
>>
>>109631966
Anything specific that you find useless in it? It seems to generate on my 5080 at around 7.2s/it for a 5 second clip which is pretty much the speed of a barebones h3 image2vid.
>>
>>109631979
All of the KJnodes that also install that hideous laggy interface obviously, but that's just personal preference, I dislike node bloat when they're clearly not needed.
>Cache nodes
>fp16 acc.
>upscale 2x using meme method
>upscale with image upscale modes
so fucking useless, shoot him for that one
>upscale rtx with a meme node instead of the original one
>watermark

These are imo useless features that are bloating the workflow and make many of the things hard to understand.
I'm not directly saying the workflow is bad or that it produces bad results, simply that it's disgusting to look at with a bunch of useless stuff, while hiding important things like "reference settings" seemingly away, or I cant see it without the nodes
>>
>>109632024
make a pr that deletes all the useless shit
>>
>>109632024
Ah i see. I have no idea in regards to cache nodes and whether they do anything for h3. I might try making a clean workflow while still using the loras and kitchen attention.
>>
File: 16746.png (162 KB, 502x477)
162 KB PNG
i pulled.....
>>
>>109632034
I already have my own workflow so I don't really need to take another one and change it.
>>109632041
Like I said, the workflow is perfectly fine in function as long as you don't use the useless stuff, you don't need to change anything if you're happy with the results and you're not agitated looking at that or by the fact you're potentially loading a bunch of nodes you don't really need each time you start up comfy.
I'd say if you like the workflow, keep it as is
Also it's hard to compare gen times but this 1MP 10sec video took me 442sec on my 5090, I've set shift_video to 6 as a test
https://litter.catbox.moe/yxlwwd.mp4
>>
Now that the dust has settled, what is the best fl2v Turbo lora for H3?
>>
8 steps turbo lora, 11 minutes 1 mp, not bad
I hate agorist fags and zogbot libertarians btw
https://files.catbox.moe/b22ko5.mp4
>>
>>109632079
What did he mess up this time?
>>
>>109632099
>https://huggingface.co/larryvrh/MiniMax-H3-Turbo-Lora/tree/main
This one of course.
>>
>>109632119
video extensions
>>
>>109632123
>Big dick Larry 600
yessir
>>
>>109632088
Man i wish i got a 5090. I got a 5080 standing in line at launch and couldve taken a 5090 but i figured id never need it. Now the damn thing costs 5k, and its not like i coudnt afford the 2k at the time...
>>
anyone have a good video ref 2 va wf?
What should the input video resolution be?
I used 1M video force 24fps, and the gen time was terrible. Forcing the resolution by half seemed to be better and I really didn't see quality loss
>>
File: 95964949484.jpg (1.6 MB, 1900x1268)
1.6 MB JPG
>>
>>109632123
Which one of them?
what about ref2v?
>>
>ref2v takes a fucking long time
>i2v doesnt give sufficient enough results
>fl2v i dont really care for, no use case for my slop
>lf2v is interesting and no use case for my slop but ill try it for fun
>t2v is not coherent enough but kino for insane chaotic slop
im tired bros
>>
>>109632137
comfy org has a very simple workflow that works
>>
Does the avatarfag that hates Ani and Debo still here?
>>
File: Untitled.png (96 KB, 2508x557)
96 KB PNG
>TFW someone smarter than you goes out of their way to explain shit to you in a really concise step-by-step way and you're still too stupid to get it.
H...haha...ha..
>>
>>109632146
>https://comfy.org/workflows/b34841f6789c-b34841f6789c/

this one? but I'm looking for video ref, not image. I assume they are the same but video ref is different beast for me.
>>
>>109632132
Besides the extra Vram I don't think the difference is that big, the actual regret is not getting a rtx 6000 pro with 96gb back when it was still 6-7k

>>109632143
Try spectrum if you're not using it already, might cut your time down by another 10-30%, haven't noticed a drawback yet

>>109632137
>>109632146
This, try that one first then expand

Also, same prompt same seed same everything but shift video 9 this time (took only 414s this time, probably because models were loaded)
https://litter.catbox.moe/173ec8.mp4
shift 6 comparison:
https://litter.catbox.moe/yxlwwd.mp4
>>
>>109632162
>igpu
the rest is solid advice but it depends so much on which igpu and what resources that fucker is using. you dont even necessarily need those specific methods to set it up. the ck attention is useful to me though but whether its better than flash attention is hard to say
>>
File: 5234sdffs.jpg (95 KB, 645x880)
95 KB JPG
>>109632172
Anon, comfy has its on workflow browser. Nothing external
>>
>>109632180
I'm a retard. I think I'll finally be able to use local AI when it gets made into a single executable.
>>
>>109632175
>Try spectrum
i have it already but i do not like the way it sacrifices quality for speed
>>
File: Scail-2.mp4 (3.4 MB, 1674x1920)
3.4 MB
3.4 MB MP4
Scail 2... whatever happened there?
>>
>>109632198
im retarded too, im using amd. https://github.com/CS1o/Stable-Diffusion-Info/wiki/Webui-Installation-Guides#amd-comfyui-with-rocm this was written by a german ESL, and has some bad choices in my opinion but it works and theres 2 methods, one is the "download a fuckin zip and run it" and "if you want the speedups"
>>
>>109632210
>H3ppened
>>
>>109632210
It was great, been replaced by h3 though.
>>
>been checking for months to see if someone backed up a lora on civitarchive
>nothing
im going to have to track down the guy who made the lora and ask him for it.

https://civitaiarchive.com/models/2454555?modelVersionId=2782555
>>
>>109632232
>>109632230
can H3 replace people like that?
>>
>>109632192
nigga, read, I don't need image ref 2VA. I need video ref 2VA workflow.
They are kinda the same but not really; if I put 1M input video, my computer'll explode.
>>
>>109632247
you are too dumb to gen a video
>>
>>109632238
Have you been living under a rock?
>https://files.catbox.moe/d6ebyp.mp4
>>
Ok now I got a full set
video_shift 12:
https://litter.catbox.moe/279cz0.mp4
video_shift 9:
https://litter.catbox.moe/173ec8.mp4
video_shift 6:
https://litter.catbox.moe/yxlwwd.mp4
very clear difference

>>109632247
>nigga, read
Ironic that the retard who cannot read is telling me to read. I have to assume this is either bait or you're mentally challenged
>>
>>109632238
Yup.
>>
>>109632172
>>109632247
just add a Load Video (Upload) node to ref_video_0
I uploaded the image ref workflow to chatgpt and it did the rest
>>
>>109632271
>just add a Load Video (Upload) node to ref_video_0
anon...I did that. I'm trying to say what the difference between 0.5 and 1M video input
because I don't see much quality change but gen time is massive
>>
File: Untitled.png (259 KB, 1355x1663)
259 KB PNG
>>109632225
I took a look at it. There's no command line for my GPU. It only goes up to 6900, I have a 7900.
>>
>>109632310
whats your gpu
>>
File: file.png (42 KB, 1367x161)
42 KB PNG
>>109632310
Those instructions are for older gpu models, the support you need is baked in already. Read it carefully
>>
>>109632318
7900xt.
>>
>>109632323
Now, In my defense, I did say I was retarded.
>>
>>109632310
Have you tried literally just downloading the comfyui folder made specifically for amd and running the "run_amd_gpu.bat" file?
Or is your problem that you want all of the speedups like sageattention triton and all this other bullshit?
>>
>>109632333
Just follow the instructions there and you can skip the part that isnt relevant to you. Just make sure youre prepared for issues somewhere
>>
>>109632341
>>109632352
I basically am just paranoid, and just want to make sure I'm not skipping something. This will install the thing for me to play around with it and shit, right? Or am I completely missing the part where I had to install the thing and this is just the GUI for the thing and I needed to install something else prior?
>>
>>109632362
my man just fuckin read the instructions its that easy
>>
>increase reference images from 2 to 3
>the time per step increases from 2 minutes to 10 minutes
nani?
>>
File: Untitled.png (248 KB, 3835x2024)
248 KB PNG
>>109632368
Oh my goodness, the new frontier. I have no idea what I'm doing but I'm in.
I'm probably slightly less retarded than I'm putting on, but to be fair my confidence is at 0 generally as a character trait so I am ALWAYS afraid I'll blow something up with everything no matter how thoroughly I read the instructions.
>>
>>109632362
To test it you don't have to install anything.
Simply download
"ComfyUI_windows_portable_amd.7z "
unpack it... then double click "run_amd_gpu.bat"
You'll instantly see if you can run comfyui or not, there is nothing to install and you're most likely just being retarded rn

>>109632372
maybe your third reference was much higher in resolution than the others, or you're running out of vram, check in task manager
>>
Is it bad for my GPU if I generate batches and leave it running non stop for many hours? I only got into vid genning since Minimax was released. Will vids destroy my GPU
>>
>>109632379
yeah the images are 4k
I'm genning at 0.6
>>
>>109632379
Yeah I ran it, works and all, see >>109632378
Thanks for being patient with me. Now that I'm in I'll try to figure out how to use it from youtube tutorials.
>>
>>109632384
your gpu will die in a year
>>
>>109632384
Throttle or cap its power and make sure the temps stay cool.
>>
>>109632378
you know maybe i was wrong to not go to school and be truant, maybe i was wrong to be terminally online and learn most of the shit i know from a computer for 20 years, maybe i was wrong to not go to college (i was right on not going into debt though), but honest to god, taking out computer literacy courses in schools is the biggest mistake we ever made
>>
>>109632385
Have you set ref_image_size to "match" to avoid throwing 4k images at your puny 0.6MP video?
Alternatively try downsampling them with a "scale image by" node down to half and see if that fixes your times for you. Super alternatively try restarting comfyui, that clears the vram cache and also could solve your issue.
Omega alternatively, try putting two references into one picture with a cut out background and describe them well.

>>109632384
yes GPU's are made for roughly 2 weeks of genning before you notice a 30% decrease in performance

>>109632392
Now the sky is the limit, don't forget to check out civitai.red (nsfw) to see what models are currently the latest and greatest and try something out if you don't already have something specific in mind.
Also I recommend the >>109632192 template browser, it usually has a workflow for everything and massive giant notes that explain everything
>>
>>109632401
I think it's a matter of curiosity, I am thankful I have been curious enough to learn everything I needed to know and more (except being social)
>>
File: devilish.jpg (47 KB, 470x595)
47 KB JPG
i see a new lora just released for something i like
>>
>>109632401
On average, though you might not believe, I'm actually more tech-literate than an average person. It's just whenever it comes to the coding side or anything related to it, it turns out as I said in that post 0 confidence is a hell of a thing.
>>
>>109632414
have you tried reading to get started like the rest of us hopefully did? im not asking you to setup a networked linux setup that you remote into to make slop, but i am expecting you to read basic instructions. you will have to let go of being spoonfed or handheld eventually. the only reason i decided to help a little bit is because you made the unfortunate choice to go amd like i did, and that makes you automatically in hard zone for this shit.
>>
i dont trust any lora for h3. i believe any lora attempt should feed data in the same prompt structure and detail as regular h3
>>
>>109632438
Not the same guy. It's the person that just made an anecdotal remark. I already got it working and said thank you for the help, I think postmortem frustration with me might be a bit excessive but it's not like you didn't earn the right. Still if you missed me saying it, thanks for the help.
>>
File: dena_chroma_00007_.png (2.37 MB, 1881x1433)
2.37 MB PNG
>>
>>109632460
i deal with retards that cant open zips all day and theyre like 25. my frustration comes from a very important place in my heart
>>
>>109632468
Honestly it wasn't even an issue of a zip file. The few times I dealt with github files in the past I had to download the software and the GUI separately, so when the AI itself is called "ComfyUI" in my mind it flagged as "This is the GUI, but it doesn't say anywhere how to get the core files." So I felt like I was coming into the final phase of a process I missed the initial steps for.
>>
>>109632447
certain motions don't work for t2v unless you have the lora for it
>>
>>109632411
>>109632400
temp is 73C after several hours. I think my 5070 Ti has good longevity because of lower power draw
>>
please stop botting this thread
>>
whats the point of civitaiarchive if i still need to log in to civit to download something?
>>
>>109632489
>not getting the 5090
isnt like 8k right now? i shudder to think what the 6090 will cost if they even bother to make one.
>>
>>109632489
use ref2v turbo lora, and dont generate at anything crazy like 5mp

4080, 16gb, at 8 steps the gens are good and take a few minutes tops, but use the lightx2v ref turbo lora at 8 steps not 4 (4 is shit audio/blur)
>>
5090 is a gaming card, and they get hot af, its sad that this is our best option
>>
>>109632259
Which of the shifts did you like best?
>>
File deleted.
is this considered nsfw for 4chan blue boards? I know this shit is allowed on instagram which is for kids
if it's nsfw I'll delete it okay?
>>
Be thankful that your RAM sticks don't have bad cells.
>>
>>109632521
>nipples
fuck off
>>
>>109632513
shell out 16k for the blackwell then
>>
>>109632494
No need to be shy anon, here is some more.
>https://files.catbox.moe/kocgxx.mp4
>>
>>109632526
what's wrong with nipples?
>>
>>109632520
Shift 12 is the only one that stayed coherent.
Shift 6 had a really nice spin.
It's clear that in this case the regular setting of shift 12 is best, but different videos may utilize the increased motion
>>
File: 1779773043038015.png (88 KB, 1326x342)
88 KB PNG
>>109632510
140s for this test at 0.6mp, same setup, 8 steps total

https://files.catbox.moe/nhygn1.mp4
>>
>>109632536
what part of "blue board" and "sfw" are you not understanding
>>
>>109632541
but it's allowed on IG
>>
>>109632542
wow thats great this is 4chan not instagram. fuck off right back to that god forsaken normie site
>>
>>109632528
>https://files.catbox.moe/kocgxx.mp4
Pure comedy
>>
I'd say that's a really solid result and time, I recommend checking out the larry lora though and comparing it to the og 4step lora, might improve results further
>>
>>109632542
IG is not a SFW site. Just because kids browse it doesn't mean much.
>>
File: 1764990111595408.jpg (172 KB, 1280x1280)
172 KB JPG
>>109632548
I don't use it, just got images from there
>>
>>109632548
>fuck off right back to that god forsaken normie site
instagram is less normie than 4chan at this point
>>
>>109632554
>i dont use it
>admits to using it to get images
>frogposting
honestly just nuke the internet, put is back to 1954
>>
>>109632538
>>109632551
missed the reply button apparently
>>
File: me in africa.gif (1.71 MB, 292x219)
1.71 MB GIF
>>109632561
what's wrong with frogs
>>
>>109630523
>ldg lazy getting started
>stolill links to the heavily outdated "Local SOTA Models Meta" as its first resource, no update since June 2025
>still linking comfy 1girl guide, which doesn't even mention anima
>>
>>109632539
>>109632551
>>109632565
Am I completely retarded or being trolled, sorry for the spam but this reply for that post...

Also the thread schizo finally mentioned the guides again lmao, don't engage him
>>
i heard minimax is super fucking slow even on a 5090, how true is that? like a few minutes or 10-20 minutes?
>>
>>109630982
>goes to ai board
>gets angry at gacha rolls
>>
File: MiniMax_H3__00296.mp4 (3.67 MB, 640x832)
3.67 MB
3.67 MB MP4
>>109632573
my gens are taking about 3 1/2 minutes for 15 second clips at .5 MP
>>
>>109632494
Not him but dryhumping is my ultimate fetish and this is doing it for me
>>
>>109632550
>Pure comedy
Just like any other hentai.
And this already has better quality than Queenbee could ever hope to have.
>>
>>109632578
i see, mind catboxing a gen so i can yoink the minimax WF you're using for that result? i'd like to try out minimax
>>
>>109632578
5090?
>>
>>109632578
that's piss btw
>>
stop saying mp. say something like 720p instead
>>
>>109632162
I am that Anon; glad you saw it at the end of the thread.

Here's a quicker alternative install method:
1) Install latest AMD Adrenalin
2) Download and try the latest portable AMD release of ComfyUI:
https://github.com/Comfy-Org/ComfyUI/releases
or
https://docs.comfy.org/installation/comfyui_portable_windows#amd-gpu
>Double click run_amd_gpu.bat to launch ComfyUI.
>Normally, ComfyUI will automatically open your default browser and navigate to http://127.0.0.1:8188. If it doesn’t open automatically, please manually open your browser and visit this address.

That lets you quickly try out how well it runs on your system.

The only thing is it's on a slightly older ROCm version (7.2), but it still works. The one advantage of 7.14 for me is that it makes cudnn/MIOpen worth enabling for VAE decode* on my machine, instead of crippling. If you don't use COMFYUI_ENABLE_MIOPEN = 1, then I think there's not as much difference.

(Oh, and it looks like ComfyUI enables dynamic vram by default for ROCm 7.14 now, so that saves me from having to type --enable-dynamic-vram when launching.)

*: VAE decode is the last step of generating, and is basically like an upscaling step. It makes VRAM usage spike, which can sometimes freeze your machine if your pic is too big. Enabling MIOpen greatly reduces the spike (after recent bugfixes; YMMV). Another workaround is to use tiled VAE decode to split up the upscaling job.

>>109632180
True, I can't say how well certain settings will work on other AMD hardware. Generally I'd expect newer hardware to have better support. I'm on Strix Point, aka gfx1150, a Ryzen AI HX 370 with 890M.
>>
>>109632578
>3 1/2 minutes for 15 second clips at .5 MP
Yeah but it also look ultra deep fried from all your cope nodes.
>>
my GPU shows 0% work and 70 degrees when genning
why does it show 0%?

Also how do I prompt in minimax to swap clothes?
Every time I try it swaps the girls as well
>>
File: MiniMax_H3_00351_.webm (3.83 MB, 768x1376)
3.83 MB
3.83 MB WEBM
I think I'm starting to get away with just 4 steps now. Results are less sloppa
6-8 steps are better, but it's 100s/it already.
>>
>>109632610
That's what got me interested in trying, and a couple very patient anons guided me through absolute retardation on my part, and I got it working (Me) >>109632378 >>109632480 Now i'm here in the blank UI. Trying to figure out how to start getting it done. Seems all the templates require me to download massive files and I don't mind but I honestly don't know what I'm doing. I'll just work on it slowly with time now that I have it installed.
>>
>>109632610
>890M.
Im so fucking sorry for you anon. Like genuinely.
t. 7900 XTX and 7700S owner
>>
>>109632591
i would but every time i upload to catbox it tells me invalid uploader
>>
>>109632601
yeah 5090
>>
>>109632591
https://litter.catbox.moe/vs8acw.mp4
>>
>>109631429
This one is excellent
>>
>>109632633
thanks
>>
>>
>>109632641
yeah, does it load in comfy like that? i didn't realize you could load an mp4 as a workflow?
>>
I wonder how Anima even knows this artist @monkechrome
>>
>>109632659
yeah, it loaded fine. mp4 files can contain metadata so there should be no issues.
>>
>>109632664
For being trained off a robot world model it does pretty good.
>>
File: ice cream.mp4 (3.36 MB, 832x640)
3.36 MB
3.36 MB MP4
my magnum opus got ruined
good night
>>
File: MiniMax_H3_00352_.webm (2.44 MB, 544x960)
2.44 MB
2.44 MB WEBM
>>109632622
0.5M with 6 steps,50s/it is also alright
save about 2':30s vs 1M 4 steps
But there is video-audio desync somewhere; I'm to tired to figure it out
>>
>>109632608
>720p
And how many pixels is that? 100x720? 320x720? 900x720? 2156x720? Saying 0.5mp means it's 500000 pixels in any aspect ratio.
>>
>>109632624
You're on the right track with the templates, and yeah, model files are big.
A few suggestions:
>In the upper-right of the templates browser, you can click the "Runs on" dropdown and check "ComfyUI" to filter out templates that run on the cloud.
>The current state-of-the-art are something like Krea 2 for general images, Anima for anime gens with booru tags, and Minimax H3 for video. For editing, Flux 2 Klein 9B KV. Z-Image Turbo is also good for general images, and I think that's the default template that shows up when you close all your open workflows.
>SDXL-based models are older, but run faster (about 3x faster than Anima). For anime, there's Illustrious-based variants like NoobAI or WaiIllustriousSDXL (for easymode slopping; it's how I started).
>>
which disposable email lets me make discord accounts? i found the commit that brought a regression but i have to go into their discord server to report it
>>
>>109632695
Just use your email and phone number
It's not like they aren't keeping tabs on you anyway
>>
File: Untitled.jpg (1.13 MB, 3833x1972)
1.13 MB JPG
>>109632693
>and Minimax H3 for video
Oh yeah, that's the first one I saw and figured I don't know what the fuck I'm doing so I might as well just grab the first thing I see. It's downloading a bunch of files now.
>>
>>109632699
no thanks glownigger
>>
>>109632684
What's your prompt for this?
I haven't had success with replacing/editing with h3 ref2v. Often I just get the video ref regenerated on its own.
>>
>>109632681
CRUNCH
CRUNCH
>>
>>109631869
don't know if you're still around anon, but i just wanted to say thanks for clueing me in to the importance of using timestamps.
>>
File: SHEET_00008_.jpg (739 KB, 2700x2374)
739 KB JPG
>>109632711
It's a video edit workflow. Original video from /kpop/ on /gif/
prompt is from LLM and is not very good. It's too long to post on 4chin. You tell clankers to replace <subject 2> from <video 1> with < subject 1> from <picture 1>
>>
>>109631822
>>109631869
Why can't we just type in "humping, pumping, semen dumping." and have ti work? The world is not fair.
>>
>>109632758
It's Christian Helmsworth
>>
>>109632762
*T'Chaka O'Dinson, the white ape
>>
File: 1784081673621821.png (2.04 MB, 1920x1080)
2.04 MB PNG
>ensure the first frame of the video is <Picture 1>.
editing fun:

https://files.catbox.moe/m4ea5t.mp4
>>
File: SHEET_00026_.jpg (515 KB, 1798x2374)
515 KB JPG
>>
sick to death of the word "then"
>>
>>109632823
anon what the fuck why did you post a photo of me
>>
>>109632736
No problem
>>109632759
That's ideally what loras are for. Or someone would need to fuse the core model with nsfw that has those trigger words for poses.
>>
>>109632840
replace it with a comma
>>
>>109632840
Didn't you learn anything from Trey Parker? You don't write "and then" you write "therefore" or "but".
>>
File: 1783536039668358.jpg (19 KB, 474x474)
19 KB JPG
>mfw made the perfect pov cowgirl r2v workflow with 1:1 ref accuracy, complete with fast plapping sounds and a thumb in mouth [shot 2] without crunching granola or slurping ramen sfx
its literally over for my benis
>>
File: pjimage-3-3350351529.jpg (50 KB, 400x476)
50 KB JPG
>>109632909
Try finger and then therefor try but whole.
>>
>>109632968
i'd like to make a fully t2v pipeline for a long video, but the massive generation times are too risky since the first video in the chain could very well end up being a dud
>>
No doubt I'm stupid but I downloaded all the models that were missing but the thing still says missing models, how do I, you know. Do I put them in a specific directory or is there a menu option to direct it to the files?
>>
>>109633006
you need to mount the files to your tensor environment
>>
>>109632968
>and he won't share it either
>>
What is better for H3 prompt writing, Gemma 4 Heretic or Qwen3.8 Heretic?

I heard Qwen is more suited for agent workflow?
>>
im bored, give me ideas
>>
>>109633048
1girl, standing,
>>
File: MiniMax_H3_00620.mp4 (3.64 MB, 800x1152)
3.64 MB
3.64 MB MP4
>>
File: folders.png (89 KB, 554x840)
89 KB PNG
>>109633006
The template has notes on the left that show which folders the model files go in.
>>
>>109633051
damn... thats a good one
>>
>>109633052
she's got nice perchlorates
>>
>>109633052
imagine the smell
>>
>>109633006
>>109633056
Oh, and press the "R" key to refresh ComfyUI's file detection. It should be able to see the new files once you do that.
>>
the kinoplexatorium will be open shortly
>>
>>109633112
i see you fucker tell robert to eat a dick
>>
>>109633128
who is robert?
>>
>>109633056
>>109633104
You're a champ. As someone who's spend decades in tight-knit hobbies, I know how frustrating it is when faggots like me come in and can't figure out shit you've explained to motherfuckers a million times already. I fully appreciate your help.
>>
>my gens are too stiff until i use the shift node
>when i turn off the shift node, no matter how i prompt it never comes out the way i want
i hate this
>>
>>109633148
Bet it's not the only things that's stiff, eh champ, what kind of filth you generating?
>>
>>109633171
one man standing in an open field yelling slurs at various looney tunes characters
>>
>>109633177
Oh my how risque.
>>
>>109633181
its very important how well animated they are to me, please understand. i need like better than roger rabbit style mixing
>>
>>109633136
I just hate how much frequency this place is just comfyui tech support. The software is honestly shitty garbage. The only thing that makes it good is the backend speed compared to the rest of the diffusion uis and it pisses me off how unstable the entirety of the space is when it comes to this stuff. It's not your fault it's just that the entire space is run by script kiddies
>>
>>109633187
Oh dude, I'm not disparaging you at all. I'm just sitting here not even knowing what to start typing in so what you're doing is already magic to me.
>>
>>109633190
I get it. Growing pains of emergent tech and shit. For what it's worth, I'm sure it's just a matter of time before a lot of this shit gets consolidated and standardized and make the community mature a little more and spend less time being groundhog day QnA.
>>
>>109633192
type something kino in
>>
>>109633192
absurdities work usually
>>
>>109633112
i am nearly seated
>>
File: Untitled.png (473 KB, 2659x1582)
473 KB PNG
>>109633200
>>
>>109633218
punch it
>>
>>109633218
i will be waiting for the results
>>
>>109633218
we're about to witness true greatness
>>
>>109633219
>>109633222
>>109633230
Still at 0 percent so maybe not? How long do these things usually last?
>>
>>109633242
we're gonna be waiting a while for you since you have no idea what youre doing
>>
>>109633248
I do not deny this. I am as clueless as clueless gets.
>>
Just realized that if I want to get into creating a story with H3, I'll actually need to write a script.
>>
>>109633039
Anyone?
>>
File: seed.png (12 KB, 365x153)
12 KB PNG
>>109633136
Cheers!

Some other random tidbits:
>Reusing the same input words (prompt) can generate either an identical output if you use the same seed (a random number, see attached pic), or it can generate a different result if you use a different seed.
>ComfyUI saves the whole workflow (including prompt and seed) in the metadata of the output file. If you drag-and-drop the output file into ComfyUI, it will load the embedded workflow that was used to generate the file.
>4chan scrubs this metadata from uploaded images. People sometimes post catbox uploads of their gens if they want to share the workflow they used. (Or if they just want to share a video with sound, which most boards don't allow.)
>Getting back to seeds: the default Minimax template has a noise_seed field with a fixed value typed into it (looks like 168866841893410). Other templates may have a separate seed node (see attached pic) with a randomizer setting, so you can rerun the same prompt and get slightly different results each time.
>The text saying "control before generate" means that when you click Run, the number changes BEFORE your gen starts. I think the default is "control after generate".
>Control BEFORE is arguably more intuitive, since if you get an almost-good result that you want to tweak more, you can change from "Randomize" to "Fixed" and rerun the same seed easily with a modified prompt. If you use control AFTER, you have either ctrl+z to get the previous seed back, or dig it up from the output file. This will make more sense once you play with genning more.
>Anyway, the setting for Before/After is found under Settings -> Comfy -> Node Widget -> Widget Control Mode. For future reference.
>>
>>109630890
thanks!
>>
>>109633330
>he fell for the malware
its over
>>
>>109633337
proofs?
>>
>>109633198
>I'm sure it's just a matter of time before a lot of this shit gets consolidated
websloppers can't help but bloat. It's been getting worse every year while llm bros get binaries where you don't have to curate all this shitty bloat
>>
File: turbo switch.png (49 KB, 582x483)
49 KB PNG
>>109633242
Your hardware should be much faster, but on my humble iGPU, a 4-second 0.2 megapixel video originally took about 30 minutes. Switching to --use-ck-attention knocked that down to 20 mins, and then enabling the turbo lora knocked that down to 10 mins, since it uses fewer steps.

There is a newer version of the template that has a toggle switch for turbo mode (see attached pic). Drag this template into ComfyUI:
https://github.com/Comfy-Org/workflow_templates/blob/main/templates/video_minimax_h3_i2v.json
>>
>>109632493
some mirror to hf
descriptions and examples
to see what loras were deleted
>>
>>109633362
Oh neat,
>Switching to --use-ck-attention knocked that down to 20 mins,
I don't know what that means but all the other stuff got the progress bar moving.
>>
>>109631624
NP
I ended up checking it out and following some youtube guide. I was able to get comfy running on it and genned some videos. For reference, a 3 sec video took like 2 hours on my machine. I rented an RTX 4090 for two hours and it cost me a little less than $2 or like 75c/hour.
A lot of that time was just setting it up and troubleshooting, but once I got it running i was able to gen pics in seconds and videos in about 7 minutes. I didn't get great results but i think that's more due to my prompting and workflows than the models. I used wan2.1, but would be open to checking other models in the future.

https://litter.catbox.moe/zjih29mtqrmuivdf.mp4


I'll probably try it out again next weekend and creep these threads for good models or tips. In the future I might check out the 100 VRAM GPUs and run one of those massive models. Its only like $3 an hour.
>>
>>109633219
>>109633222
>>109633230
https://litter.catbox.moe/y07nei0mr013egxl.mp4
Don't say I have not delivered in gratitude for helping me reach this point.
>>
>>109633426
kino
>>
>>109633426
remarkable kino
>>
>>109633426
My PC just straight up shut down trying to run it again so I think I'm trying too much.
>>
>>109633426
Cute
>>
File: modelattentionbackend.png (107 KB, 818x910)
107 KB PNG
>>109633406
Use a text editor to open run_amd_gpu.bat and add --use-ck-attention to the launch command. I'm not sure what the default is, but the format should be something like:
python main.py --enable-dynamic-vram --enable-manager --disable-smart-memory --use-ck-attention

Et cetera. You won't have so many flags, but it should give you an idea of what it looks like. The order of the flags doesn't matter, I think.

There's also a GUI node that can be used to only enable Comfy Kitchen Attention for the model you hook up to it, instead of turning it on for everything; see attached pic. But figuring out where to hook that up in the Minimax template would be a bit annoying.
>>
Bake a collage or don't bake at all
>>
https://litter.catbox.moe/hjxx95sofxc78pg7.mp4

neat, can mix styles fine.

Use <Picture 1> for the physical identity of Denton, with the voice of <Audio 1>.

the setting is <Picture 2>.

medium shot of JC Denton in an office building with a "UNATCO" sign, looking exactly like the character in <Picture 1>, is looking forward and points to the camera. Denton speaks naturally in a dimly lit, gritty cyberpunk interior, saying the exact line: "Hey you, have you seen any Hatsune Miku AI generations in this location?.".

camera cuts to a shot of Jerry Seinfeld from the show Seinfeld, who says "what is a Hatsune Miku?"

camera cuts to Denton, who says "a virtual idol, a vocaloid, an idol. she sings music."

Low-poly aesthetics, moody green and blue ambient lighting, nostalgic year 2000 PC gaming look.
>>
>>109633136
We are all sirs here actually
>>
>>109633493
thank you Mark R.
>>
>>109633500
congrats you can read embedded windows/user folders from a workflow, mr hackerman
>>
>>109633504
i appreciate it, Mark R.
>>
>>109633511
>Inspection: Anyone who extracts or inspects the JSON payload from the video file using a text editor or a tool like ComfyUI-Workflow-Inspector can read those text fields and see the Windows username.

comfy is a retard for that btw, what if someone had sensitive data in that, it has no relevance to a node workflow.
>>
Do more detailed prompts increase or decrease time it takes to generate?
>>
>>109633563
not really, if you want a super fast test do 0.3mp, when you have a good prompt then bump the resolution.
>>
>>109633461
Finishing successfully once is an encouraging sign. There may be ways to stabilize it. Video gen's one of the most demanding things you can run.

>>109633563
I think a negligible increase unless you're doing text encoding on CPU.
>>
>>109633582
I think it was my GPU overdrawing power. Is there a way to cap it off to stop the gen from running it at 100%
>>
>2026
>no audio model trained on R18 ASMR
>>
>>109633594
what operating system are you using?
>>
>>109633605
Windows 11.
>>
>>109633619
>>109633619
>>109633619
MOVE
>>
>>109633612
actually, you're using amd right? i am not familiar with those cards. just look up for popular software that limits the power for amd gpus. usually it is overclocking software, but those would give you the option to reduce power as well
>>
>>109633627
Yeah underclocking/downvolting is an option, I was just hoping there might have been an app-specific setting.
>>
File: power limit.png (19 KB, 616x253)
19 KB PNG
>>109633594
Looks like AMD Adrenalin may have a power limit slider, or there are more general presets:
https://www.amd.com/en/resources/support-articles/faqs/DH3-020.html
>>
>>109632138
holy, catbox?
>>
>>109633629
>I was just hoping there might have been an app-specific setting
i don't think any GPU software is sophisticated enough to work that way. all hardware settings are global
>>
>>109633619
>>109633619
>>109633619
NEW
>>
>>109633644
Fair enough.
>>
>>109633633
>>109633644
Adrenalin says it has per-game/app profiles, but I don't know if it will work with ComfyUI/Python. If pointing it at run_amd_gpu.bat doesn't work, maybe try pointing it at python.exe?

https://www.amd.com/en/resources/support-articles/faqs/DH3-012.html#dh3-012-application
>>
>>109633676
oh. pointing it to the .bat script would definitely not work since it's just a launcher for a different program. he has to look at task manager or whatever other software that tracks GPU usage per process and use that one



[Advertise on 4chan]

Delete Post: [File Only] Style:
[Disable Mobile View / Use Desktop Site]

[Enable Mobile View / Use Mobile Site]

All trademarks and copyrights on this page are owned by their respective parties. Images uploaded are the responsibility of the Poster. Comments are owned by the Poster.